Writing effective tools for agents — with agents
Agents are only as effective as the tools we give them. We share how to write high-quality tools and evaluations, and how you can boost performance…
Every primary-source story across every tracked model. Filter by clicking a chip.
Agents are only as effective as the tools we give them. We share how to write high-quality tools and evaluations, and how you can boost performance…
Second Front’s chief data scientist sees data security and LLM explainability as fundamental to overcoming government caution over AI adoption.
Mistral AI raises 1.7B€ to accelerate technological progress with AI
A Detailed Look at One of the Leading Open-Source LLMs
Designing transparency and control into AI recall.
Le Chat now integrates with 20+ enterprise platforms—powered by MCP—and remembers what matters with Memories.
The new industry standard for secure, enterprise-ready machine translation.
Merge pull request #969 from youkaichao/rmsnorm act_quant_kernel
We're thrilled to introduce grok-code-fast-1, a speedy and economical reasoning model that excels at agentic coding.
fix act_quant_kernel (#968) Signed-off-by: youkaichao
support scale_fmt=ue8m0 (#964) * support scale_fmt=ue8m0 * keep improving Signed-off-by: youkaichao * keep improving Signed-off-by: youkaichao * add clamp min of 1e-4 Signed-off-by: youkaichao * rename…
Instituto PROA, a nonprofit organization in Brazil, has transformed its job preparation process for young candidates by leveraging Llama and Oracle Cloud Infrastructure.
Securely powering enterprise applications with exceptional reasoning performance, efficiency, and controllability.
Cohere and Government of Canada sign Memorandum of Understanding to transform the public sector with sovereign AI.
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We are excited to introduce Qwen-Image-Edit, the image editing version of Qwen-Image. Built upon our 20B Qwen-Image model, Qwen-Image-Edit…
Financial teams are unlocking a new era of compliance, efficiency, and customer trust—powered by AI agents.
Insights and best practices for using Elo scoring methods for evaluation leaderboards.
Fresh funding enables Cohere to accelerate its global expansion and build the next generation of secure enterprise and sovereign AI solutions.
DINOv3 scales self-supervised learning for images to create universal vision backbones that achieve absolute state-of-the-art performance across diverse domains, including web and satellite imagery.
Using DINOv2, the team at NASA's Jet Propulsion Laboratory built a convenient robot operating system interface for robotic tasks.
WRI and the Bezos Earth Fund used DINOv3 to develop an algorithm to accurately count individual trees from drone and satellite imagery.
Throughout 2025, we have been quietly entering Claude in cybersecurity competitions designed primarily for humans. In many of these competitions Claude did pretty well, often placing…
add 1M support (#1600) * add 1M support * Update README.md --------- Co-authored-by: Ren Xuancheng
Putting the AI in Charge
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.