Ollama's new app
Ollama's new app is now available for macOS and Windows.
Ollama's new app is now available for macOS and Windows.
Announcing Codestral 25.08 and the Complete Mistral Coding Stack for Enterprise
OpenAI launches Study Mode in ChatGPT — Socratic tutor for Free/Plus/Pro/Team users that asks guiding questions and refuses direct answers, co-designed with 40+ pedagogy partners. Apple…
WAIC closing day, and arguably the loudest single day in Chinese open-source AI of the year. Zhipu (rebranded as Z.ai) drops GLM-4.5 series open-weights under MIT…
Does process matter? We are about to find out.
WAIC Shanghai day 2 — Tencent's Hunyuan 3D World Model 1.0 stays the lead story, with full code + weights for explorable 360° 3D worlds. SenseTime…
PAPER DISCORD Introduction Reinforcement Learning (RL) has emerged as a pivotal paradigm for scaling language models and enhancing their deep reasoning and problem-solving capabilities. To scale…
WAIC 2025 opens in Shanghai — Premier Li Qiang proposes a Shanghai-headquartered "World AI Cooperation Organization". Tencent open-sources Hunyuan3D World Model 1.0 (first openly licensed interactive…
Alibaba releases Qwen3-235B-A22B-Thinking-2507 — 235B MoE (22B active) with native 262,144-token context, AIME25 92.3, LiveCodeBench v6 74.1, claiming parity with DeepSeek-R1-0528, OpenAI o3/o4-mini, Gemini 2.5 Pro…
Axios scoop: OpenAI targets an August GPT-5 launch — combining traditional and reasoning attributes, with mini and nano variants via API. Qwen3-Coder lands on Hugging Face…
DEMO API DISCORD Introduction Here we introduce the latest update of Qwen-MT (qwen-mt-turbo) via Qwen API. This update builds upon the powerful Qwen3, leveraging trillions multilingual…
Trump unveils "Winning the Race: America's AI Action Plan" at the Hill and Valley Forum / All-In Podcast summit — 23-page, three-pillar, 100+ recommendation roadmap drafted…
Update deploy_guidance.md Signed-off-by: bigmoyan
Alibaba releases Qwen3-Coder-480B-A35B-Instruct — a 480B-parameter MoE (35B active) open-source agentic coding model with native 256K context (1M via Yarn), 69.6% on SWE-bench Verified, nearly matching…
GITHUB HUGGING FACE MODELSCOPE DISCORD Today, we’re announcing Qwen3-Coder, our most agentic code model to date. Qwen3-Coder is available in multiple sizes, but we’re excited to…
Our contribution to a global environmental standard for AI
Google DeepMind officially announces Gemini Deep Think gold at IMO 2025 — 35/42, 5/6 problems solved in natural language with no tools, first for a general-purpose…
We compare the best image models for generating consistent characters from a single reference image.
We're joining forces with Amazon Web Services to announce a new program that will provide resources and support to 30 promising startups in the U.S. that…
OpenAI's IMO gold-medal claim from Saturday continues to dominate AI Twitter — Noam Brown's "o1 thought for seconds, Deep Research for minutes, this one thinks for…
OpenAI announces gold-medal performance on the 2025 International Mathematical Olympiad — an experimental general-purpose reasoning LLM scored 35/42, solving 5 of 6 problems under human time…
From DeepSeek-V3 to Kimi K2: A Look At Modern LLM Architecture Design
Meta refuses to sign the EU's voluntary General-Purpose AI Code of Practice — Joel Kaplan calls it "overreach" two weeks before the August 2 enforcement date.…
OpenAI launches ChatGPT Agent — unifies Operator + Deep Research + ChatGPT into a single agentic mode with its own virtual computer, 68.9% on BrowseComp (17.4…