Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet
submitted by /u/realmvp77 [link] [comments]
Every primary-source story across every tracked model. Filter by clicking a chip.
submitted by /u/realmvp77 [link] [comments]
Moonshot dropped Kimi K3, and as expected, it’s a absolute monster. Even with 2~4x RTX 6000 Blackwell local workstations, running a model natively is virtually impossible.…
RT サメQCUhttps://x.com/tokenbender/status/2081776672152686883"kimi stopped using positional embedding"inside the text:"kimi uses a linear attention layer which is a positional embedding for other subsequent layers"tokenbender: kimi guys noped the…
Last week, we made ChatGPT Work available to enterprises. It's an awesome product.So for ChatGPT Enterprise customers, enroll by August 21, and each teammate who tries…
RT Greg BrockmanLast week, we made ChatGPT Work available to enterprises. It's an awesome product.So for ChatGPT Enterprise customers, enroll by August 21, and each teammate…
You know what, yeah.What now?K3 is at the absolute least Opus 4.6 level in utility. Actually well beyond that. Everyone who's bought ≈$5M worth of GPUs…
Article URL: https://telnyx.com/release-notes/kimi-k3-telnyx-inference Comments URL: https://news.ycombinator.com/item?id=49076505 Points: 14 # Comments: 5
Article URL: https://github.com/humanlayer/advanced-context-engineering-for-coding-agents/blob/main/benchmarking-opus-5-on-slop-code-bench.md Comments URL: https://news.ycombinator.com/item?id=49076391 Points: 36 # Comments: 7
I ran a solo evaluation project benchmarking six current frontier models: GPT-5.4, Claude Sonnet 4.6, Claude Opus 4.7, Gemini Pro, Gemini Flash, and Grok 4.3. I…
Dario opens his newest essay with bullshitting. In "DeepSeek and Export Controls" as well as in "Two Scenarios" and in "Policy on the AI Exponential", he…
There’s been a lot of speculation about where we stand on open-weights models. We’ve outlined our views in full here: https://www.anthropic.com/news/position-open-weights-models
submitted by /u/KickLassChewGum [link] [comments]
Another Day 0 partnership. Happy to be part of getting Kimi K3 into Cursor from launch.Cursor: Kimi K3 is now in Cursor! It scores close to…
Kimi K3 is now in Cursor! It scores close to the frontier on CursorBench.It's available on US-based inference thanks to our partners Fireworks, Together, and Baseten.…
Kimi K3 is a triumph of scaling maximalism. They scaled the hell out of everything.
Moonshot AI's Kimi team and kvcache-ai open-sourced AgentENV (AENV) under MIT, as part of Kimi K3 Open Day. It runs agent sandboxes as Firecracker microVMs with…
Damn, Xi Jinping's sleeping pills are something fierce, must be the power of traditional Chinese medicine…according to Fable, Kimi K3 is roughly in the ballpark of…
The issue appears to have originated from Claude’s “share chat” feature, which allows users to create links that enable anyone with the assigned URL view a…
If you are familiar with my previous posts on model welfare for new Claude models, you can skip the Introduction and The Story So Far. Key…
Moonshot AI has released Kimi K3's model weights and made parts of its infrastructure open source. The Chinese model nearly matches Western frontier models such as…
I’ve been doing some benchmarking on my mini-PC setup (AMD Ryzen 7 6800H) to see how it handles the new Gemma 4 and Qwen 3.6 MoE…
Wanted to let you know that Kimi K3 is now viewable on hfviewer.com! In addition to the full graph at multiple granularity levels, we also include…
I just managed to get it running on windows and this thing is fucking insane. I get around 550-720t/s depending on task at hand. Previously to…
OpenAI analyzed over 800,000 work-related ChatGPT messages and found that 43.5 percent of job-specific queries involve tasks from other professions. The company calls this "task crossover."…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.