> After using DeepSeek v4 Flash for a day, its Harness is much better than Kimi K3!!! Intriguing
> After using DeepSeek v4 Flash for a day, its Harness is much better than Kimi K3!!!IntriguingYufan Sheng: 用了一天 DeepSeek v4 Flash,它的 Harness 比 Kimi K3…
Every primary-source story across every tracked model. Filter by clicking a chip.
> After using DeepSeek v4 Flash for a day, its Harness is much better than Kimi K3!!!IntriguingYufan Sheng: 用了一天 DeepSeek v4 Flash,它的 Harness 比 Kimi K3…
There had been no “preview subsidy”SiliconFlow prices are distinct from DeepSeek’s and in particular worse on cacheNevertheless they do work together and this is a minor…
TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit…
Grok can analyze any video https://grok.com/share/bGVnYWN5_8013f7a3-f604-4351-8cd7-acecf3ef165b
chatgpt for empowering your dad to build:chloe venn: my dad using gpt 5.6 sol for the 1st time:- wanted to build a webpage about indigenous south…
I deployed K3 on 32 H100s at work a couple of weeks ago and then got annoyed that there was no way to poke at it…
Article URL: https://www.wafer.ai/blog/kimi-k3-mi355x Comments URL: https://news.ycombinator.com/item?id=49141073 Points: 8 # Comments: 0
Article URL: https://philarchive.org/archive/NIEWTCv17 Comments URL: https://news.ycombinator.com/item?id=49140869 Points: 23 # Comments: 4
I spent a while debugging my local DeepSeek V4 Flash setup and wanted to share a few lessons from the process in case it saves someone…
ask chatgpt work to do any recurring taskBrett Bauman: chatgpt work is the new cron jobhttp://bbrett.com/movies
This picture implies V4-Pro GA scoring ≈57.5 on AA, a notch above Kimi, improving by 13.5 points. This is achievable… on *some* timeline. But I doubt.…
yah, GPT-5.6 cyber is really good… luckily it's a closed model with stringent safeguardsit's not like open weights models can do that yet. Phew. safeAndrew Curran:…
Wenfeng is too cultured to really act like this, but he totally feels this wayDeepSeek people have a lot of pride in what they doGorden Sun:…
Kimi K3 has set a new bar for OSS model intelligence! 2.8T params, 1M context, OpenAI-compatible API.Complete guide to running Kimi K3 on Together AI 👇👏👏…
Despite a lawsuit from xAI, a Minnesota ban on apps that allow users to “nudify” images can move forward.
Been using Qwen 3.6 35B-A3B quite extensively lately and honestly, I’m pretty happy with it. Also tried a few community improvements like Ornith 1.0, which add…
chatgpt work's cloud browser is really cool to use, lets you easily monitor what your AI is up to and also intervene with the live application…
Hi everyone!Hope you had a great day so far, and maybe its about to get just a little bit better (thanks Winter ;)So I had way…
RT FuserLuma Ray 3.2 from @LumaLabsAI is now live in Fuser.Keep the action intact. Reframe your scene in any aspect ratio.Try it now↓
Same harness, same task set, same agent scaffold, the only thing I swapped was the executor. Not a proper benchmark, no clean tok/s numbers, this is…
The biggest issue with preview was its inability to follow rules prompts and skills. It seems like no matter what you do it ignores them. I've…
OpenAI's CEO seemed excited to share a "cool use case" for parents.
Thanks to the community help I finally launched this llm. LM Studio refused to load weight onto second GPU but Unsloth Studio did so everything was…
I think people are sleeping on Gemma and local models so I built a free, very fast harness for Gemma 4 that I call Tomte. https://tomteapp.com…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.