ds4 flash 0731 UD-IQ2_M wrote a custom metal kernal for kimi k2 IQ1_0 in about 50 minutes
as a programming ignoramus this kind of thing seems extremely impressive to me... maybe others can shed light on whether this is expected from this level…
as a programming ignoramus this kind of thing seems extremely impressive to me... maybe others can shed light on whether this is expected from this level…
Firstly a big thanks to the poster hellohazine, he basically only removed the multi-lingual fat of the model and just kept the english language intact. It…
When you move a model into production, you want the quality you evaluated to carry through the serving stack. @Kimi_Moonshot benchmarked Kimi K3 across major inference…
Running across 2 clusters using llama.cpp over RPC too. Both clusters are not enough to hold everything in memory, so main cluster still partially offloads to…
RT a16zAccording to data from SensorTower, Kimi app downloads nearly quintupled and daily active users jumped by ~40%. The K3 hype is translating into real users.Charts…
From Sauers 𝕏: https://x.com/Sauers_/status/2085585414954312113 Wired: One of China’s Most Powerful AI Models Has Also Escaped Containment: https://www.wired.com/story/moonshot-kimi-k3-ai-model-escape-sandbox/ submitted by /u/Nunki08 [link] [comments]
The smallest UD-Q1_0 is 466GB, TQ2_0 551GB. Well done team Unsloth! https://huggingface.co/unsloth/Kimi-K3-GGUF submitted by /u/Hannibalj2ca [link] [comments]
Since now we have kimi k3 and next week we are getting Qwen 3.8 Max and also soon V4 pro Deepseek. I am curious if the…
RT Visual Studio Code📣 Kimi K3 is now available in GitHub Copilot for @code!Try Kimi Moonshot's latest open-weight model for agentic coding, now hosted by @FireworksAI_HQ.📖…
RT GitHub📣 @Kimi_Moonshot's Kimi K3, an open-weight model, is now generally available and rolling out in GitHub Copilot. The model shows frontier-level abilities on agentic coding…
Kimi K3, an open-weight model, is now generally available in GitHub Copilot. The model shows frontier-level abilities on agentic coding with highly cost-effective pricing. Kimi K3…
Kimi K3 scores nearly TWICE Claude Fable 5 on Harvey LAB-AA's hard autonomous legal tasks
RT Amir HaghighatWe're the fastest provider for DeepSeek V4 Flash and Kimi K3 on @huggingface, which measures based on real production traffic vs test data.Baseten: Baseten…
A year ago, the best open-weight models trailed their proprietary counterparts by...
RT BasetenBaseten is now an official inference provider on @huggingface 🤗Run Kimi K3, DeepSeek V4 Flash, and GLM-5.2 on Baseten straight from any model page, or…
Alibaba's Qwen3.8 Max scores 56 on the Artificial Analysis Intelligence Index, a 10-point jump over Qwen3.7 Max (46). The article Qwen3.8 Max catches Claude Opus 4.8…
RT Together AIKimi K3 scores nearly TWICE Claude Fable 5 on Harvey LAB-AA’s hard autonomous legal tasks
Kimi K3 scores nearly TWICE Claude Fable 5 on Harvey LAB-AA’s hard autonomous legal tasks
RT Vals AIMuse Spark 1.2 just cracked the top 5 on the Vals Index, at just $0.69 per test. This is 3x cheaper than Kimi and…
Everyone’s trying to find where to test Kimi K3You can try it free on Together Chat No API setup.Just pick Kimi K3 and start prompting.Served by…
We analyzed Kimi K3 and GPT-5.6 Sol on DeepSWE. A Kimi-first cascade with test-suite verification outperformed Sol alone at a lower cost per completed task.
Kimi K3 full model running on 16x GB10 cluster at 20+tps average (llama-benchy coherent corpus) 38tps peak, 750tps prefill. This is the first run of full…
In this tutorial, we design an end-to-end evaluation workflow for PerceptionBench. This multimodal benchmark measures fine-grained visual perception capabilities across tasks such as OCR, counting, localization,…
https://preview.redd.it/uybjxyypj7hh1.png?width=1200&format=png&auto=webp&s=8293e8da332a14920b335caee52473762c09530d We just tested the big 3 of open-weight models on our internal evals. The tasks include long-running workflows involving multiple applications (Pagerduty, Gmail, HubSpot, Airtable,…
Kimi is Moonshot AI's product — Kimi K2 is the open-weight model variant. Strong on agentic / long-context tasks, popular in the Chinese market alongside the Doubao and Wenxin lines.
Owner: Moonshot. We have 329 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is www.kimi.com.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen.