We're live with the @Kimi_Moonshot team to discuss Kimi K3! https://x.com/i/broadcasts/1mxPaaMeLDZKN
We're live with the @Kimi_Moonshot team to discuss Kimi K3! https://x.com/i/broadcasts/1mxPaaMeLDZKN
We're live with the @Kimi_Moonshot team to discuss Kimi K3! https://x.com/i/broadcasts/1mxPaaMeLDZKN
Everything you need to know about Kimi K3: A conversation with Moonshot AI and Together AI https://x.com/i/broadcasts/1mxPaaMeLDZKN
RT Kimi Developers🤗Together AI: Kimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it.He joins Jue Wang…
submitted by /u/Hannibalj2ca [link] [comments]
Moonshot AI has open-sourced MoonEP, an Expert Parallelism (EP) communication library for distributed Mixture-of-Experts (MoE) workloads. The team announced the release as a library built to…
RT zhyncsInteresting data from the OpenRouter Kimi K3 dashboard:@togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt…
Has anyone yet tried to extract experts from kimi (or GLM 5.2) per chance? There is REAP that removes experts based on routing, but I could…
we're experimenting with our own dynamic GGUF quants of kimi k3, made from the original weights with our llama.cpp fork. Q3_K_S is done and works 1114.76…
The model was quantized to 8, 4, 2, and 1 bit. Characteristics: Q8: 8-bit 1.56 TB, lossless Q4: 4-bit, 1.51 TB Q2: 2-bit: 861 GB Q1:…
Article URL: https://www.kimi.com/code/docs/en/kimi-code/models Comments URL: https://news.ycombinator.com/item?id=49101852 Points: 41 # Comments: 4
I've got better results than expected for 768gb DDR5 and 2x5090. Using fork https://github.com/pwilkin/llama.cpp/tree/kimi-k3-text and https://huggingface.co/GrEarl/Kimi-K3-GGUF Q2_K quant. Prefill speed for big prompt is 50-70 tps.…
Article URL: https://aistack.imec-int.com/blog/gpu-self-hosting Comments URL: https://news.ycombinator.com/item?id=49098130 Points: 6 # Comments: 0
RT DogeDesignerBREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and…
Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just…
submitted by /u/Hannibalj2ca [link] [comments]
In this tutorial, we configure and operate Kimi CLI as a fully non-interactive AI coding agent. We install the CLI through uv with an isolated Python…
submitted by /u/_TheWolfOfWalmart_ [link] [comments]
Article URL: https://github.com/gavamedia/deltafin Comments URL: https://news.ycombinator.com/item?id=49090233 Points: 44 # Comments: 28
We've added Kimi K3 to Perplexity and Perplexity Computer for Pro and Max subscribers.Kimi K3 in Perplexity is hosted exclusively on U.S.-based servers.
RT Philippe LemoineThe other day, I saw a ML researcher on here claim that Alibaba, http://Z.ai, Moonshot, etc. were SOEs. I'm sure this guy is very…
submitted by /u/_maverick98 [link] [comments]
RT ZainEverything you need to know about Kimi K3 with @Kimi_Moonshot and @togethercompute team.If you're looking into evaling and building with Kimi K3 I would not…
RT Together AIKimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it.He joins Jue Wang and Zain…
RT Together AIKimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it.He joins Jue Wang and Zain…
Kimi is Moonshot AI's product — Kimi K2 is the open-weight model variant. Strong on agentic / long-context tasks, popular in the Chinese market alongside the Doubao and Wenxin lines.
Owner: Moonshot. We have 330 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is www.kimi.com.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen.