What xAI's Grok Build CLI Actually Sends to xAI
Article URL: https://gist.github.com/cereblab/dc9a40bc26120f4540e4e09b75ffb547 Comments URL: https://news.ycombinator.com/item?id=48877371 Points: 41 # Comments: 11
Every primary-source story across every tracked model. Filter by clicking a chip.
Article URL: https://gist.github.com/cereblab/dc9a40bc26120f4540e4e09b75ffb547 Comments URL: https://news.ycombinator.com/item?id=48877371 Points: 41 # Comments: 11
submitted by /u/TheRealMasonMac [link] [comments]
funny that we're back to the era of Opus 3 vs GPT-4TOpus was slower, more expensive, and frankly less usefulbut it had more SOVL, wit, verbal…
Grok is the most politically neutral and objectively truth-seeking AIBrivael Le Pogam: Encore un scandale.On accuse Elon Musk sans arrêt de se servir de Grok pour…
Hard to believe that Moonshot only released their first frontier LLM a year ago. And what an LLM. They played up agentic coding, but K2 was…
Truly the greatest post-training machine on Earthbut even OpenAI can't overcome the physics of LLMsMythos is a stronger base. Fable has a higher floor. Who knows…
Article URL: https://mrzk.io/posts/qmlx-maximising-ai-psychosis-minmaxing-mac-studio/ Comments URL: https://news.ycombinator.com/item?id=48876619 Points: 5 # Comments: 0
5.5 to 5.6 is a notable improvement in partial knowledge cutoffGemini 3.5 and 3.1 are identicalthe cursed bloodline refuses to acknowledge the passage of timeApoorv Saxena:…
Decided to collect what passes for AGI manifestos of Chinese company CEOs:@jietang of Zhipu: https://x.com/bingxu_/status/2075961011816092158Zhilin Yang of Moonshot: https://mp.weixin.qq.com/s/uqUGwJLO30mRKXAtOauJGAEddie Wu of Alibaba: https://x.com/Sino_Market/status/1970675035255005544Wenfeng Liang of DeepSeek:…
WEIRD DISCLAIMER: none of this was written by an LLM until you get to the Github repo/site, which was obviously assembled by your friend and mine,…
RT Mark KretschmannGrok 4.5 by @SpaceXAI is the most neutral AI model out there. Almost perfectly balanced between the political Left and Right.No other model comes…
Article URL: https://www.androidauthority.com/claude-latest-models-pushback-bad-3683521/ Comments URL: https://news.ycombinator.com/item?id=48875494 Points: 17 # Comments: 14
RT Samuel Cardilloand yes, i would say that @elonmusk mission with @SpaceXAI is going well. the results definitively shows how Grok 4.5 training pushed even further…
Which settings would suffice to work with it ? submitted by /u/soyalemujica [link] [comments]
Built around a recent sd.cpp release, aims to expose most of what the backend can do (generate, edit, video paths, models, hardware options), Windows + Linux…
How far can i stretch the context window with Qwen 3.6 27B (using Q8_0) before it gets too unreliable? I am at 100k right now and…
OpenAI's GPT-5.6 Sol Ultra produced a proof of the Cycle Double Cover Conjecture in under an hour, using 64 subagents working in parallel. The conjecture had…
Fable noticed a cool detail about DeepSeek Legal team job posting:a LoGH referenceTeortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞): DeepSeek job postings this season go almost unbelievably…
"physicians found fewer flaws in GPT-5.6 responses than physician-written responses."Karan Singhal: ♥️ GPT-5.6 is a major step forward for health, both at the frontier and at…
Upon request, here's an updated version with Grok 4.5 and Meta's Muse Spark 1.1.Grok 4.5 seems to sit at the Pareto frontier. Good bang for the…
Muse Spark 1.1 is surprisingly close to Grok 4.5 on many high-signal evalsThis is the current top on CritPTArtificial Analysis: Meta's Muse Spark 1.1 scores 51…
Trust and security has always been crucial to large enterprises. With AI advances, it's never been more important.The partnership between Cohere, @Nvidia, & @CoreWeave ensures reliability…
ChatGPT is hiring a dedicated product manager to build experiences for families, caregivers, and older adults, according to a job posting.
Hey guys, we’ve built fastest speculative decoding for Qwen at least. To run in sglang you can use our fork Hf link Would appreciate your feedback…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.