Qwen3.8-27b on RTX 3090 – 82 tps single request, up to 672 tps peak
Hi, After a long night of optimizations, I believe I have made the fastest inference engine for Qwen3.6-28B on a 3090. Quick metrics: - 250w power…
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
Hi, After a long night of optimizations, I believe I have made the fastest inference engine for Qwen3.6-28B on a 3090. Quick metrics: - 250w power…
Hello guys, hoping you're doing well. I bring this discussion since I have noticed on internet, be USA or EU, RTX 6000 PROs at 16000USD or…
Did a quick local test because I wanted to see what is actually usable on my 12GB laptop GPU. I tested the newer Qwen3.8 27B dense…
I have a small coding test, where I ask a model to implement a simple CLI from a spec file. Qwen3.6-27b can do it in ~50k…
Article URL: https://w4g1.dev/blog/models-are-getting-dumber-on-purpose Comments URL: https://news.ycombinator.com/item?id=49322695 Points: 59 # Comments: 19
Tl;dr: Our timelines haven’t changed much (they got slightly shorter) but our modeling and evidence base have noticeably improved, so we feel somewhat more confident.SummaryWe intend…
IMO OLMo 1-3 still underrated in their contributions to science. I wish we could’ve spent more time amplifying this while building them.
Article URL: https://buf.build/blog/protobuf-lsp Comments URL: https://news.ycombinator.com/item?id=49322573 Points: 22 # Comments: 4
Intro and tl;drI'm Esa Koskinen, working as a volunteer director of AI Safety Tokyo.Epistemic status: estimates by an interested party; I run one of the organizations…
On the Baker-Anthropic Conversation, Power Acquisition and Escape VelocityWill there be an AI hegemon, i.e. an entity that, by wielding superintelligence, acquires so much economic or…
Article URL: https://math-ai-org.github.io/mathcode/ Comments URL: https://news.ycombinator.com/item?id=49322330 Points: 6 # Comments: 1
Article URL: https://www.lorekit.io/blog/give-your-agent-a-memory Comments URL: https://news.ycombinator.com/item?id=49322206 Points: 5 # Comments: 1
https://developer.nvidia.com/blog/serve-qwen3-8-2-4t-a95b-a-2-4t-parameter-model-with-configurable-reasoning-on-nvidia-gb300-nvl72/ 4k tokens per second per GPU of which there are 72. 350 tokens per second per user "Without additional model tuning, the model achieves a…
A few hours ago I switched my nameservers to Cloudflare in order to enable R2 bucket serving through my own subdomain, and I found out that…
https://x.com/i/status/2088993948983246906 Not tested by me in any way :) submitted by /u/jacek2023 [link] [comments]
The Doomsday argument was proposed in a 1983 lecture by Brandon Carter and elaborated on in John Leslie's 1996 book The End of the World: The…
What happens when humans put AIs in charge of civilisationally important decisions? A frontier AI company might hand over internal decisions (R&D, safety, deployment) or external…
Article URL: https://www.sciencedaily.com/releases/2026/08/260811052857.htm Comments URL: https://news.ycombinator.com/item?id=49321783 Points: 11 # Comments: 5
Article URL: https://rvembedded.com/blog_post/12/ Comments URL: https://news.ycombinator.com/item?id=49321717 Points: 5 # Comments: 2
RT Gokul RajaramIdeas are the new bottleneck@akshaynathan_, Core Product Engineering, @OpenAI, interviewed by @swyx and @Vibhu (@LatentSpacepod)Summary: Akshay Nathan runs the productivity pillar at OpenAI, the…
Including the rationalisation for the data below - this is a more robust version of an earlier post I did similar to this - explaining below:…
Dario Amodei is pushing back against the idea that he's been painting an overly pessimistic picture of AI.
I was looking into the story a bit further earlier. Very interesting. Couldn’t have done it without him submitted by /u/on_line187 [link] [comments]
Article URL: https://www.science.org/content/article/nih-ending-key-grant-budding-clinical-researchers Comments URL: https://news.ycombinator.com/item?id=49321353 Points: 4 # Comments: 0