b10342
model : Granite-Switch Architecture (#25107) granite-switch: add llama.cpp backend (POC, CPU) New "granite-switch" architecture: a dense, all-attention Granite-4.1 model with N embedded LoRA adapters selected per-token…
model : Granite-Switch Architecture (#25107) granite-switch: add llama.cpp backend (POC, CPU) New "granite-switch" architecture: a dense, all-attention Granite-4.1 model with N embedded LoRA adapters selected per-token…
RT Jack RaeFun demo with Muse Glimmer: ask the model to deploy itself to the HuggingFace inference endpoint and optimize the inference efficiencyBen Burtenshaw: Meta is…
Day 0 support submitted by /u/jacek2023 [link] [comments]
RT Alexey FateevI reproduced an @NVIDIAAI paper in real vLLM on 2xRTX 3090s. A small model does the prefill, a completely different bigger model answers from…
Spark is the big news and is a good model. Not quite at the frontier of open models from China, and still well behind the closed…
Which galaxy will you choose?
«Slavs worship strength and brutality + are totally devoid of empathy to anyone but themselves»TrueAs opposed to Kamil, who felt sympathy for Ukrainians ( but only…
Article URL: https://bobdahacker.com/blog/tldv-hack Comments URL: https://news.ycombinator.com/item?id=49242739 Points: 14 # Comments: 3
Article URL: https://twitter.com/thdxr/status/2086599224674681242 Comments URL: https://news.ycombinator.com/item?id=49242728 Points: 10 # Comments: 2
Article URL: https://squeak.org/release_notes/6.1/ Comments URL: https://news.ycombinator.com/item?id=49242653 Points: 5 # Comments: 0
I continue to believe we should pause frontier AI development. Any discussion of alternative strategies should be thought of as planning for contingencies. A unifying driver…
I've been building a personal project and wanted to check with the community on multi-modal inputs since I can't find a lot of material around this…
But my worst fears about MAGA are not "disenfranchisement of women". If I were an American, that wouldn't be my top 5 fear about MAGA. First,…
Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can address misuse,…
RT Sebastian MajstorovicAre open OCR models good enough to unlock historical knowledge? @vanstriendaniel and I built a leaderboard to find out: 14 open models scored on…
Activation Oracles (AOs) are language models trained to answer natural-language questions about another model's internal activations. They offer a flexible interface for reading hidden information from…
Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.
Discovered Materials raised $9 million to fund the hunt for more novel materials to build more efficient chips.
RT Ali Zaidmy weekend hackathon is officially completebuilt superstage in ~18 hours for @swyx's "help kill my saas" hackathon...it's live, submitted for review, and i'm not…
To be clear Qwen 3.6 27B dates back to Apr 21, it's the same generation as V4-PreviewGemma is earlier in AprilSo most of this is just…
RT DogeDesignerBREAKING: Lufthansa Airlines just released new pictures of Starlink being installed on its aircraft. The first Starlink-equipped Lufthansa Airbus A320neo will enter service on August…
https://preview.redd.it/xxjh11f38jih1.png?width=1852&format=png&auto=webp&s=76850ed51e29a8bc86c2ca718d4320075eed4363 Just wanted to share a user report that I found to be very interesting. Some person with an intriguing name manu69x managed to run 1M…
Article URL: https://lwn.net/Articles/1034703/ Comments URL: https://news.ycombinator.com/item?id=49242297 Points: 13 # Comments: 0
Article URL: https://bytecode.news/posts/2026/08/because-it-s-not-fun-enough Comments URL: https://news.ycombinator.com/item?id=49242245 Points: 10 # Comments: 0