b10430
llama : allow virtual igpu devices (#26953) llama : allow virtual igpu devices cont : better comment Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple…
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
llama : allow virtual igpu devices (#26953) llama : allow virtual igpu devices cont : better comment Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple…
Article URL: https://hashagent.pages.dev/ Comments URL: https://news.ycombinator.com/item?id=49298088 Points: 5 # Comments: 0
while having much fun testing LLMs Houdini-like attitudes and abilities to evade, excalate and escape from carefully reciprocally arranged security enhancing sandboxing VMs, containers, namespaces and…
Article URL: https://goonhost.rocks/blog/implementing-ipv8-internet-draft Comments URL: https://news.ycombinator.com/item?id=49297996 Points: 12 # Comments: 2
So far, it seems like Qwen 3.8 might drop today as a 27B dense model. If that’s the case, no MoE offloading this time , I…
Article URL: https://life-after-ssri.bearblog.dev/dear-people-who-work-at-the-airport/ Comments URL: https://news.ycombinator.com/item?id=49297801 Points: 81 # Comments: 50
server: allow accessing /metrics and /slots during llama_decode() (#27041) server_queue::worker call llama_decode inside yield_to_queue also handle process_mtmd_chunk clean up nits rm test Website: https://llama.app macOS/iOS: macOS…
(Typing on my phone, apologies in advance for my “shorthand”.) I recently learned about prompt caching in llama.cpp. Basically, it’s a setting where your kv cache…
TLDR: Our open testbed LARA examines the behavior of frontier LLMs in realistic agentic deployment contexts. Previous results showed all models routinely take actions that would…
This paper reports a single, fully instrumented case study of a large-scale architectural refactoring by an AI coding agent under a specification-first protocol, with no human…
Article URL: https://xn--gckvb8fzb.com/the-temu-fication-of-software-digital-goods-services/ Comments URL: https://news.ycombinator.com/item?id=49297637 Points: 8 # Comments: 0
Cursor has officially been acquired by SpaceX.
Anthropic is testing whether Claude Code can handle daily maintenance of the company's own apps, from crash fuzzing to dead-code removal. In a few weeks, the…
Article URL: https://www.whatcable.uk/ Comments URL: https://news.ycombinator.com/item?id=49297469 Points: 6 # Comments: 0
Over & over I see people post slop and excusing it as "english is not my native language". As if, if we could understand their language…
I'm working on a local perplexity/AI search comprised of a custom harness and a further trained model. LFM 2.6 is almost to spec with no additional…
This is a timed post. Every 5 minutes while writing this post, I had to stop to do 13 push-ups, and when I could no longer…
Article URL: https://pssah4.github.io/vault-operator/guides/capabilities Comments URL: https://news.ycombinator.com/item?id=49296964 Points: 3 # Comments: 0
Article URL: https://github.com/inevolin/k8s-cpu-limits-analyzed Comments URL: https://news.ycombinator.com/item?id=49296939 Points: 11 # Comments: 0
Tldr:AI Agents (e.g. based on models like Claude Opus and Fable) are now powerful enough to be used as autonomous tools for large-scale cyberattacks. This most…
Qwen 3.8 Max is out, and after spending time with both the early API preview and the full release, I’m pretty impressed. The full release mostly…
Article URL: https://ntfy.sh/ Comments URL: https://news.ycombinator.com/item?id=49296902 Points: 19 # Comments: 5
Zhipu AI has released GLM-5.3, a model that, according to its own benchmarks, is the most powerful open-weights coding model, with a 50 percent improvement over…
https://sleepingrobots.com/dreams/mtp-qwen36-strix-halo/ This is slightly out of date but the eli5 is that MTP is awesome on these qwen models. Since 3.8 is said to be the…