Ling 3.0 Flash on Strix Halo
vLLM ROCm/HiP, 4 bit compressed-tensors (int4) Not a fair comparison, but Qwen-122b on the most optimized format possible I have run (rocmFP4) does not touch Ling…
vLLM ROCm/HiP, 4 bit compressed-tensors (int4) Not a fair comparison, but Qwen-122b on the most optimized format possible I have run (rocmFP4) does not touch Ling…
Meta has been trailing competitors. Zuckerberg thinks he's found a way forward.
Hi HN!I run a few Claude Code sessions in parallel and kept cmd-tabbing around just to find out one of them had been sitting on a…
Personally, I believe it would be helpful for the alignment community to somehow quantify how much of a given piece of research goes directly into alignment…
Benchmarking video-language models has largely focused on short clips and single-sentence metrics, leaving open whether current systems can generate accurate long-form, paragraph-level descriptions. We introduce CLIP-CC-Bench,…
ggml : require contiguous src for ROLL on CUDA and Metal (#25928) ggml_roll only asserts nb[0] == ggml_type_size, so a permuted src is a valid input,…
Turn-taking is a central component of full-duplex interaction. Which turn-taking behaviors are appropriate varies with the scenario, yet current models apply a single norm regardless of…
Recently, a man I was rock climbing with told me about how he'd used AI to make a motivational poster for himself, which he'd hung on…
Based on: Potosnak, W., Wolff, M., Cao, M., Ma, R., Konstantinova, T., Efimov, D., Mahoney, M.W., Oreshkin, B., & Olivares, K.G. "Forking-Sequences: Statistically and Computationally Efficient…
In the wake of the letter calling on us to prepare to potentially Pace the Frontier, there has been much discussion of when pacing the frontier…
Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT-5.6-LunaEven worse: it seems…
A few months ago, I had the idea of making a LLM from scratch as a personal project (for learning and partly for improving my resume).…
RT Paolo RossonGot Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far.Fastest right now: Ollama's…
Article URL: https://support.claude.com/en/articles/16266773-how-claude-marks-ai-generated-content Comments URL: https://news.ycombinator.com/item?id=49250109 Points: 14 # Comments: 5
Article URL: https://discuss.google.dev/t/trusted-automation-with-google-antigravity-scaling-secure-finance-integrations-from-40-days-to-5/383313 Comments URL: https://news.ycombinator.com/item?id=49249986 Points: 3 # Comments: 0
Little less smart than Qwen, but way fewer tokens per task. submitted by /u/NoFaithlessness951 [link] [comments]
Epistemic status: Experimental framework created over a period of ~2-3 days during a hackathon at my home, and fairly heavily vibe coded. Expect some of this…
RT Nick Sortor🚨 JUST NOW: Anthony Fauci's successor, NIH Director Jay Bhattacharya, BLASTS Fauci for COVERING UP the alleged 82% MISCARRIAGE RATE for pregnant women taking…
RT ClineThe successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model…
On Monday, Mark Zuckerberg published a 6,500 word manifesto about personal AI, largely about the possibilities for the "personal superintelligence" systems Meta AI is building.
Article URL: https://arachnoid.com/lutusp/sailbook.html Comments URL: https://news.ycombinator.com/item?id=49249555 Points: 15 # Comments: 4
Article URL: https://www.massaschadeconsument.nl/collectieve-acties/playstation/ Comments URL: https://news.ycombinator.com/item?id=49249481 Points: 32 # Comments: 8
Is this model even abliterated? https://huggingface.co/DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF Because I asked it something and it literally said no. And no, it’s not gooner toons lmao. It’s for pentesting…
Amazon announces first off-the-grid data center in race to reap AI profits.