Qwen3.8-27B-int4-AutoRound (18GB) – with working MTP spec decode
submitted by /u/BusinessMud9586 [link] [comments]
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
submitted by /u/BusinessMud9586 [link] [comments]
Article URL: https://scholar.google.com/scholar?q=%22kidney+disappointment%22 Comments URL: https://news.ycombinator.com/item?id=49319389 Points: 5 # Comments: 0
Article URL: https://ei3lh.eu/2026/08/16/a-true-telnet-bbs-on-a-casio-calculator/ Comments URL: https://news.ycombinator.com/item?id=49319349 Points: 17 # Comments: 2
Article URL: https://www.ycombinator.com/companies/gooseworks/jobs/UJ4vH2F-founding-engineer Comments URL: https://news.ycombinator.com/item?id=49319215 Points: 0 # Comments: 0
This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more on AI safety, follow Robert Hart. The Stepback…
Article URL: https://www.heise.de/en/news/Access-to-telemetry-data-Automotive-industry-criticizes-intelligence-reform-11414994.html Comments URL: https://news.ycombinator.com/item?id=49319170 Points: 8 # Comments: 1
https://preview.redd.it/wkg27e152qjh1.png?width=853&format=png&auto=webp&s=2e3f8b11ea6393041f501e95c5835f9bea0245dd So ive been trying different harnesses and coding agents with the new qwen 3.8 , and after trying out many ive been mostly impressed by…
After: https://www.reddit.com/r/LocalLLaMA/comments/1vckcue/ds4_flash_0731_acquarium_panel_failure_q3_k_xl/ start C:llmllamaM6buildbinllama-server.exe --model "F:modelsQwen3.8-27B-UD-Q8_K_XL.gguf" --temp 1.0 --top-p 0.95 --top-k 20 --min-p 0.0 --presence-penalty 0.0 --reasoning-preserve --repeat-penalty 1.0 --ct
I wanted to do a study of whether prediction markets are well-calibrated aka when the market predicts a P% chance of some event, do such events…
Article URL: https://anthony.dev.profullstack.com/blog/012-post.html Comments URL: https://news.ycombinator.com/item?id=49319010 Points: 5 # Comments: 0
submitted by /u/juanviera23 [link] [comments]
A study involving Google researchers shows that when chatbots are trained not to claim consciousness, it also changes their stance on animal rights, religion, and life…
chat: refactor handling supports_string_content / supports_typed_content (#27130) better supports_string_content cap detect test: add "skip" messages_inp_normalizer Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64,…
I sadly was not able to follow as much as I wanted the new advancements. Care to share your optimised setups ? Vllm, llama.cpp or any…
submitted by /u/juanviera23 [link] [comments]
I did a short test of the different reasoning efforts, since on default xhigh the model thinks a lot. Not very scientific, just a quick "generate…
This is intended as a question about the current phase of the LLM hype cycle, and, at the same time, as a reality check about whether…
RT SpaceweaselMy thoughts on deepseek harness, from a very specific point of view:I an working on my own software delivery and hosting platform for agents, focusing…
Article URL: https://peterbloem.nl/blog/craft-coding Comments URL: https://news.ycombinator.com/item?id=49318735 Points: 27 # Comments: 8
vibeslop /vībˈslŏp/ noun a vibecoded project or mini-project so disposable it deserves its own special term. built purely for the fun of it. fun to show…
I see everyone gushing over 3.8, I get the impression people find it drastically better than previous Qwen models, but I can't believe it could be…
Only in America. Tired of winning?Reginald: Giving birth in Europe vs in America 😂
I'm just thinking, ever since microsoft announced bitnet, this sub (and myself) has been hoping for massive ternary models. In the last month alone, prismML dropped…
Modern LMs are trained on everything at once, so it is hard to tell whether a new skill was learned or merely elicited. We constrain the…