DavidAU somehow managed to improve Qwen 3.6 27B
I know DavidAU gets a bad rap, and rightfully so. I've tried some of his fine tunes in the past and they have been... Interesting. I…
I know DavidAU gets a bad rap, and rightfully so. I've tried some of his fine tunes in the past and they have been... Interesting. I…
Given the expansion of the fed/state/local/corporate surveillance net, do you find yourself expressing yourself less freely on the internet than you did five years ago? Comments…
Hi everyone, We've open-sourced Tri-Net v2, the official implementation accompanying our recently published Scientific Reports (Nature Portfolio) paper: "Tri-Net: Unified Deep Learning for Skin Lesion and…
It can only run the 1_0 quant in LM Studio. Only 3.8GB and runs on low-end hardware. Even a phone! https://prismml.com/news/bonsai-27b https://huggingface.co/prism-ml/Bonsai-27B-gguf 1 upvote submitted by…
I got my ML model training pipeline to go from 36 steps/minute to 47 steps/minute by optimizing how they're stored on disk. When training larger ML…
Hey everyone, Back when MTP came available on llama.cpp, it seemed like the common consensus was that MTP didn't matter much for MoE models. After spending…
I came across this study on X and tested it myself; the findings align with the results: FSRS is more recognizable than Jarrett Ye among LLMs.…
Nearly a year later, we've had safeguard fine-tunes but no general-purpose base model or successor. Will Kimi, Qwen and GLM force their hand? submitted by /u/prescorn…
https://huggingface.co/Motif-Technologies/Motif-3-Beta Motif-Technologies is one of the tech company participated South Korea's AI Foundation Model project. Upstage(Solar Series), LG AI Research(EXAONE Series), and SKT(A.X Series) are the…
Article URL: https://www.aclu.org/news/privacy-technology/tracking-alpr-cameras/flock-safety-credibility-lost-as-it-repeatedly-lies-to-city-councils-police-departments-and-public-across-the-country Comments URL: https://news.ycombinator.com/item?id=48986731 Points: 41 # Comments: 2
https://preview.redd.it/wluwwp6s4heh1.png?width=1248&format=png&auto=webp&s=6e95d963645c5a0ef750bf81324a6d4dcbc0389e Spent the last few days measuring speculative decoding on Qwen3.6-27B (dense, NVFP4) on one RTX PRO 6000 Max-Q, comparing vLLM and SGLang across MTP, DFlash,…
Four technologies used by Kimi K3 to make it SOTA: KimiDeltaAttention (used in Kimi Linear, essentially a more general gated delta net) AttnRes - described in…
I can't take credit for this, someone else described the approach, and I just copy/pasted it into opencode/GLM 5.2 to get it working. It appears to…
Trying to see what I can scrounge together bare minimum hardware requirements to get up to that rough speed. Right now I'm running an RX6600XT and…
Google hasn't shipped a model recently that is capable of competing with Sol or Fable. The previous models were pretty disappointing and unreliable, it seems the…
Article URL: https://forum.jellyfin.org/t-project-leadership-changes Comments URL: https://news.ycombinator.com/item?id=48986091 Points: 42 # Comments: 12
Article URL: https://words.filippo.io/passkey-record/ Comments URL: https://news.ycombinator.com/item?id=48985956 Points: 14 # Comments: 0
Kimi-K3’s release, while impressive, is still months behind the closed-source frontier, so all the “it’s over for Anthropic” talk feels overblown. According to Artificial Analysis, though,…
BackgroundThe AI safety field has spent a decade building tools for systems trained to think and digitally act. The next decade will likely deliver widely deployed…
In this post, I propose adapting banking risk management frameworks (specifically capital adequacy requirements like Basel III) to frontier AI labs. By forcing them to hold…
I completed this project over 2 weeks as part of a BlueDot Project cohort. It was my first solo project and I learned a lot! Feedback…
AI models are increasingly trained to be “safe,” meaning they refuse harmful requests. But what does it truly mean for a model to be safe? Today,…
The effectiveness of amphetamine (e.g. Adderall) and methylphenidate (e.g. Ritalin) for the symptoms of ADHD, that is, lack of executive function, has been demonstrated beyond reasonable…
(Slides because why not) I've been building a serverless post-quantum group-encryption protocol for the last few weeks, mostly with Opus 4.8 and Fable, with GPT-5.6 Sol…