Gemma4-26B-A4B & 31B-QAT Uncensored Balanced are out with MTP (35% & 53% speed boost)!
First of all, I'm stoked to announce we are almost at 20 million downloads on HF! (counted only on my own account, no duplicates/quants/finetunes/etc) and almost…
First of all, I'm stoked to announce we are almost at 20 million downloads on HF! (counted only on my own account, no duplicates/quants/finetunes/etc) and almost…
Hey fellow localites, Has anybody used DwarfStar with DeepSeek V4 Flash on 1x DGX Spark yet? What are your thoughts? Based on what I read, with…
Heres an educational resource to learn about LLM attention mechanism using very simple analogies with agents. The agents live inside a conway game of life style…
Hey all, So I have a Nvidia DGX Spark and an AMD Strix 395, both have 128GB of unified memory. The Spark has 200Gbit network and…
TL;DR: the recipe's image-build mods aren't actually public – I reconstructed them from the public kernels (with Claude) – and you have to build vLLM at…
Paper: https://arxiv.org/abs/2606.13894 GitHub: https://github.com/ndvbd/Gefen submitted by /u/indicava [link] [comments]
When I am talking with Chat GPT or Claude about abstract concepts, I am often surprised by how they seem kind of dumb... like they aren't…
I needed simple local image generation without the usual setup. No virtual environments, no ComfyUI with a complex graph and installation as an exe. So i…
I got one of these refurbished units last year. I have nothing but good things to say about it. It works great. It has heft to…
So Microsoft gives you GPT-4 for free in Copilot. They just don't give you an API for it. So I made one. It logs into your…