I built a local LLM NPC backend focused on NPC-to-NPC conversations
I just released a research project I did last year as open source. It is a fully local speech-to-speech backend for LLM NPCs. So speech-to-text, local…
I just released a research project I did last year as open source. It is a fully local speech-to-speech backend for LLM NPCs. So speech-to-text, local…
submitted by /u/pscoutou [link] [comments]
I want to up my game, with RAG. I tested it ages ago, but haven't found a usecase. I play with coding, projects, and light sysadmin…
Application Screenshot I made some updates to my agentic harness targeting legacy machines (Windows XP on .NET 4.0), figured I'd share them. Reasoning level is now…
Hi yall, just wanted to share that - i'm out. I had a blast reading through most of the threads daily, i waited with you hoping…
For context, this week they struck a deal to buy Nvidia chips and run local models for their enterprise clients. So in this video he is…
If you generate structured output from an LLM and validate against a schema, you know the failure mode: usually fine, occasionally a missing field or an…
Every time I need to deploy a model I end up manually testing 3-4 quant formats to see what actually holds up on my hardware VRAM,…
is it possible to run a reasonable quant of qwen 3.6 (or I guess even ornith/3.5) 35B A3B on a laptop with 32gb of lpddr5 (7500…
I’m pretty jaded like most of y’all. I don’t really get excited by new models much anymore. Last few weeks have been kinda meh to be…