RWKV is an RNN with great LLM performance and parallelizable like a Transformer.
submitted by /u/yogthos [link] [comments]
submitted by /u/yogthos [link] [comments]
https://thinkingmachines.ai/news/introducing-inkling/ submitted by /u/WhyLifeIs4 [link] [comments]
LLMs handle speech well once you run speech-to-text. They don't hear the rest: a bird outside, a glass breaking two rooms away, a smoke alarm two…
The full quote: I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as…
MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-A3B) 🔍 Long-horizon Search BrowseComp…
Disclosure: I work at Pluralis Research, the lab that built this. Code is open, and I'm happy to answer questions. TL;DR: As far as we can…
submitted by /u/yogthos [link] [comments]
https://youtu.be/oxpGq5FITgA?si=nkHWLReGCDYe7QfL I got Hermes running in the native Debian Terminal in Graphene OS and its really slick. Voice dictation works amazingly. Im using a remote Hermes…
Don't get me wrong, all the big models are amazing, and every contribution to open source models is great. But I'm GPU poor and I can't…
I set out to make Parakeet speech-to-text faster on CPU. Not a GPU demo. Just a real baseline and a keep gate. I did not want…