Made a new 350M model to compete with lfm2.5 but with an open license
I liked the idea of a nano llm, but decided to actually challenge myself with developing one. Keep in mind I developed this model, do your…
I liked the idea of a nano llm, but decided to actually challenge myself with developing one. Keep in mind I developed this model, do your…
Credit to /u/Superb-Translator236 for original posting on another sub - this sub doesn’t allow cross-posting. submitted by /u/Thrumpwart [link] [comments]
Arxiv paper: https://arxiv.org/abs/2606.03811 submitted by /u/Thrumpwart [link] [comments]
https://scalingintelligence.stanford.edu/blogs/hipkernels/ submitted by /u/Superb-Translator236 [link] [comments]
https://github.com/ggml-org/llama.cpp/pull/25222 Another win for Intel ARC users (all 4 of us). The community keeps improving llama.cpp for Intel ARC. This time, the hero from that Pull…
Hey folks. I've been frustrated by how difficult it is to get an idea of how good each new model (or fine-tune) is, and I've not…
With July 4th coming up, there will be idiots shooting off fireworks. That combined with everything being bone dry where I live, I generally stay up…
submitted by /u/9gxa05s8fa8sh [link] [comments]
I've been a lurker for a while and have been building my own home lab with P40's and MI50's. I've learned so much from the community…
Hey r/LocalLLaMA, Wanted to share a narrow fine-tune I've been working on and get some technical feedback from people who've done similar domain-specific work if possible.…