r/LocalLLaMA
· Communities
For Local, what are your minimum good or usable tokens per second, for both promp processing and text generation?
Hello guys, hoping you're doing fine. Lately with all the new models, and how popular is offloading, what are your min good or usable t/s for both PP and TG? Speaking on my case, I think PP about 300-350t/s for min, and for TG, about 9 t/s min. What about yours? submitted by /u/panchovix [link] [comments]