r/LocalLLaMA
· Communities
DeepSeek-V4-Flash-0731 Q8_K_XL sometimes stops mid-task in OpenCode – anyone else seeing this?
Hey everyone, I've been experimenting with the new DeepSeek-V4-Flash-0731 release locally using the Unsloth Studio Q8_K_XL GGUF with OpenCode. Overall, it's been working really well, but I've noticed a strange behavior during longer agentic coding sessions. Once the context gets above ~100K tokens, the model will somet