MindControl – llama.cpp fork to guide the reasoning process via injection during sampling
The primary driver of this project is that I'd become frustrated with the reasoning behavior of smaller local models such as Qwen3.6-27B (i believe particularly at…