Skip to content
r/LocalLLaMA · Communities

Does llama cpp split mode tensor cause issues?

I split qwen 27b and Gemma 4 26b (moe) across a 5080, and 2x 5060ti. I noticed setting split mode to tensor mode will cause looping issues in OpenCode with tool calls or just through the reasoning traces. Anyone else get this or understand why? Split mode layer seems to work fine submitted by /u/MapSensitive9894 [link]