Skip to content
r/LocalLLaMA · Communities

Intern S2 Mobius

A Qwen3.5-35B derived model with an interesting architectural difference that results in larger throughput and less token consumption (allegedly): https://huggingface.co/internlm/Intern-S2-Mobius submitted by /u/Miserable-Dare5090 [link] [comments]