r/LocalLLaMA
· Communities
Are there any interesting local versions of the OpenRouter "Fusion" or Sakana Fugu methods, being worked on, as of yet? I’m curious if these setups will enable us to use, say, a panel of various ~30B-ish local models together to get outputs closer to GLM quality without needing as much memory.
Given how much of a boost in output quality these methods seem to enable when it comes to the big cloud models of having several smaller, cheaper models give output qualities on par with or higher than the strongest, Fable-class model, I am curious about how these things, of something similar to the OpenRouter Fusion s