r/LocalLLaMA
· Communities
promptchain: a Python library + MCP server for swapping local LLMs on a single GPU
a few weeks ago i posted about promptchain a streamlit app that handles model loading and unloading for various backend and after seeing some feedback i have decided to make it into an small Python library for driving local LLMs (LM Studio, Ollama, any OpenAI-compatible server) with unified streaming across local and c