Skip to content
r/LocalLLaMA · Communities

promptchain: a Python library + MCP server for swapping local LLMs on a single GPU

a few weeks ago i posted about promptchain a streamlit app that handles model loading and unloading for various backend and after seeing some feedback i have decided to make it into an small Python library for driving local LLMs (LM Studio, Ollama, any OpenAI-compatible server) with unified streaming across local and c