Superfluid is an open-source local LLM server built for multi-agent inference
The Apache-2.0 server speaks the OpenAI, Anthropic and Ollama APIs and runs GGUF via llama.cpp or MLX models. superfluid launch claude --model <gguf> points Claude Code at a model running on your own machine.