litellm-alternative — independent software & tools
-
ollama-mesh
The control plane for self-hosted AI inference. Warm-state GPU routing, multi-runtime orchestration across Ollama, vLLM, llama.cpp, and TGI. Single Go binary. Apache-2.0.
-
nenya
A lightweight, highly secure AI API Gateway/Proxy written in Go. Acts as transparent middleware between local AI coding clients (OpenCode/Pi/Cursor) and upstream LLM providers (Gemini, DeepSeek, Zhipu z.ai).
-
opencode-local-provider
Connect OpenCode to local LLM servers like Ollama, vLLM, and LM Studio using one provider with automatic runtime model detection.