llm-gateway — independent software & tools
-
ollama-mesh
The control plane for self-hosted AI inference. Warm-state GPU routing, multi-runtime orchestration across Ollama, vLLM, llama.cpp, and TGI. Single Go binary. Apache-2.0.
-
aurora
Open-source AI gateway — route LLM traffic across OpenAI, Anthropic, Gemini, Groq & 30+ providers via a single API. Self-hosted, no vendor lock-in.
-
helm-api
Self-hosted LLM gateway — route OpenAI/Anthropic/Gemini traffic by config, not code. Fallback, protocol translation, telemetry.
-
IHUI-AI
Eight-platform full-stack AI operating system - unifies 176 LLMs via LangGraph + MCP + A2A. Multi-tenant RLS over 340 tables, RAG knowledge base, agent marketplace. Web/API/CLI/Desktop/Extension/Mobile/Miniapp. Apache 2.0.
-
litellm-multi-gateway
🚀 Multi-backend AI gateway built on LiteLLM — route Claude Code / Hermes / any OpenAI-compatible client to multiple LLM providers (ARK/Anthropic/Zhipu). Admin UI, per-client usage tracking, auto vision-to-text for text-only backends.
-
Agent402
500+ pay-per-call web tools + skill packs for AI agents — every one tested, priced, and settled on-chain. Paid in USDC over x402 on 10 chains, or free via proof-of-work on the pure-CPU tools. Self-hostable, MCP-native, deterministic. Plus the open x402 index (Find / Route / Leaderboard).
-
bifrost-plugin
Claude Code + Claude Desktop plugin for any Bifrost MCP gateway: one-command setup, skill discovery, agent-driven memory, signed admin policy, and OAuth 2.1 for Desktop.
-
StackFerry
A focused desktop provider manager and router for AI coding tools.
-
dgx-spark-llm-platform
Self-hosted multi-user LLM platform for the NVIDIA DGX Spark: OpenAI-compatible API (vLLM + LiteLLM), per-user keys & budgets, self-service portal with playground and an AI support assistant.
-
tokenpanel
Open-source AI reseller panel with customer API keys, prepaid balances, usage limits, model pricing, and profit analytics.
-
switchback
One Rust binary for explainable AI provider routing: multi-provider fallback, encrypted credentials, budgets, quotas, and metadata-only traces — no client code changes.
-
nenya
A lightweight, highly secure AI API Gateway/Proxy written in Go. Acts as transparent middleware between local AI coding clients (OpenCode/Pi/Cursor) and upstream LLM providers (Gemini, DeepSeek, Zhipu z.ai).
-
unified-ai-system
Terminal-first, self-hosted AI gateway with 8 MCP tools for Codex, Cursor, and Cline. One Docker command, no API key.
-
DIO
Drop-in OpenAI- and Ollama-compatible LLM gateway that learns each backend's latency online and routes vLLM / SGLang / TGI / Ollama with SLO-aware admission. No engine patches.
-
opencode-local-provider
Connect OpenCode to local LLM servers like Ollama, vLLM, and LM Studio using one provider with automatic runtime model detection.
-
ollabridge
OllaBridge transforms your laptop or workstation into a production-grade, OpenAI-compatible LLM provider. It features self-healing setup, built-in security, and automatic tunneling—making your local hardware ready for real applications in 60 seconds.