llama — independent software & tools
-
dgx-spark-ai
Run GPT-OSS 120B on NVIDIA DGX Spark with vLLM, build an API server, and create a local AI coding assistant
-
tera-pilot
Self-hosted, vendor-neutral coding agent for private repos and local models. TUI-first, BYOK across 16 providers or fully offline with Ollama/LM Studio. Every agent action is sandboxed, logged, and signed — built to be verified, not just trusted.
-
spark-ai-assistant-api
🚀 Run 120B AI Models on Spark 2026 - vLLM API & Coding Assistant
-
cline
Autonomous coding agent CLI - capable of creating/editing files, running commands, using the browser, and more
-
node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
-
fastapi-gen
Build LLM-enabled FastAPI applications without build configuration.
-
PDF-QA-RAG-System
📄 Transform your PDF documents into actionable insights with this RAG-based Question-Answering App for efficient and accurate responses.
-
local-ai
Run Qwen3.5 AI models locally on your iPhone without internet, cloud, or subscriptions, matching GPT-4o quality on mobile devices.
-
rag_service
-
gaia
Gaia is a command-line interface (CLI) tool for interacting with language models via a local API. It features a beautiful terminal UI, robust configuration management, and multiple interaction modes (default, describe, code, shell) for versatile assistance with programming, system administration, and more.
-
llamaindex
<p align="center"> <img height="100" width="100" alt="LlamaIndex logo" src="https://ts.llamaindex.ai/square.svg" /> </p> <h1 align="center">LlamaIndex.TS</h1> <h3 align="center"> Data framework for your LLM application. </h3>