llama-cpp
Here are 212 public repositories matching this topic...
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text, Stable Diffusion, tool calling, and local-network servers. Runs on your CPU, GPU, or NPU. No account, no API key, zero data leaves your device.
-
Updated
Aug 30, 2026 - TypeScript
Maid is a free and open source application for interfacing with llama.cpp models locally, and with Anthropic, DeepSeek, Ollama, Mistral and OpenAI models remotely.
-
Updated
Aug 31, 2026 - TypeScript
Atomic Agent is a local-first AI agent. Runs open-weight models on your own machine via llama.cpp.
-
Updated
Aug 31, 2026 - TypeScript
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
-
Updated
Aug 11, 2026 - TypeScript
Build and run AI agents using Docker Compose. A collection of ready-to-use examples for orchestrating open-source LLMs, tools, and agent runtimes.
-
Updated
Jun 4, 2026 - TypeScript
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
-
Updated
Aug 31, 2026 - TypeScript
On-device memory layer for AI agents. Claude Code, OpenClaw and Hermes. Hooks + MCP server + hybrid RAG search.
-
Updated
Aug 18, 2026 - TypeScript
Murasaki 系列模型官方推理前端。原生支持CoT(思维链)与长上下文的次时代ACGN 翻译引擎,一键高质量翻译轻小说、字幕与游戏和漫画文本。
-
Updated
Mar 20, 2026 - TypeScript
InferrLM - On-device AI for iOS & Android
-
Updated
Jun 29, 2026 - TypeScript
Self-hosted AI workspace where chat becomes visual workflows, multi-agent operations, and reviewable automations. Local memory; local or cloud models
-
Updated
Jul 20, 2026 - TypeScript
Off Grid AI — private, on-device AI. Run open models (text, vision, image, voice) locally through one OpenAI-compatible gateway. No cloud, no accounts, no API keys.
-
Updated
Aug 31, 2026 - TypeScript
Autonomous LLM-driven Among Us simulation with memory, social reasoning, pathfinding, sabotage, emergent behavior, and real-time React/PixiJS visualization.
-
Updated
Dec 3, 2025 - TypeScript
Private, on-device AI desktop app — GGUF (llama.cpp) & MLX, running Muse-Glimmer and Qwen3.8 with their native reasoning-effort ladders wired in as real controls. Local coding agent, RAG knowledge base, Deep Research, vision and voice. 100% offline, no account, no telemetry. Windows & macOS.
-
Updated
Aug 31, 2026 - TypeScript
A TUI around llama.cpp for running, managing, and benchmarking local GGUF models and launching the pi coding agent against your local server.
-
Updated
Aug 31, 2026 - TypeScript
Local AI coding agent for your terminal. Open-source, offline alternative to Claude Code, Cursor & Copilot — powered by Ollama and any local LLM. Private by default, free forever.
-
Updated
Jul 13, 2026 - TypeScript
Complete guide to running local AI on AMD RX 580 8GB via Vulkan — llama.cpp, Ollama, OpenWebUI, Stable Diffusion. No CUDA. No cloud. Free.
-
Updated
Jun 23, 2026 - TypeScript
Add this topic to your repo
To associate your repository with the llama-cpp topic, visit your repo's landing page and select "manage topics."