Acp
49 items across the graph — tagged with Acp.
From the graph · 49
Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.
Open-source 24/7 Cowork app for OpenClaw, Hermes, Claude Code, Codex, OpenCode and 20+ more CLI Agent | Customize your assistants | Team them up|Star if you lik…
Production ready toolkit to run AI locally
✨ AI Coding, Vim Style
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top!
State of the Art Natural Language Processing
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
RamaLama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production,…
Instant, controllable, local pre-trained AI models in Rust
Bonsai Demo
Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, e…
The easiest way to use Ollama in .NET
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer. Join our Discord: https://discord.com/invit…
AI pair programming in your terminal — one static binary, sub-ms startup, any model
Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Build apps powered by on-device AI
Python SDK for ACP clients and agents.
llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.
llama.cpp (GGUF LLMs) and llava.cpp (GGUF VLMs) for ROS 2
🚢 Yet another operator for running large language models on Kubernetes with ease. Powered by Ollama! 🐫
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Private on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometr…
Zero-overhead, terminal-native local-LLM runtime manager. Launches, supervises, and routes local models behind one OpenAI-compatible endpoint.
Just a open Claude Code Remote Control
Synthetic Autonomic Mind - An AI assistant for everyone.
Go implementation of AI coding agent
General-purpose agent in one static Go binary. ReAct loop, ACP server for IDEs, OpenAI-compatible REST API with embedded web UI, Telegram gateway, cron schedule…
lcpp is a dart implementation of llama.cpp used by the mobile artificial intelligence distribution (maid)
VT.ai - multimodal AI chat app with dynamic conversation routing
ezlocalai is an easy to set up local artificial intelligence server with OpenAI Style Endpoints.
Ruby FFI bindings for llama.cpp to run open-source LLMs such as GPT-OSS, Qwen 3.5, Gemma 4, and Llama 3 locally with Ruby.
AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless lo…
A curated list of awesome agentic commerce resources — protocols, MCP servers, tools, apps, APIs and services for AI agents that shop, sell and transact. For st…
Multi-Agent Mesh · Connect · Orchestrate · Scale
Inference Hub for AI at Scale
User friendly GUI for configuring and launching llama.cpp
A memory-first AI agent that remembers why decisions were made — not just the last message. Runs local (Ollama), cloud (Claude · OpenAI · Gemini), or decentrali…
One window for all your AI coding agents. Run Claude, Codex, OpenCode, Gemini, Antigravity, Cursor, and Copilot side-by-side. Terminal and chat, any layout.
Advanced code editor using local AI
Self-hosted meeting transcription portal — speech-to-text, speaker diarization, LLM-corrected transcripts, structured summaries and Word minutes, on your own GP…
A web-based memory usage and performance calculator for Huggingface GGUF models
Данный проект основан на llama.cpp и компилирует только RPC-сервер, а так же вспомогательные утилиты, работающие в режиме RPC-клиента, необходимые для реализаци…
Rust SDK for building AI agents with local OpenAI-compatible servers (LMStudio, Ollama, llama.cpp, vLLM). Features streaming, tools, hooks, retry logic, and com…
A POSIX shell with AI in the grammar: type ?? to drive Claude Code, Cursor CLI, Gemini CLI, or Ollama over ACP. Local-first, no login, no telemetry.
AI workspace and runtime for projects, chats, providers, memory, approvals, and supervised agents.
Local LLM Server Manager + LlaMA.cpp + Chat
Think first, then code. Where Claude Code and Codex jump straight to writing, this stays review-first: a code reviewer, security scanner, root-cause debugger, a…
