Acp
49 items across the graph — tagged with Acp.
From the graph · 49
Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.
Open-source 24/7 Cowork app for OpenClaw, Hermes, Claude Code, Codex, OpenCode and 20+ more CLI Agent | Customize your assistants | Team them up|Star if you lik…
Production ready toolkit to run AI locally
✨ AI Coding, Vim Style
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top!
State of the Art Natural Language Processing
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
RamaLama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production,…
Instant, controllable, local pre-trained AI models in Rust
Bonsai Demo
Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, e…
The easiest way to use Ollama in .NET
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer. Join our Discord: https://discord.com/invit…
AI pair programming in your terminal — one static binary, sub-ms startup, any model
Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Build apps powered by on-device AI
Python SDK for ACP clients and agents.
llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.
llama.cpp (GGUF LLMs) and llava.cpp (GGUF VLMs) for ROS 2
🚢 Yet another operator for running large language models on Kubernetes with ease. Powered by Ollama! 🐫
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Private on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometr…
Zero-overhead, terminal-native local-LLM runtime manager. Launches, supervises, and routes local models behind one OpenAI-compatible endpoint.
Just a open Claude Code Remote Control
Synthetic Autonomic Mind - An AI assistant for everyone.
Go implementation of AI coding agent
General-purpose agent in one static Go binary. ReAct loop, ACP server for IDEs, OpenAI-compatible REST API with embedded web UI, Telegram gateway, cron schedule…
lcpp is a dart implementation of llama.cpp used by the mobile artificial intelligence distribution (maid)
VT.ai - multimodal AI chat app with dynamic conversation routing
ezlocalai is an easy to set up local artificial intelligence server with OpenAI Style Endpoints.
AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless lo…
Ruby FFI bindings for llama.cpp to run open-source LLMs such as GPT-OSS, Qwen 3.5, Gemma 4, and Llama 3 locally with Ruby.
A curated list of awesome agentic commerce resources — protocols, MCP servers, tools, apps, APIs and services for AI agents that shop, sell and transact. For st…
Multi-Agent Mesh · Connect · Orchestrate · Scale
Inference Hub for AI at Scale
User friendly GUI for configuring and launching llama.cpp
A memory-first AI agent that remembers why decisions were made — not just the last message. Runs local (Ollama), cloud (Claude · OpenAI · Gemini), or decentrali…
One window for all your AI coding agents. Run Claude, Codex, OpenCode, Gemini, Antigravity, Cursor, and Copilot side-by-side. Terminal and chat, any layout.
Advanced code editor using local AI
Self-hosted meeting transcription portal — speech-to-text, speaker diarization, LLM-corrected transcripts, structured summaries and Word minutes, on your own GP…
A web-based memory usage and performance calculator for Huggingface GGUF models
Rust SDK for building AI agents with local OpenAI-compatible servers (LMStudio, Ollama, llama.cpp, vLLM). Features streaming, tools, hooks, retry logic, and com…
Данный проект основан на llama.cpp и компилирует только RPC-сервер, а так же вспомогательные утилиты, работающие в режиме RPC-клиента, необходимые для реализаци…
AI workspace and runtime for projects, chats, providers, memory, approvals, and supervised agents.
A POSIX shell with AI in the grammar: type ?? to drive Claude Code, Cursor CLI, Gemini CLI, or Ollama over ACP. Local-first, no login, no telemetry.
Think first, then code. Where Claude Code and Codex jump straight to writing, this stays review-first: a code reviewer, security scanner, root-cause debugger, a…
Local LLM Server Manager + LlaMA.cpp + Chat
