Local Ai
44 items across the graph — tagged with Local Ai.
From the graph · 44
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)
Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
A curated list of awesome platforms, tools, practices and resources that helps run LLMs locally
Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs + ComfyUI 100% offline. One installer, no Docker,…
A desktop app for running Large Language Models locally.
QVAC - Local AI SDK and libraries for building private, cross-platform, peer-to-peer AI applications. Run LLMs, speech-to-text, translation, and more locally on…
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.
AI semantic search for Zotero, with a built-in MCP server for AI agents (Claude Code, Codex). Find papers by meaning. 100% local and private.
Open-source local AI workspace — advancing on-device inference.
Zero-overhead, terminal-native local-LLM runtime manager. Launches, supervises, and routes local models behind one OpenAI-compatible endpoint.
Run large language models locally on Intel Macs with AMD GPUs - native macOS app with Metal acceleration
A small utility to use Ollama models in LM Studio (and vice versa via symlinking) without duplicating disk space
A complete private computer on a USB stick. PAI is a bootable Debian 12 Linux distro that runs Ollama locally on any x86_64 or ARM64 machine — full desktop, loc…
AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless lo…
A personal AI agent that can work safely on your machine, remember useful context, and keep its data under your control.
Intelligent Cost-Optimizing Model Router for OpenClaw
System monitor with offline AI (82 intents, 9-layer routing) that learns YOUR PC. TURBO optimization, thermal baselines, voltage SPC, ghost driver detection, 37…
A local-first workspace for chat, private knowledge, artifacts, and isolated model-driven work. Self-hosted. Provider-flexible. Apache 2.0.
Obsidian vault intelligence
Press ⌥C on anything. Transform with AI, scripts, and shortcuts on any selection. 100% local and private.
🐬 A pink qFlipper fork with LOTEI — a 100% local AI dolphin (Ollama) that chats, talks, watches your Flipper's screen, and one-click-installs custom firmware.
Native macOS menu bar app for realtime dictation with optional LLM polishing. Connects to any OpenAI Realtime-compatible backend — fully local on Apple Silicon…
Scryptian - inline text editing via Ctrl + Alt
Off Grid AI — private, on-device AI. Run open models (text, vision, image, voice) locally through one OpenAI-compatible gateway. No cloud, no accounts, no API k…
An AI agent that lives on your computer and does the work you ask for, in plain words — from writing code to running your business busywork. Local, one command,…
Multi-role AI orchestration runtime for Windows. Quality-gated pipeline (Brain → Miranda → Pappy → Benson) with MCP support, SQLite persistence, and a self-impr…
Lightweight AI agent runtime for homelabs, built around deterministic skills and local LLMs.
An LLM-powered social simulation engine
DeepSeek-V4-Flash-0731 284B inference in ~30 GB of RAM on any M-series MacBook
🎙️ Your privacy-first voice-to-text tool; Local Whisper transcription with optional LLM enhancement so your audio never leaves your computer 💜
Powerful no-code LLM fine-tuner: upload data → train → deploy in minutes. Unsloth 2-5× acceleration · QLoRA/DPO/RLHF/PPO/ORPO · Reward Model training · GGUF exp…
Off Grid AI — private, on-device AI. Run open models (text, vision, image, voice) locally through one OpenAI-compatible gateway. No cloud, no accounts, no API k…
the full AI stack. one browser tab. zero backend.
Local translator powered by Ollama and TranslateGemma
Make local LLM inference faster with chunk-level KV cache reuse
Este repositório visa ensinar o usuário a como usar IAs localmente da forma correta e encontrar a melhor configuração pro modelo escolhido de forma automática e…
🦞 ClawBox — Your private AI assistant on NVIDIA Jetson. Setup wizard, dashboard, and 580+ skills. Plug in, scan QR, done.
Unified KV cache compression for LLM inference — TurboQuant, IsoQuant, PlanarQuant, TriAttention. 10 methods, GPU-validated, multi-GPU planner. Compress KV cach…
A blazingly fast, privacy first (work with local LLMs), memory-efficient terminal-based AI coding assistant written in Rust
