Topic

Local Ai

44 items across the graph — tagged with Local Ai.

From the graph · 44

repo
Mintplex-Labs/anything-llm

Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience

repo
screenpipe/screenpipe

YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)

repo
qualcomm/GenieX

Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code

repo
drumih/turbo-fieldfare

Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook

repo
Blaizzy/mlx-vlm

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

repo
Osmantic/ODS

Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.

repo
rafska/awesome-local-llm

A curated list of awesome platforms, tools, practices and resources that helps run LLMs locally

repo
PurpleDoubleD/locally-uncensored

Plug-and-play local AI studio: uncensored chat, image & video generation, coding agent. Runs abliterated LLMs + ComfyUI 100% offline. One installer, no Docker,…

repo
lone-cloud/gerbil

A desktop app for running Large Language Models locally.

repo
tetherto/qvac

QVAC - Local AI SDK and libraries for building private, cross-platform, peer-to-peer AI applications. Run LLMs, speech-to-text, translation, and more locally on…

repo
inclusionAI/AReno

An easy-to-use, fast toolkit to scale up RL post-training on a single node.

repo
hogeheer499-commits/strix-halo-guide

Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.

repo
introfini/ZotSeek

AI semantic search for Zotero, with a built-in MCP server for AI agents (Claude Code, Codex). Find papers by meaning. 100% local and private.

repo
zeraix/zeraix

Open-source local AI workspace — advancing on-device inference.

repo
llamastash/llamastash

Zero-overhead, terminal-native local-LLM runtime manager. Launches, supervises, and routes local models behind one OpenAI-compatible endpoint.

repo
engeldlgado/toshllm

Run large language models locally on Intel Macs with AMD GPUs - native macOS app with Metal acceleration

repo
qaribhaider/ollama-to-lmstudio-symlinks

A small utility to use Ollama models in LM Studio (and vice versa via symlinking) without duplicating disk space

repo
nirholas/PAI

A complete private computer on a USB stick. PAI is a bootable Debian 12 Linux distro that runs Ollama locally on any x86_64 or ARM64 machine — full desktop, loc…

repo
noumena-labs/Sipp

AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless lo…

repo
dosco/aithy

A personal AI agent that can work safely on your machine, remember useful context, and keep its data under your control.

repo
atharv404/ClawRoute

Intelligent Cost-Optimizing Model Router for OpenClaw

repo
HuckleR2003/PC_Workman_HCK

System monitor with offline AI (82 intents, 9-layer routing) that learns YOUR PC. TURBO optimization, thermal baselines, voltage SPC, ghost driver detection, 37…

repo
libre-webui/libre-webui

A local-first workspace for chat, private knowledge, artifacts, and isolated model-driven work. Self-hosted. Provider-flexible. Apache 2.0.

repo
cybaea/obsidian-vault-intelligence

Obsidian vault intelligence

repo
cai-layer/cai

Press ⌥C on anything. Transform with AI, scripts, and shortcuts on any selection. 100% local and private.

repo
DUNKINKKD/lotei-qflipper

🐬 A pink qFlipper fork with LOTEI — a 100% local AI dolphin (Ollama) that chats, talks, watches your Flipper's screen, and one-click-installs custom firmware.

repo
T0mSIlver/localvoxtral

Native macOS menu bar app for realtime dictation with optional LLM polishing. Connects to any OpenAI Realtime-compatible backend — fully local on Apple Silicon…

repo
adrianium/Scryptian

Scryptian - inline text editing via Ctrl + Alt

repo
off-grid-ai/OGAD

Off Grid AI — private, on-device AI. Run open models (text, vision, image, voice) locally through one OpenAI-compatible gateway. No cloud, no accounts, no API k…

repo
Miosa-osa/OSA

An AI agent that lives on your computer and does the work you ask for, in plain words — from writing code to running your business busywork. Local, one command,…

repo
junkyard22/Orca

Multi-role AI orchestration runtime for Windows. Quality-gated pipeline (Brain → Miranda → Pappy → Benson) with MCP support, SQLite persistence, and a self-impr…

repo
evgenii-engineer/openLight

Lightweight AI agent runtime for homelabs, built around deterministic skills and local LLMs.

repo
mistval/yozakura

An LLM-powered social simulation engine

repo
yanun0323/deepseek_ssd

DeepSeek-V4-Flash-0731 284B inference in ~30 GB of RAM on any M-series MacBook

repo
app-vox/vox

🎙️ Your privacy-first voice-to-text tool; Local Whisper transcription with optional LLM enhancement so your audio never leaves your computer 💜

repo
Yog-Sotho/LLM-fine-tuner

Powerful no-code LLM fine-tuner: upload data → train → deploy in minutes. Unsloth 2-5× acceleration · QLoRA/DPO/RLHF/PPO/ORPO · Reward Model training · GGUF exp…

repo
off-grid-ai/off-grid-ai-desktop

Off Grid AI — private, on-device AI. Run open models (text, vision, image, voice) locally through one OpenAI-compatible gateway. No cloud, no accounts, no API k…

repo
ixchio/n0x

the full AI stack. one browser tab. zero backend.

repo
dwain-barnes/local-translator

Local translator powered by Ollama and TranslateGemma

repo
pythongiant/KVBoost

Make local LLM inference faster with chunk-level KV cache reuse

repo
allanschramm/local-model-autotuning

Este repositório visa ensinar o usuário a como usar IAs localmente da forma correta e encontrar a melhor configuração pro modelo escolhido de forma automática e…

repo
ID-Robots/clawbox

🦞 ClawBox — Your private AI assistant on NVIDIA Jetson. Setup wizard, dashboard, and 580+ skills. Plug in, scan QR, done.

repo
aivrar/multi-turboquant

Unified KV cache compression for LLM inference — TurboQuant, IsoQuant, PlanarQuant, TriAttention. 10 methods, GPU-validated, multi-GPU planner. Compress KV cach…

repo
jyjeanne/crustly

A blazingly fast, privacy first (work with local LLMs), memory-efficient terminal-based AI coding assistant written in Rust

Related topics