On Device Ai
14 items across the graph — tagged with On Device Ai.
From the graph · 14
Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code
26m agentic model for tiny devices
On-device LLM execution in React Native with Vercel AI SDK compatibility
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache compression, MACOS + i…
An on-device LLM understands your entire life, then proactively offers to get your work done through computer use.
Build apps powered by on-device AI
电子鹦鹉 / Toy Language Model
Open-source local AI workspace — advancing on-device inference.
A ternary, zero-heap tiny language model that runs inside a $2 microcontroller — bit-exact Python C99 Cortex-M3 (QEMU) parity. Apache-2.0.
Neutral, reproducible benchmark for local LLMs on Apple Silicon (Mac · iPhone · iPad) — MLX, llama.cpp, CoreML, Apple Foundation Models
High-performance on-device LLM inference for React Native, powered by LiteRT-LM and Nitro Modules
Run LLMs, VLMs, ASR, TTS, diarization and more fully on-device with Apple's Core AI framework (iOS/macOS 27) — one line of Swift per model, 53 models pinned to…
DeepSeek-V4-Flash-0731 284B inference in ~30 GB of RAM on any M-series MacBook
