Topic

Mlx

25 items across the graph — tagged with Mlx.

From the graph · 25

repo
AlexsJones/llmfit

Hundreds of models & providers. One command to find what runs on your hardware.

repo
jundot/omlx

LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar

repo
osaurus-ai/osaurus

Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built in Swift. Fully offline…

repo
Blaizzy/mlx-vlm

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

repo
maziyarpanahi/openmed

Local-first healthcare AI: clinical NER & HIPAA PII de-identification that runs 100% on-device. 2,200+ medical models, 21 languages, Apple MLX + Python, no clou…

repo
PrismML-Eng/Bonsai-demo

Bonsai Demo

repo
jjang-ai/vmlx

vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont Batching + etc!

repo
kennss/SiliconScope

Sudoless Apple Silicon system monitor (native SwiftUI GUI) with ANE / Media Engine / memory-bandwidth tracking

repo
SharpAI/SwiftLM

⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache compression, MACOS + i…

repo
Fabric-Project/Fabric

Node Creative Coding / 3D / Image Processing tool inspired by Quartz Composer

repo
Gremble-io/Detto

On-device meeting capture, dictation, and vault-native knowledge layer for macOS. Nothing leaves your Mac.

repo
mudler/vllm.cpp

a community oriented 1:1, vLLM-alike (Continuous batching, paged KV) engine in C++ with additional features (for example, RadixAttention, Cache-aware scheduling…

repo
Goekdeniz-Guelmez/MLX-LoRA-Studio

A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.

repo
jjang-ai/jangq

JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon

repo
Mcourtyard/m-courtyard

M-Courtyard: Local AI Model Fine-tuning Assistant for Apple Silicon. Zero-code, zero-cloud, privacy-first desktop app powered by Tauri + React + mlx-lm.

repo
mlx-node/mlx-node
repo
rayl15/OpenVision

Open-source iOS app connecting Meta Ray-Ban smart glasses to AI — 5 backends (on-device MLX models, Apple Intelligence, OpenAI, Gemini Live, OpenClaw), on-devic…

repo
manjunathshiva/turboquant-mlx

Extreme weight + KV cache compression for LLMs on Apple Silicon (MLX implementation of Google's TurboQuant)

repo
john-rocky/apple-silicon-llm-bench

Neutral, reproducible benchmark for local LLMs on Apple Silicon (Mac · iPhone · iPad) — MLX, llama.cpp, CoreML, Apple Foundation Models

repo
zouyee/dmlx

Big models. Small Macs. Zero excuses.

repo
T0mSIlver/localvoxtral

Native macOS menu bar app for realtime dictation with optional LLM polishing. Connects to any OpenAI Realtime-compatible backend — fully local on Apple Silicon…

repo
outsourc-e/bench-loop

Local-first CLI for benchmarking LLMs on real hardware — quality, speed, reliability, and a real multi-turn agent loop.

repo
marzukia/qMLX

qMLX: Custom inference engine for Qwen 3.5 122B on Apple Silicon, extending MLX with hybrid attention support, SSD-backed KV cache, and RYS layer duplication fo…

repo
wydrox/ppmlx

Run local LLMs on Apple Silicon with an OpenAI-compatible API powered by MLX.

repo
matthewdcage/llm-swarm-router

Run the LLM Swarm Router on machines to distribute Local Ai to the Swarm - More Machines - MORE SPEED

Related topics