Cpp
50 items across the graph — tagged with Cpp.
From the graph · 50
Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.
FinceptTerminal is a modern finance application offering advanced market analytics, investment research, and economic data tools, designed for interactive explo…
Learn OpenCV : C++ and Python Examples
⚡ Automatically decrypt encryptions without knowing the key or cipher, decode encodings, and crack hashes ⚡
Open3D: A Modern Library for 3D Data Processing
Production ready toolkit to run AI locally
Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allreduce, reductions, learnin…
Flower: A Friendly Federated AI Framework
High-performance AI pipeline engine with a C++ core and 50+ Python-extensible nodes. Build, debug, and scale LLM workflows with 13+ model providers, 8+ vector d…
A flexible, high-performance serving system for machine learning models
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
A curated List of Coding Questions Asked in FAANG Interviews
中国的Quant相关资源索引
Demystify AI agents by building them yourself. Local LLMs, no black boxes, real understanding of function calling, memory, and ReAct patterns.
Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top!
State of the Art Natural Language Processing
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
RamaLama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production,…
C++ DataFrame for statistical, financial, and ML analysis in modern C++
General purpose GPU compute framework built on Vulkan to support 1000s of cross vendor graphics cards (AMD, Qualcomm, NVIDIA & friends). Blazing fast, mobile-en…
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.
Instant, controllable, local pre-trained AI models in Rust
A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows
Bonsai Demo
A high performance anime upscaler
High-performance automatic differentiation of LLVM and MLIR.
Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, e…
Visual Studio Code client for Tabnine. https://marketplace.visualstudio.com/items?itemName=TabNine.tabnine-vscode
The easiest way to use Ollama in .NET
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer. Join our Discord: https://discord.com/invit…
React Native binding of llama.cpp
Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM
Sudoless Apple Silicon system monitor (native SwiftUI GUI) with ANE / Media Engine / memory-bandwidth tracking
Header-Only C++ Library for Graph Representation and Algorithms
Tree-Boosting, Gaussian Processes, and Mixed-Effects Models
A zero-dependency ML framework in C with a modern Python API for full control over execution and memory.
A library to train, evaluate, interpret, and productionize decision forest models such as Random Forest and Gradient Boosted Decision Trees.
oneAPI Data Analytics Library (oneDAL)
Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Neural Network TensorFlow C API
Fast, offline OCR for Node.js & C++. PP-OCRv6 with Core ML / WebGPU hardware acceleration — recognize text in images with confidence scores & coordinates. npm:…
无需 ROOT 的开源 Android 屏幕实时翻译工具,适合游戏、视觉小说和漫画。支持端侧与云端 OCR、离线 LLM、多种翻译服务和文字朗读(TTS),译文可直接显示在画面上。Open-source no-root Android real-time screen translator for games, vis…
Build apps powered by on-device AI
a community oriented 1:1, vLLM-alike (Continuous batching, paged KV) engine in C++ with additional features (for example, RadixAttention, Cache-aware scheduling…
Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and raw evidence.
llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.
llama.cpp (GGUF LLMs) and llava.cpp (GGUF VLMs) for ROS 2
🚢 Yet another operator for running large language models on Kubernetes with ease. Powered by Ollama! 🐫
