Training
50 items across the graph — tagged with Training.
From the graph · 50
本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)
PArallel Distributed Deep LEarning: Machine Learning Framework from Industrial Practice (『飞桨』核心框架,深度学习&机器学习高性能单机、分布式训练和跨平台部署)
The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build custom intelligence faste…
Build, Manage and Deploy AI/ML Systems
A Cloud Native Batch System (Project under CNCF)
Democratizing Reinforcement Learning for LLMs
General technology for enabling AI capabilities w/ LLMs and MLLMs
Fast ML inference & training for ONNX models in Rust
OpenLake is a high performance storage engine for efficient LLM inference and GPU Training
Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch
Nvidia GPU exporter for prometheus using nvidia-smi binary
Extension for Scikit-learn is a seamless way to speed up your Scikit-learn application
Neural Network Compression Framework for enhanced OpenVINO™ inference
Evaluate and improve models and agents using environments
The Hitchhiker's Guide to Data Science for Social Good
Train speculative decoding models effortlessly and port them smoothly to SGLang serving.
“AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。
oneAPI Data Analytics Library (oneDAL)
Official Codebase for "Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights" (ICML 2026 Spotlight)
A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
TorchX is a universal job launcher for PyTorch applications. TorchX is designed to have fast iteration time for training/research and support for E2E production…
Run Slurm in Kubernetes
A plug-and-play compiler that delivers free-lunch optimizations for both inference and training.
A modular, scalable, high-performance training framework for LLMs, VLMs, diffusion, and embodied models.
电子鹦鹉 / Toy Language Model
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
Foundation models based medical image analysis
Project Tapestry aims to give every nation and participant frontier AI they can call their own — uniting a global consortium to train a shared frontier model fr…
NNtrainer is Software Framework for Training and Inferencing Neural Network Models on Devices.
Full-stack open-source AI engine for building language models — tokenizer training, transformer architecture, cognitive reasoning and chat pipeline.
Open-source observability for PyTorch training runs.
A curated list of papers, tools, and resources on Multi-Token Prediction (MTP) and related techniques in Large Language Models (LLMs), Speech-Language Models (S…
This is the Docker container based on open source framework XGBoost (https://xgboost.readthedocs.io/en/latest/) to allow customers use their own XGBoost scripts…
从 MiniMind 源码读起,再延伸到现代大模型技术体系的中文学习笔记。主线逐行精读预训练 / SFT / DPO / PPO / GRPO 与训练机制;附录 17 篇进阶卷覆盖量化、投机解码、RLHF 全景、模型代际史等 MiniMind 没涉及、但进阶绕不开的主题。
Experimental playground for benchmarking language model (LM) architectures, layers, and tricks on smaller datasets. Designed for flexible experimentation and ex…
Run a documented subset of verl-style OPD on one consumer GPU—typed config, Parquet prompts, and PEFT scale-out artifacts.
Machine learning library, Distributed training, Deep learning, Reinforcement learning, Models, TensorFlow, PyTorch
GenAssist combines orchestration, runtime, analytics, and learning — in one open platform.
CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents
Biologically-grounded adversarial training platform: cyclic Wake/Dream/Nightmare/Compress phases that accumulate model robustness without catastrophic forgettin…
Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents
Procedural data generators for verifiable reasoning, synthetic pretraining, post-training, evaluation, and RL.
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
Parameter-Efficient Fine-Tuning For Edge-Cloud Collabrative Computing
Personalized content creation assistant for youtubers. Discord: https://discord.gg/k9sZcq2gNG
Qelm - Quantum Enhanced Language Model
Research platform for model training, evaluation, and experimentation across architectures, benchmarks, and recipes.
Local-first Python CLI and MCP server for versioned, cited context from web and document sources.
Fine-tune open-source models with Tinker from inside Pi — managed improve loops, data prep, evals, smoke tests, deploy snippets, and checkpoint chat.
