Train
50 items across the graph — tagged with Train.
From the graph · 50
本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)
PArallel Distributed Deep LEarning: Machine Learning Framework from Industrial Practice (『飞桨』核心框架,深度学习&机器学习高性能单机、分布式训练和跨平台部署)
The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build custom intelligence faste…
Build, Manage and Deploy AI/ML Systems
A Cloud Native Batch System (Project under CNCF)
Democratizing Reinforcement Learning for LLMs
Chronos: Pretrained Models for Time Series Forecasting
General technology for enabling AI capabilities w/ LLMs and MLLMs
Hugging Face model with 4144 likes. Tags: transformers, pytorch, safetensors, mistral, text-generation, pretrained, mistral-common, en, arxiv:2310.06825, licens…
The official CLI and Python client for the Hugging Face Hub.
OplaPlanner has moved to https://github.com/apache/incubator-kie-drools. This repository is archived. OptaPlanner is an AI constraint solver in Java to optimize…
Fast ML inference & training for ONNX models in Rust
OpenLake is a high performance storage engine for efficient LLM inference and GPU Training
Instant, controllable, local pre-trained AI models in Rust
Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch
The open source Solver AI for Java and Kotlin to optimize scheduling and routing. Solve the vehicle routing problem, employee rostering, task assignment, mainte…
Pre-trained Neural Network models in Axon (+ 🤗 Models integration)
Nvidia GPU exporter for prometheus using nvidia-smi binary
Extension for Scikit-learn is a seamless way to speed up your Scikit-learn application
Lightweight optimization with local, global, population-based and sequential techniques across mixed search spaces
Neural Network Compression Framework for enhanced OpenVINO™ inference
Evaluate and improve models and agents using environments
Train speculative decoding models effortlessly and port them smoothly to SGLang serving.
The Hitchhiker's Guide to Data Science for Social Good
“AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。
oneAPI Data Analytics Library (oneDAL)
Official Codebase for "Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights" (ICML 2026 Spotlight)
Get started with Timefold quickstarts here. Optimize the vehicle routing problem, employee rostering, task assignment, maintenance scheduling and other planning…
A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
A unified, high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.
TorchX is a universal job launcher for PyTorch applications. TorchX is designed to have fast iteration time for training/research and support for E2E production…
Run Slurm in Kubernetes
Awesome_Multimodel is a curated GitHub repository that provides a comprehensive collection of resources for Multimodal Large Language Models (MLLM). It covers d…
A plug-and-play compiler that delivers free-lunch optimizations for both inference and training.
电子鹦鹉 / Toy Language Model
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
Historical Ultralytics HUB repository — HUB shut down on July 31, 2026 and was replaced by Ultralytics Platform.
Foundation models based medical image analysis
Project Tapestry aims to give every nation and participant frontier AI they can call their own — uniting a global consortium to train a shared frontier model fr…
NNtrainer is Software Framework for Training and Inferencing Neural Network Models on Devices.
Full-stack open-source AI engine for building language models — tokenizer training, transformer architecture, cognitive reasoning and chat pipeline.
Open-source performance diagnostics for PyTorch training runs.
A hands-on course for building modern LLMs from scratch in PyTorch, with 26 runnable Jupyter Notebooks covering tokenizers, attention, MoE, RLHF, inference, eva…
A curated list of papers, tools, and resources on Multi-Token Prediction (MTP) and related techniques in Large Language Models (LLMs), Speech-Language Models (S…
This is the Docker container based on open source framework XGBoost (https://xgboost.readthedocs.io/en/latest/) to allow customers use their own XGBoost scripts…
从 MiniMind 源码读起,再延伸到现代大模型技术体系的中文学习笔记。主线逐行精读预训练 / SFT / DPO / PPO / GRPO 与训练机制;附录 17 篇进阶卷覆盖量化、投机解码、RLHF 全景、模型代际史等 MiniMind 没涉及、但进阶绕不开的主题。
ChinaTravel: A Real-World Benchmark for Language Agents in Chinese Travel Planning
