repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
luojieLLMaaS/haxiv
Deep learning framework for LLMs (Llama/Gemma/Qwen) on CPU, Apple MLX, Metal, CUDA. Load PyTorch/ONNX/TF/GGUF with zero conversion. PyTorch alternative with native Apple Silicon, ONNX Runtime alternative with autograd, llama.cpp alternative with backprop. FlashAttention, tensor parallel, INT4 quant.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 55%Instead of decentralized training effort we should build the “One dataset” →
- PossiblePossibly related (embedding) · 54%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 53%H64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P] →
- PossiblePossibly related (embedding) · 52%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 52%How're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D] →
- PossiblePossibly related (embedding) · 50%DynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning Research →
- PossiblePossibly related (embedding) · 46%DeepSeek v4 Flash on 4090 + DDR5, my experience →
- PossiblePossibly related (embedding) · 53%Has anyone created a "Local LLM Survival Kit"? →
Covers
newsInstead of decentralized training effort we should build the “One dataset”newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsH64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P]newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsHow're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D]
Covers (incoming)
newsDynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning ResearchnewsDeepSeek v4 Flash on 4090 + DDR5, my experiencenewsHas anyone created a "Local LLM Survival Kit"?newsDeepSeek v4 Flash on 5090 in llama.cpp with 1 Million contextnewsBenchmarks: TensorSharp vs. llama.cppnewsOpen Source Local LLM Training Tool (for consumer hardware)
Related across the graph
newsOpen Source Local LLM Training Tool (for consumer hardware)newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsOpenAI and Broadcom unveil LLM-optimized inference chipnewsDynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning ResearchnewsHow're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D]newsH64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P]newsBenchmarks: TensorSharp vs. llama.cppnewsInstead of decentralized training effort we should build the “One dataset”newsDeepSeek v4 Flash on 5090 in llama.cpp with 1 Million contextnewsHas anyone created a "Local LLM Survival Kit"?newsDeepSeek v4 Flash on 4090 + DDR5, my experience
