Helix-7B
A 7B general model with strong math and code benchmarks.
Repos7
Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.
repoSagargupta16/claude-cost-optimizerSave 30-60% on Claude Code costs -- proven strategies, real benchmarks, copy-paste configs, and interactive tools
repoOskarsEzerins/llm-benchmarksPopular LLM benchmarks for ruby code generation
repoopen-compass/VLMEvalKitOpen-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
repohelixml/helix♾️ Private Agent Fleet with Spec Coding. Each agent gets their own GPU-accelerated desktop. Run Claude, Codex, Gemini an
repomodelscope/evalscopeA streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmark
repodeepmodeling/DeePTBDeePTB: A deep learning package for tight-binding Hamiltonian with ab initio accuracy.
News3
<!-- SC_OFF --><div class="md"><p>Everyone keeps asking if the 1-bit models are actually usable for agents, so I ran the
newsBest coding model for 3x Spark setup?<!-- SC_OFF --><div class="md"><p>Hi, our company has dedicated 3x Asus Ascent GX10 (GB10) to run a coding model for our
newsDeepSWE: new benchmark looking at how well today's frontier models can actually write code [R]<table> <tr><td> <a href="https://www.reddit.com/r/MachineLearning/comments/1ue0hlp/deepswe_new_benchmark_looking_at_how
