newsReddit r/LocalLLaMATrust 52 Β· CommunityPublished 1mo agoLive Β· 1mo ago
I RL-trained Qwen3.6-35B-A3B to RL-train small task-specific Qwen models. Fully open source! π€
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it β so bad links are debuggable.
- PossiblePossibly related (embedding) Β· 48%When Model Merging Rivals Joint Multi-Task Reinforcement Learning: A Task-Vector Geometry Analysis β
