Sparse frontier-model training factory (Reflection Beam direction)
python reinforcement-learning deep-learning fault-tolerance sandbox moe zero-dependency distributed-training mixture-of-experts llm grpo training-infrastructure
-
Updated
Oct 6, 2026 - Python