10T Mixed RLVR Collection Selected Snowball 67B-A2B checkpoints from mixed-task RLVR training. • 4 items • Updated about 13 hours ago
open-athena/Snowball-67B-A2B-10T-Mixed-RLVR2-Async-Step146 Reinforcement Learning • 67B • Updated about 13 hours ago • 10
open-athena/Snowball-67B-A2B-10T-Mixed-RLVR2-Async-Step146 Reinforcement Learning • 67B • Updated about 13 hours ago • 10
10T Mixed RLVR Collection Selected Snowball 67B-A2B checkpoints from mixed-task RLVR training. • 4 items • Updated about 13 hours ago
open-athena/Snowball-67B-A2B-10T-Mixed-RLVR2-Sync-Step116 Reinforcement Learning • 67B • Updated about 14 hours ago • 13
open-athena/Snowball-67B-A2B-10T-Mixed-RLVR2-Sync-Step116 Reinforcement Learning • 67B • Updated about 14 hours ago • 13
open-athena/Snowball-67B-A2B-Mixed-RLVR-Experiment-Artifacts Viewer • Updated about 13 hours ago • 12 • 543
open-athena/Snowball-67B-A2B-Mixed-RLVR-Experiment-Artifacts Viewer • Updated about 13 hours ago • 12 • 543
open-athena/Snowball-67B-A2B-5.7T-Mixed-RLVR-Step38 Reinforcement Learning • 67B • Updated about 15 hours ago • 14
5.7T Mixed RLVR Collection Checkpoints trained with mixed RLVR from the 5.7T-token Snowball agentic SFT lineage. • 1 item • Updated about 15 hours ago
open-athena/Snowball-67B-A2B-5.7T-Mixed-RLVR-Step38 Reinforcement Learning • 67B • Updated about 15 hours ago • 14
10T Mixed RLVR Collection Selected Snowball 67B-A2B checkpoints from mixed-task RLVR training. • 4 items • Updated about 13 hours ago
2.7T Math RL Collection Research artifacts from Snowball 2.7T math RL. Raw mutable-router exports may collapse; use only models marked repaired or verified frozen. • 17 items • Updated 3 days ago