On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification Paper • 2508.05629 • Published Aug 7, 2025 • 191
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated 4 days ago • 90
nvidia/nemo-nano-codec-22khz-1.89kbps-21.5fps Feature Extraction • 39.3M • Updated 9 days ago • 25.5k • 19