@wbx_life
π₯ The era of ultra-efficient reasoning is here: our Nemotron-Cascade-2-30B-A3B currently trending at #1 on huggingface π€ π₯ Gold Medal-level performance on IMO 2025, IOI 2025, and ICPC World Finals 2025 -- all from a model with only 3B active parameters. π€― β‘ SOTA alignment and instruction following capabilties even compared with larger LLMs The secret sauce? Cascade RL. 𧬠1οΈβ£ Cascade RL not only pushes the model limits on each domain, but also generates elite "teachers" for every expert domain. 2οΈβ£ Multi-domain on-policy distillation uses expert teachers to keep the student model sharp, mitigating domain shifts and matching expert-level performance. π» Read the blog: https://t.co/mqmSMXFsMx π Check the paper: https://t.co/DRTloSE1h0 π€ Get the weights: https://t.co/LvIo0WWgI5 #AI #MachineLearning #NVIDIA #LLM