Cosmos-Predict2.5 Collection ā ļø This collection is archived. š https://huggingface.co/collections/nvidia/cosmos3 ⢠2 items ⢠Updated Aug 11 ⢠24
Grounding World Simulation Models in a Real-World Metropolis Paper ⢠2603.15583 ⢠Published Mar 16 ⢠154
Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text Paper ⢠2601.22975 ⢠Published Jan 30 ⢠113
DAComp: Benchmarking Data Agents across the Full Data Intelligence Lifecycle Paper ⢠2512.04324 ⢠Published Dec 3, 2025 ⢠160
Nex-N1: Agentic Models Trained via a Unified Ecosystem for Large-Scale Environment Construction Paper ⢠2512.04987 ⢠Published Dec 4, 2025 ⢠86
view article Article makeMoE: Implement a Sparse Mixture of Experts Language Model from Scratch AviSoori1x ⢠May 7, 2024 ⢠125
view article Article M2.1: Multilingual and Multi-Task Coding with Strong Generalization MiniMaxAI ⢠Jan 5 ⢠42
view article Article Why Did MiniMax M2 End Up as a Full Attention Model? MiniMax-AI ⢠Oct 30, 2025 ⢠82
view article Article Aligning to What? Rethinking Agent Generalization in MiniMax M2 MiniMax-AI ⢠Oct 30, 2025 ⢠43
ToolRM Collection ToolRM: Towards Agentic Tool-Use Reward Modeling ⢠4 items ⢠Updated Mar 2 ⢠4
view article Article How to Build an MCP Server with Gradio abidlabs, ysharma ⢠Apr 30, 2025 ⢠203
PIPer: On-Device Environment Setup via Online Reinforcement Learning Paper ⢠2509.25455 ⢠Published Sep 29, 2025 ⢠38
𦫠PIPer Collection All the resources for our paper "PIPer: On-Device Environment Setup via Online Reinforcement Learning"! ⢠9 items ⢠Updated Oct 1, 2025 ⢠3
view article Article Jupyter Agents: training LLMs to reason with notebooks +1 baptistecolle, hannayukhymenko, lvwerra ⢠Sep 10, 2025 ⢠67
FunReason-MT Technical Report: Overcoming the Complexity Barrier in Multi-Turn Function Calling Paper ⢠2510.24645 ⢠Published Oct 28, 2025 ⢠11