view article Article Welcome Inkling by Thinking Machines +2 burtenshaw, merve, pcuenq, ariG23498 ⢠6 days ago ⢠103
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver Paper ⢠2604.08377 ⢠Published Apr 9 ⢠295
view article Article NVIDIA Cosmos Reason 2 Brings Advanced Reasoning To Physical AI nvidia ⢠Jan 5 ⢠64
view article Article We Got Claude to Fine-Tune an Open Source LLM burtenshaw, evalstate ⢠Dec 4, 2025 ⢠630
SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models Paper ⢠2504.11468 ⢠Published Apr 10, 2025 ⢠30
SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training Paper ⢠2501.17161 ⢠Published Jan 28, 2025 ⢠125
Cosmos-Predict2 Collection ā ļø This collection is archived. š https://huggingface.co/collections/nvidia/cosmos-predict25 ⢠10 items ⢠Updated 4 days ago ⢠37
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers Paper ⢠2508.20453 ⢠Published Aug 28, 2025 ⢠63
view article Article Small Language Models (SLM): A Comprehensive Overview jjokah ⢠Feb 22, 2025 ⢠167
Wan: Open and Advanced Large-Scale Video Generative Models Paper ⢠2503.20314 ⢠Published Mar 26, 2025 ⢠64
Physical AI Collection Collection of open, commercial-grade datasets for physical AI developers ⢠53 items ⢠Updated 4 days ago ⢠173
AceReason Collection Math and Code reasoning model trained through reinforcement learning (RL) ⢠7 items ⢠Updated 4 days ago ⢠21
Reward Models 06-2025 Collection Nemotron reward models. For use in RLHF pipelines and LLM-as-a-Judge ⢠8 items ⢠Updated 4 days ago ⢠24
Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms Paper ⢠2310.07161 ⢠Published Oct 11, 2023 ⢠1
Qwen2.5-1M Collection The long-context version of Qwen2.5, supporting 1M-token context lengths ⢠3 items ⢠Updated Dec 31, 2025 ⢠127
OpenReasoning-Nemotron Collection Collection of models for OpenReasoning-Nemotron which are trained on 5M reasoning traces for Math, Code and Science. ⢠6 items ⢠Updated 4 days ago ⢠47