view article Article Welcome Inkling by Thinking Machines +2 burtenshaw, merve, pcuenq, ariG23498 • 6 days ago • 106
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare Paper • 2602.06717 • Published Feb 6 • 76
Scalable Visual Pretraining for Language Intelligence Paper • 2607.09657 • Published 11 days ago • 55
Bridging the Agent-World Gap: Text World Models for LLM-based Agents Paper • 2606.09032 • Published Jun 8 • 8
Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Paper • 2606.12191 • Published Jun 10 • 70
SWE-FastContext Collection A family of code-search models powering the Explore subagent for coding agents.(It will be made public later) • 3 items • Updated 21 days ago • 16
Materials Collection Welcome to IBM’s multi-modal foundation model for materials, FM4M, designed to support and advance research in materials science and chemistry. • 6 items • Updated Jan 28, 2025 • 15
view article Article Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL +6 aminediroHF, qgallouedec, kashif, lewtun, edbeeching, albertvillanova, lvwerra, sergiopaniego • May 27 • 43
The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence Paper • 2605.26494 • Published May 26 • 41
Look Before You Leap: Autonomous Exploration for LLM Agents Paper • 2605.16143 • Published May 15 • 10
📊 DNA benchmarks Collection Zero-shot DNA benchmarks for Variant Effect prediction, Sequence Recovery and Perturbation tasks. • 5 items • Updated May 19 • 13
Laguna XS.2 Collection Designed for agentic coding and long-horizon work on a local machine. Apache 2.0. • 5 items • Updated 22 days ago • 28
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 23 items • Updated 4 days ago • 337
ISO-Bench: Can Coding Agents Optimize Real-World Inference Workloads? Paper • 2602.19594 • Published Feb 23 • 3
Structured Distillation of Web Agent Capabilities Enables Generalization Paper • 2604.07776 • Published Apr 9 • 23
Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Paper • 2601.14004 • Published Jan 20 • 49
💧 LFM2.5 Collection Collection of post-trained and base LFM2.5 models. • 14 items • Updated 25 days ago • 179
Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence Paper • 2604.24954 • Published Apr 27 • 26