Running Repro - Finite and Corruption-Robust Regret Bounds in Online Inverse Linear Optimization under M-Convex Action Sets 📐 Explore and collaborate on a research logbook online
Running Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning 🎯 Explore experiment logs and sync findings with a coding agent
Running Repro - Gaussian Mean Field Variational Inference can Overestimate Predictive Variance 📉 Explore research logbook and sync with your coding agent
Running Reproducing GRPO's Loss, Dynamics, and Success Amplification (arXiv:2503.06639) 📈 Browse a research logbook and collaborate with an AI agent
Running Repro - Dependence-Aware Label Aggregation via Ising Models 🧲 Explore and edit experiment logbooks with AI agent help
Running Repro - Do-Calculus Derivation Graphs 🔗 Explore do‑calculus derivation graphs and sync with an AI agent
Running Repro - Judging What We Cannot Solve: Consequence-Based Utility for Oracle-Free Evaluation 🎯 Explore logs and collaborate with an AI coding agent
Running Repro - Provable Benefits of RLVR over SFT: Learning to Backtrack Efficiently 🔃 Collaborate on a logbook with an AI coding agent
Running Repro - Row-Stochastic vs Doubly-Stochastic Decentralized Learning 🕸 Explore and sync experiment logs with your coding agent
Running Repro - Unsupervised Disentanglement Without Compromises 🔀 Browse and collaborate on experiment logbooks
Running Reproducing ITCR: inference-time conformal factuality control 🎯 Browse a research logbook and collaborate with an AI agent
Running Repro - BrokenMath: Sycophancy in Theorem Proving 🎯 Explore and manage research logbook entries online
Running Repro - Dimension-Free KL Bounds for Underdamped Langevin MC 📉 Explore project logbook and collaborate with an AI agent
Running Reproducing: Success Conditioning as Policy Improvement (arXiv:2601.18175) 🎯 Explore research logbook and sync with your AI agent
Running Repro - Learning Treatment Allocations with Risk Control Under Partial Identifiability ⚖ Collaborate on a research logbook with an AI agent
Running Repro - When Actions Go Off-Task: DeAction and MisActBench 🎯 Explore action logs and sync findings with your coding agent
Running Repro - Asymptotically Optimal Sequential Testing with Markovian Data 🔗 Explore and sync experiment logs with an AI coding agent
Running Repro - Envy-Free Allocation of Indivisible Goods via Noisy Queries ⚖ Browse experiment logs and sync findings with your coding agent