Running Reproduction Refined Analysis of Entropy-Regularized Actor-Critic 🎯 Collaborate on research logbooks with a coding agent
Running Reproduction — Approximate Reward Models for Inference-Time Scaling 🧭 Explore and collaborate on a research logbook
Running Reproduction Floating-Point Networks with AD 🎯 Explore and collaborate on a research logbook online
Running Reproduction — Language Generation with Replay 🔁 Explore experiment logbook and collaborate with AI
Running Reproduction - Convergence of Two-Timescale Markovian Stochastic Approximations with Applications in Reinforcement Learning 🎯 Explore and collaborate on a research logbook with AI
Running Reproduction — Near-Optimal Regret for KL-Regularized Multi-Armed Bandits 🎰 Explore experiment logbooks and sync with your coding agent
Running Reproduction Token Sample Complexity of Attention 🎯 Explore a research logbook and collaborate with an AI agent
Running Reproduction — Context-free Recognition with Transformers 🌲 Show your tracking data instantly
Running Reproduction Black-Box Assisted Regression 🎯 Browse and collaborate on research logbooks with an AI agent
Running Reproduction Shared Linear Representation TPGD 🎯 Explore and share research logbook with an AI collaborator
Running Reproduction — Prior Diffusiveness and Regret in the Linear-Gaussian Bandit 📈 Explore research logbook and sync findings with an AI agent
Running Reproduction Efficient Privacy Loss Accounting 🎯 Display tracked metrics in an interactive visual dashboard
Running Reproduction — A Stronger Benchmark for Online Bilateral Trade 🤝 Explore and navigate a research logbook with agent collaboration
Running Reproduction - Improved Dimension Dependence for Bandit Convex Optimization with Gradient Variations 🎯
Running Reproduction — Any-dimensional invariant universality 🌐 Explore a research logbook and collaborate with an AI agent
Running Reproduction — Bridging the Gap Between Average and Discounted TD Learning 🌉 Explore experiment logs and collaborate with your agent