Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation Paper • 2609.08798 • Published 4 days ago • 89
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL Paper • 2608.28476 • Published 15 days ago • 28
TinyCast: Probabilistic Zero-Shot Forecasting with Computed Periodicity Paper • 2608.15767 • Published 27 days ago • 9
SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Interaction Feedback Paper • 2608.13120 • Published about 1 month ago • 32
4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published 23 days ago • 83
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 23 days ago • 275
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning Paper • 2608.14290 • Published 29 days ago • 34
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents Paper • 2608.08389 • Published Aug 9 • 12
Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events Paper • 2608.06485 • Published Aug 6 • 4
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Paper • 2608.05000 • Published Aug 6 • 63
ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Paper • 2608.04436 • Published Aug 5 • 61
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published Aug 3 • 184
Can AI agents conduct open-ended AI research? Early evidence from two case studies Paper • 2607.27191 • Published Jul 29 • 20
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space Paper • 2607.25675 • Published Jul 28 • 69
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search Paper • 2607.24280 • Published Jul 27 • 84