SEER: Self-Evolving Event Reasoning and Retrieval for Time Series Forecasting Paper • 2610.04109 • Published 5 days ago • 15
Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite Paper • 2610.02826 • Published 5 days ago • 90
UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement Paper • 2609.38721 • Published 7 days ago • 293
TGRL: Temperature-Grouped Reinforcement Learning for Efficient Exploration in LLMs Paper • 2609.33589 • Published 10 days ago • 8
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 12 days ago • 160
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 16 days ago • 223
Designer-RSI: Evolving Procedural Memory from User Traffic for Agentic Graphic Design Paper • 2609.22086 • Published 19 days ago • 34
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 23 days ago • 251
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 27 days ago • 174
Recursive Code World Models: Building Complex Worlds through Recursive Scene Programs Paper • 2609.11499 • Published 27 days ago • 33
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience Paper • 2609.03241 • Published Sep 3 • 54
RISE: Recursive Improvement via Self-Extrapolating Policy Distillation Paper • 2609.05295 • Published Sep 4 • 19
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published Sep 3 • 188
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation Paper • 2608.29846 • Published Aug 30 • 15