Papers

Filtered to recursive self-improvement · clear filter

Browse by term

continual learning 86reinforcement learning 48benchmarking 13large language models 12benchmarks 11vision-language models 10language models 9robotics 7world models 7natural language processing 6recursive self-improvement 6generative models 5on-policy distillation 5video generation 5attention mechanisms 4coding agents 4LLMs 4multi-agent systems 4multimodal learning 4multimodal models 4self-distillation 4self-supervised learning 4transformers 4vision-language-action models 4world modeling 4agent-based systems 3agentic models 3agentic search 3agents 3autonomous systems 3

Matching papers

NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness

331 upvotes · 8 SEP 2026 · NeoHorse Team, Guoliang Cao, Guohao Dai et al.

This paper proposes a method for recursive self-improvement in AI systems, where a model can learn from its own performance and use that knowledge to improve itself. Practitioners might care about this approach because it could lead to more efficient and effective AI systems.

Dream-RSI: Recursive Self-Improvement through Evolving Worlds

256 upvotes · 14 SEP 2026 · Tong Zheng, Xidong Wu, Zheng Zhang et al.

This paper introduces Dream-RSI, a framework for recursive self-improvement in exploration, which helps autonomous AI agents discover high-value solutions more efficiently by using a replay simulator to provide low-cost feedback. Practitioners might care because effective exploration is crucial for AI progress, and Dream-RSI can improve discovery quality and reduce costs.

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

168 upvotes · 30 JUL 2026 · Junlin Yang, Che Jiang, Yu Fu et al.

This paper trains an AI model to improve itself in the process of building AI, with a focus on machine learning engineering, and shows promising results in various benchmarks. Practitioners may care about this research as it could lead to more efficient and autonomous AI development.

SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness

94 upvotes · 17 SEP 2026 · Haozhe Liu, Tian Ye, Sensen Gao et al.

This paper develops a method to efficiently scale agent research loops, allowing for more effective self-improvement and reusable improvements across diverse environments. Practitioners might care about this research because it could lead to significant cost savings and improved performance in automated code completion and generation tasks.

RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments

70 upvotes · 14 SEP 2026 · Sibo Zhu, Shicheng Fan, Xinyue Wang et al.

This paper introduces a new framework called RSIAgent that helps digital agents adapt to new environments without needing to be retrained. A practitioner might care about this because it allows for more efficient and effective AI systems that can learn and improve on their own.

Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement

59 upvotes · 11 SEP 2026 · Hongyao Tang, Yi Ma, Pengyi Li et al.

This paper proposes a single framework that can describe both iterative policy improvement and recursive self-improvement, which are key concepts in AI and machine learning, and helps analyze and design new systems that can learn and improve over time.