Submitted:
09 August 2026
Posted:
14 August 2026
You are already at the latest version
Abstract
Keywords:
1. Introduction

2. Related Work
3. Formal Model of Bounded Recursive Improvement
4. Materials and Methods
4.1. Experimental Environment
4.2. Frozen Baseline and Protocol
4.3. Cycle Procedure
4.4. Preregistration
4.5. Sandbox and Authority Constraints
4.6. Provenance and Integrity
4.7. Use of AI in the Research and Manuscript
5. Results
5.1. Overall Trace
5.2. Cycles 001-002
5.3. Cycle 003 Failure
5.4. Forensic Analysis and Boundary v1.1
5.5. Cycle 004
5.6. Cycle 005
5.7. Cycles 006-007
5.8. Final Freeze
6. Discussion
7. Threats to Validity and Limitations
8. Reproducibility, Data and Artifact Availability
9. Conclusion
| Event | Parent | Outcome | State consequence / principal observation |
| Cycle 001 | RLC0 | PROMOTED | RLC1; executable frozen-reference validation |
| Cycle 002 | RLC1 | PROMOTED | RLC2; inherited-state / lineage validation |
| Cycle 003 | RLC2 | FAILED_NOT_PROMOTED | RLC2 retained; integrity pause; hypothesis not met |
| Boundary v1.1 | - | PROTOCOL_BOUNDARY_REVISION | Prospective integrity semantics revised after forensics |
| Cycle 004 | RLC2 | PROMOTED | RLC3; boundary-aware inheritance gate |
| Cycle 005 | RLC3 | PROMOTED | RLC4; deterministic resource/effect evidence reconciliation |
| Cycle 006 | RLC4 | NO_CHANGE | No qualifying deficiency; no candidate; RLC4 retained |
| Cycle 007 | RLC4 | NO_CHANGE | Independent repeat; no candidate; RLC4 retained |
AI-Assisted Technology Disclosure
Supplementary Materials
Author Contributions
Funding
Institutional Review Board Statement
Informed Consent Statement
Data Availability Statement
Conflicts of Interest
References
- Madaan, A.; Tandon, N.; Gupta, P.; et al. Self-Refine: Iterative Refinement with Self-Feedback. Adv. Neural Inf. Process. Syst. arXiv 2023, arXiv:2303.17651. [Google Scholar]
- Shinn, N.; Cassano, F.; Berman, E.; Gopinath, A.; Narasimhan, K.; Yao, S. Reflexion: Language Agents with Verbal Reinforcement Learning. Adv. Neural Inf. Process. Syst. arXiv 2023, arXiv:2303.11366. [Google Scholar]
- Huang, J.; Gu, S.S.; Hou, L.; Wu, Y.; Wang, X.; Yu, H.; Han, J. Large Language Models Can Self-Improve. arXiv 2022, arXiv:2210.11610. [Google Scholar]
- Bai, Y.; Kadavath, S.; Kundu, S.; et al. Constitutional AI: Harmlessness from AI Feedback. arXiv 2022, arXiv:2212.08073. [Google Scholar]
- Schmidhuber, J. Gödel Machines: Self-Referential Universal Problem Solvers Making Provably Optimal Self-Improvements. In arXiv; Technical Report IDSIA-19-03; IDSIA: Manno-Lugano, Switzerland, 2003. [Google Scholar]
- Zelikman, E.; Lorch, E.; Mackey, L.; Kalai, A.T. Self-Taught Optimizer (STOP): Recursively Self-Improving Code Generation. arXiv 2023, arXiv:2310.02304. [Google Scholar]
- Yin, X.; Wang, X.; Pan, L.; Wan, X.; Wang, W.Y. Goedel Agent: A Self-Referential Agent Framework for Recursive Self-Improvement. arXiv 2024, arXiv:2410.04444. [Google Scholar]
- Lu, C.; Lu, C.; Lange, R.T.; Foerster, J.; Clune, J.; Ha, D. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery. arXiv 2024, arXiv:2408.06292. [Google Scholar]
- Shirbint, E. The Reinforcement Learning Contour: From Document Production to Governed Epistemic Transformation in AI-Assisted Scientific Inquiry. 2026, 202608.0212.v1. [Google Scholar] [CrossRef]
- Popper, K.R. The Logic of Scientific Discovery; Routledge: London, 1959. [Google Scholar]
- Lakatos, I. The Methodology of Scientific Research Programmes; Cambridge University Press: Cambridge, 1978. [Google Scholar]
- Argyris, C.; Schoen, D.A. Organizational Learning II: Theory, Method, and Practice; Addison-Wesley: Reading, MA, 1996. [Google Scholar]
- Shumailov, I.; Shumaylov, Z.; Zhao, Y.; Gal, Y.; Papernot, N.; Anderson, R. AI models collapse when trained on recursively generated data. Nature 2024, 631, 755–759. [Google Scholar] [CrossRef] [PubMed]
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2026 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).