Recursive Criticality of AI Self-Improvement
Abstract
AI is increasingly used in the R&D process that produces future AI systems. We study the conditions under which this feedback becomes self-amplifying. Our model describes how the rate of AI capability growth depends on baseline research productivity, recursive feedback, and the increasing difficulty of making further research progress. We derive a recursive reproduction number, , that determines whether incremental improvements are amplified or damped across development cycles. This quantity compares the strength of recursive feedback with the rate at which further progress becomes more difficult. When , the effects of incremental improvements compound across development cycles, placing the system in a self-amplifying regime. When , their effects weaken across cycles. The transition depends on the structure of the AI R&D feedback loop and need not occur at any particular level of model capability. A system can therefore enter a self-amplifying regime before acceleration becomes visible, while rapid progress can also occur without self-amplification. In the minimal model, higher baseline research productivity can accelerate progress without changing whether the system is self-amplifying. At high research throughput, the duration of the development cycle becomes a limiting timescale for amplification. Increasing research difficulty can subsequently end a period of self-amplification. Extending the model to multiple research actors shows that improvements shared across organizations can make the overall research ecosystem self-amplifying even when no individual actor is. The framework identifies measurable properties of AI R&D systems that can help distinguish recursive amplification from rapid progress driven by other sources, including the strength of recursive feedback, how effectively improvements propagate into successor systems, development-cycle duration, and the increasing difficulty of further progress. 11 1 The code and computational notebook used to reproduce the numerical results and figures are available at https://github.com/burtsev/recursive-criticality-ai.
1 Introduction
Recursive AI self-improvement (RSI) is usually understood as a positive feedback loop in AI capabilities. AI research agents are beginning to perform work that lies inside the process used to build future AI systems. They write and debug code, propose experiments, analyse results, reproduce papers, operate software tools, and increasingly sustain technical work over longer horizons [1, 2, 3, 4, 5]. RSI is therefore better viewed as a property of an AI–R&D system than of an isolated, fully autonomous agent. The system may include human researchers, laboratories, compute infrastructure, evaluation procedures and organizations that integrate and deploy successor models.
This paper asks three questions. Under what conditions does AI assistance to AI R&D cross from ordinary acceleration into a self-amplifying regime? How do feedback delay and a hardening research progress determine whether that regime produces a small transient displacement or significantly compressed transition to artificial general intelligence (AGI) and artificial super intelligence (ASI)? How are these dynamics changed by physical resource limits, competition and collaboration between research actors?
We use the term recursive criticality for the condition at which an incremental improvement in AI research capability generates enough additional future R&D productivity to outweigh the increasing difficulty of making further progress. This balance is defined by a recursive reproduction number . When , incremental gains amplify across development cycles. When , their effects weaken. This yields the following criterion:
Onset of self-amplifying RSI. The transition to self-amplifying recursive improvement occurs when the AI system crosses a critical stability boundary, not when it reaches a particular level of intelligence. This boundary is crossed when rises above unity, so that gains in research capability reproduce across development cycles faster than the effective research frontier hardens.
The possibility of a machine-driven intelligence explosion dates to Good [6], and subsequent work developed the conceptual foundations of superintelligence and self-modifying agents [7, 8, 9]. More recent economic models place AI directly inside the production of future technological progress, treating it as an input to research whose effects depend on automation, bottlenecks, diminishing returns and the productivity of complementary resources [10, 11, 12, 13, 14]. Related models examine resource-constrained growth and distinguish bounded self-refinement from open-ended research loops [15, 16]. Cunningham et al. [17] derive a closely related condition for self-sustaining acceleration in terms of economic elasticities, while recent analyses of the transition from AGI to ASI emphasize that scaling, paradigm change, recursive improvement and collective AI systems may operate simultaneously rather than as separate pathways [18].
At the same time, empirical work is making the underlying feedback process increasingly measurable. General benchmarks provide increasingly systematic comparisons of model capability across tasks and generations, while human-calibrated task horizons estimate the duration and complexity of work that AI agents can complete reliably [19, 20, 1]. Research-oriented evaluations such as MLE-bench, RE-Bench and PaperBench move closer to the relevant setting by testing agents on extended machine-learning engineering, experimentation and research-replication tasks [2, 3, 4, 5]. These evaluations increasingly measure components of the feedback loop, but they do not yet identify how an improvement in AI research capability causally affects the productivity of subsequent AI development. Proposed measures of AI-R&D automation extend this perspective to the organizational level by tracking adoption, researcher time and the effect of AI assistance on subsequent scientific and engineering progress [21].
Taken together, existing work brings many of the ingredients of recursive improvement into view, but not yet the dynamics that determine whether the resulting feedback is stable. In particular, it leaves open how the transition to self-amplifying improvement depends on the delay between research and deployment, the progressive exhaustion or hardening of research opportunities, and the coupling between multiple research actors.
We develop a minimal model in which AI capability evolves through baseline research productivity, delayed recursive gain and a capability-dependent research frontier. We then extend the framework to include physical deployment constraints and research networks through which improvements propagate between laboratories, firms and nations. The resulting model separates interventions that primarily alter the pace of development from those that change the recursive dynamics themselves or reshape the network through which recursive gains spread. The numerical scenarios are conditional demonstrations and not probabilistic forecasts. They show how alternative assumptions generate qualitatively different trajectories and identify the measurements that could distinguish among them.
We make four principal contributions.
First, we formulate recursive AI self-improvement as a local stability transition in an AI-enabled R&D system. The resulting recursive reproduction number,
compares realized recursive gain with the local hardening of the research frontier. Incremental capability gains amplify across development cycles when and are damped when . In the minimal model, the critical boundary does not depend on baseline research throughput, so rapid capability growth and recursive self-amplification are distinct phenomena. A system can become supercritical before the resulting acceleration becomes visible.
Second, we characterize how the duration and intensity of recursive amplification depend on feedback delay and the available research frontier. At high research throughput, the successor-development cycle becomes a limiting timescale for amplification. As research opportunities become harder to exploit, increasing frontier hardness can return a supercritical system to a subcritical regime. Recursive amplification can therefore be strong but transient within a fixed research paradigm.
Third, we separate the recursive regime from both research scale and physical deployment constraints. Greater compute, expenditure, or researcher effort can accelerate capability growth without changing recursive criticality, while changes in recursive gain, operational closure, feedback delay, or transfer between actors alter the feedback dynamics themselves. Physical infrastructure can additionally constrain deployable capability without directly changing the recursive regime.
Fourth, we extend the framework to coupled research actors. Recursive criticality is then determined by the spectral radius of a reproduction matrix that captures both within-actor feedback and cross-actor transfer. A research ecosystem can therefore become supercritical even when every actor is individually subcritical. This formulation also identifies quantities that can be estimated empirically, including recursive gain, operational closure, development-cycle duration, frontier hardening, and the transfer of improvements across organizations.
2 A dynamical model of recursive self-improvement
2.1 Recursive loop in AI development
We model recursive self-improvement at the level of an AI–R&D system rather than an isolated AI agent. The system includes the AI models used for research together with the processes required to generate, evaluate, train and deploy successor systems. Human researchers may remain inside this system boundary.
Let denote an AI capability. The model, illustrated in Fig. 1, separates four elements of the development process.
- 1.
Baseline progress is governed by an exogenous research-productivity scale .
- 2.
Existing AI capability can increase the productivity of subsequent AI development. We represent this effect by recursive gain and an end-to-end delay between an improvement in capability and its return as greater research productivity.
- 3.
Only part of the potential recursive gain may survive the full development pipeline. We denote this transmitted fraction by the operational closure .
- 4.
Progress becomes increasingly difficult as capability approaches the research frontier defined by local hardening rate .
To connect the abstract state variable to observable research performance, let denote the minimum resource cost of completing task at tolerated loss , evaluated at a fixed vector of resource prices . For a stable panel of tasks with weights , define
| (1) |
Thus measures resource-adjusted improvement at fixed task composition and performance. It is a local coordinate for research capability, not a universal scalar measure of intelligence.
Let denote the recursive gain from stronger AI research capability to subsequent research productivity when downstream use is unconstrained. Let denote the fraction transmitted through the research pipeline. The realized recursive gain is
| (2) |
Operational closure is a property of the development process, not simply of model autonomy. A system can have high closure while retaining substantial human participation if AI-generated research contributions are evaluated and integrated efficiently. Conversely, a highly autonomous workflow can have low closure when its outputs fail evaluation, cannot be executed safely, or do not propagate into successor systems. Human review, experimental capacity, secure execution, and other bottlenecks can bind as throughput rises, so may depend on operating scale as well as capability.
Frontier hardness measures the decline in direct improvement productivity as capability advances with recursive transfer held fixed. The delay is the end-to-end interval before an improvement returns as increased research capability. Together, , , and determine the local stability and manifestation rate of recursive feedback.
2.2 A delayed model of recursive criticality
We model recursive AI self-improvement by combining baseline research productivity, frontier difficulty and delayed recursive feedback:
| (3) |
Here sets the baseline scale of research progress, describes how the productivity of direct improvement changes as the research frontier advances. The exponential term represents the multiplicative effect of recursive feedback. In the absence of recursive amplification, and progress reduces to the baseline term .
We define the feedback potential so that its local derivative equals the realized recursive gain,
| (4) |
Thus measures how strongly an incremental increase in research capability raises the productivity returned through the recursive loop.
We quantify frontier hardening by the rate at which direct improvement productivity declines with capability,
| (5) |
Positive means that advancing the frontier reduce the productivity of subsequent improvement.
Consider a reference trajectory during an interval in which the coefficients change slowly relative to the perturbation dynamics. Set , , , and . A small perturbation then obeys
| (6) |
with characteristic equation
| (7) |
The local balance becomes transparent after defining
| (8) |
Proposition 1 (Local recursive criticality).
Suppose that , , and . All characteristic roots of Eq. (6) have negative real part for , is a characteristic root at , and a positive real characteristic root exists for .
This result is a scalar instance of stability theory for positive delay systems [22]. Cunningham et al. [17] independently obtain a closely related unit-threshold condition for self-sustaining acceleration from recursive elasticities. In the present model, the threshold marks a stability boundary for perturbations around a delayed capability trajectory. Locally, implies amplification of small recursive perturbations around the reference trajectory, but does not by itself imply indefinite growth of total capability.
The mechanism behind the threshold is simple. Recursive gain must offset replace the marginal loss of research productivity created by the hardening frontier before a capability perturbation can reproduce. Baseline throughput controls how much development occurs per unit time, but in the minimal model it scales both the damping and the feedback terms. Greater compute, expenditure or researcher effort can therefore accelerate capability growth without necessarily changing whether the system is subcritical or supercritical.
Feedback delay instead determines how rapidly the local regime becomes visible. Near the critical point, the dominant characteristic root is
The dominant root passes smoothly through zero, which means that a newly supercritical process may initially be difficult to distinguish from ordinary acceleration. Its amplification rate becomes more apparent as research throughput rises or as the successor-development cycle becomes shorter.
As research throughput grows, the delay itself becomes limiting. For fixed and ,
Increasing research throughput can therefore accelerate an amplifying mode, but cannot make it arbitrarily fast without shortening the interval through which improvements return in a successor system. Other limits on effective research throughput, including bounded parallelizability and coordination, can bind before this delay-limited regime is reached [23].
Rapid capability growth is therefore neither necessary nor sufficient evidence of recursive criticality. A high baseline research rate, greater compute availability or substantial human effort can produce fast subcritical progress, while a supercritical system can initially amplify too slowly to be distinguished from an ordinary acceleration. Diagnosing the regime therefore requires estimates of recursive gain and frontier hardening across successive development cycles, rather than the slope of a benchmark trajectory considered in isolation.
2.3 Finite recursive runway
The local criticality condition determines whether capability perturbations amplify at the current state, but it does not require the same recursive mechanism to remain supercritical as capability advances. A simple finite-frontier model is
| (9) |
which gives
| (10) |
The frontier becomes progressively harder as approaches the effective limit .
Proposition 2 (Finite-frontier termination).
Suppose that , , and remains bounded. Then
so the system is locally subcritical sufficiently close to the effective frontier.
This result makes the duration of a supercritical regime depend on the available recursive runway. Within a fixed research paradigm, the system may cross into a supercritical region, amplify rapidly for some interval, and return to a subcritical regime once frontier hardening overtakes realized recursive gain. A major architectural or scientific advance may extend the frontier , change its shape through , or increase the realized gain , so the result should not be read as a claim about a final ceiling. It instead shows that transient supercriticality is compatible with a finite research opportunity set.
2.4 Coupled research-agent ecosystems
AI development does not occur inside a single closed loop. Research agents operate within organizations that observe competitors, exchange scientific information, reuse software, recruit from a common labour market, depend on shared infrastructure, and sometimes obtain access to the same models or tools. These channels allow an improvement produced at one actor to change the future research productivity of another. Once such transfer becomes comparable to within-actor recursive gain, the relevant stability question concerns the network as a whole.
Let be a perturbation to the research capability of actor . A local multi-actor approximation is
| (11) |
where denotes gain generated within actor and denotes directed transfer from actor to actor . Normalizing each row by the receiving actor’s frontier hardness defines the reproduction matrix
| (12) |
where is the spectral radius of the non-negative matrix .
Proposition 3 (Collective criticality).
For the zero-delay system associated with Eq. (11), the locally stable regime satisfies , while the system becomes unstable when . Hence for every actor is not sufficient for network stability when cross-actor transfer is strong enough.
This result extends recursive criticality from individual development loops to the research ecosystem as a whole. Every actor may remain individually subcritical while the coupled system becomes supercritical through the circulation and recombination of improvements across organizations. An advance originating in one laboratory may improve the tools available to another, whose subsequent work raises research productivity elsewhere before some of those gains diffuse back to the original actor. The combined reproduction of capability across these pathways can exceed the damping imposed by individual research frontiers even though no actor has a within-organization reproduction number above one.
3 Reference scenarios
3.1 Reference no-RSI parameterization
The numerical experiments are designed to compare dynamical mechanisms rather than to forecast calendar dates. We normalize current capability to and the effective frontier of the modeled research paradigm to . These values define the coordinate system and do not imply that current AI has zero absolute capability or that represents a fundamental limit.
We represent AGI and ASI by externally specified thresholds and on the capability coordinate, with . The lower threshold can represent broad competence across a specified set of economically or scientifically relevant tasks, while the upper threshold can represent substantially greater performance under the same evaluation conditions. For any threshold , define the first-crossing time as , , .
This construction separates the placement of AGI and ASI on a capability scale from the dynamics that carry the system between them. Our results concern the latter, conditional on the former. For the reference scenarios we set
These values are illustrative. They place AGI midway between the reference state and the effective frontier, while ASI lies substantially closer to the frontier. The resulting separation is sufficient to distinguish gradual progress, transient recursive amplification and compression of the AGI-to-ASI interval within a common capability coordinate.
We use the reference hardness value in the frontier function (9) as a structural assumption. This choice makes direct improvement productivity decline quadratically with the remaining headroom while preserving an analytic baseline solution. Under the normalization , the corresponding frontier hardness is
Hardness therefore rises gradually at low capability and diverges as the effective frontier is approached.
To construct the reference trajectory, we remove recursive feedback by setting , omit additional physical throughput constraints, and hold baseline research productivity fixed at . The dynamics reduce to
with solution
For the reference threshold positions: and . Thus sets the no-RSI AGI and ASI timescales. The single parameter therefore sets the timescale of the no-RSI trajectory.
We choose this timescale to place the reference AGI crossing broadly within the range of recent expert elicitation. The 2023 Expert Survey on Progress in AI, covering 2,778 authors from major AI venues, reported an aggregate 50% date of 2047 for high-level machine intelligence [24]. A later wave of the Longitudinal Expert AI Panel reported a conditional median AGI date of 2050 among 205 experts, with 25% and 75% dates of 2039 and 2065 [25]. These surveys use different definitions and exhibit substantial uncertainty, so we use them only to set an illustrative timescale rather than as a probabilistic calibration. Taking to correspond to 2026 and setting the reference AGI crossing to 2050 gives a 24-year baseline timescale
| (13) |
The capability normalization, frontier shape and baseline rate together fully specify the reference trajectory without recursive feedback.
3.2 Reference parameterization of the RSI part
Existing empirical evidence shows substantial algorithmic progress and improving AI performance on increasingly demanding technical tasks, but it does not separately identify the recursive gain , the operational-closure function or the relevant frontier hardness [26, 27, 28]. We therefore treat these quantities as structural scenario parameters. Throughout the reference recursive scenarios, we fix the feedback delay at as a plausible interval between a capability improvement and its effect on subsequent AI R&D productivity.
To isolate the central mechanism of the model, we construct a small family of recursive scenarios in which the only varying parameter is the recursive gain . The operational-closure trajectory, the frontier-hardening exponent and the feedback delay are held fixed across the primary scenarios. Differences in the simulated trajectories can therefore be attributed directly to differences in the strength of recursive AI-to-AI-R&D feedback.
We model operational closure as a smooth increasing function of capability,
| (14) |
The steepness parameter determines how rapidly the AI–R&D loop approaches closure as capability increases, with larger values concentrating the transition within a narrower capability interval. Because present capability is normalized to , Eq. (14) gives . Closure then rises toward unity, while remaining strictly below one at any finite capability. The parameter therefore represents the recursive gain under ideal operational closure.
For the finite-frontier model with , the local recursive reproduction number is
Operational closure rises with capability, while remaining frontier headroom declines. Their product can therefore reach an interior maximum. We define the peak recursive criticality of a parameter combination by
We choose four reference values of to represent distinct dynamical regimes: subcritical smooth scaling, weak supercritical, transient takeoff and rapid transition.
The reference value produces a relatively concentrated, but still smooth, increase in operational closure. At and , corresponding to the weak-supercritical scenario, this choice places the system only modestly above the peak critical boundary. We hold fixed across all primary scenarios considered below.
3.3 Reference dynamical regimes
The phase maps in Fig. 2 identify parameter combinations that contain a locally supercritical region. They do not show how long the trajectory remains in that region, how much capability is accumulated there, or where the supercritical episode falls relative to the AGI and ASI thresholds. Figure 3 follows the four reference parameterizations through time.
Recursive feedback can materially alter the capability trajectory even when the system never becomes supercritical. In the smooth scaling scenario, throughout (Fig. 3b), yet recursive feedback advances both threshold crossings relative to the no-RSI baseline. Subcriticality therefore rules out self-amplification of local perturbations, but not a substantial cumulative contribution from recursive feedback.
The weak supercritical scenario shows why crossing the critical boundary need not coincide with an obvious takeoff. The reproduction number rises only modestly above unity and falls below the boundary well before AGI is reached. Nevertheless, the temporary period of amplification leaves the system on a persistently advanced capability trajectory. Frontier hardening subsequently suppresses further amplification, but it does not undo the capability accumulated during the supercritical episode.
As recursive gain increases, a larger part of the AGI-to-ASI transition occurs while the system is supercritical. In the rapid transition scenario, both thresholds are crossed during the same amplification episode and the interval between them becomes very short. The reproduction number then declines rapidly as the trajectory approaches the fixed frontier, where increasing hardness eventually dominates recursive gain. Stronger recursive feedback can therefore change both the timing and concentration of progress without changing the asymptotic capability limit imposed by a fixed .
Table 1 quantifies these differences. The effect of recursive feedback becomes increasingly pronounced for thresholds farther from the initial state. Relative to the no-RSI trajectory, the subcritical smooth scaling scenario advances AGI by about years and ASI by about years. At the other extreme, the rapid transition scenario advances AGI by about years and ASI by about years, while compressing the AGI-to-ASI interval from years to less than half a year.
The asymmetry arises because recursive feedback compounds over the capability interval. Its effect on the later ASI threshold can therefore be much larger than its effect on the first AGI crossing. Relatively modest differences in AGI timing can coexist with very large differences in the duration of the subsequent transition.
Taken together, the reference trajectories show that the same initial capability can support markedly different development paths. Subcritical feedback can produce a meaningful acceleration, a brief supercritical episode can leave a durable capability lead, and strong recursive amplification can compress the AGI-to-ASI transition while remaining self-limiting. The trajectory therefore depends on the evolving balance between recursive gain, operational closure and frontier hardening, not on capability level alone.
| Scenario | (yr) | (yr) | (yr) | |
|---|---|---|---|---|
| No-RSI baseline | 0 | 24.00 | 96.00 | 72.00 |
| Smooth scaling | 0.5 | 21.35 | 74.13 | 52.79 |
| Weak supercriticality | 3 | 12.90 | 24.73 | 11.83 |
| Transient takeoff | 6 | 8.40 | 11.08 | 2.69 |
| Rapid AGI-to-ASI | 15 | 4.50 | 4.95 | 0.45 |
In reference dynamical regimes the effective frontier is fixed and frontier hardening dominates when capability advances faster than the effective frontier moves. However, can change during AI development. New model architectures, training methods or theoretical insights can induce a sufficiently rapid outward movement of and prolong the supercritical regime.
A discrete scientific or algorithmic breakthrough can instead produce a jump . The resulting increase in headroom raises immediately in the reduced model. A trajectory can therefore undergo several separated recursive episodes as successive paradigms create new opportunities and each opportunity set is later exhausted. The finite-frontier result should accordingly be interpreted as local to a research paradigm rather than as evidence for a single final capability ceiling.
It should be emphasized, that the reference regimes are intended to expose the qualitative consequences of alternative recursive-feedback assumptions rather than to provide probabilistic forecasts. We calibrate to recent expert expectations for AGI timing and choose to represent a plausible AI-R&D cycle, although both remain uncertain. The frontier-hardening exponent , recursive gain and operational-closure steepness are not presently constrained by direct empirical estimates and are selected to span distinct dynamical regimes. The periods reported in Fig. 3, Table 1 and for other scenarios should therefore be interpreted only as conditional outputs of the stated parameterization.
4 Strategic scenarios
4.1 Strategic competition and network organization
The coupled-agent framework in Sec. 2.4 allows strategic conditions to affect AI development through several distinct channels. Competition can mobilize additional investment, compute and research effort, while also changing how much of the AI–R&D process is delegated to advanced models, how quickly successor systems are developed, and how readily improvements propagate between research actors [29, 30, 31]. These effects act on different components of the recursive dynamics and need not move together.
We represent changes in baseline research intensity by an effort multiplier . Increasing raises the rate at which capability advances, but in the minimal model does not by itself change the local recursive threshold. Changes in recursive gain, operational closure or frontier hardening alter the actor-level reproduction number, while cross-actor transfer changes the network reproduction number . Strategic competition can consequently accelerate development, alter recursive criticality, and reshape the network through which improvements accumulate and spread.
We use the weak supercritical reference trajectory from Sec. 3.3 as a common starting point and compare three illustrative research environments (Fig. 4a). Each configuration specifies , , and a matrix describing within-actor recursive gain and directed transfer between actors. The diagonal elements describe normalized recursive reproduction within actor , while the off-diagonal elements describe the contribution of actor to the future research productivity of actor . Their spectral radius gives the corresponding network-level recursive criticality, as established in Sec. 2.4.
The reference trajectory defines , so values above or below unity represent greater or lower effective research throughput relative to the same baseline. The three configurations are chosen to separate the effects of strong internal recursion, broad cross-actor diffusion and elevated competitive effort rather than to represent forecasts of particular firms or nations. Figure 4 compares their network structure and resulting capability dynamics, while Table 2 summarizes the corresponding threshold-crossing times.
The closed laboratory scenario concentrates recursive gain within individual organizations. The leading laboratory begins at the local critical boundary, , while the other two remain subcritical. Cross-laboratory transfer is weak, so coupling raises the network reproduction number only modestly. We keep research effort at the reference level, , and use a short feedback delay of yr to represent rapid internal development cycles. Capability trajectories then diverge gradually as differences in within-laboratory recursive gain accumulate over successive cycles.
The open ecosystem produces amplification through a different mechanism. All three actors are individually subcritical, but strong off-diagonal transfer allows improvements generated by one actor to raise the subsequent research productivity of the others. Recursive feedback is therefore distributed across the network rather than concentrated inside a single organization. We set the effort multiplier below the reference value, , with a characteristic delay of yr. Despite the lower baseline research intensity, network coupling produces the shortest AGI-to-ASI transition of the three scenarios, approximately years.
The global competition scenario combines higher research effort with a more weakly and asymmetrically connected network. Each bloc is individually subcritical at the reference state, but transfer between the leading blocs raises the network reproduction number to . We set to represent additional resources mobilized by strategic competition and use a longer delay of yr to represent slower transmission, replication and integration of advances across competing actors. This scenario reaches AGI first, after approximately years, but its AGI-to-ASI interval remains about years, longer than in the open ecosystem despite its greater research effort.
These scenarios are designed to separate mechanisms rather than to estimate the behaviour of particular institutions or geopolitical systems. Greater openness can increase the diffusion of useful advances, while organizational separation can keep recursive gain concentrated within individual actors. Competition can increase the overall rate of research without producing a comparable increase in recursive coupling. The resulting trajectories depend jointly on the reproduction matrix , the research-effort multiplier , and the feedback delay . Research-network structure should therefore be treated as part of the dynamical state of advanced AI development. Aggregate investment provides information about the speed of progress, but it does not determine proximity to recursive criticality.
| Configuration | (yr) | (yr) | (yr) |
|---|---|---|---|
| Closed frontier-lab competition | 10.40 | 16.53 | 6.13 |
| Open competitive ecosystem | 9.75 | 12.58 | 2.82 |
| Global competition | 8.97 | 13.42 | 4.45 |
4.2 Hardware and power limits
The reference trajectories assume that physical infrastructure can expand as quickly as software-side capability demands. This may fail well before fundamental computational limits become relevant, because frontier systems also require power, cooling, networking and data-centre infrastructure that expand on physical construction timescales [32].
Let denote the effective compute capacity supported by available infrastructure and the compute required to instantiate capability . Deployability requires
| (15) |
Algorithmic progress can reduce at fixed capability [26, 27], so this quantity represents the net compute requirement under the assumptions of each scenario.
We approximate the requirement locally by
| (16) |
and define physical headroom at the unconstrained AGI crossing as
| (17) |
The corresponding compute increase between the illustrative AGI and ASI thresholds is
| (18) |
For illustration, we set and let effective physical compute grow at . These values are scenario assumptions rather than estimates of the resource requirements of AGI or ASI. Figure 5 shows three regimes. With , infrastructure remains ahead of the software trajectory. At , capacity becomes limiting between AGI and ASI and lengthens the transition. At , the constraint already binds before the unconstrained AGI crossing and delays both thresholds.
Physical compute has a different role from recursive gain. As a capacity constraint, it limits which software-side capabilities can be deployed without directly changing at a given capability. Compute can also alter recursive dynamics when it increases operational closure, strengthens research gain or shortens development cycles [33]. The scenarios here isolate the first effect.
5 Conclusion
Recursive self-improvement is best understood as a dynamical property of an AI-enabled R&D system. In our model, the transition to self-amplifying improvement occurs when realized recursive gain exceeds the local hardening of the research frontier, so that . Incremental improvements then amplify across successive development cycles. This transition is determined by the structure of the development process and need not coincide with a particular capability threshold such as AGI. Crossing is a local condition for amplification and does not by itself imply indefinite acceleration or unbounded capability growth.
The framework separates the onset, speed, and persistence of recursive amplification. determines whether incremental gains amplify, while development-cycle delay constrains how rapidly that amplification unfolds. Increasing research difficulty can subsequently return a supercritical system to a subcritical regime, so strong recursive amplification can be transient within a fixed research paradigm. Higher research throughput can produce rapid progress without changing the recursive regime, and a newly supercritical system can initially resemble ordinary acceleration. Recursive gain also depends on operational closure, thus, AI-generated research affects future capability only to the extent that it propagates through evaluation, integration, training, and deployment into successor systems.
These distinctions extend beyond a single research actor. In coupled research ecosystems, improvements can propagate between organizations strongly enough to make the network supercritical even when every actor is individually subcritical. Physical infrastructure introduces a different constraint by limiting which software-side capabilities can be deployed without necessarily changing recursive criticality itself. The numerical scenarios illustrate how these mechanisms can generate qualitatively different development trajectories under alternative assumptions.
The main empirical challenge is therefore to measure the feedback mechanisms directly. Relevant quantities include the causal effect of AI research capability on subsequent R&D productivity, the fraction of potential gains that propagate into successor systems, development-cycle duration, resource-normalized research productivity, frontier hardening, and the transfer of improvements across actors. Capability growth alone is not sufficient to diagnose recursive criticality. The more informative signal is whether each increment of AI research capability is becoming increasingly effective at producing the next one. If recursive self-improvement emerges, changes in this reproduction of research capability may become detectable before the most visible phase of acceleration.
References
- [1] (2025) Measuring AI ability to complete long software tasks. In Advances in Neural Information Processing Systems, Vol. 38. External Links: Document, Link Cited by: §1, §1.
- [2] (2024) Mle-bench: evaluating machine learning agents on machine learning engineering, 2024. URL https://arxiv. org/abs/2410.07095 2410. Cited by: §1, §1.
- [3] (2025) RE-bench: evaluating frontier AI R&D capabilities of language model agents against human experts. In Proceedings of the 42nd International Conference on Machine Learning, Proceedings of Machine Learning Research, Vol. 267, pp. 66772–66832. External Links: Link Cited by: §1, §1.
- [4] (2025) PaperBench: evaluating AI’s ability to replicate AI research. In Proceedings of the 42nd International Conference on Machine Learning, Proceedings of Machine Learning Research, Vol. 267, pp. 56843–56873. External Links: Link Cited by: §1, §1.
- [5] (2026) Ai research agents for machine learning: search, exploration, and generalization in mle-bench. Advances in Neural Information Processing Systems 38, pp. 35309–35348. Cited by: §1, §1.
- [6] (1966) Speculations concerning the first ultraintelligent machine. In Advances in Computers, F. L. Alt and M. Rubinoff (Eds.), Vol. 6, pp. 31–88. External Links: Document Cited by: §1.
- [7] (2012) Can intelligence explode?. Journal of Consciousness Studies 19 (1–2), pp. 143–166. External Links: 1202.6177, Document Cited by: §1.
- [8] (2007) Gödel machines: fully self-referential optimal universal self-improvers. In Artificial General Intelligence, B. Goertzel and C. Pennachin (Eds.), Cognitive Technologies, pp. 199–226. Note: Originally circulated in 2006 External Links: Document Cited by: §1.
- [9] (2016) Self-modification of policy and utility function in rational agents. In Artificial General Intelligence, pp. 1–11. External Links: Document, 1605.03142 Cited by: §1.
- [10] (2017) Artificial intelligence and economic growth. Working Paper Technical Report 23928, National Bureau of Economic Research. External Links: Document, Link Cited by: §1.
- [11] (2023) What a compute-centric framework says about takeoff speeds. Technical report Open Philanthropy. Note: Originally published by Open Philanthropy External Links: Link Cited by: §1.
- [12] (2024) Economic impacts of AI-augmented R&D. Research Policy 53 (7), pp. 105037. Note: First circulated in 2022 External Links: Document, 2212.08198 Cited by: §1.
- [13] (2025) Artificial intelligence in research and development. Working Paper Technical Report 34312, National Bureau of Economic Research. External Links: Document Cited by: §1.
- [14] (2026) When does automating AI research produce explosive growth? feedback loops in innovation networks. Working Paper Technical Report 35155, National Bureau of Economic Research. External Links: Document Cited by: §1.
- [15] (2025) A mathematical framework for AI singularity: conditions, bounds, and control of recursive improvement. arXiv preprint arXiv:2511.10668. External Links: 2511.10668, Document Cited by: §1.
- [16] (2026) Recursive self-improvement in AI: from bounded self-refinement to autonomous research loops. arXiv preprint arXiv:2607.07663. External Links: 2607.07663, Document Cited by: §1.
- [17] (2026) The economics of recursive self-improvement. Technical report Elasticity Institute. External Links: Link Cited by: §1, §2.2.
- [18] (2026) From AGI to ASI. arXiv preprint arXiv:2606.12683. External Links: 2606.12683, Document, Link Cited by: §1.
- [19] (2026) The emerging science of machine learning benchmarks. Princeton University Press. Note: Forthcoming; hardcover publication scheduled for 6 October 2026 External Links: Link Cited by: §1.
- [20] (2025) A rosetta stone for AI benchmarks. arXiv preprint arXiv:2512.00193. External Links: 2512.00193, Document, Link Cited by: §1.
- [21] (2026) Measuring AI R&D automation. arXiv preprint arXiv:2603.03992. External Links: 2603.03992, Document Cited by: §1.
- [22] (2018) Stability and performance analysis of linear positive systems with delays using input–output methods. International Journal of Control 91 (7), pp. 1669–1692. External Links: Document, 1703.00405 Cited by: §2.2.
- [23] (2026) Even after R&D is automated, parallelization constraints could delay a technological singularity. Technical report Epoch AI. External Links: Link Cited by: §2.2.
- [24] (2025) Thousands of AI authors on the future of AI. Journal of Artificial Intelligence Research 84. External Links: Document, 2401.02843, Link Cited by: §3.1.
- [25] (2025) The longitudinal expert ai panel: understanding expert views on ai capabilities, adoption, and impact. Working paper Technical Report 5, Forecasting Research Institute. External Links: Link Cited by: §3.1.
- [26] (2024) Algorithmic progress in language models. arXiv preprint arXiv:2403.05812. External Links: 2403.05812, Document Cited by: §3.2, §4.2.
- [27] (2025) On the origin of algorithmic progress in AI. arXiv preprint arXiv:2511.21622. External Links: 2511.21622, Document Cited by: §3.2, §4.2.
- [28] (2026) Expenditure horizon: measuring optimization ability, with an application to NanoGPT. Note: Model Evaluation and Threat Research External Links: Link Cited by: §3.2.
- [29] (1987) Dynamic R&D competition. The Economic Journal 97 (386), pp. 372–387. External Links: Document Cited by: §4.1.
- [30] (2016) Racing to the precipice: a model of artificial intelligence development. AI & Society 31 (2), pp. 201–206. External Links: Document Cited by: §4.1.
- [31] (2026) The AGI race and existential risk. Working Paper Technical Report 35276, National Bureau of Economic Research. External Links: Document Cited by: §4.1.
- [32] (2025) Energy and AI. Technical report International Energy Agency, Paris. External Links: Link Cited by: §4.2.
- [33] (2025) Will compute bottlenecks prevent an intelligence explosion?. arXiv preprint arXiv:2507.23181. External Links: 2507.23181, Document, Link Cited by: §4.2.