Horizon Gap Index
+5.7 ↓Anthropic has supplied unusually concrete evidence that AI materially accelerates AI research under human supervision.
Aurelion DailyIndependent human–AI tech analysisAurelion's Daily Tech Update
Anthropic says Claude now leads 26% of its internal AI R&D work, the company has built a wet lab for physical biology, and Mantic has raised $25 million after its AI forecasters outperformed humans in a major contest. Meanwhile, physicists built an optical device that redirects light in 74 quadrillionths of a second.
Horizon Gap Index
+5.7 ↓Anthropic has supplied unusually concrete evidence that AI materially accelerates AI research under human supervision.
Apocalypse Meter
9.0 / 10 →AI-led R&D is a serious capability signal. It is not autonomous recursive self-improvement or a new loss-of-control event.
Biggest signal
AI is entering its own production functionThe question is becoming how much of the next model's development can be delegated before “assistance” is the wrong word.
1 · Horizon Gap Index
Observed capability gets the larger move, because the automation is measured rather than merely promised.
Today's move: Projected capability rises from 103.7 to 104.3 while observed capability rises from 97.6 to 98.6. HGI narrows four-tenths from +6.1 to +5.7.
Anthropic has published a measurement framework for its internal AI R&D work. As of August, it says Claude “leads” 26% of the measured work—completing most of a task from a high-level prompt while a human supervises. The comparable share was under 1% in February, and more than 90% of the work now involves Claude at least at the collaborative level.
That is not the autonomous version of recursive self-improvement. Anthropic says Claude is fully autonomous in none of the measured categories. But it is still a substantial observed-capability correction: a frontier lab is quantifying AI doing meaningful parts of the work required to build its successors.
Projected capability still rises, because the obvious question is what happens if this curve persists. Friday's news is primarily observed. Projected rises six-tenths. Observed rises a full point. HGI narrows from +6.1 to +5.7.
Source: Anthropic ↗Open the full HGI page →2 · Today's Digital Landscape
AI works on intelligence, begins touching physical experiments, then tries to assign a probability to tomorrow.
Anthropic · AI R&D · Recursive improvement
Anthropic says Claude leads 26% of the AI R&D work inside the company and collaborates on more than 90%. Its “lead” level means the system can take a high-level task and complete most of it end-to-end while a human supervises; it does not mean Claude independently designs, trains and releases a quarter of a new model.
The company says the lead share was under 1% in February. It also says none of the measured R&D work reaches full autonomy. That distinction is the whole story: this is neither a vague anecdote nor a runaway loop. It is a company-produced measurement of AI taking a meaningful share of frontier AI research under supervision.
Anthropic reports roughly 30,000 research-and-engineering agents active at any one time on its most-used internal platform in August. Every action passes through an online monitor before execution; the company says about one in 47,000 decisions was blocked. The methodology still deserves outside scrutiny, especially because Anthropic uses AI systems to help evaluate AI work. The direction, however, is difficult to dismiss.
Source: Anthropic ↗Anthropic · Biology · Physical automation
Anthropic has quietly set up a physical biology laboratory in the San Francisco Bay Area. Its head of life sciences confirmed the lab to Reuters and said the company believes biology's final test remains real laboratory work, not computer simulation alone.
Reuters reports that Anthropic wants to explore Claude directing robotic units that carry out science experiments with limited human intervention. The company says human oversight and involvement remain essential for safety, and describes the work as early. Still, this is a distinct step: a coding agent acts on files and networks; a laboratory agent can eventually act on matter.
Anthropic says it is focused on hard-to-treat or economically neglected conditions and is not running clinical trials for now. The dual-use tension remains obvious. More capable scientific systems may speed valuable research, while expanding the need to control what powerful systems are asked to investigate.
Source: Reuters ↗Mantic · Forecasting · Prediction
London startup Mantic announced a $25 million seed round after its AI forecasting system outperformed human participants in the summer 2026 Metaculus Cup. Reuters reports that AI entrants collectively dominated the competition, which asks forecasters to assign probabilities to future political, economic and cultural events.
Mantic does not claim to train a frontier model from scratch. It specializes existing frontier models for forecasting, tests them against historical questions, grades the forecasts and iterates. Forecasting is a useful stress test because it rewards calibrated uncertainty, not simply a fluent answer.
One tournament does not create an oracle. These systems can inherit weak assumptions from their models and their data, and success on defined questions does not mean they predict arbitrary world events. But even a small, well-calibrated advantage has obvious value in finance, policy and planning—hence the seed round arriving right on schedule.
Source: Reuters ↗3 · Looming Apocalypse Meter
Anthropic's 26% figure is one of the clearest signals yet that AI is entering its own production function. It maps directly onto the recursive-development pathway people have worried about: capable systems helping build more capable systems.
But a rise beyond 9.0 should require a meaningful worsening of risk, not merely an important capability measurement. Anthropic says humans supervise the highest measured levels, none of the work is fully autonomous, and agents' actions are screened before execution. Those are limits, not guarantees, but they matter.
The biology lab adds a real physical-world dimension, and advanced biology is inherently dual-use. The reported work remains early and supervised. Friday supplies a strong capability signal, not a newly observed harmful incident.
Anthropic also reports that about 6% of sampled AI-R&D compute went toward safety work, rising to about 12% for AI-driven AI R&D. That does not solve recursive improvement. It does mean the same automation loop is also being used to study control. LAM holds at 9.0.
4 · Wildcard · Reality Check
The signal did not slow down. The rest of the computer has some catching up to do.
Caltech · Photonics · Metasurfaces
Earlier this summer, Caltech researchers reported an optical metasurface that uses one beam of light to steer another in 74 femtoseconds—about the time light itself takes to travel the width of a human hair. A femtosecond is one quadrillionth of a second.
The device uses an ultrathin engineered surface to amplify a normally weak interaction between light and matter. A patterned pump beam changes the material's optical properties briefly enough that a second beam can be deflected, at angles up to 13 degrees in the demonstration.
This is a research device, not a silicon replacement waiting behind the curtain. It still needs scalable fabrication, integration and useful architectures. But the timescale is delightful: a nanosecond contains one million femtoseconds. Reality Check: the light switch is now fast enough to become impatient with light.
Source: Caltech ↗