{"id":14507,"date":"2026-08-07T14:28:49","date_gmt":"2026-08-07T14:28:49","guid":{"rendered":"https:\/\/futureknowledge.in\/?p=14507"},"modified":"2026-08-07T14:28:49","modified_gmt":"2026-08-07T14:28:49","slug":"inside-the-race-to-make-ai-build-itself","status":"publish","type":"post","link":"https:\/\/futureknowledge.in\/?p=14507","title":{"rendered":"Inside the Race to Make AI Build Itself"},"content":{"rendered":"<p>Jack Clark, Anthropic\u2019s co-founder, left on paternity leave last November. When he returned in February, he was surprised to learn that colleagues hardly wrote code anymore. They managed five or six copies of the company\u2019s AI, Claude, which sometimes managed several more Claudes.\u00a0<\/p>\n<p>To Clark, this looked like an early form of something the field has anticipated and feared for decades: recursive self-improvement, or the point at which AI begins to accelerate its own development. First, the thinking goes, models make researchers faster, but as each improvement feeds the next, the models take over more of the research cycle.\u00a0Years of progress compress into months\u2014leaving society with little time to absorb the consequences, from job disruption to engineered pathogens. Taken to its limit, AI could improve itself without humans,\u00a0triggering a runaway loop long known as an &quot;intelligence explosion,&quot; where machines rapidly advance beyond human understanding, and possibly beyond human control.<\/p>\n<p>Clark believed the world needed to confront this prospect. He posted a flurry of blog posts on recursive self-improvement and flew home to England to deliver a talk on the topic in May, then led an Anthropic report in June titled \u201cWhen AI Builds Itself,\u201d arguing the technology is already accelerating its development. The volume of code produced per person at Anthropic has increased eight-fold, with Claude writing 80%, the report noted. \u201cWe&#039;re trying to help substantiate this concept now &#8230; before it becomes something that is politicized or otherwise gains some valence that makes talking about it difficult,\u201d Clark says.<\/p>\n<p>It worked out as Clark feared. \u201cAnthropic is trying to strike terror into everyone\u2019s hearts,\u201d wrote AI skeptic Gary Marcus, adding \u201call they have really shown is just faster coding.\u201d Skeptics point out that technological progress has always compounded. Oil is used to drill oil. Why, in AI\u2019s case, should the curve suddenly bend upward? For a company betting on continued advances, they argued, the claim is plainly self-serving.<\/p>\n<p>Even Clark concedes coding volume is a crude yardstick. Claude\u2019s code can be long-winded. But the trouble runs deeper. Neither Claude nor any large language model is written in code at all. Researchers set growth conditions\u2014deciding the size and shape of a neural network, then pour an internet\u2019s worth of text through it, letting it adjust itself billions of times until abilities to answer questions, write code, and hold a conversation emerge. They are cultivated, the way one grows a plant by tending the soil and the light without ever deciding where a single leaf will go.\u00a0<\/p>\n<p>Progress, therefore, depends on trial and error. If Claude could take over that cycle, designing, running, and analyzing experiments, would progress accelerate gradually\u2026 or suddenly explode? And if it did, could anyone pump the brakes? The uncomfortable truth is that the people building the technology are nearly as much in the dark as everyone else.<\/p>\n<p>The change Clark had walked into was not entirely unexpected. Fellow Anthropic co-founder and chief science officer Jared Kaplan had long feared that AI would eventually accelerate research, perhaps outpacing safety efforts. In early 2025, he folded a warning into the company\u2019s Responsible Scaling Policy, its plan for managing AI\u2019s growing dangers. Back then, Claude was no good at running experiments. But he believed that would change one day, and Anthropic would need to be ready.\u00a0<\/p>\n<p>To find out whether that day was coming, Anthropic built a series of tests\u2014tasks that would take a human expert hours. Could Claude train a smaller AI model from scratch? Could it program a virtual robot dog? No single measure would settle it, but together they offered a snapshot.\u00a0<\/p>\n<p>In one test, Claude had to rewrite a piece of code to use GPUs\u2014the chips for training AI\u2014more efficiently. In spring, it momentarily got them running seven times faster, then broke the code. By summer, a newer Claude pushed the same speedup from seven times faster to 73 without introducing errors. \u201cWe started seeing these tasks fall over,\u201d says Daniel Freeman, a member of Anthropic\u2019s frontier red team who designed the evaluations.\u00a0<\/p>\n<p>As Claude began completing such tasks too reliably to reveal much, Freeman devised a test closer to the messiness of the real world. Two human teams competed in a series of challenges, involving a robotic dog and a beach ball. One team could use Claude, the other could not. After narrowly losing the first challenge, the team using Claude finished the second nearly two hours ahead. With time to spare, they trotted their robotic dog around the warehouse\u2014until a miscalculation sent it springing toward the other team, still hunched over their laptops. An overseer caught it just in time.<\/p>\n<p>The November release of Claude Opus 4.5 was a tipping point. Where previous generations tended to stall partway, the model could carry a researcher&#039;s experiment through to the end, freeing them to run more at once. &quot;Before, I&#039;d have eight ideas and I&#039;d try one of them,&quot; Kaplan told TIME in February. &quot;Now I ask Claude to just try all eight.&quot;\u00a0<\/p>\n<p>Ahead of a February release, Anthropic surveyed 16 of its researchers, asking whether it could replace an entry-level colleague. Five thought it might. Asked to reflect on their answer, all five walked it back. That they even entertained the idea\u2014that they were already largely redundant\u2014is perhaps more revealing than any benchmark.\u00a0<\/p>\n<p>If Anthropic\u2019s tests fell, outside measures have similarly reached their limit. The METR graph, perhaps the best-known independent measure of AI software engineering ability, tracks the complexity of tasks models can complete based on how long they would take a human expert. But in May, Claude exceeded the benchmark\u2019s upper limit.<\/p>\n<p>This spring, Anthropic let the robot dogs out again, but this time, an improved Claude worked alone. On every task it could attempt, it was at least 10 times faster than the Claude-assisted humans had been months earlier. The researchers saw the same pattern they had found in cybersecurity. First AI helps humans do better; then it largely takes over.<\/p>\n<p>Jack Clark speaks during the Hill &amp; Valley Forum at the U.S. Capitol Visitor Center Auditorium in Washington, DC, on April 30, 2025. Brendan Smialowski\u2014AFP\/Getty Images<\/p>\n<p><em>Source: <a href='https:\/\/time.com\/article\/2026\/08\/07\/ai-recursive-self-improvement-anthropic-openai\/' target='_blank'>Read the original article on time.com<\/a><\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Jack Clark, Anthropic\u2019s co-founder, left on paternity leave last November. When he returned in February, he was surprised to learn that colleagues hardly wrote code anymore. They managed five or six copies of the company\u2019s AI, Claude, which sometimes managed several more Claudes.\u00a0 To Clark, this looked like an early form of something the field [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":14508,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[4,36,3],"tags":[12,29,33],"class_list":["post-14507","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-important","category-share-suggestions","category-technology","tag-impact-intc","tag-signal-avoid","tag-stage-stage-4"],"_links":{"self":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/14507","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=14507"}],"version-history":[{"count":0,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/14507\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/media\/14508"}],"wp:attachment":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=14507"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=14507"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=14507"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}