Ptechhub
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs
No Result
View All Result
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs
No Result
View All Result
PtechHub
No Result
View All Result

I Built a Self-Improving AI, and So Can You

By Wired by By Wired
July 8, 2026
Home AI & ML
Share on FacebookShare on Twitter


These days, the frontier AI labs are all racing to build self-improving models. Some believe it’s the surest route to superintelligence—as AI improves itself in a mind-melting loop, the thinking goes, it will eventually surpass human comprehension (and perhaps even control).

That’s all well and good, but I have a newsletter to produce. I wondered if recursive self-improvement might also be useful for me. Could I use AI to train and continually improve a model that automates some of this newsletter’s busywork?

After a week or so of experimenting, the answer appears to be a resounding—and surprising—hell yes. What’s more, dabbling with self-improving models shows a different vision for how AI might unfold—one that doesn’t center on a handful of companies that control the whole industry.

I started by trying out a simple self-improving loop

To get my feet wet, I experimented with training a small language model from scratch—by which I mean I dumped all the hard work on Claude’s plate.

I installed AutoResearch, which helps an off-the-shelf AI model build and improve a smaller model. AutoResearch is the brainchild of Andrej Karpathy, a superstar AI researcher who helped found OpenAI, led AI work at Tesla, and recently joined Anthropic.

I fired up Claude and gave it the recommended instruction: “Hi, have a look at program.md and let’s kick off a new experiment!” While Claude did the hard stuff, I provided silicon (an Nvidia DGX, a desktop “supercomputer” designed for AI experimentation), the electricity (running hot for a few days straight), and a possibly ill-advised willingness to let the model skip all the usual permission checks in order to do its thing (let him cook!)

I checked in on the AutoResearch project every few hours and marveled as Claude adjusted parameters and training regimes, looked at how this changed the smaller model’s output, and went on refining it further.

Here’s what an early version of that smaller language model produced when I prompted it to complete the phrase “In the beginning …”

“In the beginning of the beginning of the end of the end of the end end of end end end end end end end end beginning end end end end…”

Not so brilliant. But later models, improved autonomously by Claude, got more coherent and less prone to insane, endless repetition. It’s hardly GPT-5, but it showed a promising path toward continual improvement.

My journey continued with something more complex—and useful

I already use an agent that relies on Claude to help me find noteworthy research papers, so I decided to see whether it was possible to build something that went beyond that.

I turned to a tool from a startup called Prime Intellect, which uses AI to train a custom model for a specific task. I collected 100 or so previous “Elsewhere on the frontier of AI” entries—the bits and bobs of research that follow the main essay in my newsletter. Then, I created a Prime Intellect training environment and asked Claude to help me build my own model, which it dubbed Frontier_Paper_Curator, to find and summarize interesting papers.

Claude found more papers and generated a bunch of synthetic data to help with training. It then tapped yet another model to assess Frontier_Paper_Curator’s output, while the training environment also improved the model with reinforcement learning.



Source link

Tags: Agentic AIai labanthropicArtificial Intelligencechatgptclaudeopenairesearch
By Wired

By Wired

Next Post
Messi and Ronaldo Are Building Tech Portfolios. Mo Salah Is Playing a Different Game

Messi and Ronaldo Are Building Tech Portfolios. Mo Salah Is Playing a Different Game

Recommended.

This bargain fintech stock is stuck in a five-year rut. A turnaround is coming

This bargain fintech stock is stuck in a five-year rut. A turnaround is coming

March 27, 2026
HPE Sells Nearly B Stake In H3C Tech, Plans To Sell Remaining Stake For 0M

HPE Sells Nearly $1B Stake In H3C Tech, Plans To Sell Remaining Stake For $370M

May 14, 2026

Trending.

AWS, Google, Oracle, Microsoft Top Gartner’s Cloud AI Infrastructure List For 2026

AWS, Google, Oracle, Microsoft Top Gartner’s Cloud AI Infrastructure List For 2026

July 29, 2026
Cloud Market Share Q1 2026: AWS, Microsoft, Google Battling In AI Era

Cloud Market Share Q1 2026: AWS, Microsoft, Google Battling In AI Era

May 4, 2026
The 50 Coolest Software-Defined Storage Vendors: The 2026 Storage 100

The 50 Coolest Software-Defined Storage Vendors: The 2026 Storage 100

April 13, 2026
Anthropic lost control of Claude in latest AI cyber blunder | Computer Weekly

Anthropic lost control of Claude in latest AI cyber blunder | Computer Weekly

July 31, 2026
IDCA datacentres report: Global concentration and the Goldilocks zone | Computer Weekly

IDCA datacentres report: Global concentration and the Goldilocks zone | Computer Weekly

May 12, 2026

PTechHub

A tech news platform delivering fresh perspectives, critical insights, and in-depth reporting — beyond the buzz. We cover innovation, policy, and digital culture with clarity, independence, and a sharp editorial edge.

Follow Us

Industries

  • AI & ML
  • Cybersecurity
  • Enterprise IT
  • Finance
  • Telco

Navigation

  • About
  • Advertise
  • Privacy & Policy
  • Contact

Subscribe to Our Newsletter

  • About
  • Advertise
  • Privacy & Policy
  • Contact

Copyright © 2025 | Powered By Porpholio

No Result
View All Result
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs

Copyright © 2025 | Powered By Porpholio