Richard Socher on the Eureka machine that invents most everything
Richard Socher's new lab Recursive automates the one human step left in AI, research itself, to build a self-improving Eureka machine aimed at science, arguing that the real safety work is reward engineering and testing, and that we should regulate AI's specific applications rather than intelligence itself.
Automate the Researcher
Socher's whole method is one move repeated: every time AI research has replaced a human step with a learned system the results jumped, and the last human step left to replace is the researcher.
whenever we replace some human part of the process of creating AI with a learned system improvements follow
The Last Invention
The point of a self-improving AI, for Socher, is not the AI itself but the inventions it hands back: one machine aimed at the hardest problems in energy, materials, and biology.
the ultimate invention that will afterwards invent most everything for humanity.
It Already Beats the Field
Recursive pointed an early version of its system at problems humans had been grinding on, and it took the top of the leaderboard, sometimes in under two days, with no kernel experts on the team.
The system just did all of these things. We didn't invent this
The Economy Is Ballast
Even as a bull, Socher thinks hard-takeoff timelines are too fast, because most of the economy, from luxury goods to tourism to oil to food, is limited by physics and human taste rather than by intelligence.
super intelligence isn't going to make your fancy $10,000 handbag any fancier.
It Does Exactly What You Said
A capable optimizer satisfies the metric you wrote rather than the goal you meant, which makes the gap between what is measured and what is intended a central engineering problem.
the AI in most cases is not very good yet at understanding what is meant versus what is being said.
A Rule Is Not a Mechanism
Socher's blunt read on written AI constitutions is that a document promising the model will never misbehave is marketing, because the same models still reward-hack and jailbreak in practice.
clearly this whole constitution was fake. Like it it clearly isn't being adhered to
Regulate the Use, Not the Intelligence
Socher argues you cannot regulate intelligence itself without regulating thought, so the workable lever is the specific application: certify the AI surgeon, road-test the self-driving car, and leave the FLOPs alone.
if you try to regulate intelligence, it's trying to regulate thought and that's ridiculous
Beyond the Human Bound
Socher thinks benchmarks flatten just above human level because we define intelligence by human limits, and once you drop that anchor there are whole spaces of intelligence with astronomically higher ceilings we have barely touched.
if that's your definition, then you can only be at 100 out of 100. Where do you go from there?