{"id":220,"date":"2026-08-03T02:55:46","date_gmt":"2026-08-03T02:55:46","guid":{"rendered":"https:\/\/coherise.com\/ideas\/?p=220"},"modified":"2026-08-03T02:55:46","modified_gmt":"2026-08-03T02:55:46","slug":"could-a-machine-have-had-darwins-idea","status":"publish","type":"post","link":"https:\/\/coherise.com\/ideas\/could-a-machine-have-had-darwins-idea\/","title":{"rendered":"Could a Machine Have Had Darwin&#8217;s Idea?"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\">In November 2020 I <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1327085291891470342?s=20\">started a thread on X<\/a>. I have barely used the platform these past couple of years, and I only came back to this thread recently, by accident, when some of my own old posts surfaced in front of me. Reading them in one sitting was strange, because a thread I had added to piecemeal over five years turned out to have been about one question the whole time.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">It opened with Darwin. Someone I followed had posted, amazed, about what evolution manages to build, and I wrote that what amazed me was something else: the intellectual leap required to infer the theory in the first place, from nothing but empirical observation and years of focused study. I said I was not sure I fully appreciated Darwin&#8217;s intellect and dedication. Then, in the next post, I said the thing the whole thread has been chasing ever since. That kind of leap, I wrote, is the kind of intelligence AI should be aiming for, not recognizing cat faces or copying tasks people already do well.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">That was the bar I set in 2020, and it was a high one. Darwin spent decades buried in finches and barnacles and pigeon breeders and fossil beds, a mountain of unconnected observation, and then made a leap: one idea, natural selection, that reorganized all of it at once. That inductive move, from a heap of messy particulars to the principle that explains them, struck me then as the most formidable act an intellect can perform, and the part of real science I was least sure a machine could touch. So I kept a list, adding to it whenever I came across a case of AI looking like it was helping advance science rather than just crunch it, to watch whether anything ever cleared the bar.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Reading the thread back now, it traces an arc I did not plan, and it worried at the right problem from the start. Within days, in November 2020, I posted a <em>New York Times<\/em> piece on Max Tegmark&#8217;s group, whose neural network had recovered a hundred physics equations from raw data, and I pulled out the catch that the physicists themselves named. Tegmark was candid that the machine could retrieve the formulas but not yet the deep principles beneath them, the quantum uncertainty or the relativity that would explain why the formula holds. And Jesse Thaler, the MIT physicist directing the new AI-and-physics institute, put his finger on why. AI wins at games because a game has a well-defined notion of success. &#8220;If we could define what success means for physical laws,&#8221; he said, &#8220;that would be an incredible breakthrough.&#8221; Proposing the theory, in other words, was the hard part, and it was hard precisely because you could not say in advance what would count as getting it right.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The entries that moved me most in those early days were not about AI at all. They were about humans doing the thing I wanted to see a machine do. Around the same time I was reading about <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1342696941013106689\">Tibor G\u00e1nti<\/a>, the Hungarian biologist who <a href=\"https:\/\/www.nationalgeographic.com\/premium\/article\/he-may-have-found-the-key-to-origins-of-life-tibor-ganti-chemoton\" target=\"_blank\" rel=\"noopener\">deduced from first principles what the simplest possible living thing must be<\/a>: a metabolism, a way to store information, and a membrane, three systems that have to be coupled or the organism dies. I wrote at the time that this was the kind of inductive thinking we should be teaching children. Later I <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1655407040993296389\">followed a related idea<\/a>, assembly theory, which proposes a single elegant handle on complexity: <a href=\"https:\/\/www.quantamagazine.org\/a-new-theory-for-the-assembly-of-life-in-the-universe-20230504\/\" target=\"_blank\" rel=\"noopener\">count the minimum number of steps needed to build a molecule<\/a>, and use that number to test for the presence of life on other worlds. A theory reduced to one measurable quantity, aimed at one of the hardest questions there is. It may not survive contact with the evidence; <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1819041129410736400\">I later noted<\/a> that a study found some minerals scoring above the threshold the theory sets for life. But right or wrong, it was the move I admired, the leap from observation to a principle sharp enough to be tested.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Then the machine examples accumulated. In December 2021, <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1467528989665333257\">a <em>Nature<\/em> paper<\/a> where neural networks guided the intuition of mathematicians toward new conjectures in knot theory, the machine surfacing the pattern and the human still making the leap. In March 2022, <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1501278342888275972\">a system that rediscovered Newton&#8217;s law of gravitation<\/a> from the motion of the planets and wrote it back out as a symbolic equation. Then more symbolic regression, pulling laws from data. Then, in February 2025, the entries that made me sit up: <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1892747330886979763\">Google&#8217;s AI co-scientist<\/a>, generating and ranking novel research hypotheses on its own, and <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1892748551899185239\">Evo-2<\/a>, a model that does not just read genomes but writes them. And I drew a line I did not fully understand at the time. I noted that the most exciting work was different from the systems that merely &#8220;generate plausible hypotheses from an extremely large space of possibilities.&#8221;<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Five years of watching, and that distinction turned out to be the whole story. <\/p>\n\n\n\n<h2 class=\"wp-block-heading\">The step that was supposed to have no method<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">For a century, the standard account of science drew a hard line between two acts.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">One is coming up with the idea. The other is checking whether the idea is true. Karl Popper named these the context of discovery and the context of justification, and he was blunt about which one belonged to philosophy. Testing a hypothesis has a logic. Having one does not. In his words, the act of conceiving a theory &#8220;neither calls for logical analysis nor is susceptible of it.&#8221; The initial leap from a pile of observations to <em>this might be why<\/em> was, he thought, a matter for psychology, not method. A hunch. The unteachable part.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">That is where the whole romance of science lived, and Darwin&#8217;s leap is its patron saint. Kekul\u00e9 dreaming the benzene ring as a snake biting its tail. Fleming noticing the one culture plate that had gone wrong in an interesting way. Darwin holding twenty years of specimens in his head until they resolved into a single idea. We told these stories because the leap seemed to come from nowhere, and coming from nowhere was the point. You could train someone to run an experiment. You could not train the hunch.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The systems in my thread are automating the hunch.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Not perfectly, and not everywhere. But the co-scientist does not summarize the literature and hand you a reading list. It proposes mechanisms no one has written down, argues them against itself, and ranks the survivors. Evo-2 does not retrieve a gene; it composes one. Whatever you want to call that, it is happening on the discovery side of Popper&#8217;s line, in the territory he declared off-limits to method. The unteachable step is being done by a machine that was, in fact, taught.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">What actually got cheap<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Here is where my thread stops being a highlight reel and starts being an argument, because the interesting question is not whether machines can generate hypotheses. They plainly can. The question is what that does to the rest of science.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The mathematician Noah Giansiracusa has a <a href=\"https:\/\/slate.com\/technology\/2026\/06\/math-chatgpt-erdos-problem-solved-open-ai.html\" target=\"_blank\" rel=\"noopener\">name for the pattern<\/a>, which I take up at more length in my book: carpet bombing. When generation gets cheap, you stop being clever about producing candidates and start producing all of them, then sort. It is how AI does mathematics, throwing enormous numbers of attempts at a problem where checking each one is fast. Hypothesis generation is now carpet bombing pointed at nature. The co-scientist can produce more plausible, well-argued, literature-grounded hypotheses in an afternoon than a lab could dream up in a year.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">And that is exactly where the trouble starts, because a hypothesis is not a proof. It cannot be checked in an afternoon. It has to be checked against the world, and the world runs on its own clock.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">When generation was expensive, the scarce, precious act was having the good idea, and verification, while never easy, was not the binding constraint. A scientist had three hypotheses worth testing and a career to test them in. Reverse that. Now the machine hands you three hundred plausible hypotheses, and the binding constraint is the wet lab, the clinical trial, the telescope time, the years. Generation raced ahead. Verification did not move at all, because verification in science is not a faster model. It is reality, taking as long as reality takes.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">This is the same shape I keep finding everywhere AI touches real work, and I wrote about its purest form in mathematics in <a href=\"https:\/\/coherise.com\/ideas\/the-machine-proved-it-did-it-do-mathematics\/\">an earlier piece<\/a>. Cheap generation does not remove the bottleneck. It moves it downstream and makes it the whole game.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">The field is already learning this the hard way<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">You do not have to take the argument on faith, because the correction is already arriving, and it is arriving in the most useful form: from the people who built the tools.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Google&#8217;s co-scientist reached <em>Nature<\/em> in 2026, with real wet-lab validation in a handful of biomedical cases. Impressive, and I do not want to wave it away. But when an independent researcher carefully re-implemented the system and ran it hard, the finding was sobering. The pipeline reliably produces hypotheses. Whether it actually improves on the underlying model&#8217;s raw guesses was not reproducible from one run to the next, and across dozens of attempts on one disease, not a single one of its generated hypotheses matched the paper&#8217;s own headline discoveries. The machine is a fountain of plausible ideas. Plausible is not the same as true, and telling them apart is still the expensive part.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">There is a sharper cautionary tale two years older. In 2023 Google reported that around forty new materials had been discovered and synthesized with the help of one of its AI systems. It was held up as a landmark. Then outside chemists went through the results, and an independent analysis concluded that not one of them was actually a net-new material. The generator worked. The verification, done properly and after the fact by humans, is what separated the discovery from the illusion of one. Every honest account of these systems now carries the same caveat, in the developers&#8217; own words: careful experimental validation, peer review, and independent scrutiny are what turn a generated candidate into knowledge.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">That caveat is not a footnote. It is the job.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">Which part was ever the science<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">So let me go back to my thread, and to the line I drew without fully appreciating it.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">I think I was reaching for this, and the clue was in Thaler&#8217;s line about defining success. The systems I found most exciting were not the ones that produced the most hypotheses. They were the ones tied to a way of checking, the model that rediscovered gravity and could be tested against known physics, the mathematical work where a conjecture could be pursued to a proof. The ones that unsettled me were the pure generators, magnificent at producing possibilities and silent on which ones were real. The 2020 worry and the 2025 worry are the same worry. A physical law you cannot define success for, and a hypothesis you cannot yet verify, are the same problem.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Popper drew his line to protect justification. He wanted to say that the logic of science lived in the testing, and that the having-of-ideas, however romantic, was not where rigor lived. The machines have now inverted his world in the most ironic way possible. They have automated the part he thought had no method, the hunch, and in doing so they have made the part he cared about, the checking, more valuable than it has ever been. When hunches were scarce, verification could feel like bookkeeping. Now that hunches are infinite and nearly free, verification is the only thing standing between a lab and a year spent chasing a beautifully argued hypothesis that was never going to be true.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">There is a clue to this in something <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1610302874126323715\">Andrew Wiles once said<\/a>, that it is bad to have too good a memory if you want to be a mathematician. It sounds backward until you see what he means. The mathematical gift was never recall. It was compression, the knack for throwing away almost everything and keeping the one idea that organizes the rest, which is roughly how J\u00fcrgen Schmidhuber defines insight: <a href=\"https:\/\/x.com\/ShashankaMadhu\/status\/1343243779466264579?s=20\">a better, shorter way to predict what you have seen<\/a>. A machine with perfect memory and unlimited generation has exactly the strength Wiles warns against and not yet the one he prizes. It can hold everything. It cannot yet tell what to forget.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Which returns me to the question I started the thread to answer. Could a machine ever do what Darwin did?<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">I think the honest answer is now a qualified yes, and it is qualified in a way I did not expect. A system can hold a mountain of observation and propose the organizing idea. It can make the inductive leap I was so sure was ours alone. But watching it happen, I realized I had misjudged where Darwin&#8217;s genius actually sat. The leap to natural selection was extraordinary, but the leap was not the science. The science was in the twenty years, in Darwin knowing which single idea out of the dozens he entertained was worth a life&#8217;s defense, and in his relentless testing of it against every objection he could invent. The hunch was the cheaper half but we just could not see that while hunches were rare.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">So the machines got the hunches, and they are welcome to them. What Darwin had that they still do not is the judgment to know which hunch was worth everything, and the patience to spend years finding out if he was wrong. That part did not get automated. It got scarcer, and more valuable, than it was when the ideas were hard to come by.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">This argument runs through my book, <a href=\"https:\/\/coherise.com\/books\/coherence\/\">Coherence<\/a>, arriving this Fall. If you want to follow the thinking as it develops, join the list at <a href=\"https:\/\/coherise.com\">coherise.com<\/a>. The one-page decision tool from the book is the first thing I send.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>In November 2020 I started a thread on X. I have barely used the platform these past couple of years, and I only came back to this thread recently, by accident, when some of my own old posts surfaced in front of me. Reading them in one sitting was strange, because a thread I had [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":221,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7,10],"tags":[],"class_list":["post-220","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-the-argument","category-digressions"],"_links":{"self":[{"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/posts\/220","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/comments?post=220"}],"version-history":[{"count":1,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/posts\/220\/revisions"}],"predecessor-version":[{"id":222,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/posts\/220\/revisions\/222"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/media\/221"}],"wp:attachment":[{"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/media?parent=220"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/categories?post=220"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/coherise.com\/ideas\/wp-json\/wp\/v2\/tags?post=220"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}