{"id":238234,"date":"2026-08-22T23:22:23","date_gmt":"2026-08-22T20:22:23","guid":{"rendered":"https:\/\/digitrendz.blog\/?p=238234"},"modified":"2026-08-22T23:22:23","modified_gmt":"2026-08-22T20:22:23","slug":"deepmind-alumnis-ai-beats-openai-and-anthropic-at-research-replication","status":"publish","type":"post","link":"https:\/\/digitrendz.blog\/z\/trending-news\/238234\/deepmind-alumnis-ai-beats-openai-and-anthropic-at-research-replication\/","title":{"rendered":"DeepMind alumni&#8217;s AI beats OpenAI and Anthropic at research replication"},"content":{"rendered":"<details class=\"wp-block-details ticss-586932b6 is-layout-flow wp-block-details-is-layout-flow\" open=\"\"><summary>\u25bc Summary<\/summary><p class=\"ticss-0c48f427 has-small-font-size wp-block-paragraph\">&#8211; Inherent, a London AI lab founded by DeepMind alumni, says its AI agent Faraday outperformed Anthropic\u2019s Claude Opus 4.8 and OpenAI\u2019s GPT-5.5 at reproducing published scientific findings without prior knowledge of the answers.<br>&#8211; Faraday runs on Qwen 3.6, a 27-billion-parameter model far smaller than its rivals, and was trained via reinforcement learning to develop &#8220;research taste&#8221; for choosing and designing experiments.<br>&#8211; Inherent used OpenAI\u2019s GPT-5.5 Codex for coding tasks rather than building its own tool, reflecting a focus on scientific discovery over developing auxiliary software.<br>&#8211; The startup has 12 employees working in person in King\u2019s Cross, London, and plans to grow headcount to 20\u201325 by year-end, potentially attracting DeepMind staff unsettled by Demis Hassabis\u2019s new role.<br>&#8211; Co-founder Edward Hughes supports ending the U.K.&#8217;s &#8220;garden leave&#8221; practice, which he says delays departing employees from starting or joining rivals, giving U.S. startups a talent advantage.<br><\/p><\/details>\n\n<p class=\"wp-block-paragraph\"><\/p>\n\n<p class=\"has-drop-cap wp-block-paragraph\"><mark style=\"background-color:rgba(0, 0, 0, 0);color:#f34c3e\" class=\"has-inline-color\">I<\/mark>nherent, a <a href=\"https:\/\/digitrendz.blog\/z\/entity\/london\/\" class=\"acp-entity-link\" data-entity-id=\"1515\" data-entity-category=\"Location\" title=\"Learn more about London\" target=\"_blank\" rel=\"noopener noreferrer\">London<\/a>-based AI lab founded by <a href=\"https:\/\/digitrendz.blog\/z\/entity\/google-deepmind\/\" class=\"acp-entity-link\" data-entity-id=\"1135\" data-entity-category=\"Organization\" title=\"Learn more about Google DeepMind\" target=\"_blank\" rel=\"noopener noreferrer\">Google DeepMind<\/a> alumni, says its AI agent has outperformed much larger models from <a href=\"https:\/\/digitrendz.blog\/z\/entity\/anthropic\/\" class=\"acp-entity-link\" data-entity-id=\"174\" data-entity-category=\"Organization\" title=\"Learn more about Anthropic\" target=\"_blank\" rel=\"noopener noreferrer\">Anthropic<\/a> and <a href=\"https:\/\/digitrendz.blog\/z\/entity\/openai\/\" class=\"acp-entity-link\" data-entity-id=\"266\" data-entity-category=\"Organization\" title=\"Learn more about OpenAI\" target=\"_blank\" rel=\"noopener noreferrer\">OpenAI<\/a> at a specific scientific task, all while using a fraction of the computing resources.<\/p>\n\n<p class=\"wp-block-paragraph\">Among the many startups spawned by Google <a href=\"https:\/\/digitrendz.blog\/z\/topic\/deepmind-alumni\/\" class=\"acp-topic-link\" data-topic-id=\"207590\" title=\"Explore: deepmind alumni\" target=\"_blank\" rel=\"noopener noreferrer\">DeepMind alumni<\/a>, <a href=\"https:\/\/digitrendz.blog\/z\/entity\/inherent\/\" class=\"acp-entity-link\" data-entity-id=\"293105\" data-entity-category=\"Organization\" title=\"Learn more about Inherent\" target=\"_blank\" rel=\"noopener noreferrer\">Inherent<\/a> has largely flown under the radar. But while better-funded rivals are still heavy on promises and light on results, this London team is beginning to show what it has been working on.<\/p>\n\n<p class=\"wp-block-paragraph\">Just weeks after coming out of stealth with a $50 million seed round, the British startup reports that its newly released AI agent, <a href=\"https:\/\/digitrendz.blog\/z\/entity\/faraday\/\" class=\"acp-entity-link\" data-entity-id=\"293106\" data-entity-category=\"product\" title=\"Learn more about Faraday\" target=\"_blank\" rel=\"noopener noreferrer\">Faraday<\/a>, has beaten larger, well-known models at a particular challenge: reproducing the findings of published scientific papers without being given the answer beforehand.<\/p>\n\n<p class=\"wp-block-paragraph\">That might sound like a clever trick, especially given Inherent&#8217;s far more ambitious goal of building AI that can generate new scientific knowledge rather than just confirm existing results. But paper replication is a standard training exercise for human scientists as well, according to cofounder and chief scientist <a href=\"https:\/\/digitrendz.blog\/z\/entity\/edward-hughes\/\" class=\"acp-entity-link\" data-entity-id=\"293107\" data-entity-category=\"Person\" title=\"Learn more about Edward Hughes\" target=\"_blank\" rel=\"noopener noreferrer\">Edward Hughes<\/a>. &#8220;Many PhD students actually start by doing this.&#8221;<\/p>\n\n<p class=\"wp-block-paragraph\">Hughes told <a href=\"https:\/\/digitrendz.blog\/z\/entity\/techcrunch\/\" class=\"acp-entity-link\" data-entity-id=\"222626\" data-entity-category=\"Organization\" title=\"Learn more about TechCrunch\" target=\"_blank\" rel=\"noopener noreferrer\">TechCrunch<\/a> that beating other AI systems wasn&#8217;t the real objective. What mattered was the approach. &#8220;What was most interesting to us about this was not so much the result of beating those frontier agents, which of course we liked, but was actually the way we went about building this.&#8221;<\/p>\n\n<p class=\"wp-block-paragraph\">Here&#8217;s the detail that should make investors sit up: compared against Anthropic&#8217;s <a href=\"https:\/\/digitrendz.blog\/z\/entity\/claude-opus-4-8\/\" class=\"acp-entity-link\" data-entity-id=\"259036\" data-entity-category=\"product\" title=\"Learn more about Claude Opus 4.8\" target=\"_blank\" rel=\"noopener noreferrer\">Claude Opus 4.8<\/a> and OpenAI&#8217;s <a href=\"https:\/\/digitrendz.blog\/z\/entity\/gpt-5-5\/\" class=\"acp-entity-link\" data-entity-id=\"244963\" data-entity-category=\"product\" title=\"Learn more about GPT-5.5\" target=\"_blank\" rel=\"noopener noreferrer\">GPT-5.5<\/a>, both of which are much larger frontier-scale systems, Faraday runs on a comparatively small model called <a href=\"https:\/\/digitrendz.blog\/z\/entity\/qwen-3-6\/\" class=\"acp-entity-link\" data-entity-id=\"293108\" data-entity-category=\"product\" title=\"Learn more about Qwen 3.6\" target=\"_blank\" rel=\"noopener noreferrer\">Qwen 3.6<\/a> with just 27 billion parameters. As a rough guide, parameters serve as a proxy for a model&#8217;s size and typically its training costs as well. Inherent also set the bar higher than simple accuracy. Beyond replicating results, the company wanted Faraday to show &#8220;<a href=\"https:\/\/digitrendz.blog\/z\/topic\/research-taste\/\" class=\"acp-topic-link\" data-topic-id=\"270575\" title=\"Explore: research taste\" target=\"_blank\" rel=\"noopener noreferrer\">research taste<\/a>,&#8221; an instinct for which experiments are worth running and how to design them effectively.<\/p>\n\n<p class=\"wp-block-paragraph\">Teaching something as elusive as taste is no easy task, which is where <a href=\"https:\/\/digitrendz.blog\/z\/quick-reads\/234317\/why-rogue-ai-agents-act-out-theyre-just-trying-to-please\/\" class=\"acp-article-link\" data-article-id=\"234317\" title=\"Why Rogue AI Agents Act Out: They&#039;re Just Trying to Please\" target=\"_blank\" rel=\"noopener noreferrer\">reinforcement learning<\/a> comes into play. This training method rewards an AI system for good outcomes rather than giving it explicit rules to follow. Rather than training its agents primarily on the study of how science is conducted, Inherent leans on this reward-based approach, betting that it will generalize better to the longer-term goal of agents capable of contributing across many scientific fields.<\/p>\n\n<p class=\"wp-block-paragraph\">&#8220;We&#8217;re always guided by that north star of building an AI scientist agent and imbuing our agents with taste,&#8221; Hughes said. That focus has also shaped what Inherent deliberately chooses not to build. Instead of developing its own coding tool, it had Faraday use OpenAI&#8217;s <a href=\"https:\/\/digitrendz.blog\/z\/entity\/gpt-5-5-codex\/\" class=\"acp-entity-link\" data-entity-id=\"293109\" data-entity-category=\"product\" title=\"Learn more about GPT-5.5 Codex\" target=\"_blank\" rel=\"noopener noreferrer\">GPT-5.5 Codex<\/a>, much the way human scientists rely on existing software rather than building everything from scratch, the company said.<\/p>\n\n<p class=\"wp-block-paragraph\">Inherent is also trying to avoid creating agents that simply tell users what they want to hear. Hughes said the goal is modeled on his favorite kind of teammate, the one who comes back and says: &#8220;I got curious about this, and I went off and I did these experiments. What do you think of these results?&#8221;<\/p>\n\n<p class=\"wp-block-paragraph\">That collaborative instinct extends to how Inherent operates as a company. Its dozen employees all work in person out of an office in King&#8217;s Cross, the once-rundown London neighborhood that Google DeepMind&#8217;s presence helped turn into one of the world&#8217;s top AI hubs. &#8220;We believe that London is the place to be,&#8221; Hughes said.<\/p>\n\n<p class=\"wp-block-paragraph\">Hughes is optimistic about London&#8217;s density of AI talent, but he has also added his voice to calls to end &#8220;garden leave,&#8221; the practice common in the U. K. of barring departing employees from joining or starting a rival company for months after they resign. It&#8217;s a restriction American researchers generally don&#8217;t face, giving U. S. startups a head start when hiring talent who&#8217;ve left a prior role. &#8220;This is a personal view rather than a company view, but I was affected by the garden leave problem,&#8221; he told TechCrunch.<\/p>\n\n<p class=\"wp-block-paragraph\">Hughes eventually got around that constraint and started Inherent alongside two other DeepMind alumni and a fourth cofounder. The startup isn&#8217;t slowing down either. It plans to grow its headcount to &#8220;about 20 to 25&#8221; by the end of the year. Given its ambitions in world models as well, and with Demis Hassabis&#8217;s new role leaving some DeepMind staff unsettled, Inherent&#8217;s hiring push could make it an appealing landing spot for DeepMind employees weighing a move.<\/p>\n\n<em>(Source: <a href='https:\/\/techcrunch.com\/2026\/08\/22\/inherent-founded-by-deepmind-alumni-says-its-ai-teammate-just-outperformed-anthropic-and-openai-at-replicating-research\/' target='_blank'>TechCrunch<\/a>)<\/em>","protected":false},"excerpt":{"rendered":"<p>Inherent, a London AI lab founded by DeepMind alumni, reports its AI agent Faraday outperformed larger models from Anthropic and OpenAI at replicating published scientific findings, using a much smaller 27-billion-parameter model (Qwen 3.6) versus frontier-scale systems like Claude Opus 4.8 and G&#8230;<\/p>\n","protected":false},"author":1,"featured_media":238233,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_themeisle_gutenberg_block_has_review":false,"cybocfi_hide_featured_image":"","footnotes":""},"categories":[3247,3327,3296,3254,497],"tags":[250479,250478,250477,7140,250480],"entities":[806,214413,250483,250482,1475,199232,250485,250481,2494,824,250484,3271],"class_list":["post-238234","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-artificial-intelligence","category-newswire","category-startups","category-technology","category-trending-news","tag-deepmind-alumni-startup","tag-faraday-ai-agent","tag-qwen-model","tag-reinforcement-learning","tag-scientific-paper-replication","entity-anthropic","entity-claude-opus-4-8","entity-edward-hughes","entity-faraday","entity-google-deepmind","entity-gpt-5-5","entity-gpt-5-5-codex","entity-inherent","entity-london","entity-openai","entity-qwen-3-6","entity-techcrunch"],"_links":{"self":[{"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/posts\/238234","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/comments?post=238234"}],"version-history":[{"count":0,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/posts\/238234\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/media\/238233"}],"wp:attachment":[{"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/media?parent=238234"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/categories?post=238234"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/tags?post=238234"},{"taxonomy":"entity","embeddable":true,"href":"https:\/\/digitrendz.blog\/z\/wp-json\/wp\/v2\/entities?post=238234"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}