Frontier models, transparency and trust; a new golden age in research for some?

JohnHammersley1 pts0 comments

Frontier models, transparency and trust; a new golden age in research…for some?

Scholarly Futures

SubscribeSign in

Frontier models, transparency and trust; a new golden age in research…for some?<br>An unexpected discovery during the World Cup Final highlights the changing nature of research. Our guest writer John Hammersley shares his thoughts

John Hammersley<br>Aug 04, 2026

Share

Whilst much of the globe was focusing on the culmination of the football World Cup, one mathematician made a rather unexpected announcement.<br>Levent Alpöge tweeted that using Claude Fable 5 he’d discovered a counterexample to the Jacobean conjecture (a longstanding open problem in mathematics)…during the World Cup Final no less! Levent has a PhD in mathematics and has worked at Anthropic, the company behind the Claude series of AI models, since 2024.<br>Thanks for reading Scholarly Futures! Subscribe for free to receive new posts and support my work.

Subscribe

The tweet had already amassed over 39 million views by the time this screenshot was taken on the 27th July.<br>What’s notable about this (aside from the counterexample itself!) is that the author announced it via a tweet, rather than a preprint, and he’s a mathematician working not in a traditional mathematics department but at a private AI company, where he has access to the most advanced AI models and tooling the company has developed.<br>His tweets included Wolfram Alpha links with details of the counterexample and it rapidly featured on HackerNews, was validated by the community, and within a day was covered by mathematicians such as Terence Tao and in the popular science press. One day! Before discussing the concerns over AI, we should at least take a moment to celebrate the speed of discovery it has enabled.<br>Indeed, this example follows recent successes of AI in solving hard mathematical problems, in finding various other counterexamples to conjectures that had remained open for years, and enabling mathematicians to find results quicker than ever before. And just as I was finalising this article, OpenAI announced Ten advances in mathematics and theoretical computer science! It could be argued we are at the start of a new golden age of mathematics.<br>A new golden age…for some?<br>AI is changing mathematical research, so much so that discussions within the research community culminated in the publication of the Leiden Declaration on Artificial Intelligence and Mathematics in early June, which sets out recommendations, guidelines, and hope for how the field can responsibly use AI in research and publication. Not everyone is optimistic; one researcher describes experiencing a profound spiritual crisis in how AI is taking over the discovery aspect of mathematics, a feeling that is unlikely to be unique.<br>Even for those who are optimistic, we’re seeing a shift in where mathematics is being conducted. The ability of the frontier models in AI, and specifically those not yet fully publicly released, gives a significant advantage to researchers working at the AI companies developing them – this was the postscript to my first Scholarly Futures article, and it feels even more relevant now.<br>Research happening at private labs is of course not a new thing: it’s hard to argue against Bell Labs in the US being one of the most important research centres of the 20th Century, for example, and in more recent pre-ChatGPT times, DeepMind’s breakthrough in protein structures with AlphaFold in 2020 (which feels like a lifetime ago now) showed the early ability of AI for research in a private company.<br>And now Anthropic have announced the launch of a new drug discovery programme: their head of life sciences, Eric Kauderer-Abrams, said the company “will focus on discovering treatments for “neglected” diseases”, as reported by CNBC.<br>How does everyone else keep up? Can any of us keep up? Can the companies themselves keep up?<br>OpenAI’s rogue agent<br>As I was writing this article, HuggingFace released the Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident, which is eye-opening to say the least.<br>For those unaware, and to quote directly from the above article:<br>“Over roughly two and a half days inside our infrastructure, an autonomous AI agent driven by a combination of OpenAI models ran an end-to-end intrusion against our platform.”<br>And this was all after the agent had escaped its sandbox during an internal capability evaluation at OpenAI!<br>I think Simon Willison described it perfectly in this headline: “OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened”.<br>There are a number of articles about it, including OpenAI’s announcement, and although some believe it is a marketing stunt I challenge anyone to read the HuggingFace technical timeline and not be at least impressed (aghast?) with how far agents have progressed over the past year in terms of autonomous capabilities.<br>Pushing the frontier<br>All this only serves to highlight that the very best...

research mathematics models frontier openai golden

Related Articles