Recent milestones — AI systems reportedly settling long-unsolved math problems with machine-checkable proofs — raise a genuinely exciting question: is AI now doing real research? The honest answer is nuanced, and worth getting right, because both the hype and the dismissal miss what's actually happening.
The case that it's real
When a system produces a valid, machine-verified proof of a problem that resisted human researchers for a decade, something real happened. The result is correct (the checker confirms it), novel (it was open), and useful (it advances knowledge). That's a meaningful contribution by any reasonable standard — regardless of whether the system "understands" in a human sense. The proof stands on its own.
A verified solution to an open problem is a real contribution, full stop. Whether the machine "understood" it is a separate — and less important — question than whether the proof checks out.
The case for caution
But context matters. These wins concentrate where the problem can be formalized and verified — mathematics, formal logic. Much of research isn't like that: asking the right questions, designing experiments, interpreting messy data, forming theories. AI is a powerful tool within a formalizable slice, not (yet) an autonomous scientist across the board. And these systems still work best guided by humans who frame the problem and judge what matters.
The realistic picture
The truth is a spectrum. AI has become a genuine research accelerator and, in verifiable domains, a genuine contributor — settling problems, generating and checking proofs, exploring vast spaces faster than people can. It is not replacing the human scientist's role of asking what's worth solving and why. The most productive frame: AI as an extraordinarily capable collaborator that expands what researchers can attempt.
Why it matters
Getting this right avoids two errors: dismissing real contributions as "just pattern matching," and overclaiming that science is solved. AI is doing real, verifiable work in formalizable domains — a big deal — while the broader creative and judgment-laden parts of research remain human-led. That combination is likely to accelerate discovery dramatically. Watching where AI contributes, and where it doesn't, is one of the most interesting stories in the field.