arXiv

Stop Worrying About AI Overfitting. Your Benchmarks Are the Real Problem.

We’ve all feared that AI is just a giant lookup table, memorizing answers without understanding. But ML research agents break this rule. They don’t overfit because they don’t live in static datasetsโ€”they explore dynamic worlds where the act of searching changes the questions. Overfitting is a flaw in the exam, not the model.

That ‘Groundbreaking’ AI Paper You Just Read? It Was Written by an AI. And That’s a Problem.

An AI-generated paper claiming a 6M-token window on a single GPU went viral on Hacker News โ€” but it was empty noise. This is the new normal: AI-written research papers that look legitimate but contain nothing. Here’s how to spot them and why it matters for the future of science.