A load-bearing term with nothing under it
Artificial general intelligence appears in company mission statements, in contracts, in national policy documents, and in arguments about whether the current trajectory is dangerous. It is doing a great deal of work.
And until recently there was no operational definition anyone agreed on. Ask ten researchers and you get answers ranging from a system that can do any task a human can, to a system that can do most economically valuable work, to a system that can improve itself, to a system that would pass as human in unrestricted conversation.
Those are not variations on one idea. They pick out different systems, arrive at different times, and imply different policies.
The practical consequence is that claims about AGI cannot be checked. If someone says a system is close to AGI, or that AGI is decades away, there is no measurement that would settle it, so the disagreement is unfalsifiable in both directions.
This is the gap a paper published in October 2025, with Dan Hendrycks, Dawn Song, Christian Szegedy, Yarin Gal, Erik Brynjolfsson and more than thirty other authors, sets out to close.

