Story: state of ai/leading models
Context: The Verge, reporting the same day, noted the claimed margins over GPT-4 were 'mostly very close' — the 30-of-32 tally is technically Google's own benchmark table, and the practical gap it implied did not match the headline framing.