Will the Next 'GPT Moment' Be Visual? The Harvard, Stanford and Google Study Challenging LLMs
A white paper bringing together researchers from Google, Stanford, Harvard, and Princeton proposes something that caught me off guard: maybe AGI won't come from text, but from video models like Veo 3. I explain why that hypothesis isn't as crazy as it sounds.