Source: Marcus on AI
OpenAI's Astra multimodal model performs real tasks with video input and real-time reasoning, but the company's marketing conflates narrow demonstration wins with genuine AGI progress. Showing a model handle a specific, curated interaction—like reading code from a screen—gets presented as evidence of human-level reasoning, when it's pattern matching against training data without understanding underlying principles. The gap between what Astra can demonstrate in controlled conditions and what it can reliably do in the wild matters because it shapes how enterprises allocate billions in AI infrastructure spend. This pattern inflates investor expectations while obscuring what the system actually does and cannot do.