Multimodal AI

DeepSeek’s Vision Model Exposes the Dirty Secret of AI: AGI Is a Distraction

DeepSeek’s founder said they’d go text-only to achieve AGI. Then they shipped a vision model. This isn’t hypocrisy—it’s the industry’s dirty secret: even the most principled AI labs are forced by market demand to build practical multimodal tools. The real race isn’t about abstract intelligence; it’s about reading screenshots reliably.