Mark RadarMARK RADAR
About
EN
Sign in

DeepSeek Launches Vision Model to Challenge Anthropic

1 reports · First detected 2026-08-24 · Last active 2026-08-24

Chinese artificial-intelligence startup DeepSeek is extending its challenge to Western model developers from text-based reasoning into visual understanding. Multimodal systems can interpret text alongside images and interface screenshots, a capability increasingly central to AI agents that navigate software, process documents and complete complex workflows. The move brings DeepSeek into more direct competition with premium offerings from Anthropic and other leading AI developers.

DeepSeek unveiled the experimental V4-Flash-Vision-Exp model with support for image and screenshot analysis. The company said performance improved sharply on multimodal agent benchmarks and approached Anthropic’s high-end Claude Opus 4.8, though it did not disclose scores or detailed comparisons. As of Aug. 24, 2026, the model was available to developers through DeepSeek’s API platform, allowing customers to begin testing its visual capabilities in applications.

All Coverage

1 original reports

The Backstory

The history behind this event

No historical echoes for this signal

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)