Synthszr Charts — die großen AI-Marken im Wettkampf ums Podium
synthszr charts
deepseek

DeepSeek V4-Flash-Vision

#8 en Modèles multimodaux

DeepSeek · v4 · flash vision · 8× · vu le 03 sept. 2026

66
Momentum

DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal AI model from DeepSeek built on the DeepSeek-V4-Flash architecture and extended with visual modules. It was initially released as an API model on August 21, 2026, and subsequently made available as Open Source weights under MIT license on Hugging Face on August 31, 2026. The model achieves performance comparable to DeepSeek-V4-Flash on text tasks (agents, Reasoning, world knowledge) and demonstrates significant progress on multimodal agent Benchmarks, which according to the manufacturer approaches Anthropic's Claude Opus-4.8. It is based on a Mixture-of-Experts architecture with 284 billion total parameters and approximately 13 billion active parameters, along with a 1-Million-Token context window.

Historique du momentum
05.06.03.09.

Fonctionnalités

Key Benchmark (%)ApexBench (Pass@1): 36.5% (vs. 26.2% for text-only V4-Flash)
Context Window (Tokens)1,048,576 tokens (1M), max 384,000 tokens output
LicenseMIT License (model weights on Hugging Face)
MultimodalityText and image (JPEG, PNG, GIF, WebP) input, text output; native vision integrated into V4-Flash architecture
PlatformDeepSeek API (OpenAI- and Anthropic-compatible), Hugging Face (open weights), inference via vLLM/SGLang
Price per 1M TokensSame as V4-Flash: $0.22 input (cache miss) / $0.66 output, off-peak; images capped at 384 tokens each
Release DateAPI launch: August 21, 2026; open weights: August 31, 2026

Plus de produits dans cette catégorie: Modèles multimodaux

Preuves (8)

Subscribe free. Unsubscribe the second it sucks.

High-signal news across AI, business, UX, and tech. Every morning.