DeepSeek V4 Flash Vision Exp
DeepSeek · text, image → text
DeepSeek’s officially released new multimodal visual-understanding model, DeepSeek‑V4‑Flash‑Vision‑Exp, is experimental in nature and supports multimodal inputs. In pure-text capabilities (agents, reasoning, world knowledge, etc.), DeepSeek‑V4‑Flash‑Vision‑Exp is on par with the official DeepSeek‑V4‑Flash release.(The model name deepseek-v4-flash-vision-exp remains callable, but the official corresponding model has been taken offline; requests will be served by the DeepSeek-V4.1-Flash model and billed at Flash pricing.)
Input$0.155 /M
Output$0.62 /M
Cache read$0.0031 /M