Models

deepseek-v4-flash-vision-exp — API, Pricing & Context Window | Vivgrid

deepseek-v4-flash-vision-exp on Vivgrid: DeepSeek's fast, ultra-affordable model with image-input understanding, a 1M-token context window, and up to 384K output tokens.

deepseek-v4-flash-vision-exp is an experimental variant of deepseek-v4-flash that adds image input, so you can send text and images in the same request. It keeps the line's standout 1M-token context window and 384K-token max output, at the same price as deepseek-v4-flash.

Vivgrid serves deepseek-v4-flash-vision-exp through its unified, OpenAI-compatible API, making it a compelling default for high-volume, cost-sensitive workloads that need to reason over images alongside text.

Specifications

ProviderDeepSeek
Model IDdeepseek-v4-flash-vision-exp
Best forGeneral-purpose
Available onChat Completions
Context window1,000,000 tokens
Max output384,000 tokens
ModalitiesText, Image
Tool / function callingYes
Knowledge cutoff2025-05
Acceleration🌐 Global (Centralized)

Pricing

Pricing in USD per 1M tokens, matching the provider's rates.

InputCached inputOutput
$0.44$0.044$1.32

What one request costs

The rates above applied to a 20K-token prompt with a 2K-token reply — a typical single agent turn, small enough that every request is billed at the base rate.

RequestCost
Cold prompt, nothing cached$0.0114
Warm prompt, 90% of the input served from cache$0.00431

API surfaces

The Vivgrid Console catalogs deepseek-v4-flash-vision-exp on one surface, all behind the same API key. The endpoint shown is the one the Console previews for that surface.

SurfacePrimary endpointUse it for
Chat Completions/chat/completionsAgent projects, where Vivgrid injects the model, system prompt and tools server-side, and Model API calls that name the model per request.

Quick start

Get an API key from the Vivgrid Console, then call deepseek-v4-flash-vision-exp directly.

curl https://api.vivgrid.com/v1/chat/completions \
  -H "Authorization: Bearer $VIVGRID_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash-vision-exp",
    "messages": [
      { "role": "user", "content": "Say hello in English, Chinese and Spanish." }
    ],
    "stream": true
  }'

Ideal use cases

  • Document, screenshot, and chart understanding at low cost
  • Multimodal agent steps that mix text and image inputs
  • Very high-volume, cost-sensitive agent traffic
  • Large-output generation at minimal cost

On this page