# deepseek-v4-flash-vision-exp — API, Pricing & Context Window | Vivgrid

> Source: https://vivgrid.com/docs/models/deepseek-v4-flash-vision-exp

deepseek-v4-flash-vision-exp on Vivgrid: DeepSeek's fast, ultra-affordable model with image-input understanding, a 1M-token context window, and up to 384K output tokens.

`deepseek-v4-flash-vision-exp` is an experimental variant of `deepseek-v4-flash` that adds image input, so you can send text and images in the same request. It keeps the line's standout **1M-token context window** and **384K-token max output**, at the same price as `deepseek-v4-flash`.

Vivgrid serves `deepseek-v4-flash-vision-exp` through its unified, OpenAI-compatible API, making it a compelling default for high-volume, cost-sensitive workloads that need to reason over images alongside text.

## Specifications

| Field | Value |
| --- | --- |
| Provider | DeepSeek |
| Model ID | `deepseek-v4-flash-vision-exp` |
| Best for | General-purpose |
| Available on | Chat Completions |
| Context window | 1,000,000 tokens |
| Max output | 384,000 tokens |
| Modalities | Text, Image |
| Tool / function calling | Yes |
| Knowledge cutoff | 2025-05 |
| Acceleration | Global (centralized) |

## Pricing

Pricing in USD per 1M tokens, matching the provider's rates.

| Input | Cached input | Output |
| --- | --- | --- |
| $0.44 | $0.044 | $1.32 |

### What one request costs

The rates above applied to a 20K-token prompt with a 2K-token reply — a typical single agent turn, small enough that every request is billed at the base rate.

| Request | Cost |
| --- | --- |
| Cold prompt, nothing cached | $0.0114 |
| Warm prompt, 90% of the input served from cache | $0.00431 |


## API surfaces

The Vivgrid Console catalogs `deepseek-v4-flash-vision-exp` on one surface, all behind the same API key. The endpoint shown is the one the Console previews for that surface.

| Surface | Primary endpoint | Use it for |
| --- | --- | --- |
| Chat Completions | `/chat/completions` | Agent projects, where Vivgrid injects the model, system prompt and tools server-side, and Model API calls that name the model per request. |

## Quick start

Get an API key from the Vivgrid Console (https://console.vivgrid.com), then call `deepseek-v4-flash-vision-exp` directly.

```bash
curl https://api.vivgrid.com/v1/chat/completions \
  -H "Authorization: Bearer $VIVGRID_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash-vision-exp",
    "messages": [
      { "role": "user", "content": "Say hello in English, Chinese and Spanish." }
    ],
    "stream": true
  }'
```


## Ideal use cases

- Document, screenshot, and chart understanding at low cost
- Multimodal agent steps that mix text and image inputs
- Very high-volume, cost-sensitive agent traffic
- Large-output generation at minimal cost

## Related models

- [gpt-6-astra](/docs/models/gpt-6-astra) — OpenAI's frontier gpt-6 coding model
- [deepseek-v4-flash (0731)](/docs/models/deepseek-v4-flash) — the text-only, identically priced base model
- [deepseek-v4-pro-0813](/docs/models/deepseek-v4-pro-0813) — the latest flagship V4 release
- [deepseek-v4-pro](/docs/models/deepseek-v4-pro) — the flagship V4 model
