DeepSeek-V4.1-Flash
DeepSeekReleased Sep 10, 2026
- providers
- 1
- context
- 1.0M
- max out
- 393.2K
- input /M
- $0.30
- output /M
- $1.20
- cached input /M
- $0.006
- reasoning /M
- $0.60
deepseek/deepseek-flashcurrentlyDeepSeek V4.1 Flashcatalogue row pending
Source:vendor documentation
About
- knowledge cutoff
- May 1, 2025
- open weights
- yes
- best for
- extractioncodingcompute
Capabilities
What the model accepts and returns, where a probe through our own gateway proved it.
- streaming
- tools
- json mode
- vision
- audio
Providers
What you pay through us for each provider, and its own list price where that differs.
DeepSeek
deepseek/deepseek-flash- input /M
- $0.30
- output /M
- $1.20
- cache read /M
- $0.006
Cache read $0.006/M
Current: peakPeak input list price /M$0.30/MPeak output list price /M$1.20/M
Off-peak prices and peak hours
Off-peak input list price /M$0.15/MOff-peak output list price /M$0.60/M
Peak Mon to Fri 01:00 to 04:00 and 06:00 to 10:00 UTC, off-peak otherwise
The provider's own published list prices for each UTC tier, not what a request is charged.
1.0M context393.2K max out
Speed
- fastest over 24 hours
- 76.4 t/s
Fastest: DeepSeek
- DeepSeek334 ms to first token76.4 t/smedian of 24
Served
- served in the last 24 hours
- 100.00%
Across 1 provider, 47 of 47 requests
- DeepSeek100.00% · 47 of 47
Rebuilt once a day from the requests we sent, so it is a daily figure and not a live one.
Not yet. These figures come from Artificial Analysis, which the publish pipeline reads once a day.
Integrations
What the model holds, rather than a compatibility grade: nothing here measures how well a particular tool gets on with it, so nothing here scores one.
- answers on
- chat completions
- accepts
- textimage
- returns
- text
Use this model
Works with the OpenAI SDK. Point it at our base URL and use this id.
curl https://api.hopscotchlabs.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek/deepseek-flash",
"messages": [{"role": "user", "content": "Hello"}],
"max_tokens": 256
}'