Loading
LoadingModel record / DeepSeek
Access deepseek-v4-flash from DeepSeek through one OpenAI-compatible API with live market-linked pricing and usage-based billing.
Observed reliability
99.55
Output speed
—
This model averaged 38.8% below comparable direct provider prices across the selected range. Its strongest observed window was 7:00 PM–1:00 AM in UTC.
Average saved
38.8%
Cheapest time
7:00 PM–1:00 AM
Hourly coverage
24/24
Signal quality
High
Hover a bar for savings per token and observation counts.
Source buckets are UTC; labels are converted to UTC.
Last 24 hours
#3
of 60 ranked models
Last 7 days
#3
of 60 ranked models
Last 30 days
#3
of 60 ranked models
p95 latency
68,862
Evidence strength
—
curl https://api.cheaperinference.com/v1/chat/completions \ -H "Authorization: Bearer ci_live_YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "deepseek-v4-flash", "messages": [{"role": "user", "content": "Hello!"}] }'