Frontier Models
DeepSeek-V4-Pro GA Delivers 1.7T Agentic Model Under MIT License With Pricing Tiers
The August 13 rollout adds Responses API compatibility and flexible reasoning controls while shifting to demand-based rates that halve during off-peak hours starting August 16.
DeepSeek-V4-Pro is a 1.7 trillion parameter frontier model released by DeepSeek under the MIT license that delivers premium agentic performance with general availability on app, web, and API.
The general availability rollout on August 13, 2026, extends the model to production users across mobile app, web dashboard, and programmatic API endpoints while keeping the model identifier deepseek-v4-pro unchanged for existing integrations.
What background led to the V4-Pro general availability announcement?
DeepSeek previously released earlier variants including DeepSeek-V4-Flash-0731 that focused on speed optimizations. The V4-Pro line targets higher capability thresholds for agentic tasks that require extended reasoning chains and tool use. Company documentation indicates the new release consolidates prior experimental features into a stable offering suitable for enterprise workloads.
The MIT license on the 1.7T parameter weights hosted at Hugging Face allows downstream modification and commercial redistribution without restrictive terms. This licensing choice aligns with broader industry trends toward open weights for large-scale models while maintaining API revenue through hosted inference.
What new capabilities arrive with the V4-Pro general availability?
The model introduces native compatibility with the OpenAI Responses API format that enables direct migration for applications already built around Codex-style completions. Developers can activate one-click setup within supported environments. Three discrete thinking effort settings allow runtime adjustment: low for routine queries, high for standard agent loops, and max for multi-step planning scenarios.
Benchmark results released alongside the announcement include HLE scores of 42.7 without tools rising to 60.0 when tools are available. Terminal Bench 2.1 reaches 87.9 under the reported evaluation protocol. These figures position the model among leading open-weight systems for agent-oriented workloads.
What technical specifications define DeepSeek-V4-Pro-0813?
The architecture supports a 1 million token context window that accommodates large codebases or extended conversation histories. Concurrency is capped at 500 simultaneous requests for the V4-Pro tier to maintain service stability. The model name for API calls remains deepseek-v4-pro following the GA transition.
| Benchmark | Score Without Tools | Score With Tools |
|---|---|---|
| HLE | 42.7 | 60.0 |
| Terminal Bench 2.1 | 87.9 | N/A |
How does the new peak and off-peak pricing structure operate?
Effective at 16:00 UTC on August 16, 2026, the API adopts a two-tier rate schedule. Off-peak periods receive a 50 percent discount relative to peak rates. The peak input price for cache misses on DeepSeek-V4-Pro-0813 stands at 1.32 dollars per million tokens according to the published pricing table.
The adjustment aims to balance compute demand across daily cycles. Users running continuous agent fleets can schedule non-urgent tasks into off-peak windows to reduce operational costs while preserving performance during business hours.
What market implications follow from the open weights and pricing changes?
Availability of the full 1.7T weights under MIT terms lowers barriers for organizations that prefer on-premises or private-cloud deployments. At the same time the hosted API with differentiated pricing provides a consumption-based alternative for teams without dedicated infrastructure.
Stakeholders in regulated sectors gain an additional option that combines high benchmark scores with transparent licensing. The concurrency limit and context window support concurrent multi-agent systems without immediate scaling constraints for mid-size deployments.
What reactions have emerged from industry observers?
We’re launching DeepSeek-V4-Pro today! 🚀 Major Agent upgrades with strong production gains! Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, max for complex tasks. Native OpenAI Responses API support, optimized for Codex with one-click setup.DeepSeek
Analysts note that the combination of open weights and API pricing tiers creates dual pathways for adoption. Enterprises can prototype on the hosted service before deciding on self-hosted implementations using the released checkpoints.
What developments are expected next for the DeepSeek model family?
Future updates may extend the thinking effort controls to additional model variants. Continued benchmark reporting will clarify performance on new agent evaluation suites. The pricing schedule will be monitored for adjustments based on observed utilization patterns after the August 16 transition.
- Review current API usage patterns to identify off-peak scheduling opportunities.
- Test Responses API integration using the one-click Codex setup path.
- Evaluate benchmark results against internal agent workloads before full migration.
- Monitor the official changelog for any post-GA parameter or pricing refinements.
Frequently asked
When does the peak and off-peak pricing take effect?
New prices apply at 16:00 UTC on August 16, 2026. Off-peak rates equal 50 percent of the corresponding peak rates for the same token category.
Does the model name change for existing API users?
The model identifier deepseek-v4-pro remains unchanged after the general availability rollout on August 13, 2026.
Sources
- DeepSeek — The GA release of DeepSeek-V4-Pro has been rolled out on the APP, Web, and API.
- Hugging Face — DeepSeek-V4-Pro-0813 is the official release with 1.7T params under MIT License and benchmarks matching GA announcement.
- DeepSeek — DeepSeek-V4-Pro-0813 pricing table with peak/off-peak rates; off-peak half of peak; context 1M; concurrency 500.
- DeepSeek — The GA release of DeepSeek-V4-Pro has been rolled out on the APP, Web, and API. ... API Pricing Adjustment ... new prices will take effect at 16:00 (UTC Time) on August 16, 2026.