DeepSeek-V4 Pro now available on Together AI
Quick Answer
DeepSeek-V4 Pro is now available on Together AI, offering 1.6T-parameter MoE with 512K context for serverless inference.
Quick Take
It supports three reasoning modes and reduces costs for repeated long-context queries, with input pricing at $2.10 per million tokens.
Key Points
- DeepSeek-V4 Pro features a 1.6T-parameter MoE architecture.
- Offers three reasoning modes: Non-Think, Think High, and Think Max.
- Input pricing is $2.10 per million tokens, with cached input at $0.20.
- Supports workloads like code agents, document intelligence, and research synthesis.
- Cached input pricing reduces costs by 90% for reused contexts.
Source Excerpt
DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence, and research synthesis.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from Together AI
See more →
Open, convenient and predictable: Introducing Provisioned Throughput
Together AI introduces Provisioned Throughput, offering guaranteed inference capacity for MiniMax M3 and GLM-5.2 at $0.05 per PTU per minute, achieving costs up to 90% lower than Claude Opus 4.8. This new model provides predictable pricing and a 99% uptime SLA, catering to companies transitioning to open weight models for production workloads.

