Together AI · Together AI platform · LLM

deepseek-ai/DeepSeek-V4-Flash-0731

Capability changeCapability change announced Aug 13, 2026

2 notices

Capability change

Capability change announced Aug 13, 2026

Together AI documents a capability change: “​ New dedicated endpoint models The following models are now available for deployment on dedicated endpoints : deepseek-ai/DeepSeek-V4-Flash-0731.”

Hosting route
Together AI platform
Affected scope
Not stated in the notice
Announced
Aug 13, 2026
First seen by ModelClock
Oct 7, 2026

What the provider published

docs.together.ai ↗

​ New dedicated endpoint models

The following models are now available for deployment on dedicated endpoints :

  • deepseek-ai/DeepSeek-V4-Flash-0731.

Read from the provider's text; the quote is the provider's exact lines.

Revision history

Loading…
Raw history JSON
API availability

Available in the API since Aug 3, 2026

Together AI announced API availability for this model on 2026-08-03. serverless

This date applies to API availability. Rollouts in consumer products or other hosting routes can happen on different dates.

Hosting route
Together AI platform
Affected scope
serverless
Announced
Aug 3, 2026
First seen by ModelClock
Oct 7, 2026

What the provider published

docs.together.ai ↗

​ New serverless models

The following models are now available on serverless :

  • deepseek-ai/DeepSeek-V4-Flash-0731: 1,000,000 context length, FP4 quantization. Pricing: $0.14 input / $0.28 output / $0.03 cached input (per 1M tokens).

Read from the provider's text; the quote is the provider's exact lines.

Revision history

Loading…
Raw history JSON