Model Price Index
Updated 2026-09-15 · Changelog

DeepSeek V4 Flash 0731

Together AI

DeepSeek's V4 Flash July 2025 snapshot on Together AI in FP4, with tool calling, structured outputs, and a 1M token context window

$0.14

Input

$0.28

Output

per 1M tokens

Ratesper 1M tokens

Input$0.14
Output$0.28
Cache Read$0.03

Estimate Your Monthly Cost

A request with input and output tokens, requests / month at a % cache hit rate $12.60/mo

Compare all models at this workload →

Excludes cache writes and batch discounts.

Specifications

API Model IDdeepseek-ai/DeepSeek-V4-Flash-0731
Context Window1M tokens
Modalitiestext
Capabilitiestool-use · streaming · json-mode · reasoning · code

More Together AI Text / Chat Models

LFM2-24B-A2B · LFM2.5-8B-A1B · MiniMax M2.7 · MiniMax M3 · Qwen 2.5 7B Instruct Turbo · Qwen3 235B A22B Instruct 2507 · Qwen3.5 397B A17B · Qwen3.5 9B ·