Model Price Index
Updated 2026-07-29 · Changelog

Inkling

Together AI

Thinking Machines' NVFP4 chat model with a 524K context window, tool calling, structured outputs, and cached-input pricing

$1.00

Input

$4.05

Output

per 1M tokens

Ratesper 1M tokens

Input$1.00
Output$4.05
Cache Read$0.17

Estimate Your Monthly Cost

A request with input and output tokens, requests / month at a % cache hit rate $141.25/mo

Compare all models at this workload →

Excludes cache writes and batch discounts.

Specifications

API Model IDthinkingmachines/Inkling
Context Window524.288K tokens
Modalitiestext
Capabilitiestool-use · streaming · json-mode

More Together AI Text / Chat Models

LFM2-24B-A2B · MiniMax M2.7 · MiniMax M3 · Qwen 2.5 7B Instruct Turbo · Qwen3 235B A22B Instruct 2507 · Qwen3.5 397B A17B · Qwen3.5 9B · Qwen3.6 Plus ·