Trade AI inference. Lock your price.

Buy AI inference tokens at spot, or lock the price ahead with monthly forwards. Open 24/7.

Spot Market

ModelInput / MOutput / MLatencyUptime
DeepSeek V4 Flash$0.20$0.40~918ms100%
Qwen3.8 Max$2.00$6.00~50ms100%
Kimi K2.7 Code$0.95$4.00~67ms100%
GLM-5.2$1.40$4.40~576ms100%
Kimi K3$3.00$15.00~456ms100%

Forward Market

GLM-5.2 · 20260731 · daily

$3.03/ 1M OST+1.44%

Delivery Period: Aug 01, 2026 - Aug 31, 2026

OST · Output-Standardized Tokens

1.00OST=1 output token
0.32OST=1 input token
0.06OST=1 cached token
Live benchmarksUpdated 00:00:00 UTC

Watch the spread
between silicon and intelligence.

GPU hours are the input; tokens are the output. The gap between them is where AI margins live. Both indices update daily from live rate cards — published as benchmarks, not yet tradeable.

TPI · Token Price Index

87 models

Cost of intelligence

Blended $/1M tokens across frontier models

$5.42-46.25%

30D Daily Updated

  • US TPI$5.42-46.25%
  • CN TPI$1.11+11.15%
  • Coding TPI$2.67-30.65%
VIEW INDEX

GPI · GPU Price Index

40 providers

Cost of compute

On-demand $/GPU/hr across 30+ providers

$3.86/hr+4.56%

30D Daily Updated

  • H100$3.86/hr+4.56%
  • H200$4.54/hr+4.17%
  • B200$6.44/hr-1.21%
  • B300$8.51/hr+1.10%
VIEW INDEX
Supply sideIndex · Spot · Forward

Every name here prices, sells, and delivers.

The company that publishes a rate card is the company on the hook when a forward expires — quoting and delivering sit with the same entity. No scraped estimates, no stale quotes.

Listed providers

  • DeepSeek
  • Zz.ai
  • MiniMax
  • Alibaba
  • Yotta Labs