Trade AI inference. Lock your price.

Buy AI inference tokens at spot, or lock the price ahead with monthly forwards. Open 24/7.

Spot Market

ModelInput / MOutput / MLatencyUptime
GLM-5.3$1.40$4.40~1022ms100%
DeepSeek V4 Pro$0.66$1.98~1560ms100%
Qwen3.8 Max$2.00$6.00~78ms100%
Kimi K2.7 Code$0.95$4.00~63ms100%
GLM-5.2$1.40$4.40~677ms100%
Kimi K3$3.00$15.00~571ms100%

Forward Market

GLM-5.2 · 20260731 · daily

$3.03/ 1M OST+1.44%

Delivery Period: Aug 01, 2026 - Aug 31, 2026

OST · Output-Standardized Tokens

1.00OST=1 output token
0.32OST=1 input token
0.06OST=1 cached token
Live benchmarksUpdated 00:00:00 UTC

Watch the spread
between silicon and intelligence.

GPU hours are the input; tokens are the output. The gap between them is where AI margins live. Both indices update daily from live rate cards — published as benchmarks, not yet tradeable.

TPI · Token Price Index

110 models

Cost of intelligence

Blended $/1M tokens across frontier models

$4.01-20.06%

30D Daily Updated

  • US TPI$4.01-20.06%
  • CN TPI$1.37+44.34%
  • Coding TPI$2.26+2.80%
VIEW INDEX

GPI · GPU Price Index

40 providers

Cost of compute

On-demand $/GPU/hr across 30+ providers

$3.83/hr-0.12%

30D Daily Updated

  • H100$3.83/hr-0.12%
  • H200$4.50/hr+0.66%
  • B200$6.37/hr+0.03%
  • B300$8.89/hr+5.45%
VIEW INDEX
Supply sideIndex · Spot · Forward

Every name here prices, sells, and delivers.

The company that publishes a rate card is the company on the hook when a forward expires — quoting and delivering sit with the same entity. No scraped estimates, no stale quotes.

Listed providers

  • DeepSeek
  • Zz.ai
  • MiniMax
  • Alibaba
  • Yotta Labs