LIVE
Models: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:10 PMModels: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:10 PM
Marketplace
Providers Models
B

BaseTen

AGGREGATEDINFERENCE
N/A
Uptime
N/A
Rating

30-Day Uptime

100%
2026-07-212026-08-19

Inference Latency

OpenAI: gpt-oss-120b229ms TTFT · 195 TPS
DeepSeek: DeepSeek V4 Flash 0731547ms TTFT · 59 TPS
NVIDIA: Nemotron 3 Ultra602ms TTFT · 84 TPS
MoonshotAI: Kimi K2.6825ms TTFT · 18 TPS
Thinking Machines: Inkling195ms TTFT · 98 TPS
DeepSeek: DeepSeek V4 Pro 0813939ms TTFT · 98 TPS
Z.ai: GLM 5.21893ms TTFT · 45 TPS
DeepSeek: DeepSeek V4 Pro 0423404ms TTFT · 57 TPS
Z.ai: GLM 5.21827ms TTFT · 49 TPS
MoonshotAI: Kimi K33017ms TTFT · 41 TPS

Inference Models

ModelInput $/MOutput $/MTTFTTPS
OpenAI: gpt-oss-120b$0.10$0.50229ms195
DeepSeek: DeepSeek V4 Flash 0731$0.13$0.26547ms59
NVIDIA: Nemotron 3 Ultra$0.60$2.40602ms84
MoonshotAI: Kimi K2.6$0.95$4.00825ms18
Thinking Machines: Inkling$1.00$4.05195ms98
DeepSeek: DeepSeek V4 Pro 0813$1.32$3.96939ms98
Z.ai: GLM 5.2$1.40$4.401893ms45
DeepSeek: DeepSeek V4 Pro 0423$1.74$3.48404ms57
Z.ai: GLM 5.2$2.10$6.601827ms49
MoonshotAI: Kimi K3$3.00$15.003017ms41

Community Reviews

4.5★★★★★(2 reviews)
clouduser42
★★★★★2025-06-15

Reliable service, great API documentation.

mlresearcher
★★★★2025-06-10

Good performance but support could be faster.