LIVE
Models: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:11 PMModels: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:11 PM
Marketplace
Providers Models
O

OpenInference

AGGREGATEDINFERENCE
N/A
Uptime
N/A
Rating

30-Day Uptime

100%
2026-07-212026-08-19

Inference Latency

OpenAI: gpt-oss-120b (free)991ms TTFT · 15 TPS
Google: Gemma 4 31B1265ms TTFT · 13 TPS
DeepSeek: DeepSeek V4 Flash 07314055ms TTFT · 9 TPS

Inference Models

ModelInput $/MOutput $/MTTFTTPS
OpenAI: gpt-oss-120b (free)$0.00$0.00991ms15
MiniMax: MiniMax M2.5 (free)$0.00$0.00
Google: Gemma 4 31B$0.08$0.351265ms13
DeepSeek: DeepSeek V4 Flash 0731$0.08$0.184055ms9

Community Reviews

4.5★★★★★(2 reviews)
clouduser42
★★★★★2025-06-15

Reliable service, great API documentation.

mlresearcher
★★★★2025-06-10

Good performance but support could be faster.