LIVE
Models: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:12 PMModels: —+Providers: —+Cheapest H100: $2.49/hrUpdated: 01:12 PM
Marketplace
Providers Models
F

Friendli

AGGREGATEDINFERENCE
N/A
Uptime
N/A
Rating

30-Day Uptime

100%
2026-07-212026-08-19

Inference Latency

Google: Gemma 4 31B1027ms TTFT · 33 TPS
MiniMax: MiniMax M2.5416ms TTFT · 55 TPS
Z.ai: GLM 5.2505ms TTFT · 107 TPS
Z.ai: GLM 5.1266ms TTFT · 106 TPS

Inference Models

ModelInput $/MOutput $/MTTFTTPS
Google: Gemma 4 31B$0.14$0.401027ms33
MiniMax: MiniMax M2.5$0.30$1.20416ms55
DeepSeek: DeepSeek V3.2$0.50$1.50
Z.ai: GLM 5.2$1.40$4.40505ms107
Z.ai: GLM 5.1$1.40$4.40266ms106

Community Reviews

4.5★★★★★(2 reviews)
clouduser42
★★★★★2025-06-15

Reliable service, great API documentation.

mlresearcher
★★★★2025-06-10

Good performance but support could be faster.