AI Models & Performance Benchmarks
One OpenAI-compatible endpoint with automatic prompt caching. Real model pricing, billed per token.
| MODEL | CONTEXT | INPUT / 1M | CACHED | OUTPUT / 1M | SPEED (TPS) | LATENCY |
|---|---|---|---|---|---|---|
DeepSeek-V4-FlashPOPULAR DeepSeek AI · FP8 MoE | 1M / 384K | $0.10 | $0.014 | $0.18 | ~257 t/s | ~0.4s |
MiniMax-M2.7 MiniMaxAI · FP8 MoE | 204K / 131K | $0.25 | $0.02 | $1.00 | ~117 t/s | ~0.3s |
Ready to connect to Fusion Gateway?
Get your unified Fusion API key in seconds with local MMK payment support.