AI Models & Performance Benchmarks

One OpenAI-compatible endpoint with automatic prompt caching. Real model pricing, billed per token.

MODELCONTEXTINPUT / 1MCACHEDOUTPUT / 1MSPEED (TPS)LATENCY
DeepSeek-V4-FlashPOPULAR
DeepSeek AI · FP8 MoE
1M / 384K$0.10$0.014$0.18~257 t/s~0.4s
MiniMax-M2.7
MiniMaxAI · FP8 MoE
204K / 131K$0.25$0.02$1.00~117 t/s~0.3s

Ready to connect to Fusion Gateway?

Get your unified Fusion API key in seconds with local MMK payment support.

View Subscription Plans