Llama 5 Maverick
NEWCHEAPFASTMeta · meta/llama-5-maverick
Open-weight 600B+ MoE flagship with 2M context — longest of any frontier model. Only ~40-80B params activate per token, making inference ultra-efficient.
Input / 1M tokens
৳18
Output / 1M tokens
৳72
Context length
2M
Max output
64K
Placeholder rates. Final pricing is published per model before launch.
Suitable for
- Fast
- Low Cost
- Coding
- Vision
- Long context
Example request
curl https://padmarouter.com/v1/chat/completions \
-H "Authorization: Bearer $PADMA_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta/llama-5-maverick",
"messages": [{"role": "user", "content": "Hello from Dhaka!"}]
}'Performance
Throughput145 tok/s
Time to first token0.5s
StreamingSupported
Tool callingSupported
Measured from Dhaka over the last 24h of prototype traffic.
Routing
Primary routeMeta direct
FallbackSecondary region
Prompt retentionNot retained