Llama serves 2 of the models we track. We do not support bring-your-own-key for this provider yet.
| Model | Made by | Input / 1M | Output / 1M | Context |
|---|---|---|---|---|
| Llama 4 Maverick 17B 128E Instruct FP8 | Azure | $0 | $0 | 128K |
| Llama 3.3 70b Instruct | NVIDIA | $0 | $0 | 128K |
Llama serves 2 of the models we track, across 1 model type.
Llama 4 Maverick 17B 128E Instruct FP8 at $0 per million input tokens.
Not yet — Llama is not currently one of the providers Serenities supports for bring-your-own-key.