Should you use Llama-4-Maverick-17B-128E-Instruct-FP8?
Llama-4-Maverick-17B-128E-Instruct-FP8 is listed for chat workloads and supports a 256K context window.
Use it when GitHub Models's free tier is enough for evaluation, demos, or light production traffic.