Models
llama-3.1-8b-instant
Meta's smaller Llama 3.1 model with 8 billion parameters, available on Groq for fast, low-cost responses; runs at ~560 tokens per second with a 131,072-token context window.
Alternate definitions 1
Definition 2Models
A small, fast Llama model on Groq suited to high-volume tasks where speed and a higher request cap matter more than reasoning depth.
Models · curriculum materials