Models

llama-3.1-8b-instant

Meta's smaller Llama 3.1 model with 8 billion parameters, available on Groq for fast, low-cost responses; runs at ~560 tokens per second with a 131,072-token context window.

Alternate definitions 1

Definition 2Models

A small, fast Llama model on Groq suited to high-volume tasks where speed and a higher request cap matter more than reasoning depth.

Models · curriculum materials

Return to all terms