provider · Groq Inc.
Groq
Groq is a hardware + inference company that runs open-weight models (Llama 3, Gemma 2, Mistral, Whisper) on custom LPU (Language Processing Unit) chips at exceptional speed — typically 500-800 tokens per second, compared to 50-100 on GPU-based inference. The speed difference is visceral: a 500-word response appears in under a second instead of five. For biology students the practical implication is fast iteration: testing a prompt, seeing the result, adjusting, and re-running takes seconds rather than minutes. Groq is especially valuable for batch-style pipelines in n8n where you need to process 50+ papers quickly: at 500 tok/s you can summarise 50 abstracts in under two minutes. Free tier available at console.groq.com with generous daily limits for the 8B models; paid tier for heavier use. Important: use your own free API key — the shared course key is restricted.
What it can help with
- fast llm inference
- high token throughput
- hardware acceleration
- speech-to-text
- batch summarization
- open-weight models
- low latency api
- real-time annotation