Groq
Fastest AI inference powered by custom LPU hardware, delivering ultra-low latency for text generation tasks.
Groq
Fastest AI inference powered by custom LPU hardware, delivering ultra-low latency for text generation tasks.
Authentication
API KeyAdd your Groq API key from Organization → Integrations → AI Vendors. Keys are encrypted at rest and scoped to your organization.
Available Models
27 enabled models across 2 categories.
Text23
allam-2-7bchatModel, languageModel, completionModel
groq/compoundchatModel, languageModel, completionModel
functionCalling, reasoning
groq/compound-minichatModel, languageModel, completionModel
functionCalling, reasoning
deepseek-r1-distill-llama-70bchatModel, languageModel, completionModel
functionCalling, reasoning
gemma2-9b-itchatModel, languageModel, completionModel
functionCalling
openai/gpt-oss-120bchatModel, languageModel, completionModel
functionCalling, reasoning
openai/gpt-oss-20bchatModel, languageModel, completionModel
functionCalling, reasoning
moonshotai/kimi-k2-instructchatModel, languageModel, completionModel
functionCalling
moonshotai/kimi-k2-instruct-0905chatModel, languageModel, completionModel
functionCalling
llama-3.1-8b-instantchatModel, languageModel, completionModel
functionCalling
llama-3.3-70b-versatilechatModel, languageModel, completionModel
functionCalling
llama3-70b-8192chatModel, languageModel, completionModel
functionCalling
llama3-8b-8192chatModel, languageModel, completionModel
functionCalling
meta-llama/llama-4-maverick-17b-128e-instructchatModel, languageModel, completionModel
functionCalling, imageInput, vision
meta-llama/llama-4-scout-17b-16e-instructchatModel, languageModel, completionModel
functionCalling, imageInput, vision
llama-guard-3-8bchatModel, languageModel, completionModel
meta-llama/llama-guard-4-12bchatModel, languageModel, completionModel
imageInput, vision
meta-llama/llama-prompt-guard-2-22mchatModel, languageModel, completionModel
meta-llama/llama-prompt-guard-2-86mchatModel, languageModel, completionModel
mistral-saba-24bchatModel, languageModel, completionModel
functionCalling
qwen/qwen3-32bchatModel, languageModel, completionModel
functionCalling, reasoning
qwen-qwq-32bchatModel, languageModel, completionModel
functionCalling, reasoning
openai/gpt-oss-safeguard-20bchatModel, languageModel, completionModel
functionCalling, reasoning
Sound4
canopylabs/orpheus-arabic-saudispeech
canopylabs/orpheus-v1-englishspeech
whisper-large-v3transcription
whisper-large-v3-turbotranscription