Create a minimal working Groq chat completion example. Use when starting a new Groq integration, testing your setup after installing the SDK, or learning the basic Groq API request/response pattern before building something larger. Trigger with phrases like "groq hello world", "groq example", "groq quick start", "simple groq code".
68
84%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No findings from the security scan
Build a minimal chat completion with Groq's LPU inference API. Groq uses an OpenAI-compatible endpoint, so the API shape is familiar -- but responses arrive 10-50x faster than GPU-based providers. This skill gets you from an installed SDK to a working, verified request; deeper variants (streaming, Python, model selection) live in references/.
groq-sdk installed (npm install groq-sdk)GROQ_API_KEY environment variable setgroq-install-auth setupUse Write to create the example file, then run it to confirm your key and SDK work. Start with the single basic request below; reach for the reference variants only once this succeeds.
import Groq from "groq-sdk";
const groq = new Groq();
async function main() {
const completion = await groq.chat.completions.create({
model: "llama-3.3-70b-versatile",
messages: [
{ role: "system", content: "You are a helpful assistant." },
{ role: "user", content: "What is Groq's LPU and why is it fast?" },
],
});
console.log(completion.choices[0].message.content);
console.log(`Tokens: ${completion.usage?.total_tokens}`);
}
main().catch(console.error);Once Step 1 returns text, extend it with the moved-out variants:
A successful run prints the assistant's reply text followed by the total token count, e.g.:
Groq's LPU (Language Processing Unit) is a deterministic, single-core
inference chip... [assistant response continues]
Tokens: 142The underlying API returns an OpenAI-compatible ChatCompletion object: the text is at choices[0].message.content, and usage carries token counts plus four Groq-specific timing fields (queue_time, prompt_time, completion_time, total_time). Full response shape: references/models-and-response.md.
| Error | Cause | Solution |
|---|---|---|
401 Invalid API Key | Key not set or invalid | Check GROQ_API_KEY env var |
model_not_found | Typo in model ID or deprecated model | Check model list at console.groq.com/docs/models |
429 Rate limit | Free tier: 30 RPM on large models | Wait for retry-after header value |
context_length_exceeded | Prompt + max_tokens > model context | Reduce prompt size or set lower max_tokens |
GROQ_API_KEY.stream: true loop that writes tokens to stdout as they arrive.groq-local-dev-loop for development workflow setup.8585d6d
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.