Back to Catalog
🚀

Groq

AI & LLMs

Ultra-fast LLaMA & DeepSeek R1 inference at 800+ tokens/sec

Groq LPU Inference Engine powers open models like LLaMA 3.3 70B, DeepSeek R1, and Mixtral at world-record speed. Fully OpenAI API compatible.

Details

Auth Type
API Key (Bearer token)
Rate Limit
30 req/min (free), 14,400 req/day
Pricing
Free tier available; low-cost metered token pricing
Full Docs
Step 1: Save your provider key

This is NOT your Callio key. Enter the API key from the provider's dashboard (e.g. OpenAI/SendGrid).

API Key (Bearer token)

1. Sign up at https://console.groq.com 2. Navigate to API Keys and generate key 3. Paste key into Callio setup

Get API Credentials

Getting Started

1

Try It Instantly

Click "Try It" above to test the API in the playground

2

Add to Your Agent

Click "Add to Agent" to get your API key and integrate

Common Use Cases

Ultra-fast real-time chat
Agent tool calling
Sub-second classification
Code generation

💻 Code Examples

Get started quickly with these code examples in your favorite language

curl -X GET \
'https://www.callio.dev/api/proxy/groq/forward?target=https%3A%2F%2Fapi.groq.com%2Fopenai%2Fendpoint' \
-H 'Authorization: Bearer YOUR_CALLIO_KEY' \
-H 'Content-Type: application/json'

💡 Tip: Replace YOUR_CALLIO_KEY with your actual Callio API key from the dashboard.

Ready to integrate Groq?

Test endpoints live or generate your API key and start building in minutes

Browse More APIs