Groq
SağlayıcıLow-latency inference API for open LLMs and Whisper
- Fiyat
- Free tier, then paid
- Erişim
- API key
Hakkında
GroqCloud runs open models such as GPT-OSS, Qwen and Whisper on Groq's inference cloud behind an OpenAI-compatible API. It is for apps that need very fast token output. Its model list is short compared with larger hosts.
Neler yapabilirsin
- Run GPT-OSS and Qwen models with very low latency
- Transcribe audio fast with Whisper Large v3 Turbo
- Switch from OpenAI by changing the base URL to Groq
Başlarken
- Create an API key in the GroqCloud console
- Set GROQ_API_KEY in your environment
Örnek kod
curl -X POST https://api.groq.com/openai/v1/responses \
-H "Authorization: Bearer $GROQ_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "openai/gpt-oss-20b",
"input": "Explain the importance of fast language models"
}'Ayrıntılar
- Barındırma
- Hazır hizmet (bulut)
- Kullanılabildiği yerler
- Tüm dünya
- Resmi SDK'lar
- Python, JavaScript/TypeScript
- MCP sunucusu
- Yok