Groq

Sağlayıcı

Low-latency inference API for open LLMs and Whisper

Fiyat
Free tier, then paid
Erişim
API key

Hakkında

GroqCloud runs open models such as GPT-OSS, Qwen and Whisper on Groq's inference cloud behind an OpenAI-compatible API. It is for apps that need very fast token output. Its model list is short compared with larger hosts.

Neler yapabilirsin

  • Run GPT-OSS and Qwen models with very low latency
  • Transcribe audio fast with Whisper Large v3 Turbo
  • Switch from OpenAI by changing the base URL to Groq

Başlarken

  1. Create an API key in the GroqCloud console
  2. Set GROQ_API_KEY in your environment

Örnek kod

curl -X POST https://api.groq.com/openai/v1/responses \
  -H "Authorization: Bearer $GROQ_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-oss-20b",
    "input": "Explain the importance of fast language models"
  }'

Ayrıntılar

Barındırma
Hazır hizmet (bulut)
Kullanılabildiği yerler
Tüm dünya
Resmi SDK'lar
Python, JavaScript/TypeScript
MCP sunucusu
Yok

Görevler

Alternatifler

Aynı görevler için başka araçlar.

Son kontrol: .