APISpotlight
中文
← Back to home

Platform / groq

Groq API

A high-speed inference API for open models, offering an OpenAI-compatible developer experience.

API capabilities

  • · High-speed text inference
  • · OpenAI-compatible API
  • · Open model serving

Recommended use

  • · Low-latency demos
  • · Real-time interactive prototypes
  • · Open model comparison

Important restrictions

  • · The free tier is noticeably rate limited
  • · The model list and rate limits are adjusted over time
  • · Not suitable for high-concurrency production workloads without evaluation

Free tier

  • · The official console provides a free-tier entry point; dynamic limits are not guessed at as a fixed token allowance.
  • · Free limits: Rate limits such as RPM/TPM are configured per model; see the official Rate Limits page for the specific values.
  • · Credit card: Existing records show a free sign-up path that does not require a credit card; if the policy changes, refer to the official page.
  • · Sign-up: Requires sign-up and creating an API key.

Benchmark observation

Current public measurement: status Online, latency 133 ms, success rate 100%. These values come directly from the existing platform data and are not duplicated in the content layer.

See the full benchmark leaderboard →