← Back to directory
Groq logo
AI APIs & TokensVERIFIED

Groq

Ultra-fast LLM inference (Llama 3.3, DeepSeek-R1) reaching 500+ tokens/second via custom LPU chips.

llminferencelpuapi

Pricing plans

From $0.00
Pay-as-you-go
Free
  • Qwen 3.6 27B
  • Llama 3.3 70B & Whisper Large V3
  • OpenAI-compatible API

Active coupons

This tool has no active coupons right now.

What's included

Language model inference API with market-leading speed. Custom LPU chips deliver token generation up to 10x faster than GPUs, with a generous free tier for testing and instant integration via an OpenAI-compatible REST API.