US

Groq

Groq Inc., US

⭐⭐⭐⭐⭐ 4.6/5

Editorial estimate compiled from public vendor data — not a user-review average. How we rate

Visit Website →

📊 Quick Facts

Pricing ModelFreemium
Categorycoding data
UsersThousands of developers
Data Snapshot2026-09-18

📝 About Groq

Groq is revolutionizing AI inference with its custom LPU (Language Processing Unit) chip architecture that delivers deterministic, ultra-low-latency performance. Unlike GPU-based solutions, Groq's LPU is purpose-built for inference, achieving ~500 tokens/second. LPX architecture enables next-gen LPU + GPU hybrid inference. GroqCloud developer platform provides instant access to models (Llama, Mistral, Gemma). $650M raised in 2026 to scale inference cloud globally. Hundreds of megawatts of capacity under deployment.

💰 Pricing

GroqCloud free tier; pay-per-token

✨ Key Features

LPU chip: custom architecture for deterministic ultra-low-latency inferenceGroqCloud: developer platform for instant model access

⚖️ Pros & Cons

✓ Pros

  • Fastest inference speed in the industry
  • Deterministic latency — consistent performance
  • Free GroqCloud tier for developers
  • Innovative LPU+GPU hybrid with LPX

✗ Cons

  • Limited to inference only; no training
  • Smaller model selection than GPU platforms
  • Proprietary LPU creates single-vendor dependency
  • Scaling capacity still being built out

🧠 Models

  • GPT OSS 20B
  • GPT OSS 120B
  • Llama 3.3 70B
  • Llama 3.1 8B
  • Qwen 3.6 27B

ℹ️ How this page is built

This page is compiled by an automated pipeline from publicly available vendor information. It is not a hands-on review, and no editor reviews it before publication.

Ratings are editorial estimates derived from public vendor data (features, pricing, ecosystem coverage). They are not user-review averages, and we do not publish user ratings.

Data Snapshot shows when this entry was last refreshed from our directory dataset. Pricing and model line-ups change frequently, so confirm critical pricing on the vendor’s own site.

Read our directory methodology · Report an error