LLM APIS / PRODUCT PROFILE
Official website Groq Cloud
LPU Inference Engine delivering unmatched token output speeds.
Start with your workload. These examples use the category defaults. Customize this comparison
Supported plans
Official pricing$10.43/ month equivalent
Base $0 / month equivalent. Metered charges and allowances are included at the default workload.
Conditions and limitations
- Industry-leading 250+ tokens per second inference speeds on dedicated LPU hardware.
- Context prompt caching discounts, vision inputs on this model.
Pricing checked 2026-10-06. Inspect source
Add this plan to a stackEVIDENCE, NOT GUESSWORK
1 recent recordsPricing observations
These dates mark when a price model was checked. They do not establish when a vendor changed its price. One observation is not a historical trend.
Llama 3.3 70B (Groq) · Base $0 / month equivalent · Monthly billing
Official source ↗ · Usage charges are additional where modeled.
HEAD-TO-HEAD COMPARISONS
Compare Groq Cloud with Competitors
Groq Cloud vs OpenAI APIShowdown
Industry-standard foundation models with automatic prompt caching and vision.
Compare head-to-head ↗Groq Cloud vs Anthropic ClaudeShowdown
Leading coding and analytical models with 90% prompt cache read discounts.
Compare head-to-head ↗Groq Cloud vs DeepSeek APIShowdown
Ultra cost-effective open-weights models with deep reasoning and context caching.
Compare head-to-head ↗