Groq pulls Llama 4 Scout, Qwen3 32B, and the Browser Automation tool
Groq removed Llama 4 Scout and Qwen3 32B from its public rate card and docs catalog, dropped the /bin/bash.08/hour Browser Automation tool, and swapped enterprise-only Minimax M2.5 for M2.7. Retained model prices held stable.
Eight priced LLMs on groq.com/pricing including Llama 4 Scout (17Bx16E) at /bin/bash.11//bin/bash.34 and Qwen3 32B at /bin/bash.29//bin/bash.59; five Compound built-in tools including Browser Automation at /bin/bash.08/hour; enterprise-only Minimax M2.5 and Qwen3-VL 32B.
Six priced LLMs (GPT OSS 20B, GPT OSS Safeguard 20B, GPT OSS 120B, Llama 3.3 70B Versatile, Llama 3.1 8B Instant, Qwen 3.6 27B); four Compound built-in tools (basic search, advanced search, visit website, code execution); one enterprise-only model, Minimax M2.7.
Between the 2026-07-14 and 2026-07-21 captures of groq.com/pricing, Groq shrank its published serverless catalog rather than repricing it. Llama 4 Scout (17Bx16E) 128k and Qwen3 32B 131k disappeared from both the pricing page rate card and the Supported Models page in the docs, where they had been listed under Production and Preview models respectively. The enterprise-only column also thinned: Qwen3-VL 32B is gone and Minimax M2.5 was replaced by a newer Minimax M2.7, still at contact-sales.
The agentic tool rate card lost a line too. Browser Automation, which had launched the previous week at /bin/bash.08 per hour, is no longer listed under Built-In Tools (Compound); that table is back to basic search at per 1,000 requests, advanced search at per 1,000, visit website at per 1,000, and code execution at /bin/bash.18 per hour. The separate GPT-OSS tool table (browser search per 1,000, visit website per 1,000, Python code execution /bin/bash.18 per hour) is unchanged.
Every retained price held: Llama 3.1 8B Instant at /bin/bash.05//bin/bash.08, GPT OSS 20B and Safeguard 20B at /bin/bash.075//bin/bash.30, GPT OSS 120B at /bin/bash.15//bin/bash.60, Llama 3.3 70B Versatile at /bin/bash.59//bin/bash.79, Qwen 3.6 27B at /bin/bash.60/.00, Whisper at /bin/bash.111 and /bin/bash.04 per hour transcribed, and Orpheus TTS at 2.00 and 0.00 per 1M characters. The change is catalog contraction, not a price move — and it is a reminder that Groq’s docs explicitly warn that Preview models “may be discontinued at short notice,” so a model you priced a workload around can leave the rate card inside a week.
One week after expanding the rate card, Groq contracted it without repricing. Llama 4 Scout (17Bx16E, $0.11/$0.34) and Qwen3 32B ($0.29/$0.59) left both the pricing page and the docs model catalog, the Browser Automation built-in tool ($0.08/hour) was pulled from Built-In Tools (Compound) seven days after launch, Qwen3-VL 32B was dropped from the enterprise-only list, and Minimax M2.5 was replaced by M2.7. Every retained price held exactly, so the cost impact lands as forced substitution rather than as a rate move — the nearest remaining Qwen, Qwen 3.6 27B at $0.60/$3.00, costs 2.1x more on input and 5.1x more on output than the delisted Qwen3 32B.