AI Devtools Radar

Groq

Inference platformsgroq.com1 published change

2 sources under watch

Change timeline

otherhigh

Groq pricing page completely redesigned to focus on company positioning and fundraise announcement rather than pricing details

Users visiting the pricing page for pricing information will no longer find the detailed pricing tables for models, tokens, and tools. Instead, they'll see marketing messaging about Groq's $650 million fundraise and company positioning. This significantly impacts developers trying to compare costs or integrate Groq's API.

Evidence
Smart, Fast, and Affordable Unmatched Price Performance Fast responses, scalable performance, and costs you can plan for. Start Building Large Language Models *Approximate number of tokens per $ AI Model Current Speed(Tokens per Second) Input Token Price(Per Million Tokens) Output Token Price(Per Million Tokens) [Detailed pricing tables for GPT OSS 20B 128k, GPT OSS Safeguard 20B, GPT OSS 120B 128k, Llama 3.3 70B Versatile 128k, Llama 3.1 8B Instant 128k, Qwen 3.6 27B 131k, and other models with specific prices]
+Announcing our $650 million fundraise to scale global inference Read more> Every customer served. Every product sold. Every commit merged. Every agent task completed. That's inference. Training creates the possibility. Inference creates the value. As AI does more, inference multiplies. And inference is becoming the bottleneck. Groq was built for this. We pioneered the LPU. Now, with LPX, it works alongside NVIDIA's next-generation GPUs to deliver unparalleled inference capability, reliably, affordably, at scale. Fast or affordable is no longer a tradeoff. AI keeps training. Now it needs a better way to work. Groq makes inference work at scale.

groq.com