Inference provider

Groq

Low-latency hosted inference on GroqCloud.

Models tracked2
Route-ready2
Input from$0.075
Output from$0.30

Routing contract

Modedirect
Protocolopenai
API basehttps://api.groq.com/openai/v1
Last observedAug 28, 2026
Routing groundwork only

These public route targets identify where and how a model can be requested. Model Markets does not hold provider keys or send inference traffic from this interface.

Provider records describe a serving surface, not the underlying model. Pricing, quantization, region, service tier, retention policy, and measured performance remain endpoint-specific.

Catalog coverage

Model endpoints

Provider website ↗
ModelRoute modelContextInput / 1MOutput / 1MRegionEvidence
gpt-oss-20bRoute-readyopenai/gpt-oss-20b128K$0.075$0.30Provider-reported
gpt-oss-120bRoute-readyopenai/gpt-oss-120b128K$0.15$0.60Provider-reported

Questions

Groq questions

Which models does Groq offer?+

The current source snapshot includes 2 model endpoint for Groq: gpt-oss-20b, gpt-oss-120b. Browse all models

How much does Groq cost?+

For gpt-oss-20b, the cited source reports $0.075 input and $0.30 output per million tokens.

Is Groq the fastest provider?+

Model Markets has not run approved independent probes for this delivery, so it does not make a fastest-provider claim.

How current is Groq data?+

This source snapshot was observed Aug 28, 2026 and is marked non-live until production collectors are activated.

How does Model Markets verify Groq?+

Catalog and price claims retain their source classification. Independent telemetry will use disclosed-region standardized probes and will be presented separately. Read the methodology