“We love using Groq hosting and models. They are a core part of our chat and widget agents. The models are lightning fast and the Groq team is fantastic.”
LLM APIMaker says so +1 · source ↗
Groq
Very fast inference for open models on custom LPU hardware.
Groq is an inference cloud. It runs open-weight models made by others, such as OpenAI's GPT-OSS, Qwen, Whisper and Orpheus text-to-speech, on its own LPU chips instead of GPUs, and sells access through the GroqCloud API. The point is low latency and high output speed, which matter for voice agents, live autocomplete and multi-step agent loops where every call adds wait time.
The developer model is an OpenAI-style HTTP API at api.groq.com/openai/v1. You can keep the OpenAI client libraries and change only the key and base URL, or use Groq's own Python and TypeScript SDKs. It supports Chat Completions and a Responses API, tool use with built-in browser search and code execution, remote MCP tools, structured outputs and prompt caching.
For a small team the relevant pieces are a free plan to start, a Batch API and Flex processing on the paid Developer plan, spend limits, and data controls an admin can set to zero retention. Groq runs it as a hosted service from data centers in North America, Europe, the Middle East and Australia; stored customer data sits in US Google Cloud buckets.
The limit is the catalog. You get only the models Groq has chosen to host, a few of them, such as the Llama 3 models, now listed as enterprise-only. Preview models can be withdrawn at short notice, and free-plan rate limits are low.
Where it fits
How Groq itself is built
4 tools, from its own code, website and Product Hunt page.
Who uses it
39 makers' products, each linked to the source that shows it, and 62 open-source projects that declare it in their code.
“You can select models hosted on Groq to power the AI agents you build on MindPal!”
LLM APIMaker says so · source ↗“One of the inference providers we use”
LLM APIMaker says so · source ↗“Super fast and globally stable! Sped up our workflow so much <3”
LLM APIMaker says so · source ↗“Remarkable speed and reliability.”
LLM APIMaker says so · source ↗“Fastest inference on LLaMa models”
LLM APIMaker says so · source ↗“We use Groq for ultra-fast inference when analyzing millions of contact records and enriching them with AI. It enables us to run deep research and structured reasoning at speeds that would be impossible on standard GPU setups. This level of performance lets us deliver intelligent outputs in real time, even at scale. We're grateful to the Groq team for building the kind of infrastructure that makes this possible.”
LLM APIMaker says so · source ↗“Groq Chat enhances our LPU's performance significantly, enabling faster inference and improved user interaction.”
LLM APIMaker says so · source ↗“Blazingly fast Large Language Model for code generation”
LLM APIMaker says so · source ↗“YapIt only exists because of Groq. We needed instant transcription and rapid LLM formatting to make the app feel magical. Groq's whisper model and AI inference speeds are unmatched—they make our voice-to-text pipeline feel like it has zero latency.”
LLM APIMaker says so · source ↗“Neat tool for inference”
LLM APIMaker says so · source ↗“The fastest Realtime latency is not possible without Groq inference”
LLM APIMaker says so · source ↗“Super reliable and fast so our users don't have to wait long”
LLM APIMaker says so · source ↗“Magine runs on Groq's LLM inference for parallel context optimization.”
LLM APIMaker says so · source ↗“Super low latency STT and LLM inference for agent brains”
LLM APIMaker says so · source ↗“Unparalleled speed that enables complex agentic workflows to be instant.”
LLM APIMaker says so · source ↗“It's lightning-fast! Perfect for our motto”
LLM APIMaker says so · source ↗“Built with Groq for blazing-fast inference NBot is powered by Groq’s inference infrastructure, allowing us to process large volumes of content, run real-time summarization, and deliver high-quality AI responses at low latency. As we scale AI-powered feeds, chat, and synthesis across the internet, Groq’s performance and reliability have been critical to making the experience feel fast, responsive, and production-ready. Huge thanks to the Groq team for supporting us as an early partner.”
LLM APIMaker says so · source ↗“Great models, very accurate, very fast”
LLM APIMaker says so · source ↗“Groq has a generous free tier and many models to choose from.”
LLM APIMaker says so · source ↗“Shoutout to Groq — the lightning fast inference engine running Voicr's AI under the hood. Without Groq's speed, that under 3 seconds promise wouldn't be possible.”
LLM APIMaker says so · source ↗“Super speed LLM inference for extracting product titles”
LLM APIMaker says so · source ↗Open source: a project that declares Groq as a dependency in its public code — verifiable, but not necessarily a live product.
What makers say
17 makers on why they use Groq, in their own words on Product Hunt.
Built with Groq for blazing-fast inference NBot is powered by Groq’s inference infrastructure, allowing us to process large volumes of content, run real-time summarization, and deliver high-quality AI responses at low latency. As we scale AI-powered feeds, chat, and synthesis across the internet, Groq’s performance and reliability have been critical to making the experience feel fast, responsive, and production-ready. Huge thanks to the Groq team for supporting us as an early partner.
NBotSep 2026 ↗We use Groq for ultra-fast inference when analyzing millions of contact records and enriching them with AI. It enables us to run deep research and structured reasoning at speeds that would be impossible on standard GPU setups. This level of performance lets us deliver intelligent outputs in real time, even at scale. We're grateful to the Groq team for building the kind of infrastructure that makes this possible.
graph8Sep 2026 ↗YapIt only exists because of Groq. We needed instant transcription and rapid LLM formatting to make the app feel magical. Groq's whisper model and AI inference speeds are unmatched—they make our voice-to-text pipeline feel like it has zero latency.
Yap-ItSep 2026 ↗Shoutout to Groq — the lightning fast inference engine running Voicr's AI under the hood. Without Groq's speed, that under 3 seconds promise wouldn't be possible.
Voicr for MacSep 2026 ↗We love using Groq hosting and models. They are a core part of our chat and widget agents. The models are lightning fast and the Groq team is fantastic.
VectorizeSep 2026 ↗Loved and watch-outs
Themes that recur in makers' words and Hacker News comments, each linked to what it summarises, with how Product Hunt tags its reviews.
Who switches
Public pull requests on GitHub since Oct 2024 whose title says "X to Y" — real code changes moving a project from one tool to another, by developers in general. Open a row to see the pull requests.
Gemini API → Groq152 PRs
- College MVP: switch LLM provider from Gemini to Groq (sole provider)akshat2685/pathmind · 2026-10-01
- Switch idea rating/roadmap generation from Gemini to GroqLaKhWaN/startup-game · 2026-09-29
- Switch in-app AI provider from Gemini to Groqjubayerjuhan/cognivo · 2026-09-24
- Migrate from Gemini API to Groq APIHushnudbek-s-organisation/ScholarBridgeAi · 2026-09-23
- feat(ai): switch AI diagnostic engine from Google Gemini to GroqSHOEBILL04/Garij · 2026-09-21
- feat(gateway): switch the non-technical summary from gemini to groqMicroTodoSuite/microservice-app-slack-approval-gateway · 2026-09-21
- Use Gemini as Jarvis's reply engine, fall back to Groq on failuredeepseatrader18/deepsea-dashboard · 2026-09-19
- changed the model provider from google gemini to groqRit2002/placeintel · 2026-09-19
- Switch AI Assistant from Gemini to Groqevery1hatestaha-png/busniessOS · 2026-09-16
- Switch formatting/ front & back matter conversion from Gemini to Groqlndat18/production-legal-qa-rag · 2026-09-16
OpenAI API → Groq50 PRs
- Switch eval pipeline LLM provider from OpenAI to Groqgsarthakdev/llm-eval-pipeline · 2026-09-15
- Migrate LLM backend from OpenAI GPT-4 to Groq (openai/gpt-oss-120b)Blandskron/pentest-agent · 2026-09-09
- Switch LLM provider from OpenAI to GroqkarMareal1/WhyUs · 2026-09-04
- feat(clips): switch Whisper and chat from OpenAI to Groqasgard02/vyrll · 2026-08-16
- changed from openai to groq because no way I'm paying to openai for a…ubturja/Is-It-For-Real · 2026-08-15
- Revert AI assistant LLM provider from OpenAI back to Groq (free tier)mudone1/Tdislogistics · 2026-08-10
- Swap transcription provider from OpenAI to GroqF-NAN-AI-ENGINEERING-RESIDENCY/compass · 2026-08-03
- Rename OpenAI provider to Groq providerKutlay07/ai-assistant-platform · 2026-07-31
- Task 1 done Migrate completely from OpenAI to GroqQuantumLogicsLabs/RepoMind · 2026-07-27
- Switch AI provider from OpenAI to GroqShiveshDragon/Benefit_Bridge · 2026-07-25
Claude → Groq15 PRs
- Swap festival AI content generation from Claude to Groqchaitanya-cloud08/nbt-personalised-feed · 2026-09-07
- Switch classification/topic tagging from Claude to Groqangeloremedy/remedy-pulse · 2026-09-05
- Switch AI-powered chat assistant from Claude API to Groqjlibranda/Project2 · 2026-08-29
- refactor: switch AI provider from Anthropic Claude to Groq (free tier)BorysCzapski/NewProject · 2026-07-03
- Swap AI summary provider from Claude to Groq (Llama 3.3 70B)amrit2611/AuRIS · 2026-06-29
- docs(claude): update LLM provider to Groq, add PR workflowltanafranca1004/lens · 2026-06-25
- Migrate chatbot backend from Anthropic Claude to Groq Chat CompletionsBinayakc155/NeuroDesk · 2026-05-08
- Migrate chatbot LLM integration from Claude to Groq Chat CompletionsBinayakc155/NeuroDesk · 2026-05-08
- feat: swap AI from Claude to Groq (free tier for MVP)wkliwk/FormPilot · 2026-03-28
- Refactor: Migrate Anthropic/Claude references to Groq APIajmejia02/Salud-Conecta-IA · 2026-03-27
DeepSeek → Groq9 PRs
- Update leftover DeepSeek references to Groqlgf2111/phish-report · 2026-09-18
- Switch contextual translation from DeepSeek to Groqgroz387/football-backup · 2026-09-04
- feat(chef): switch chef agent from DeepSeek to Groq gpt-oss-120bjordangaston/harvest · 2026-09-01
- Switch AI review & translation from DeepSeek to GroqRedo-San/RedoSan-Authenticity · 2026-05-24
- feat: Deepseek as peer to Groq with env-controlled primary + model taggingdgtalquantumleap-ai/ebenova-reddit-monitor · 2026-04-28
- feat: Made transition from using deepseek as LLM API provider to Groq.ristovljupcho/llm_orchestrator · 2025-11-26
- feat: add deepseek-r1-distill-llama-70b to Groqlangflow-ai/langflow · 2025-01-28
- Add DeepSeek R1 70B model to Groq extensionraycast/extensions · 2025-01-27
- feat: add deepseek-r1-distill-llama-70b to groq providerstackblitz-labs/bolt.diy · 2025-01-27
Mistral AI → Groq5 PRs
- This is the rag branch we have switched from mistral ai to groq ai fo…Adityaz23/Video-Transcriber · 2026-06-15
- feat: migrate AI Provider from Mistral to Groqshockerqt/discord-ai-bot · 2026-04-11
- feat: M6 migrate from Mistral to Groq + Ollama for LLM and embeddingsdarth-dodo/ledger · 2026-04-01
- Add Mistral Voxtral STT as alternative to Groq Whispermafaltti/caab-whatsapp-router · 2026-02-16
- Javascipt loading + prompt change + dynamic sources + changed from mistral to groq(llama3.1) + DDOS changesDrAlzahraniProjects/csusb_fall2024_cse6550_team1 · 2024-12-02
Groq → Gemini API64 PRs
- Fall back from Groq to Gemini text when the model is retired or rate-limitedcyangjr/marketplace-scout · 2026-10-01
- Switch Lumi chat from Groq to Google Gemini 2.0 FlashThabisoCollinSengane/Pulsify · 2026-09-20
- Switch deepeval CI judge from Groq to Gemini 3.1 Flash-Litetruongpx396/agent-core-demo · 2026-09-17
- fix: route Site Agent plain Groq calls to Gemini fallbackakamanim/khasroy · 2026-09-13
- Move the AI features from Groq to Geminibogdan0089/fastapi-ecommerce-backend · 2026-09-11
- Switch AI grading/recommendation backend from Groq to Geminicrabb-beltran/de-dojo · 2026-08-31
- feat(songs): switch screenshot extraction from Groq to Geminialesmo30/song-shift · 2026-08-24
- switched groq api call to gemini apiNafisaTasnimR/Keepify · 2026-08-23
- Migrate LLM backend from Groq to Geminimuizzusman/youtube-content-engine · 2026-08-21
- Switch LLM and transcription from Groq to Geminiabhijeetmishra2104/sonicscribe-app1 · 2026-08-18
Groq → OpenAI API25 PRs
- Switch SAP prompt improver from Groq to OpenAIabhishekashishde-byte/SAP-Brain · 2026-09-22
- refactor: migrate from Groq to OpenAI gpt-5.4-nano for ISL translationkarchit1128/Sanket · 2026-09-22
- Change Groq LLM models: from Llama 3.3 to OpenAI GPTpattames/vet_chatbot_v2 · 2026-09-20
- Changed AI used from groq to openaiReDILABMunich/Helferei · 2026-09-16
- Restore 2026-03-22 QuantSage build and switch AI provider from Groq to OpenAIkabwefrancis13-web/KABWE · 2026-09-15
- fix(groq): migrate deprecated llama models to openai gpt-ossDushmilan/CodeCoach-AI · 2026-08-24
- feat: migrate ForgePair agents from Groq to OpenAIbfrpaulondev/dev-agent-lab · 2026-08-21
- fix(ai): never send Groq catalog ids to OpenAIRevealUIStudio/revealui · 2026-08-18
- Migrate chatbot from Groq to OpenAIhari9255/app_service_demo · 2026-08-10
- Migrate receipt extraction from Groq to OpenAI (gpt-4.1-mini)IlyaShynkevich/grocery-buddy · 2026-08-01
Groq → Claude10 PRs
- Move all AI test generation from Groq to the Claude Code CLIarthurbohan/automation-practice-project · 2026-08-31
- Move ai:generate test generation from Groq to the Claude Code CLIarthurbohan/automation-practice-project · 2026-08-31
- Switch the LLM backend from Groq to the Claude APIPsycho-Kinesis/RASA_VOICE_ASSISTANT · 2026-08-30
- Move failure analysis from Groq to the Claude Code CLIarthurbohan/automation-practice-project · 2026-08-24
- Switch backend LLM from Groq to the Anthropic Claude API4izan/aitutor · 2026-07-27
- Migrate KI-Buddy LLM from Groq to Anthropic Claudesmokemoney81/BaitBuddy · 2026-07-21
- Migrate from Groq to Anthropic Claude with prompt cachingsei-xu/khaos · 2026-07-12
- feat: switch intel query synthesis from Groq to Claude HaikuMkultraUSA/battle_buddy · 2026-04-18
- Migrate from groq to Claude and Change the Readme.mdOshadhaVimuB/ArchionLabs · 2026-03-16
- feat: Migrate from Groq to Claude API integration with updated config…KhaledJamalKwaik/bridgeai-backend · 2026-01-30
Groq → DeepSeek7 PRs
- Switch ARIA cloud provider from Groq to DeepSeekmuhammadaryan377/RIA · 2026-09-08
- refactor(ai): switch LLM provider from Groq to DeepSeekThomas-lab17/HBntory · 2026-08-27
- fix(swarm): route agents off Groq free-tier caps to DeepSeek + Cerebraspatriotnewsactivism/codeforge-v2 · 2026-07-06
- fix(pr-prescreen): switch LLM screener from Groq to DeepSeekjeremylongshore/tons-of-skills-marketplace · 2026-06-08
- Switch LLM provider from Groq to DeepSeekmurza4ok/Political_game_powered_by_AI · 2026-03-06
- Update keys.py change groq model to deepseek r1 llama 70bpt-boop/uptodate-manga-image-translator · 2025-04-01
- feat: migrate groq to deepseekafc163/fanyi · 2025-01-20
Groq → Mistral AI3 PRs
- Switch chat backend from Groq to Mistralleerobber/AILicious · 2026-09-13
- Switch E2E opencheck from Groq to Mistral; add snapshot truncationnlqdb/nlqdb · 2026-05-18
- feat: migrate AI research from Groq to Mistrallobinuxsoft/LinuxPlayDB · 2026-03-04
Alternatives to Groq
All alternatives by situation →Questions makers ask about Groq
Can I use the OpenAI SDK with Groq?
Yes. Set the base URL to https://api.groq.com/openai/v1 and use your Groq API key. A few fields, such as logprobs, logit_bias and messages[].name, return a 400 error, and n must be 1. source ↗
Does Groq store my prompts and outputs?
Not by default for inference. It may log them for up to 30 days to troubleshoot reliability problems or investigate abuse, and batch and fine-tuning files are kept while those features need them. Any customer can turn on Zero Data Retention in Data Controls, which disables the features that need storage. source ↗
Where is my data stored?
Customer data Groq retains is kept in Google Cloud buckets in the United States. Transfers from other countries can rely on standard contractual clauses. source ↗
How strict are the free plan's limits?
Limits are per organization and measured in requests and tokens per minute and per day. On the free plan they are low, for example 30 requests per minute and 1,000 per day for the GPT-OSS models. The Developer plan raises them and adds Batch and Flex processing. source ↗
Is there a batch mode for bulk jobs?
Yes. Batch jobs cost 50% less than synchronous calls, don't count against your normal rate limits, and run within a window of 24 hours to 7 days. source ↗
Can I rely on any model in the catalog for production?
Only on the ones Groq labels as production models. Preview models are meant for evaluation and can be discontinued at short notice. source ↗
Which languages have an official SDK?
Python and JavaScript/TypeScript. For anything else, use the REST API or an OpenAI-compatible client. source ↗
Is Groq free?
Yes — there is a free tier a small product can run on; paid use starts at Pay per token. source ↗
Is Groq open source or self-hostable?
Not open source, and hosted only.
Can AI coding agents work with Groq?
No llms.txt, official MCP server or CLI found yet.
Who uses Groq?
39 makers' products we track, each with a source, and 62 open-source projects declare it in their code. source ↗