How to integrate Groq using MCP
Ultra-fast AI inference (OpenAI-compatible). This guide creates an MCP server from Arthur's Groq template — tools preconfigured against https://api.groq.com/openai/v1 — and finishes by sharing the server's MCP swagger documentation.
Steps
Open the Secrets vault and create a secret named
GROQ_API_KEYholding your access token. Create a free API key at console.groq.com. Groq offers a generous free tier.


Go to the REST API Templates gallery and search for Groq.


Click Use template on the Groq card to open the server-creation dialog.


Review the server name and select
GROQ_API_KEYin the credential field — the dialog only accepts vault secrets, and it lists every tool that will be created.


Click Create server. Arthur creates the server, applies the authentication, and generates all the tools.





Open the server's Tools tab to review the generated tools.

Open the server's Connect section and click Share the MCP swagger documentation — Arthur generates the public documentation link and QR code for this server.



Confirm it worked
The server appears with 3 preconfigured tool(s): chat_completion, list_models, create_transcription. The Connect section offers the MCP swagger documentation — a shareable page with setup instructions for any AI client.
Good to know
- Run LLM inference at very high speed using Groq's purpose-built LPU hardware. Fully compatible with the OpenAI API format — swap in models like Llama 3, Mixtral, and Gemma.
- This integration requires a credential (bearer). Get one at https://console.groq.com.
- Official API documentation: https://console.groq.com/docs/openai
- Editing a tool later never changes the template — templates are starting points, and the server is fully yours after creation.
Related
Tutorial video