<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Groq — The AI Toolchain</title>
    <link>https://aitoolchain.io/tools/groq</link>
    <description>New releases and features in Groq, tracked by The AI Toolchain.</description>
    <language>en</language>
    <lastBuildDate>Mon, 31 Aug 2026 08:31:51 GMT</lastBuildDate>
    <atom:link href="https://aitoolchain.io/tools/groq/rss.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Groq snapshot-20260831</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Mon, 31 Aug 2026 08:31:51 GMT</pubDate>
      <description>• Both models support structured outputs and multilingual reasoning, with the 20B variant scoring 98.7% on AIME 2025 (math with tools) and 60.7% on SWE-Bench Verified, and the 120B variant scoring 90.0% MMLU and 62.4% SWE-Bench Verified.</description>
    </item>
    <item>
      <title>Groq snapshot-20260830</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Sun, 30 Aug 2026 08:32:03 GMT</pubDate>
      <description>Groq adds OpenAI GPT-OSS 20B and 120B MoE models with reasoning, browser search, and code execution at 1000+ TPS.
• Adds `openai/gpt-oss-20b` model to the `POST https://api.groq.com/openai/v1/chat/completions` endpoint — a 20B MoE model with 131K token context, 32K max output tokens, built-in browser search and code execution, structured output support, and ~1000+ TPS throughput.
• Adds `openai/gpt-oss-120b` model to the `POST https://api.groq.com/openai/v1/chat/completions` endpoint — a 120B MoE model with 131K token context, 32K max output tokens, built-in browser search and code execution, structured output support, and ~500+ TPS throughput.
• Both models support reasoning capabilities and structured outputs, with 32 and 128 experts respectively in their MoE architectures.</description>
    </item>
    <item>
      <title>Groq snapshot-20260829</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Sat, 29 Aug 2026 13:56:28 GMT</pubDate>
      <description>• Adds `openai/gpt-oss-20b` model via `POST https://api.groq.com/openai/v1/chat/completions`: a 20B MoE model running at ~1000+ TPS with a 131K token context window, 32K max output tokens, built-in browser search, code execution, and structured output support.
• Adds `openai/gpt-oss-120b` model via `POST https://api.groq.com/openai/v1/chat/completions`: a 120B MoE model (128 experts) running at ~500+ TPS with the same 131K context window, 32K max output tokens, built-in browser search, code execution, and structured output support.
• Both models carry built-in reasoning capabilities that surpass OpenAI o4-mini on several benchmarks, including 98.7% AIME 2025 (20B) and 90.0% MMLU (120B).</description>
    </item>
    <item>
      <title>Groq snapshot-20260828</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Fri, 28 Aug 2026 13:11:15 GMT</pubDate>
      <description>• Adds `openai/gpt-oss-20b` model to the `POST /openai/v1/chat/completions` endpoint — a 20B MoE model running at ~1000+ TPS with a 131K token context window and 32K max output tokens.
• Adds `openai/gpt-oss-120b` model to the `POST /openai/v1/chat/completions` endpoint — a 120B MoE model running at ~500+ TPS with the same 131K context window and 32K max output tokens.
• Both models support structured outputs, built-in browser search, and built-in code execution as native capabilities.</description>
    </item>
    <item>
      <title>Groq snapshot-20260826</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Wed, 26 Aug 2026 09:56:43 GMT</pubDate>
      <description>Groq adds OpenAI GPT-OSS 20B and 120B MoE models with reasoning, browser search, and code execution via API.
• Adds `openai/gpt-oss-20b` model: 131K token context, 32K max output, ~1000+ TPS, 32-expert MoE, with built-in browser search, code execution, and structured output support.
• Adds `openai/gpt-oss-120b` model: 131K token context, 32K max output, ~500+ TPS, 128-expert MoE, with built-in browser search, code execution, and structured output support.
• Both models are available via the existing `POST https://api.groq.com/openai/v1/chat/completions` endpoint using the `model` field.
• GPT-OSS 20B achieves 60.7% SWE-Bench Verified (coding), 98.7% AIME 2025 (math with tools), and 85.3% MMLU; GPT-OSS 120B achieves 90.0% MMLU, 62.4% SWE-Bench Verified, and 57.6% HealthBench Realistic.</description>
    </item>
    <item>
      <title>Groq snapshot-20260825</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Tue, 25 Aug 2026 13:23:34 GMT</pubDate>
      <description>• Adds `openai/gpt-oss-20b` model endpoint: a 20B MoE reasoning model with 131K token context, 32K max output tokens, built-in browser search and code execution, structured outputs support, and ~1000+ TPS throughput.
• Adds `openai/gpt-oss-120b` model endpoint: a 120B MoE reasoning model with 131K token context, 32K max output tokens, built-in browser search and code execution, structured outputs support, and ~500+ TPS throughput.
• Both models support structured outputs, enabling schema-constrained JSON responses via the Groq `POST /openai/v1/chat/completions` API.</description>
    </item>
    <item>
      <title>Groq snapshot-20260824</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Mon, 24 Aug 2026 14:50:05 GMT</pubDate>
      <description>• Adds `openai/gpt-oss-20b` model via `POST https://api.groq.com/openai/v1/chat/completions`: a 20B MoE reasoning model with 131K context window, 32K max output tokens, ~1000+ TPS, built-in browser search, code execution, and structured output support.
• Adds `openai/gpt-oss-120b` model via `POST https://api.groq.com/openai/v1/chat/completions`: a 120B MoE reasoning model with 131K context window, 32K max output tokens, ~500+ TPS, built-in browser search, code execution, and structured output support.</description>
    </item>
    <item>
      <title>Groq snapshot-20260823</title>
      <link>https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog#openai-gptoss-20b--openai-gptoss-120b</guid>
      <pubDate>Sun, 23 Aug 2026 09:42:02 GMT</pubDate>
      <description>Groq adds OpenAI GPT-OSS 20B and 120B MoE models with built-in browser search, code execution, and 131K context at 1000+ TPS.
• Adds `openai/gpt-oss-20b` model to the `POST /openai/v1/chat/completions` endpoint: 20B MoE with 32 experts, 131K token context, 32K max output tokens, ~1000+ TPS, structured outputs, built-in browser search and code execution.
• Adds `openai/gpt-oss-120b` model to the `POST /openai/v1/chat/completions` endpoint: 120B MoE with 128 experts, 131K token context, 32K max output tokens, ~500+ TPS, structured outputs, built-in browser search and code execution.</description>
    </item>
    <item>
      <title>Groq snapshot-20260820</title>
      <link>https://console.groq.com/docs/changelog</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog</guid>
      <pubDate>Thu, 20 Aug 2026 09:49:56 GMT</pubDate>
      <description>Groq adds MCP Connectors for Google Workspace, GPT-OSS-Safeguard 20B, Remote MCP Beta, new Orpheus voices, and Enterprise vision models.
• Adds MCP Connectors (Beta) with pre-built Google Workspace integrations accessible via `connector_id` field (`connector_gmail`, and equivalents for Calendar and Drive) in the `POST https://api.groq.com/openai/v1/responses` payload, exposing Gmail tools `get_profile`, `search_emails`, `get_recent_emails`, `read_email`; Google Calendar tools `get_profile`, `search`, `search_events`, `read_event`; and Google Drive tools `get_profile`, `search`, `recent_documents`, `fetch` — with OAuth 2.0 auth and zero custom MCP server setup required.
• Adds Remote Model Context Protocol (MCP) server integration (Beta) on GroqCloud, compatible with the OpenAI Responses API and OpenAI remote MCP specification, supporting models including `openai/gpt-oss-20b`, `openai/gpt-oss-120b`, `moonshotai/kimi-k2-instruct-0905`, `qwen/qwen3-32b`, `meta-llama/llama-4-maverick-17b-128e-instruct`, `meta-llama/llama-4-scout-17b-16e-instruct`, `llama-3.3-70b-versatile`, and `llama-3.1-8b-instant`.
• Adds `openai/gpt-oss-safeguard-20b` — OpenAI&apos;s 20B open-weight safety classification model with a 131K token context window, 65K max output tokens, ~1000 TPS, prompt caching (50% cost savings at $0.037/M cached vs $0.075/M uncached), `Harmony` response format for structured reasoning with `low`/`medium`/`high` effort, and support for tool use, browser search, code execution, JSON Object/Schema modes, and content moderation.
• Enables automatic prompt caching for `openai/gpt-oss-120b` with 50% cost savings on cached input tokens ($0.075/M cached vs $0.15/M uncached), lower latency, and cached tokens excluded from rate limit accounting — zero setup required.
• Enables automatic prompt caching for `openai/gpt-oss-20b` with 50% cost savings on cached input tokens ($0.037/M cached vs $0.075/M uncached) and automatic prefix matching — zero setup required.
• Adds Enterprise models `minimaxai/minimax-m2.5` (MiniMax general-purpose) and `qwen/qwen3-vl-32b-instruct` (vision-language multimodal) to GroqCloud for Enterprise customers.
• Adds two new voices (`Abdullah` — now the default, and `Aisha`) to `canopylabs/orpheus-arabic-saudi`, bringing the total to six supported voices: Abdullah, Fahad, Sultan, Lulwa, Noura, and Aisha.
• Migrates platform-wide TTS to Orpheus models (`canopylabs/orpheus-v1-english` with voices autumn, diana, hannah, austin, daniel, troy; `canopylabs/orpheus-arabic-saudi` with voices fahad, sultan, lulwa, noura), replacing the deprecated `playai-tts` and `playai-tts-arabic`.
• Python SDK v1.1.0 adds support for binary request streaming and a custom JSON encoder for extended type support.</description>
    </item>
    <item>
      <title>Groq snapshot-20260820</title>
      <link>https://console.groq.com/docs/changelog</link>
      <guid isPermaLink="true">https://console.groq.com/docs/changelog</guid>
      <pubDate>Thu, 20 Aug 2026 09:49:56 GMT</pubDate>
      <description>Groq adds MCP Connectors for Google Workspace, GPT-OSS-Safeguard 20B, Remote MCP Beta, new Orpheus voices, and Enterprise vision models.
• Adds MCP Connectors (Beta) with pre-built Google Workspace integrations accessible via `connector_id` field (`connector_gmail`, and equivalents for Calendar and Drive) in the `POST https://api.groq.com/openai/v1/responses` payload, exposing Gmail tools `get_profile`, `search_emails`, `get_recent_emails`, `read_email`; Google Calendar tools `get_profile`, `search`, `search_events`, `read_event`; and Google Drive tools `get_profile`, `search`, `recent_documents`, `fetch` — with OAuth 2.0 auth and zero custom MCP server setup required.
• Adds Remote Model Context Protocol (MCP) server integration (Beta) on GroqCloud, compatible with the OpenAI Responses API and OpenAI remote MCP specification, supporting models including `openai/gpt-oss-20b`, `openai/gpt-oss-120b`, `moonshotai/kimi-k2-instruct-0905`, `qwen/qwen3-32b`, `meta-llama/llama-4-maverick-17b-128e-instruct`, `meta-llama/llama-4-scout-17b-16e-instruct`, `llama-3.3-70b-versatile`, and `llama-3.1-8b-instant`.
• Adds `openai/gpt-oss-safeguard-20b` — OpenAI&apos;s 20B open-weight safety classification model with a 131K token context window, 65K max output tokens, ~1000 TPS, prompt caching (50% cost savings at $0.037/M cached vs $0.075/M uncached), `Harmony` response format for structured reasoning with `low`/`medium`/`high` effort, and support for tool use, browser search, code execution, JSON Object/Schema modes, and content moderation.
• Enables automatic prompt caching for `openai/gpt-oss-120b` with 50% cost savings on cached input tokens ($0.075/M cached vs $0.15/M uncached), lower latency, and cached tokens excluded from rate limit accounting — zero setup required.
• Enables automatic prompt caching for `openai/gpt-oss-20b` with 50% cost savings on cached input tokens ($0.037/M cached vs $0.075/M uncached) and automatic prefix matching — zero setup required.
• Adds Enterprise models `minimaxai/minimax-m2.5` (MiniMax general-purpose) and `qwen/qwen3-vl-32b-instruct` (vision-language multimodal) to GroqCloud for Enterprise customers.
• Adds two new voices (`Abdullah` — now the default, and `Aisha`) to `canopylabs/orpheus-arabic-saudi`, bringing the total to six supported voices: Abdullah, Fahad, Sultan, Lulwa, Noura, and Aisha.
• Migrates platform-wide TTS to Orpheus models (`canopylabs/orpheus-v1-english` with voices autumn, diana, hannah, austin, daniel, troy; `canopylabs/orpheus-arabic-saudi` with voices fahad, sultan, lulwa, noura), replacing the deprecated `playai-tts` and `playai-tts-arabic`.
• Python SDK v1.1.0 adds support for binary request streaming and a custom JSON encoder for extended type support.</description>
    </item>
  </channel>
</rss>
