<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Braintrust — The AI Toolchain</title>
    <link>https://aitoolchain.io/tools/braintrust</link>
    <description>New releases and features in Braintrust, tracked by The AI Toolchain.</description>
    <language>en</language>
    <lastBuildDate>Fri, 28 Aug 2026 18:29:53 GMT</lastBuildDate>
    <atom:link href="https://aitoolchain.io/tools/braintrust/rss.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Braintrust 1.0.0</title>
      <link>https://www.braintrust.dev/docs/openapi.yaml</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/openapi.yaml</guid>
      <pubDate>Fri, 28 Aug 2026 18:29:53 GMT</pubDate>
      <description>API surface changed: 65 modified
• DELETE /v1/acl: request body changed; response schema changed
• DELETE /v1/acl/{acl_id}: response schema changed
• DELETE /v1/ai_secret: request body changed
• DELETE /v1/function/{function_id}: response schema changed
• DELETE /v1/role/{role_id}: response schema changed
• DELETE /v1/service_token: request body changed
• DELETE /v1/view/{view_id}: request body changed; response schema changed
• GET /v1/acl: response schema changed
• GET /v1/acl/list_org: response schema changed
• GET /v1/acl/{acl_id}: response schema changed
• GET /v1/function: response schema changed
• GET /v1/function/{function_id}: response schema changed
• GET /v1/prompt: response schema changed
• GET /v1/role/{role_id}: response schema changed
• GET /v1/view: response schema changed
• GET /v1/view/{view_id}: response schema changed
• PATCH /v1/env_var/{env_var_id}: request body changed
• PATCH /v1/function/{function_id}: response schema changed
• PATCH /v1/role/{role_id}: request body changed; response schema changed
• PATCH /v1/view/{view_id}: request body changed; response schema changed
• POST /v1/acl: request body changed; response schema changed
• POST /v1/acl/batch_update: request body changed; response schema changed
• POST /v1/agent: request body changed
• POST /v1/ai_secret: request body changed
• POST /v1/dataset: request body changed
• POST /v1/dataset/{dataset_id}/feedback: request body changed
• POST /v1/dataset/{dataset_id}/insert: request body changed
• POST /v1/dataset_snapshot: request body changed
• POST /v1/env_var: request body changed
• POST /v1/experiment: request body changed
• POST /v1/experiment/{experiment_id}/feedback: request body changed
• POST /v1/experiment/{experiment_id}/insert: request body changed
• POST /v1/function: request body changed; response schema changed
• POST /v1/group: request body changed
• POST /v1/mcp_server: request body changed
• POST /v1/org_automation: request body changed
• POST /v1/project: request body changed
• POST /v1/project_automation: request body changed
• POST /v1/project_group: request body changed
• POST /v1/project_logs/{project_id}/feedback: request body changed
• POST /v1/project_logs/{project_id}/insert: request body changed
• POST /v1/project_score: request body changed
• POST /v1/project_tag: request body changed
• POST /v1/prompt: request body changed
• POST /v1/role: request body changed; response schema changed
• POST /v1/service_token: request body changed
• POST /v1/span_iframe: request body changed
• POST /v1/view: request body changed; response schema changed
• PUT /v1/agent: request body changed
• PUT /v1/ai_secret: request body changed
• PUT /v1/dataset_snapshot: request body changed
• PUT /v1/env_var: request body changed
• PUT /v1/function: request body changed; response schema changed
• PUT /v1/group: request body changed
• PUT /v1/mcp_server: request body changed
• PUT /v1/org_automation: request body changed
• PUT /v1/project_automation: request body changed
• PUT /v1/project_group: request body changed
• PUT /v1/project_score: request body changed
• PUT /v1/project_tag: request body changed
• PUT /v1/prompt: request body changed
• PUT /v1/role: request body changed; response schema changed
• PUT /v1/service_token: request body changed
• PUT /v1/span_iframe: request body changed
• PUT /v1/view: request body changed; response schema changed</description>
    </item>
    <item>
      <title>Braintrust 1.0.0</title>
      <link>https://www.braintrust.dev/docs/openapi.yaml</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/openapi.yaml</guid>
      <pubDate>Mon, 24 Aug 2026 15:04:58 GMT</pubDate>
      <description>API surface changed: +24 endpoints, 53 modified
• New endpoint DELETE /v1/agent/{agent_id}
• New endpoint DELETE /v1/org_automation/{org_automation_id}
• New endpoint DELETE /v1/project_group/{project_group_id}
• New endpoint GET /v1/agent
• New endpoint GET /v1/agent/{agent_id}
• New endpoint GET /v1/org_automation
• New endpoint GET /v1/org_automation/{org_automation_id}
• New endpoint GET /v1/project_group
• New endpoint GET /v1/project_group/{project_group_id}
• New endpoint OPTIONS /v1/agent
• New endpoint OPTIONS /v1/agent/{agent_id}
• New endpoint OPTIONS /v1/org_automation
• New endpoint OPTIONS /v1/org_automation/{org_automation_id}
• New endpoint OPTIONS /v1/project_group
• New endpoint OPTIONS /v1/project_group/{project_group_id}
• New endpoint PATCH /v1/agent/{agent_id}
• New endpoint PATCH /v1/org_automation/{org_automation_id}
• New endpoint PATCH /v1/project_group/{project_group_id}
• New endpoint POST /v1/agent
• New endpoint POST /v1/org_automation
• New endpoint POST /v1/project_group
• New endpoint PUT /v1/agent
• New endpoint PUT /v1/org_automation
• New endpoint PUT /v1/project_group
• DELETE /v1/acl: request body changed; response schema changed
• DELETE /v1/acl/{acl_id}: response schema changed
• DELETE /v1/api_key/{api_key_id}: response schema changed
• DELETE /v1/function/{function_id}: response schema changed
• DELETE /v1/project/{project_id}: response schema changed
• DELETE /v1/project_score/{project_score_id}: response schema changed
• DELETE /v1/prompt/{prompt_id}: response schema changed
• DELETE /v1/role/{role_id}: response schema changed
• DELETE /v1/service_token: response schema changed
• DELETE /v1/service_token/{service_token_id}: response schema changed
• DELETE /v1/view/{view_id}: request body changed; response schema changed
• GET /v1/acl: response schema changed
• GET /v1/acl/list_org: response schema changed
• GET /v1/acl/{acl_id}: response schema changed
• GET /v1/api_key: response schema changed
• GET /v1/api_key/{api_key_id}: response schema changed
• GET /v1/function: response schema changed
• GET /v1/function/{function_id}: response schema changed
• GET /v1/project: response schema changed
• GET /v1/project/{project_id}: response schema changed
• GET /v1/project_score: response schema changed
• GET /v1/project_score/{project_score_id}: response schema changed
• GET /v1/prompt: response schema changed
• GET /v1/prompt/{prompt_id}: response schema changed
• GET /v1/role/{role_id}: response schema changed
• GET /v1/service_token: response schema changed
• GET /v1/service_token/{service_token_id}: response schema changed
• GET /v1/view: response schema changed
• GET /v1/view/{view_id}: response schema changed
• PATCH /v1/function/{function_id}: request body changed; response schema changed
• PATCH /v1/organization/members: request body changed
• PATCH /v1/project/{project_id}: response schema changed
• PATCH /v1/project_score/{project_score_id}: request body changed; response schema changed
• PATCH /v1/prompt/{prompt_id}: request body changed; response schema changed
• PATCH /v1/role/{role_id}: request body changed; response schema changed
• PATCH /v1/view/{view_id}: request body changed; response schema changed
• POST /v1/acl: request body changed; response schema changed
• POST /v1/acl/batch_update: request body changed; response schema changed
• POST /v1/eval: request body changed
• POST /v1/function: request body changed; response schema changed
• POST /v1/function/{function_id}/invoke: request body changed
• POST /v1/project: response schema changed
• POST /v1/project_score: request body changed; response schema changed
• POST /v1/prompt: request body changed; response schema changed
• POST /v1/role: request body changed; response schema changed
• POST /v1/service_token: request body changed; response schema changed
• POST /v1/view: request body changed; response schema changed
• PUT /v1/function: request body changed; response schema changed
• PUT /v1/project_score: request body changed; response schema changed
• PUT /v1/prompt: request body changed; response schema changed
• PUT /v1/role: request body changed; response schema changed
• PUT /v1/service_token: request body changed; response schema changed
• PUT /v1/view: request body changed; response schema changed</description>
    </item>
    <item>
      <title>Braintrust 1.0.0</title>
      <link>https://www.braintrust.dev/docs/openapi.yaml</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/openapi.yaml</guid>
      <pubDate>Thu, 20 Aug 2026 10:01:57 GMT</pubDate>
      <description>Braintrust now publishes an API — 210 endpoints across 28 areas: Cors, Datasets, Experiments, …
• Cors (65 endpoints) — read
• Datasets (10 endpoints) — create, read, update, delete
• Experiments (10 endpoints) — create, read, update, delete
• Acls (7 endpoints) — create, read, delete
• Aisecrets (7 endpoints) — create, read, update, delete
• Functions (7 endpoints) — create, read, update, delete
• Datasetsnapshots (6 endpoints) — create, read, update, delete
• Envvars (6 endpoints) — create, read, update, delete
• 20 more areas: Groups, Mcpservers, Projectautomations, Projectscores, Projecttags, Prompts, Proxy, Roles, Servicetokens, Spaniframes, Views, Environments, Projects, Logs, Organizations, Apikeys, Users, Crossobject, Evals, Other</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260830</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>• Adds `bt trace setup`, `bt trace run`, and `bt trace import` subcommands to the `bt` CLI for managing tracing plugins for Claude Code, Codex, OpenCode, and pi.
• Adds `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` environment variable (Python SDK v0.32.0) to re-enable LiveKit Agents audio attachments on `agent_speaking` spans, which are now off by default.
• Adds `braintrust eval` auto-instrumentation in TypeScript SDK v3.29.0 so supported AI clients and libraries emit spans before eval files load.
• Introduces the AWS Lambda Extension for Python and TypeScript/JavaScript, providing a local handoff path for traces to reduce time spent in flush() during request handling.
• Adds Group scope to online scoring rules, letting you evaluate multi-turn trace sessions as a single unit keyed by a session of your choice, with scores written to the first trace or to every trace.
• Adds Monitoring views as named, project-scoped Dashboards with per-dashboard pages, search, starring, cloning from the built-in Cost and quality dashboard, and auto-saving of chart, filter, and grouping changes.
• Adds annotated version history for prompts, parameters, and scorers, letting you attach a note when saving a new version and view adjacent versions side by side.
• Adds support for dataset rows referencing a group of up to 64 traces, rendering each inline and flagging unavailable ones.
• Adds Kimi K3 and DeepSeek V4 Flash 0731 as built-in open-source models, requestable as `kimi-k3` and `deepseek-v4-flash-0731` through the Braintrust Gateway with no AI provider setup required.
• Adds the Harbor job plugin for syncing Harbor evaluation results to Braintrust in Python SDK v0.33.0, defaulting the project name to `Harbor` so `project_name` is optional.
• Adds Ollama and `@cloudflare/think` instrumentation, Anthropic beta sessions tracing via `anthropic.beta.sessions.turn` and `anthropic.beta.sessions.thread.turn`, Flue v2 support, and a vitest-evals span input override via `meta.eval.input` in TypeScript SDK v3.27.0.
• Adds Cloudflare Agents and Cloudflare AI Chat auto-instrumentation in TypeScript SDK v3.26.0.
Breaking changes:
• The `bt` CLI now handles authentication and trace routing for Claude Code, Codex, OpenCode, and pi plugins; older plugin-specific API key, project, tracing, and config-file settings must be migrated using the bt CLI migration guide and per-agent upgrade notes.
• LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default in Python SDK v0.32.0; set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore previous behavior.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260829</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Braintrust adds MCP write tools, bt CLI trace routing, group-scoped scoring, Azure AI Gateway, and new open-source models in August 2026.
• New `bt trace setup`, `bt trace run`, and `bt trace import` subcommands route tracing for Claude Code, Codex, OpenCode, and pi through the `bt` CLI, handling authentication and trace routing centrally.
• Adds `dataset_id` parameter to init_dataset() (Python SDK v0.35.0) to initialize a dataset by ID, taking precedence over `project`, `project_id`, and `name`.
• Adds `datasetId` parameter to initDataset() (TypeScript SDK v3.29.0) to look up a dataset by ID without specifying a project.
• Adds `apiKey` parameter to loadPrompt() in TypeScript SDK v3.29.0.
• TypeScript SDK v3.29.0: `braintrust eval` now applies auto-instrumentation so supported AI clients can emit spans before eval files load.
• Adds invoke_async() as an async counterpart to invoke() in Python SDK v0.32.0.
• Adds Vercel AI SDK for Python auto-instrumentation (enabled by default in auto_instrument()) in Python SDK v0.33.0.
• Adds Cursor SDK Python instrumentation for tracing agent runs, model turns, and tool calls in Python SDK v0.33.0.
• Adds Pipecat auto-instrumentation in Python SDK v0.32.0 for tracing real-time voice AI pipelines, including LLM turns, STT, TTS, and tool calls.
• Adds Voyage AI auto-instrumentation in TypeScript SDK v3.28.0 for embeddings, multimodal embeddings, reranking, and contextualized embeddings.
• TypeScript SDK v3.27.0 adds Ollama and `@cloudflare/think` instrumentation, Anthropic beta sessions tracing via `anthropic.beta.sessions.turn` and `anthropic.beta.sessions.thread.turn`, Flue v2 support, and a vitest-evals span input override via `meta.eval.input`.
• TypeScript SDK v3.29.0 traces Vercel AI SDK `generateImage` calls as LLM spans with generated images stored as Braintrust attachments.
• TypeScript SDK v3.29.0 records Anthropic thinking tokens and LangChain reasoning tokens as span metrics.
• The Braintrust MCP server now exposes write tools, enabling coding agents to create and update prompts, scorers, classifiers, dashboards, alerts, scheduled jobs, evals, and dataset rows.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models available through the Braintrust Gateway and provider with no AI provider setup required.
• New AWS Lambda Extension provides a local handoff path for traces in Python and TypeScript/JavaScript Lambda functions, reducing time spent in flush() on the request path.
• Monitoring views are now Dashboards with a dedicated page per dashboard, project-scoped list, search, starring, clone/duplicate/rename/delete via a Dashboard actions menu, and auto-saving chart, filter, and grouping changes.
• Prompt, parameter, and scorer version history now supports annotated saves with a description note pinned to each version, and shows each version alongside the one it replaced.
• A dataset row can now reference a group of up to 64 traces instead of a single one, with each trace rendered inline and unavailable traces flagged.
• Braintrust now supports Azure AI Gateway as an AI provider for backends behind an Azure API Management endpoint, supporting OpenAI Chat Completions, OpenAI Responses, and Anthropic Messages APIs.
• Members&apos; permission groups can now be managed directly from Settings &gt; Members, showing direct groups, inherited groups, and available groups to add.
• TypeScript SDK v3.28.0 passes `id` and `tags` of each case to scorer functions in Eval().
Breaking changes:
• `bt` now handles authentication and trace routing for Claude Code, Codex, OpenCode, and pi plugins; older plugin-specific API key, project, tracing, and config-file settings must be migrated per the bt CLI migration guide.
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default; set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore previous behavior.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260828</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Braintrust adds bt CLI tracing for coding agents, MCP write tools, new built-in models, Lambda extension, and Group-scope online scoring.
• Adds `bt trace setup`, `bt trace run`, and `bt trace import` subcommands to the `bt` CLI to install/update tracing plugins, select a Braintrust project, and manage one-off and saved-session tracing workflows for Claude Code, Codex, OpenCode, and pi.
• Exposes write tools on the Braintrust MCP server, enabling coding agents to create and update prompts, scorers, classifiers, Topics pipeline configuration, monitor views, alerts, scheduled jobs, evals, and dataset rows using the authenticated account&apos;s permissions.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models available via the Braintrust provider in playgrounds, prompts, and scorers, and requestable through the Braintrust Gateway with no external AI provider setup.
• Introduces the AWS Lambda Extension, giving Python and TypeScript/JavaScript Lambda functions a local handoff path for traces so the Braintrust SDK&apos;s flush() method spends less time in the request path.
• Adds Group scope to online scoring rules, letting you evaluate a set of related multi-turn traces as a single unit using a session key of your choice, without changing your logging.
• Adds a Summary table layout to the experiments list that shows every experiment as a column with scores and metrics as rows, including an &apos;All scores (avg)&apos; row and group-based aggregation.
• Allows dataset rows to reference a group of up to 64 traces instead of a single trace, rendering each trace inline for multi-turn sessions or related log sets.
• Adds Azure AI Gateway as a supported AI provider, enabling use of a single provider for all backends behind an Azure API Management endpoint with OpenAI Chat Completions, OpenAI Responses, or Anthropic Messages API support.
• Adds a collapsible query sidebar to the SQL sandbox with search by name, drag-to-reorder, command-bar navigation, per-query rename/duplicate/delete menu, and a &apos;Copy share link&apos; action that opens the query in a teammate&apos;s sandbox without auto-running it.
• Python SDK v0.34.0 adds Hugging Face Transformers auto-instrumentation for local pipelines covering text generation, summarization, translation, feature extraction, and question answering.
• Python SDK v0.33.0 adds Vercel AI SDK for Python auto-instrumentation (enabled by default in auto_instrument()), Cursor SDK Python instrumentation for agent runs/model turns/tool calls, and a native Harbor job plugin for syncing Harbor evaluation results to Braintrust.
• Python SDK v0.34.0 forwards eval case fields to scorer functions and defaults the Harbor plugin&apos;s Braintrust project name to `Harbor`, making `project_name` optional.
• Go SDK v0.11.1 adds `WithProvider` and `WithModel` options on Firebase Genkit&apos;s `NewMiddleware` for explicit model attribution, plus traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool` that auto-replace their untraced counterparts.
Breaking changes:
• `bt` now handles authentication and trace routing for Claude Code, Codex, OpenCode, and pi plugins; older plugin-specific API key, project, tracing, and config-file settings must be migrated per the bt CLI migration guide and each agent&apos;s upgrade notes.
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default; set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&apos;gemini&apos;` to `&apos;google&apos;`; update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): `ChatModel` span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map; embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260826</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>• Adds MCP write tools to the Braintrust MCP server so coding agents can create and update prompts, scorers, classifiers, Topics pipeline config, monitor views, alerts, scheduled jobs, evals, and dataset rows — write tools use your authenticated account permissions; configure your client to require confirmation before running destructive tools.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models served through the Braintrust Gateway alongside GLM-5.2, with no external AI provider setup required; usage draws from monthly model credits shared with Topics.
• Python SDK v0.34.0: forwards eval case fields to scorer functions.
• Python SDK v0.32.0: adds invoke_async() as an async counterpart to invoke().
• TypeScript SDK v3.28.0: scorer functions in Eval() now receive the `id` and `tags` of each case.
• Go SDK v0.11.1: Firebase Genkit gains `WithProvider` and `WithModel` options on `NewMiddleware` for explicit model attribution, plus traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool`; auto-instrumentation now replaces `genkit.DefineTool`, `genkit.DefineToolWithInputSchema`, and `genkit.DefineMultipartTool` calls with their traced equivalents.
• Go SDK v0.11.0: Anthropic spans now capture `prompt_cache_creation_5m_tokens` and `prompt_cache_creation_1h_tokens` for TTL-specific prompt caching.
• Go SDK v0.11.0: Bedrock Runtime spans now capture audio and video content blocks; `InvokeModelWithResponseStream` is now fully instrumented for Anthropic Claude models.
• Go SDK v0.11.0: Google ADK spans now include reasoning and cached token metrics.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default; set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&quot;gemini&quot;` to `&quot;google&quot;`; update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): ChatModel span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map; embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`; provider metadata is now lowercase (e.g. `&quot;openai&quot;` instead of `&quot;OpenAI&quot;`). Update trace queries that rely on the previous formats.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260825</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>• Adds write tools to the Braintrust MCP server so coding agents can create and update prompts, scorers, classifiers, Topics pipeline config, monitor views, alerts, scheduled jobs, evals, and dataset rows — not just read them.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models routable through the Braintrust Gateway with no external AI provider setup; select them under the Braintrust provider in playgrounds, prompts, and scorers.
• Version history for prompts, parameters, and scorers now accepts a text annotation when saving a new version, and shows each version side-by-side with the one it replaced.
• Dataset rows can now reference a group of up to 64 traces instead of a single trace, rendering each inline and flagging any that are no longer available.
• SQL sandbox gains a collapsible query sidebar with search-by-name, drag-to-reorder, command-bar navigation, per-query rename/duplicate/delete menu, and a &apos;Copy share link&apos; action that generates a shareable URL.
• Member permission groups are now editable directly from Settings &gt; Members, showing direct memberships, inherited memberships, and available groups in one dialog.
• Python SDK v0.34.0 adds Hugging Face Transformers auto-instrumentation for text generation, summarization, translation, feature extraction, and question answering; adds automatic Braintrust project name defaulting to `Harbor` in the Harbor plugin (making `project_name` optional); and forwards eval case fields to scorer functions.
• Python SDK v0.33.0 adds Vercel AI SDK auto-instrumentation (enabled by default in auto_instrument()) and Cursor SDK instrumentation for tracing agent runs, model turns, and tool calls; adds a native Harbor job plugin for syncing evaluation results to Braintrust.
• Python SDK v0.32.0 adds Pipecat auto-instrumentation for real-time voice AI pipelines including LLM turns, STT, TTS, and tool calls; adds invoke_async() as an async counterpart to invoke().
• TypeScript SDK v3.28.0 adds Voyage AI auto-instrumentation for embeddings, multimodal embeddings, reranking, and contextualized embeddings; scorer functions in Eval() now receive the `id` and `tags` of each case.
• TypeScript SDK v3.27.0 adds Ollama and `@cloudflare/think` instrumentation, Anthropic beta sessions tracing (`anthropic.beta.sessions.turn` and `anthropic.beta.sessions.thread.turn`), Flue v2 support, and a vitest-evals span input override via `meta.eval.input`.
• TypeScript SDK v3.26.0 adds auto-instrumentation for Cloudflare Agents, Cloudflare AI Chat, and Hugging Face Transformers.js, plus system prompt capture for Strands Agents SDK spans.
• Go SDK v0.11.1 adds Firebase Genkit `WithProvider` and `WithModel` options on `NewMiddleware` for explicit model attribution, plus traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool`; auto-instrumentation now replaces the un-traced `genkit.DefineTool`, `genkit.DefineToolWithInputSchema`, and `genkit.DefineMultipartTool` calls automatically.
• Go SDK v0.11.0 adds Anthropic span capture for `prompt_cache_creation_5m_tokens` and `prompt_cache_creation_1h_tokens` for TTL-specific prompt caching; Bedrock Runtime spans now capture audio and video content blocks and fully instrument `InvokeModelWithResponseStream` for Anthropic Claude models; Google ADK spans now include reasoning and cached token metrics.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default. Set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&apos;gemini&apos;` to `&apos;google&apos;`. Trace queries that filter on the previous provider value must be updated.
• Go SDK v0.11.1 (Eino): ChatModel span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map. Embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`. Provider metadata is now lowercase (e.g. `&apos;openai&apos;` instead of `&apos;OpenAI&apos;`). Trace queries relying on the previous formats must be updated.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260824</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Braintrust adds MCP write tools, two new open-source models, Lambda extension, group-scoped online scoring, and more.
• Adds write tools to the Braintrust MCP server, enabling coding agents to create and update prompts, scorers, classifiers, monitor views, alerts, scheduled jobs, dataset rows, and Topics pipeline configuration; configure your client to require confirmation before running write tools.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models available through the Braintrust Gateway and provider selector with no AI provider setup required.
• Adds the Braintrust Lambda Extension to give Python and TypeScript/JavaScript Lambda functions a local handoff path for traces, reducing time spent by the SDK&apos;s flush() method in the request path.
• Adds Group scope to online scoring rules, letting you evaluate a set of related multi-turn traces as a single unit using a session key of your choice without changing your logging.
• Adds a Summary table layout to the experiments list, comparing every experiment in a project as columns with scores and metrics as rows, including an &apos;All scores (avg)&apos; row.
• Adds annotated version history for prompts, parameters, and scorers, pinning a description note to each saved version and showing each version alongside the one it replaced.
• Adds trace group references to dataset rows, allowing a multi-turn session or related set of logs to become one example with up to 64 traces per row.
• Adds Azure AI Gateway as a supported AI provider, supporting models that use the OpenAI Chat Completions, OpenAI Responses, or Anthropic Messages API behind an Azure API Management endpoint.
• Adds Hugging Face Transformers auto-instrumentation in Python SDK v0.34.0 for local pipelines including text generation, summarization, translation, feature extraction, and question answering.
• Adds Vercel AI SDK for Python auto-instrumentation in Python SDK v0.33.0 (enabled by default in auto_instrument()).
• Adds Cursor SDK Python instrumentation in v0.33.0 for tracing agent runs, model turns, and tool calls.
• Adds native Harbor job plugin in Python SDK v0.33.0 for syncing Harbor evaluation results to Braintrust.
• Adds Pipecat auto-instrumentation in Python SDK v0.32.0 for tracing real-time voice AI pipelines including LLM turns, STT, TTS, and tool calls.
• Adds Voyage AI auto-instrumentation in TypeScript SDK v3.28.0 for embeddings, multimodal embeddings, reranking, and contextualized embeddings.
• Adds Ollama and `@cloudflare/think` instrumentation, Anthropic beta sessions tracing (`anthropic.beta.sessions.turn` and `anthropic.beta.sessions.thread.turn`), Flue v2 support, and a vitest-evals span input override via `meta.eval.input` in TypeScript SDK v3.27.0.
• Adds auto-instrumentation for Cloudflare Agents, Cloudflare AI Chat, and Hugging Face Transformers.js in TypeScript SDK v3.26.0.
• Adds system prompt capture for Strands Agents SDK spans in TypeScript SDK v3.26.0.
• Adds `WithProvider` and `WithModel` options on `NewMiddleware` for explicit model attribution, plus traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool` for Firebase Genkit in Go SDK v0.11.1.
• Adds `braintrust.context_json` carrying SDK name and version to all spans in Ruby SDK v0.4.1.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default; set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&quot;gemini&quot;` to `&quot;google&quot;`; update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): ChatModel span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map; embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`; provider metadata is now lowercase (e.g. `&quot;openai&quot;` instead of `&quot;OpenAI&quot;`); update trace queries that rely on the previous formats.
• Go SDK v0.11.0 (Anthropic): Span metadata no longer includes `endpoint`; the `output` field is now a single message object instead of an array; non-streaming spans no longer emit `time_to_first_token`.
• Go SDK v0.11.0 (Bedrock): Span metadata renames `stop_sequences` to `stop` and removes `additional_model_request_fields`; image, document, and tool block shapes now align with Bedrock&apos;s native wire format.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260823</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>• Adds `kimi-k3` and `deepseek-v4-flash-0731` as requestable model identifiers through the Braintrust Gateway, joining GLM-5.2 as built-in open-source models with no AI provider setup required; usage draws from monthly model credits shared with Topics.
• Adds invoke_async() as an async counterpart to invoke() in the Python SDK (v0.32.0).
• Adds `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` environment variable to the Python SDK to re-enable LiveKit Agents audio attachments on `agent_speaking` spans (disabled by default as of v0.32.0).
• TypeScript SDK v3.28.0 surfaces `id` and `tags` of each eval case to scorer functions inside Eval().
• Go SDK v0.11.0 adds Anthropic span capture for `prompt_cache_creation_5m_tokens` and `prompt_cache_creation_1h_tokens` for TTL-specific prompt caching.
• Go SDK v0.11.0 extends Bedrock Runtime spans to capture audio and video content blocks, and fully instruments `InvokeModelWithResponseStream` for Anthropic Claude models.
• Go SDK v0.11.0 adds reasoning and cached token metrics to Google ADK spans.
• Exposes MCP write tools on the Braintrust MCP server, enabling coding agents to create and update prompts, scorers, classifiers, monitor views, alerts, scheduled jobs, dataset rows, and the Topics pipeline end to end; write tools use authenticated account permissions and several replace or remove objects.
• Adds a Summary table layout to the experiments list, showing every experiment as a column with scores and metrics as rows, an &apos;All scores (avg)&apos; row for mean non-pairwise scores, and grouping to compare aggregated scores across groups.
• Adds a collapsible query sidebar to the SQL sandbox with search by name, drag-to-reorder, command-bar jump, and a &apos;Copy share link&apos; action that generates a URL opening the query in a teammate&apos;s sandbox without running it.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default. Set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&quot;gemini&quot;` to `&quot;google&quot;`. Update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): `ChatModel` span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map. Embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`. Provider metadata is now lowercase (e.g. `&quot;openai&quot;` instead of `&quot;OpenAI&quot;`). Update trace queries that rely on the previous formats.
• Go SDK v0.11.0 (Anthropic): Span metadata no longer includes `endpoint`. The `output` field is now a single message object instead of an array. Non-streaming spans no longer emit `time_to_first_token`.
• Go SDK v0.11.0 (Bedrock): Span metadata renames `stop_sequences` to `stop` and removes `additional_model_request_fields`. Image, document, and tool block shapes now align with Bedrock&apos;s native wire format.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260822</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>• Adds write tools to the Braintrust MCP server, enabling coding agents to create and update prompts, scorers, classifiers, Topics pipeline config, monitor views, alerts, scheduled jobs, evals, and dataset rows using the permissions of the authenticated Braintrust account.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models available under the Braintrust provider in playgrounds, prompts, and scorers, or via the Braintrust Gateway — no AI provider setup required.
• Adds invoke_async() as an async counterpart to the existing invoke() method in the Python SDK v0.32.0.
• Adds `WithProvider` and `WithModel` options on `NewMiddleware` and traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool` for Firebase Genkit in Go SDK v0.11.1; auto-instrumentation now replaces `genkit.DefineTool`, `genkit.DefineToolWithInputSchema`, and `genkit.DefineMultipartTool` calls with their traced equivalents.
• Adds `prompt_cache_creation_5m_tokens` and `prompt_cache_creation_1h_tokens` capture for TTL-specific prompt caching on Anthropic spans in Go SDK v0.11.0.
• Adds audio and video content block capture and full instrumentation for `InvokeModelWithResponseStream` for Anthropic Claude models in Bedrock Runtime spans in Go SDK v0.11.0.
• Adds annotated version history for prompts, parameters, and scorers, allowing a description note to be pinned to each saved version; history now shows each version side-by-side with the one it replaced.
• Adds support for trace groups in dataset rows, letting a multi-turn session or related set of logs become a single example (up to 64 traces per row), rendered inline with unavailable traces flagged.
• Adds a collapsible query sidebar to the SQL sandbox with search by name, drag-to-reorder, command-bar navigation, and a &apos;Copy share link&apos; option that opens a query in a teammate&apos;s sandbox without running it.
• Forwards eval case fields to scorer functions in Python SDK v0.34.0.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default. Set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&apos;gemini&apos;` to `&apos;google&apos;`. Update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): `ChatModel` span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map. Embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`. Provider metadata is now lowercase (e.g. `&apos;openai&apos;` instead of `&apos;OpenAI&apos;`). Update trace queries that rely on the previous formats.
• Go SDK v0.11.0 (Anthropic): Span metadata no longer includes `endpoint`. The `output` field is now a single message object instead of an array. Non-streaming spans no longer emit `time_to_first_token`.
• Go SDK v0.11.0 (Bedrock): Span metadata renames `stop_sequences` to `stop` and removes `additional_model_request_fields`. Image, document, and tool block shapes align with Bedrock&apos;s native wire format.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260821</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Braintrust adds MCP write tools, two new open-source models, AWS Lambda extension, group-scoped online scoring, and a raft of SDK instrumentation additions.
• Adds write tools to the Braintrust MCP server so coding agents can create and update prompts, scorers, classifiers, monitor views, alerts, scheduled jobs, dataset rows, and Topics pipeline configuration — not just read them; write tools use your authenticated account permissions and several replace or remove objects, so configure your client to require confirmation.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models served by Braintrust with no AI-provider setup; select them under the Braintrust provider in playgrounds, prompts, and scorers, or request them by name through the Braintrust Gateway; usage draws from monthly model credits shared with Topics.
• Adds the Braintrust Lambda Extension for Python and TypeScript/JavaScript Lambda functions, giving the Braintrust SDK&apos;s flush() method a local handoff path for traces to reduce latency in the request path.
• Adds Azure AI Gateway as a supported AI provider, routing calls to any model behind an Azure API Management endpoint that uses the OpenAI Chat Completions, OpenAI Responses, or Anthropic Messages API.
• Adds a Summary table layout to the experiments list that puts each experiment in a column, scores and metrics in rows, and an &apos;All scores (avg)&apos; row showing the mean of non-pairwise scores, with grouping support.
• Traces that share a `metadata.conversation_id` now surface as related traces automatically for end-to-end multi-turn conversation review without requiring grouping configuration; &apos;Group by&apos; and &apos;Cluster by&apos; have moved into the row-type selector in the toolbar.
• Adds trace-group references to dataset rows so a multi-turn session or related log set becomes one example, with up to 64 traces per row rendered inline.
• Python SDK v0.33.0 adds Vercel AI SDK for Python auto-instrumentation (enabled by default in auto_instrument()).
• Python SDK v0.33.0 adds Cursor SDK Python instrumentation for tracing agent runs, model turns, and tool calls.
• Go SDK v0.11.1 adds `WithProvider` and `WithModel` options on `NewMiddleware` for explicit model attribution in Firebase Genkit, plus traced tool wrappers `DefineTool`, `DefineToolWithInputSchema`, and `DefineMultipartTool`; auto-instrumentation now replaces `genkit.DefineTool`, `genkit.DefineToolWithInputSchema`, and `genkit.DefineMultipartTool` calls with their traced equivalents.
• Ruby SDK v0.4.1 adds `braintrust.context_json` carrying SDK name and version to all spans.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default. Set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.1 (Google GenAI): Provider metadata changed from `&apos;gemini&apos;` to `&apos;google&apos;`. Update trace queries that filter on the previous provider value.
• Go SDK v0.11.1 (Eino): `ChatModel` span output is now an OpenAI-compatible choices array (`[{&quot;index&quot;: 0, &quot;finish_reason&quot;: &quot;...&quot;, &quot;message&quot;: {...}}]`) instead of a flat message map. Embedding input is now `{&quot;inputs&quot;: [{&quot;content&quot;: &quot;...&quot;}]}` and output is `{&quot;count&quot;: N}`, removing `embedding_length` and renaming `embeddings_count`. Provider metadata is now lowercase (e.g. `&apos;openai&apos;` instead of `&apos;OpenAI&apos;`). Update trace queries that rely on the previous formats.
• Go SDK v0.11.0 (Anthropic): Span metadata no longer includes `endpoint`. The `output` field is now a single message object instead of an array. Non-streaming spans no longer emit `time_to_first_token`.
• Go SDK v0.11.0 (Bedrock): Span metadata renames `stop_sequences` to `stop` and removes `additional_model_request_fields`. Image, document, and tool block shapes now align with Bedrock&apos;s native wire format.</description>
    </item>
    <item>
      <title>Braintrust snapshot-20260820</title>
      <link>https://www.braintrust.dev/docs/changelog#august-2026</link>
      <guid isPermaLink="true">https://www.braintrust.dev/docs/changelog#august-2026</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Braintrust adds Kimi K3 &amp; DeepSeek V4 Flash, AWS Lambda Extension, Group-scoped online scoring, and major SDK auto-instrumentation expansions.
• Adds `kimi-k3` and `deepseek-v4-flash-0731` as built-in open-source models served through the Braintrust Gateway — no AI provider setup required; select them under the Braintrust provider in playgrounds, prompts, and scorers.
• Introduces the Braintrust Lambda Extension, giving Python and TypeScript/JavaScript Lambda functions a local handoff path for traces so the SDK&apos;s flush() method spends less time in the request path.
• Adds Group scope to online scoring rules, letting you evaluate a set of related, multi-turn traces as a single unit based on a session key of your choice, without changing logging.
• Adds a Summary table layout to the experiments list that compares every experiment in a project at a glance, with experiments as columns and scores/metrics as rows, plus an &apos;All scores (avg)&apos; row showing the mean of non-pairwise scores.
• Automatic conversation threading: traces sharing a `metadata.conversation_id` now surface as related traces automatically for end-to-end review without configuring grouping.
• Adds annotated version history for prompts, parameters, and scorers — attach a note when saving a new version and view it pinned to that version alongside the replaced version.
• Dataset rows can now reference a group of up to 64 traces instead of a single trace, rendering each inline and flagging unavailable ones.
• Adds Azure AI Gateway as a supported AI provider, supporting models using the OpenAI Chat Completions, OpenAI Responses, or Anthropic Messages API behind an Azure API Management endpoint.
• Python SDK v0.34.0 adds Hugging Face Transformers auto-instrumentation for local pipelines (text generation, summarization, translation, feature extraction, question answering).
• Python SDK v0.33.0 adds Vercel AI SDK for Python auto-instrumentation (enabled by default in auto_instrument()) and Cursor SDK Python instrumentation for tracing agent runs, model turns, and tool calls.
• Python SDK v0.33.0 adds native Harbor job plugin for syncing Harbor evaluation results to Braintrust.
• Python SDK v0.32.0 adds Pipecat auto-instrumentation for real-time voice AI pipelines including LLM turns, STT, TTS, and tool calls.
• Python SDK v0.32.0 adds invoke_async() as an async counterpart to invoke().
• TypeScript SDK v3.28.0 adds Voyage AI auto-instrumentation for embeddings, multimodal embeddings, reranking, and contextualized embeddings.
• TypeScript SDK v3.28.0: scorer functions in Eval() now receive the `id` and `tags` of each case.
• TypeScript SDK v3.27.0 adds Ollama and `@cloudflare/think` instrumentation, Anthropic beta sessions tracing (`anthropic.beta.sessions.turn` and `anthropic.beta.sessions.thread.turn`), Flue v2 support, and a vitest-evals span input override via `meta.eval.input`.
• TypeScript SDK v3.26.0 adds auto-instrumentation for Cloudflare Agents, Cloudflare AI Chat, and Hugging Face Transformers.js, plus system prompt capture for Strands Agents SDK spans.
• Go SDK v0.11.0: Anthropic spans now capture `prompt_cache_creation_5m_tokens` and `prompt_cache_creation_1h_tokens` for TTL-specific prompt caching; Bedrock Runtime spans now capture audio and video content blocks with full instrumentation for `InvokeModelWithResponseStream` on Anthropic Claude models; Google ADK spans now include reasoning and cached token metrics.
• Ruby SDK v0.4.1: all spans now carry `braintrust.context_json` with SDK name, version, instrumentation scope, and detected runtime environment; override with `BRAINTRUST_ENVIRONMENT_NAME` and `BRAINTRUST_ENVIRONMENT_TYPE`.
• Custom views now support writing scores via `trace.update`, enabling numeric feedback capture without leaving the view.
• The Edit billing information dialog now includes a &apos;Purchase order&apos; field and &apos;Tax information&apos; fields (country and tax ID type).
• Starter and Pro plan subscribers can now redeem coupon codes from Settings &gt; Billing by clicking &apos;Redeem coupon&apos; under Current plan.
• Images pointing to local/private networks or using unsafe URL schemes now require approval before loading, even in Auto-load images mode.
Breaking changes:
• Python SDK v0.32.0: LiveKit Agents audio attachments on `agent_speaking` spans are now disabled by default. Set `BRAINTRUST_CAPTURE_AGENT_AUDIO_ATTACHMENTS=true` to restore the previous behavior.
• Go SDK v0.11.0 (Anthropic): span metadata no longer includes `endpoint`; the `output` field is now a single message object instead of an array; non-streaming spans no longer emit `time_to_first_token`.
• Go SDK v0.11.0 (Bedrock): span metadata renames `stop_sequences` to `stop` and removes `additional_model_request_fields`; image, document, and tool block shapes now align with Bedrock&apos;s native wire format.</description>
    </item>
  </channel>
</rss>
