# Agno changelog

> Every feature, fix and improvement shipped in the Agno SDK, newest first, grouped by the day it shipped.

- Canonical: https://www.agno.com/changelog
- Release notes: https://github.com/agno-agi/agno/releases

## 2026-09-04

- [AG-UI now sends wire fields to remote entities and exposes streamed errors](https://www.agno.com/articles/ag-ui-now-sends-wire-fields-to-remote-entities-and-exposes-streamed-errors.md) (SDK v3.0.6): Remote entities now receive wire fields rather than the internal `RunContext` object, and run errors reach the client over the AG-UI stream instead of disappearing.
- [AgentOS now accepts .zip and .eml uploads](https://www.agno.com/articles/agentos-now-accepts-zip-and-eml-uploads.md) (SDK v3.0.6): Archives and email files are now accepted file types for AgentOS uploads, so they can go straight into a run.
- [AgentOS now serves an MCP server card](https://www.agno.com/articles/agentos-now-serves-an-mcp-server-card.md) (SDK v3.0.6): A new `GET /mcp/server-card` endpoint lists the tools your MCP server serves, with a configurable name, version and instructions, so a client can see what an instance offers before connecting.
- [Bring your own async client to Bedrock](https://www.agno.com/articles/bring-your-own-async-client-to-bedrock.md) (SDK v3.0.6): The Bedrock model and embedder now accept an async client you supply, so you control the AWS session and configuration instead of the one Agno builds.
- [Claude thinking blocks are no longer rejected on replay](https://www.agno.com/articles/claude-thinking-blocks-are-no-longer-rejected-on-replay.md) (SDK v3.0.6): Assistant turns are now replayed to Claude verbatim, so extended-thinking blocks survive the next turn instead of being altered and rejected.
- [Gemini now uses each image's real MIME type](https://www.agno.com/articles/gemini-now-uses-each-images-real-mime-type.md) (SDK v3.0.6): An image's actual MIME type is resolved before it goes to Gemini, so a PNG or WebP crosses with the right type instead of being hard-coded as JPEG.
- [Mark custom routes public without turning off auth](https://www.agno.com/articles/mark-custom-routes-public-without-turning-off-auth.md) (SDK v3.0.6): `AuthorizationConfig.excluded_route_paths` lists the custom routes that skip auth, so one public route no longer means disabling authorization across AgentOS.
- [MCPTools can now negotiate the newest MCP protocol era](https://www.agno.com/articles/mcptools-can-now-negotiate-the-newest-mcp-protocol-era.md) (SDK v3.0.6): A new `protocol_mode` argument on `MCPTools` picks which MCP protocol era the client negotiates. It defaults to `legacy`, so existing servers are unaffected.
- [RecursiveChunking no longer emits duplicate trailing chunks](https://www.agno.com/articles/recursivechunking-no-longer-emits-duplicate-trailing-chunks.md) (SDK v3.0.6): `RecursiveChunking` no longer produces a duplicated chunk at the end of a document, so chunk counts and embeddings lose the extra copy.
- [Serve MCP statelessly so any replica can answer any request](https://www.agno.com/articles/serve-mcp-statelessly-so-any-replica-can-answer-any-request.md) (SDK v3.0.6): `MCPConfig(stateless=True)` serves `/mcp` without session tracking, so a multi-instance deployment no longer needs session affinity in front of it.

## 2026-09-01

- [GandrTools adds Gandr text-to-speech to your agents](https://www.agno.com/articles/gandrtools-adds-gandr-text-to-speech-to-your-agents.md) (SDK v3.0.5): The new `GandrTools` toolkit lets an agent turn text into speech through the Gandr TTS API, built for voice agents and covering 23 languages with six voices.
- [llmman is now a supported model provider](https://www.agno.com/articles/llmman-is-now-a-supported-model-provider.md) (SDK v3.0.5): Agno now supports `llmman` for running local models, using the same OpenAI-compatible pattern as its other local-server providers.
- [MCPTools now accepts static auth headers](https://www.agno.com/articles/mcptools-now-accepts-static-auth-headers.md) (SDK v3.0.5): A new `headers` argument on `MCPTools` passes connect-time auth headers for Streamable HTTP and SSE directly, without building a `StreamableHTTPClientParams` object first.
- [We now surface embedding failures during knowledge ingestion](https://www.agno.com/articles/we-now-surface-embedding-failures-during-knowledge-ingestion.md) (SDK v3.0.5): A chunk that fails to embed is now visible and recoverable: files that only partly embedded are marked `partial`, retries are available, and embedders raise instead of returning an empty vector. Check your error handling when you upgrade.

## 2026-08-30

- [AgentOS now publishes agents, teams, workflows, and toolkits as named MCP tools](https://www.agno.com/articles/agentos-now-publishes-agents-teams-workflows-and-toolkits-as-named-mcp-tools.md) (SDK v3.0.2): AgentOS now serves each agent, team, workflow, and toolkit you pass to `MCPConfig.tools` as its own named MCP tool, so a client calls it by name instead of through a generic run tool.
- [An agent can have its own email inbox](https://www.agno.com/articles/an-agent-can-have-its-own-email-inbox.md) (SDK v3.0.2): Agno agents can now provision and operate their own email inbox through Atomic Mail, registering it themselves with no account setup and no human step.
- [CodeMode again blocks shell access on IPython 9.17](https://www.agno.com/articles/codemode-again-blocks-shell-access-on-ipython-9-17.md) (SDK v3.0.2): `CodeMode(allow_shell=False)` blocks shell execution again on IPython 9.17, which had let shell magics like `%%bash` keep running.
- [MCPConfig now rejects unknown fields](https://www.agno.com/articles/mcpconfig-now-rejects-unknown-fields.md) (SDK v3.0.2): `MCPConfig` now raises on fields it doesn't recognize, so a typo like `tool=` instead of `tools=` fails at construction instead of being silently ignored.
- [Metadata passed to run() now wins over agent metadata](https://www.agno.com/articles/metadata-passed-to-run-now-wins-over-agent-metadata.md) (SDK v3.0.2): Run metadata now resolves component, then session, then call-site, so metadata passed to `run()` overrides the same key on `agent.metadata`. A behavior change worth checking when you upgrade.
- [Reasoning detection now checks with the provider first](https://www.agno.com/articles/reasoning-detection-now-checks-with-the-provider-first.md) (SDK v3.0.2): For Gemini, Claude, Ollama, OpenRouter, and Moonshot, Agno now asks the provider whether a model supports reasoning before falling back to matching on the model id.

## 2026-08-26

- [Agents with large toolkits now run faster](https://www.agno.com/articles/agents-with-large-toolkits-now-run-faster.md) (SDK v3.0.1): Agno now derives each tool's JSON schema once and caches it across runs, removing per-run setup work for agents that carry large toolkits.
- [Gemini 3 tool-result media is nested correctly](https://www.agno.com/articles/gemini-3-tool-result-media-is-nested-correctly.md) (SDK v3.0.1): For Gemini 3 and later models, images and documents returned from tools are now nested inside the function response instead of being sent as sibling parts.
- [Long conversations stay fast as they grow](https://www.agno.com/articles/long-conversations-stay-fast-as-they-grow.md) (SDK v3.0.1): Session history now loads incrementally, one turn at a time, so response time stays roughly flat as a conversation gets longer.
- [Strict-mode tool schemas without a properties field no longer error](https://www.agno.com/articles/strict-mode-tool-schemas-without-a-properties-field-no-longer-error.md) (SDK v3.0.1): Strict mode no longer raises a `KeyError` when a tool schema leaves out the `properties` key, so schemas registered verbatim, including MCP server schemas, now work.

## 2026-08-24

- [Breaking changes in v3.0](https://www.agno.com/articles/breaking-changes-in-v3-0.md) (SDK v3.0.0): The condensed reference for every 2.x user upgrading to v3.0, covering storage, AgentOS, agents, teams and workflows, tools, knowledge, the scheduler, evals and models. The database migration is mandatory.
- [CodeMode lets an agent write Python that calls your tools directly](https://www.agno.com/articles/codemode-lets-an-agent-write-python-that-calls-your-tools-directly.md) (SDK v3.0.0): `CodeMode` gives an agent a persistent Python kernel in place of a long list of individual tool calls. The model gets `execute` and `restart`, writes real Python, and calls your tools as awaitable handles.
- [Default models updated across several providers](https://www.agno.com/articles/default-models-updated-across-several-providers.md) (SDK v3.0.0): `Cerebras` and `CerebrasOpenAI` now default to `gpt-oss-120b`, Gemini defaults to 3.7 Flash, and Groq’s deprecated `llama-3.3-70b-versatile` default is now `openai/gpt-oss-120b`. Setting a model id explicitly is unaffected.
- [Durable background execution keeps runs alive through crashes and deploys](https://www.agno.com/articles/durable-background-execution-keeps-runs-alive-through-crashes-and-deploys.md) (SDK v3.0.0): An accepted background run is committed to your database before it starts, so any replica can pick it up and finish it. The queue is backed by your database, and Redis becomes optional coordination rather than the source of truth.
- [FinanceTools unifies financial data behind one toolkit](https://www.agno.com/articles/financetools-unifies-financial-data-behind-one-toolkit.md) (SDK v3.0.0): `FinanceTools` is a single finance toolkit with swappable data providers, so an agent pulls prices, fundamentals and market data through one consistent interface while you choose the source underneath.
- [Generate video with MiniMax](https://www.agno.com/articles/generate-video-with-minimax.md) (SDK v3.0.0): New MiniMax video generation tools let an agent produce video as part of a run.
- [Media offloading stores media in object storage instead of your database](https://www.agno.com/articles/media-offloading-stores-media-in-object-storage-instead-of-your-database.md) (SDK v3.0.0): Set `media_storage` and images, audio, video and files go to object storage instead of being persisted as base64 in the session row, leaving a small `MediaReference` behind. A 113 KB JPEG drops from about 151,000 characters to 2,897.
- [Per-user isolation now covers the whole platform](https://www.agno.com/articles/per-user-isolation-now-covers-the-whole-platform.md) (SDK v3.0.0): Per-user data isolation extends beyond sessions to metrics, schedules, evals, knowledge, components, entity memory and 17 vector databases, so you can run a multi-tenant product on a single AgentOS with real boundaries between users.
- [Runs now have their own database table](https://www.agno.com/articles/runs-now-have-their-own-database-table.md) (SDK v3.0.0): Each run is now a row in a dedicated `agno_runs` table with real columns rather than being packed into the session record, so writes scale linearly and the item-size ceiling on stores like DynamoDB and Firestore is gone. A migration is required before v3.0 serves traffic.
- [Studio 3.0 adds draft-and-publish governance for production teams](https://www.agno.com/articles/studio-3-0-adds-draft-and-publish-governance-for-production-teams.md) (SDK v3.0.0): `create_*` now writes a private draft that serves no one until you publish it, published versions are immutable with rollback behind them, and compare-and-set guards reject a stale write with a typed 409 instead of clobbering a change.
- [Tool result offloading keeps large outputs out of the model's context](https://www.agno.com/articles/tool-result-offloading-keeps-large-outputs-out-of-the-models-context.md) (SDK v3.0.0): Set `offload_tool_results=True` and any tool result over 16,000 characters goes to AgentFS instead of into context, leaving a short envelope with a preview, the size, and a `result_id` the agent can read or search on demand.
- [Toolkits now have stable IDs](https://www.agno.com/articles/toolkits-now-have-stable-ids.md) (SDK v3.0.0): Every toolkit now carries a stable id that AgentOS uses to reference its tools, so a tool reference stays valid across restarts and deploys instead of depending on load order.
- [Two new ways to connect a model](https://www.agno.com/articles/two-new-ways-to-connect-a-model.md) (SDK v3.0.0): Ramp Router joins as a model provider, and the xAI SuperGrok model now authenticates with device-code OAuth so you can connect it without pasting a long-lived key.

## 2026-08-13

- [A2A stream client keeps Task-level metadata on status updates](https://www.agno.com/articles/a2a-stream-client-keeps-task-level-metadata-on-status-updates.md) (SDK v2.9.0): The A2A stream client no longer drops Task-level metadata when it breaks on a status-update event. Metadata attached to a task now survives the stream instead of being lost partway through.
- [Framework return annotations are guarded](https://www.agno.com/articles/framework-return-annotations-are-guarded.md) (SDK v2.9.0): Guarded framework return annotations against a case that could raise during introspection, so components that expose them load without error.
- [list_components now supports a name filter](https://www.agno.com/articles/list-components-now-supports-a-name-filter.md) (SDK v2.9.0): You can filter list_components by name to narrow the results to the component you're after instead of scanning the full list.
- [Paused team member runs survive a session reload](https://www.agno.com/articles/paused-team-member-runs-survive-a-session-reload.md) (SDK v2.9.0): Paused member runs are now persisted, so a team HITL resume works after the session is reloaded. An approval that arrives after a reload resolves against the run that was waiting for it.
- [StudioRunnerTools brings identity-aware dispatch to any router or team lead](https://www.agno.com/articles/studiorunnertools-brings-identity-aware-dispatch-to-any-router-or-team-lead.md) (SDK v2.9.0): New StudioRunnerTools hands out run access on its own: tools to list the components an orchestrator may run and to run one by id, with no create, edit or delete anywhere in the toolkit. Dispatched runs execute as the calling user.
- [Toolkit instructions are preserved on rehydration](https://www.agno.com/articles/toolkit-instructions-are-preserved-on-rehydration.md) (SDK v2.9.0): Loading a persisted component now keeps its toolkit instructions instead of dropping them, so a rehydrated agent behaves the same as it did before it was saved.
- [Workflow runs honor the selected version over WebSocket](https://www.agno.com/articles/workflow-runs-honor-the-selected-version-over-websocket.md) (SDK v2.9.0): A workflow run started over WebSocket now uses the version you selected rather than falling back to another. The version you pin is the version that executes.

## 2026-08-05

- [AdvisorTools brings multi-model escalation to any agent](https://www.agno.com/articles/advisortools-brings-multi-model-escalation-to-any-agent.md) (SDK v2.8.7): New AdvisorTools lets an agent run a fast, cheap model as its primary and consult heavier models only when it decides a problem needs one. Hand it a list of advisors and it can ask one by name or ask them all and compare.
- [Audio tool results are handled more reliably](https://www.agno.com/articles/audio-tool-results-are-handled-more-reliably.md) (SDK v2.8.7): Audio tool-result handling is more robust, so responses that don't match the expected shape no longer break the run.
- [Cohere now respects zero-valued sampling params](https://www.agno.com/articles/cohere-now-respects-zero-valued-sampling-params.md) (SDK v2.8.7): An explicit 0 for temperature, top_k, seed, frequency_penalty, or presence_penalty is no longer dropped before the request. Cohere calls now honor a deliberate zero instead of silently ignoring it.
- [FileSystemTools now supports a configurable toolkit name](https://www.agno.com/articles/filesystemtools-now-supports-a-configurable-toolkit-name.md) (SDK v2.8.7): You can now override the FileSystemTools toolkit name instead of using the default, which makes it practical to mount more than one filesystem toolkit on a single agent and name each for the job it does.
- [HITL confirmations propagate to tool execution on deserialization](https://www.agno.com/articles/hitl-confirmations-propagate-to-tool-execution-on-deserialization.md) (SDK v2.8.7): A top-level confirmation now carries through to tool_execution when a requirement is deserialized, so an approval resolves against the right tool after a save-and-reload round trip.
- [nltk 3.10.1 excluded to protect the unstructured import chain](https://www.agno.com/articles/nltk-3-10-1-excluded-to-protect-the-unstructured-import-chain.md) (SDK v2.8.7): Pinned installs away from nltk 3.10.1, which broke the unstructured import chain. Installs that depend on unstructured keep working.
- [OpenRouteServiceTools grounds agents in real routes and travel times](https://www.agno.com/articles/openrouteservicetools-grounds-agents-in-real-routes-and-travel-times.md) (SDK v2.8.7): New OpenRouteServiceTools swaps an LLM's distance guesses for a real routing engine. Pass plain place names and an agent gets back accurate distances, real drive times, and turn-by-turn routes from live map data.
- [Persisted components rehydrate toolkit-qualified tools correctly](https://www.agno.com/articles/persisted-components-rehydrate-toolkit-qualified-tools-correctly.md) (SDK v2.8.7): Loading a saved component no longer drops tools that are identified by their toolkit-qualified name. Persisted agents now restore their full tool set instead of coming back with pieces missing.
- [StudioTools gains component-aware scheduling and run history](https://www.agno.com/articles/studiotools-gains-component-aware-scheduling-and-run-history.md) (SDK v2.8.7): StudioTools can now schedule Studio-built components and read their past runs, with new history parameters for narrowing what you pull back.
- [Team.load no longer crashes on SQLite](https://www.agno.com/articles/team-load-no-longer-crashes-on-sqlite.md) (SDK v2.8.7): Fixed a crash when loading a team from SQLite caused by an unexpected label keyword argument. Teams stored on SQLite now load cleanly.

## 2026-07-30

- [Give your agents a voice with Smallest AI](https://www.agno.com/articles/give-your-agents-a-voice-with-smallest-ai.md) (SDK v2.8.6): New SmallestTools lets an agent generate speech with Smallest AI's Lightning models. Call text_to_speech to turn text into natural audio, or get_voices to see what's on offer.
- [Poll the status of background metrics refreshes](https://www.agno.com/articles/poll-the-status-of-background-metrics-refreshes.md) (SDK v2.8.6): AgentOS adds a GET /metrics/refresh/status endpoint so clients can watch a background metrics refresh instead of waiting blind and timing out.
- [Run knowledge search on OpenSearch, with hybrid built in](https://www.agno.com/articles/run-knowledge-search-on-opensearch-with-hybrid-built-in.md) (SDK v2.8.6): OpenSearch is now a supported vector database in Agno through agno.vectordb.opensearch. It handles vector, keyword, and hybrid search, in both sync and async variants, so you can back a knowledge base…

## 2026-07-27

- [Ask your AgentOS how it's doing, in plain English](https://www.agno.com/articles/ask-your-agentos-how-its-doing-in-plain-english.md) (SDK v2.8.5): New AgentOSTools gives an agent a read-only ops view of the AgentOS it runs on. Point it at your database and it can report on usage, latency, failures, schedules, evals, components, and pending…
- [Break down agent latency and errors in your traces](https://www.agno.com/articles/break-down-agent-latency-and-errors-in-your-traces.md) (SDK v2.8.5): Traces now report latency and error stats grouped by agent, team, workflow, or endpoint, along with tool and model call stats, so you can see exactly where time goes and where things fail rather than…
- [Delete ClickHouse records by metadata safely](https://www.agno.com/articles/delete-clickhouse-records-by-metadata-safely.md) (SDK v2.8.5): delete_by_metadata now binds metadata keys and values as query parameters rather than inlining them, so deletions run correctly and safely regardless of what the metadata contains.
- [Keep Moonshot reasoning intact across turns](https://www.agno.com/articles/keep-moonshot-reasoning-intact-across-turns.md) (SDK v2.8.5): Moonshot now preserves reasoning_content from one turn to the next, so a model's prior reasoning carries forward instead of being dropped.
- [Toggle thinking and feed files and video to Moonshot](https://www.agno.com/articles/toggle-thinking-and-feed-files-and-video-to-moonshot.md) (SDK v2.8.5): Moonshot gains a use_thinking flag to turn thinking mode on or off, so you choose between deeper reasoning and faster, cheaper responses per use case.

## 2026-07-24

- [Give your agents notes that survive the run](https://www.agno.com/articles/give-your-agents-notes-that-survive-the-run.md) (SDK v2.8.2): Agents are good at working with files, but the filesystem in most setups is a scratch directory that vanishes when the run ends.

## 2026-07-23

- [Cap learning extraction to prevent infinite loops](https://www.agno.com/articles/cap-learning-extraction-to-prevent-infinite-loops.md) (SDK v2.8.1): Learning stores add an extraction_tool_call_limit that bounds how many tool calls the extraction step can make.
- [Find the right moment in a video by describing it](https://www.agno.com/articles/find-the-right-moment-in-a-video-by-describing-it.md) (SDK v2.8.1): TwelveLabsTools now supports Marengo embeddings. Marengo embeds text into the same latent space TwelveLabs uses for video, audio, and image, so a written query and a video clip come out as vectors you…
- [Let Slack agents talk to each other](https://www.agno.com/articles/let-slack-agents-talk-to-each-other.md) (SDK v2.8.1): A new respond_to_other_apps flag lets a Slack agent respond to messages from other agents, not just people.
- [Point FileTools at the directory you choose](https://www.agno.com/articles/point-filetools-at-the-directory-you-choose.md) (SDK v2.8.1): FileTools now exposes a directory parameter, so you can set where it reads and writes instead of relying on the default location.
- [Stream sub-agent events from every context provider](https://www.agno.com/articles/stream-sub-agent-events-from-every-context-provider.md) (SDK v2.8.1): stream_sub_agent_events is now supported across all context providers, not just a subset. Any provider that runs a sub-agent can surface its events as they happen, so you get consistent real-time…

## 2026-07-20

- [Generate code files from your agents](https://www.agno.com/articles/generate-code-files-from-your-agents.md) (SDK v2.8.0): FileGenerationTools now generates code files, extending the toolkit beyond documents and data formats.
- [Page through larger Gmail result sets](https://www.agno.com/articles/page-through-larger-gmail-result-sets.md) (SDK v2.8.0): Gmail tools now support pagination, with a max_results_per_request control over how many messages each request pulls.
- [Pull market sentiment into your agents with Adanos](https://www.agno.com/articles/pull-market-sentiment-into-your-agents-with-adanos.md) (SDK v2.8.0): A new AdanosTools toolkit gives agents multi-source stock sentiment and Reddit cryptocurrency sentiment from Adanos, drawing on signals across Reddit, X, financial news, and prediction markets.
- [Your agent's own best runs are the training data](https://www.agno.com/articles/your-agents-own-best-runs-are-the-training-data.md) (SDK v2.8.0): New in Agno: a straight path from evaluation to fine-tuning data, built from a few pieces that snap together.

## 2026-07-17

- [Do more with Telegram from your agents](https://www.agno.com/articles/do-more-with-telegram-from-your-agents.md) (SDK v2.7.4): Telegram tools gain pin_message, get_chat, get_file, and react_with_emoji, so agents can manage a chat more fully rather than just posting to it.
- [Filter Tavily searches more precisely](https://www.agno.com/articles/filter-tavily-searches-more-precisely.md) (SDK v2.7.4): Tavily searches now accept domain, date range, topic, and country filters, so you can pin an agent's searches to exactly the sources, timeframe, and region that matter.
- [Get full page content as Markdown from Oxylabs](https://www.agno.com/articles/get-full-page-content-as-markdown-from-oxylabs.md) (SDK v2.7.4): Oxylabs website scraping can now return full page content as Markdown, so an agent receives clean, structured text it can actually read and reason over rather than raw HTML.
- [Give your long-running agents a sandbox that survives between calls](https://www.agno.com/articles/give-your-long-running-agents-a-sandbox-that-survives-between-calls.md) (SDK v2.7.4): New in Agno: SuperserveTools, which lets an agent write and run its own code inside a Superserve sandbox. The sandbox is a Firecracker microVM, and the part that matters is that it persists.
- [Let your agents text and call real phone numbers](https://www.agno.com/articles/let-your-agents-text-and-call-real-phone-numbers.md) (SDK v2.7.4): New PlivoTools gives an agent a phone line. It can send SMS and place voice calls through Plivo, and look up a number before it does either.
- [Pass run_context into Router and Condition logic](https://www.agno.com/articles/pass-run-context-into-router-and-condition-logic.md) (SDK v2.7.4): Router selectors and Condition evaluators now accept run_context, giving that logic access to the full run context when deciding which branch to take.
- [Scaffold projects faster with agno create](https://www.agno.com/articles/scaffold-projects-faster-with-agno-create.md) (SDK v2.7.4): agno create is now interactive, prompting you for a starter template and a project name instead of making you get the invocation exactly right up front.
- [See what your agents actually do once they're in production](https://www.agno.com/articles/see-what-your-agents-actually-do-once-theyre-in-production.md) (SDK v2.7.4): We've added an observability integration with The Context Company, so you can trace Agno agent runs and understand how they behave in the wild. Setup is one line.

## 2026-07-14

- [Extract user memory and profiles more reliably](https://www.agno.com/articles/extract-user-memory-and-profiles-more-reliably.md) (SDK v2.7.3): Memory and profile extraction from conversations is now more dependable, so agents pick up and retain the right details about a user more consistently.
- [Handle more human-in-the-loop patterns in AG-UI](https://www.agno.com/articles/handle-more-human-in-the-loop-patterns-in-ag-ui.md) (SDK v2.7.3): AG-UI now extends its support for human-in-the-loop confirmation, input, and feedback, so more of your approval and input flows work natively through the interface.
- [Manage Redmine issues, comments, and time logs from your agents](https://www.agno.com/articles/manage-redmine-issues-comments-and-time-logs-from-your-agents.md) (SDK v2.7.3): New RedmineTools lets an agent work directly in Redmine, the open source project tracker. It can find and read issues, create and update them, leave comments, and log time against them.
- [Run agents on TokenLab](https://www.agno.com/articles/run-agents-on-tokenlab.md) (SDK v2.7.3): TokenLab joins as a new OpenAI-compatible model provider, so you can point agents at TokenLab using the interface you already know.
- [Run agents on Valkey for fast in-memory storage](https://www.agno.com/articles/run-agents-on-valkey-for-fast-in-memory-storage.md) (SDK v2.7.3): ValkeyDb brings Valkey to Agno as an in-memory database for agents, teams, and workflows. Your sessions and state live in memory, so reads and writes stay quick under load.
- [Search knowledge in Valkey with vectors and keywords](https://www.agno.com/articles/search-knowledge-in-valkey-with-vectors-and-keywords.md) (SDK v2.7.3): Valkey also lands as a vector store, and it runs both vector and keyword search from the same backend. You can do semantic and lexical retrieval over one in-memory store.
- [See a clean error when an MCP server goes unreachable](https://www.agno.com/articles/see-a-clean-error-when-an-mcp-server-goes-unreachable.md) (SDK v2.7.3): When an MCP server becomes unreachable, Agno now surfaces a clear error instead of a confusing failure.
- [Skip a redundant model call when saving session context](https://www.agno.com/articles/skip-a-redundant-model-call-when-saving-session-context.md) (SDK v2.7.3): Saving session context no longer makes an extra model call it didn't need. Cutting that redundant round trip trims latency and token cost on every save, so long-running sessions stay leaner without…
- [Support MCP tools that return only structuredContent](https://www.agno.com/articles/support-mcp-tools-that-return-only-structuredcontent.md) (SDK v2.7.3): MCP tools that respond with just structuredContent and no text block now work correctly. Agno reads the structured payload as intended, so tools following that response shape integrate without special…

## 2026-07-09

- [Apply A2A scope mappings correctly under custom prefixes](https://www.agno.com/articles/apply-a2a-scope-mappings-correctly-under-custom-prefixes.md) (SDK v2.7.2): A2A scope mappings now live on the interface itself, with proper support for custom mount prefixes. Routes served under a non-default prefix get the scope checks they're supposed to, so authorization…
- [Manage multiple AgentOS connections with agno connect](https://www.agno.com/articles/manage-multiple-agentos-connections-with-agno-connect.md) (SDK v2.7.2): agno connect gains the controls you need once you're wiring up more than one target. You can select multiple targets at once, tear a connection down with disconnect, and lean on restart hints when a…
- [Run frontend tools from AG-UI](https://www.agno.com/articles/run-frontend-tools-from-ag-ui.md) (SDK v2.7.2): AG-UI now supports client tool execution, so a tool can run in the frontend rather than only on the server.
- [Secure the AgentOS MCP endpoint with standards-based OAuth](https://www.agno.com/articles/secure-the-agentos-mcp-endpoint-with-standards-based-oauth.md) (SDK v2.7.2): Set AgentOS(mcp_auth=...) to put standards-based OAuth in front of your /mcp endpoint. Instead of relying only on tokens, you can gate MCP access through a proper OAuth flow, so connecting clients…

## 2026-07-07

- [Authenticate every transport through one layer](https://www.agno.com/articles/authenticate-every-transport-through-one-layer.md) (SDK v2.7.0): A single AuthMiddleware on the parent app now covers REST, /mcp, and WebSocket transports, so auth no longer drifts from one transport to the next.
- [Authorize A2A and AGUI routes, not just authenticate them](https://www.agno.com/articles/authorize-a2a-and-agui-routes-not-just-authenticate-them.md) (SDK v2.7.0): A2A and AGUI routes now enforce authorization alongside authentication, with scope mappings merged per interface at the mount prefix.
- [Connect any AgentOS to your coding agent with one command](https://www.agno.com/articles/connect-any-agentos-to-your-coding-agent-with-one-command.md) (SDK v2.7.0): Service accounts give AgentOS proper machine identities in the form of agno_pat_... personal access tokens.
- [Discover an AgentOS's capabilities before you connect](https://www.agno.com/articles/discover-an-agentoss-capabilities-before-you-connect.md) (SDK v2.7.0): A new GET /info endpoint reports the agno_version, whether MCP is enabled and at what path, and the auth_mode.
- [Enforce scopes the same way on every path](https://www.agno.com/articles/enforce-scopes-the-same-way-on-every-path.md) (SDK v2.7.0): check_route_scopes now runs identically across JWT, service-account, and MCP paths, and a data-driven get_resource_context_from_path replaces the old hardcoded substring matching.
- [Get accurate traces for MCP-initiated runs](https://www.agno.com/articles/get-accurate-traces-for-mcp-initiated-runs.md) (SDK v2.7.0): Runs kicked off through MCP now start their own root trace instead of nesting under FastMCP's identity-less protocol span.
- [Keep MCP requests from leaking into each other](https://www.agno.com/articles/keep-mcp-requests-from-leaking-into-each-other.md) (SDK v2.7.0): MCP components now resolve through shared resolve_* helpers that make create_fresh deep copies instead of sharing singleton state, and a fresh session is minted per call whenever session_id is…
- [Manage AgentOS from the command line with agnoctl](https://www.agno.com/articles/manage-agentos-from-the-command-line-with-agnoctl.md) (SDK v2.7.0): A new CLI ships as agnoctl on PyPI and runs as agno. agno connect discovers an AgentOS, mints per-client PATs, and writes MCP config for Claude Code, Claude Desktop, Cursor, Codex, and ChatGPT, then…
- [Operate your AgentOS over MCP with a focused tool surface](https://www.agno.com/articles/operate-your-agentos-over-mcp-with-a-focused-tool-surface.md) (SDK v2.7.0): MCP Interface v2 exposes a clean eight-tool operator surface at /mcp: get_agentos_config, run_agent, run_team, run_workflow, continue_run, cancel_run, get_sessions, and get_session_runs.
- [Run eval suites and gate CI with agno.eval](https://www.agno.com/articles/run-eval-suites-and-gate-ci-with-agno-eval.md) (SDK v2.7.0): A new agno.eval layer gives you a proper suite runner built from Case and run_cases/arun_cases, plus an argparse CLI that supports team subjects and numeric judge scoring.

## 2026-07-03

- [Analyze video and embed multimodal content with TwelveLabs](https://www.agno.com/articles/analyze-video-and-embed-multimodal-content-with-twelvelabs.md) (SDK v2.6.22): A new TwelveLabsTools toolkit lets your agents analyze videos and generate multimodal text embeddings, so video becomes something an agent can search, summarize, and reason over rather than a black…
- [Reach Google, News, Images, and YouTube through SearchAPI](https://www.agno.com/articles/reach-google-news-images-and-youtube-through-searchapi.md) (SDK v2.6.22): A new SearchApiTools toolkit wires up SearchAPI's Google, News, Images, and YouTube endpoints, so an agent can run several kinds of search from one toolkit instead of stitching together separate…
- [Search, extract, and research with Sofya](https://www.agno.com/articles/search-extract-and-research-with-sofya.md) (SDK v2.6.22): A new SofyaTools toolkit gives agents search, extraction, and research in one place, so they can find sources, pull content, and dig into a topic through a single integration.
- [Stop slow tool calls from hanging your agents](https://www.agno.com/articles/stop-slow-tool-calls-from-hanging-your-agents.md) (SDK v2.6.22): The base Toolkit now takes a timeout, and HTTP timeouts are wired across the tools, with support extended to more toolkits.

## 2026-07-02

- [Let agents read local files, safely scoped to a directory](https://www.agno.com/articles/let-agents-read-local-files-safely-scoped-to-a-directory.md) (SDK v2.6.21): LocalFileSystemTools can now read files, not just write them, once you set the enable_read_file flag. By default it keeps every file operation inside your target_directory, so an agent stays within…
- [Rename StudioTool to StudioTools without breaking your code](https://www.agno.com/articles/rename-studiotool-to-studiotools-without-breaking-your-code.md) (SDK v2.6.21): StudioTool is now StudioTools, bringing it in line with the plural naming the other toolkits use. A backward-compatible alias keeps the old name working, so existing code runs unchanged while you move…

## 2026-06-26

- [Add as many quick prompts as you need](https://www.agno.com/articles/add-as-many-quick-prompts-as-you-need.md) (SDK v2.6.20): Surface as many suggested prompts as your interface calls for. AgentOS drops the three-per-entity cap on quick_prompts, so you shape the list around your users instead of an arbitrary limit.
- [Authenticate every Google toolkit the same way](https://www.agno.com/articles/authenticate-every-google-toolkit-the-same-way.md) (SDK v2.6.20): Google toolkits now share one auth base class, so authentication works consistently across all of them. You get a cleaner, more predictable foundation and integrations that are easier to maintain.
- [Get reliable structured output from LiteLLM providers](https://www.agno.com/articles/get-reliable-structured-output-from-litellm-providers.md) (SDK v2.6.20): Turn on native structured outputs and JSON schema outputs for LiteLLM per provider with supports_native_structured_outputs or supports_json_schema_outputs.
- [Give your agents web search through Scavio](https://www.agno.com/articles/give-your-agents-web-search-through-scavio.md) (SDK v2.6.20): Add the new Scavio toolkit to an agent and it gains Scavio-backed web search instantly, the same way it picks up any other Agno toolkit.
- [Handle high-volume traces with ClickHouse](https://www.agno.com/articles/handle-high-volume-traces-with-clickhouse.md) (SDK v2.6.20): Land your traces in ClickHouse and let a column store built for analytics carry the load. It ingests heavy trace volume and runs fast OLAP scans, so aggregating and slicing observability data stays…
- [Read OpenAI's web-search sources straight off the response](https://www.agno.com/articles/read-openais-web-search-sources-straight-off-the-response.md) (SDK v2.6.20): Grab the citations behind a grounded OpenAI answer directly from response.citations, now populated for the OpenAIChat and OpenAILike providers.
- [See every AgentOS route with newer FastAPI](https://www.agno.com/articles/see-every-agentos-route-with-newer-fastapi.md) (SDK v2.6.20): Run AgentOS on FastAPI >= 0.137 and get_routes() returns every registered route. You get the full picture of your deployment when you introspect it, rather than a partial list.

## 2026-06-23

- [Compose agents, teams, and workflows on the fly with StudioTool](https://www.agno.com/articles/compose-agents-teams-and-workflows-on-the-fly-with-studiotool.md) (SDK v2.6.19): A new StudioTool toolkit lets an agent dynamically compose other Agno primitives, assembling agents, teams, and workflows at runtime rather than wiring them all up ahead of time.
- [Fork and resume runs from a known-good checkpoint](https://www.agno.com/articles/fork-and-resume-runs-from-a-known-good-checkpoint.md) (SDK v2.6.19): Agno now checkpoints runs at the tool-batch level and exposes a unified /continue endpoint that handles both regenerating a run and forking it, along with support for forking sessions.
- [Import Gemini without pinning google-genai 2.0](https://www.agno.com/articles/import-gemini-without-pinning-google-genai-2-0.md) (SDK v2.6.19): GeminiInteractions now imports lazily, so pulling in Gemini no longer forces google-genai 2.0 on your environment.
- [Pass raw regex to the PII guardrail](https://www.agno.com/articles/pass-raw-regex-to-the-pii-guardrail.md) (SDK v2.6.19): The PII guardrail's custom_patterns now accepts raw regex strings and compiles them for you, so you can add a pattern inline instead of pre-compiling it yourself.
- [Query supported search types on ClickHouse and Pinecone](https://www.agno.com/articles/query-supported-search-types-on-clickhouse-and-pinecone.md) (SDK v2.6.19): ClickHouse and Pinecone vector DBs now report their supported search types through get_supported_search_types(), matching the other vector stores.

## 2026-06-18

- [Keep model connection params intact when rebuilding stored agents](https://www.agno.com/articles/keep-model-connection-params-intact-when-rebuilding-stored-agents.md) (SDK v2.6.18): Rebuilding a DB-stored agent or team now reuses the live model instance from the registry instead of reconstructing it from scratch.

## 2026-06-17

- [Load agents even when one component fails](https://www.agno.com/articles/load-agents-even-when-one-component-fails.md) (SDK v2.6.17): Loading agents and teams from the database now isolates each component, so a single bad component gets skipped instead of dropping the whole set.
- [Stop re-instantiated toolkits from piling up in the registry](https://www.agno.com/articles/stop-re-instantiated-toolkits-from-piling-up-in-the-registry.md) (SDK v2.6.17): The registry now deduplicates toolkits by their type, name, and function set, so a toolkit that gets re-instantiated collapses onto the existing entry instead of adding a duplicate.

## 2026-06-15

- [Extend the AgentOS MCP server with custom, scoped, identity-aware tools](https://www.agno.com/articles/extend-the-agentos-mcp-server-with-custom-scoped-identity-aware-tools.md) (SDK v2.6.15): The AgentOS MCP server at /mcp is now a real extension point, configured through a single MCPServerConfig object rather than custom middleware.

## 2026-06-12

- [Carry JSON instructions into follow-up prompts for json_object providers](https://www.agno.com/articles/carry-json-instructions-into-follow-up-prompts-for-json-object-providers.md) (SDK v2.6.14): For providers that use json_object structured output, Agno now passes the JSON formatting instructions into the follow-up prompt as well, not just the initial one.
- [Manage what your agents have learned with full CRUD](https://www.agno.com/articles/manage-what-your-agents-have-learned-with-full-crud.md) (SDK v2.6.14): AgentOS now exposes create, read, update, and delete endpoints for learnings, giving you direct control over what an agent has learned instead of treating that store as write-only.
- [Run Gemini safely under concurrent load](https://www.agno.com/articles/run-gemini-safely-under-concurrent-load.md) (SDK v2.6.14): Gemini no longer does a per-response cleanup that could race when multiple responses were in flight at once.

## 2026-06-10

- [Keep one failed MCP server from taking down the rest](https://www.agno.com/articles/keep-one-failed-mcp-server-from-taking-down-the-rest.md) (SDK v2.6.13): MultiMCP now handles connection failures cleanly (v2.6.13), so a single server that fails to connect no longer disrupts the others.
- [Keep same-document inserts distinct when metadata differs](https://www.agno.com/articles/keep-same-document-inserts-distinct-when-metadata-differs.md) (SDK v2.6.13): Content hashing now folds metadata into the content hash (v2.6.13), so upsert=False inserts of the same document no longer collapse into one.
- [Populate the AgentOS registry from what you've already defined](https://www.agno.com/articles/populate-the-agentos-registry-from-what-youve-already-defined.md) (SDK v2.6.13): The registry gained knowledge and managers support (v2.6.10), and the AgentOS registry now auto-populates from the agents, teams, and workflows you've defined (v2.6.13).
- [Resolve workflow approvals over a live connection](https://www.agno.com/articles/resolve-workflow-approvals-over-a-live-connection.md) (SDK v2.6.13): Paused workflows now surface approval requests and take responses over a live socket connection (v2.6.13), so a reviewer can approve or reject mid-workflow in real time rather than polling for pending…
- [See a context provider's sub-agent work as it happens](https://www.agno.com/articles/see-a-context-providers-sub-agent-work-as-it-happens.md) (SDK v2.6.13): Context providers now stream their sub-agent events through to the parent run (v2.6.10, refined in v2.6.13), so a provider that runs its own sub-agent surfaces that work as it happens instead of going…
- [Stand up the AgentOS Slack app from a ready-made manifest](https://www.agno.com/articles/stand-up-the-agentos-slack-app-from-a-ready-made-manifest.md) (SDK v2.6.13): A ready-made Slack app manifest now ships for the AgentOS Slack interface (v2.6.13), so you can create the Slack app from a known-good configuration instead of assembling scopes and settings by hand.

## 2026-06-05

- [Hand back finished DOCX and HTML files, not raw text](https://www.agno.com/articles/hand-back-finished-docx-and-html-files-not-raw-text.md) (SDK v2.6.12): FileGenerationTools adds two output formats: DOCX (v2.6.10) and HTML (v2.6.12), the latter shipping with an example app.
- [Run agents on Tuning Engines](https://www.agno.com/articles/run-agents-on-tuning-engines.md) (SDK v2.6.12): Tuning Engines joins the lineup as a new model provider, extending the set of models you can run agents on without leaving Agno.
- [Set up role-based access control with WorkOS](https://www.agno.com/articles/set-up-role-based-access-control-with-workos.md) (SDK v2.6.12): A new worked WorkOS RBAC example gives you a concrete starting point for wiring role-based access control into an AgentOS deployment, instead of assembling the pattern from scratch.
- [Trace agents in Latitude through OpenInference](https://www.agno.com/articles/trace-agents-in-latitude-through-openinference.md) (SDK v2.6.12): A new Latitude observability example sends traces via OpenInference, giving you a ready reference for piping agent traces into Latitude rather than figuring out the integration yourself.
- [Track agent state from AG-UI front-ends](https://www.agno.com/articles/track-agent-state-from-ag-ui-front-ends.md) (SDK v2.6.12): The AG-UI integration now emits state events, so a front-end built on AG-UI can follow an agent's state as it changes and react in real time rather than waiting for the run to finish.

## 2026-06-02

- [Generate Word documents straight from your agents](https://www.agno.com/articles/generate-word-documents-straight-from-your-agents.md) (SDK v2.6.10): Agno now supports DOCX file generation, so an agent can produce a finished .docx as output rather than handing back raw text for someone to format.
- [Give each AgentOS entity its own UI metadata](https://www.agno.com/articles/give-each-agentos-entity-its-own-ui-metadata.md) (SDK v2.6.11): A new Manifest adds per-entity UI metadata to AgentOS, so every agent, team, and workflow carries its own presentation details rather than sharing one generic look.
- [Persist cancelled runs across agents, teams, and workflows](https://www.agno.com/articles/persist-cancelled-runs-across-agents-teams-and-workflows.md) (SDK v2.6.10): Agents, teams, and workflows now persist cancelled runs properly, so a run that gets cancelled is recorded in the database instead of vanishing.
- [Pick up output files from the RunCompleted event](https://www.agno.com/articles/pick-up-output-files-from-the-runcompleted-event.md) (SDK v2.6.10): The RunCompleted event now carries a files field, so anything listening for run completion can grab the files a run produced directly off the event instead of fetching them separately.
- [Reference Gemini Interactions by model string](https://www.agno.com/articles/reference-gemini-interactions-by-model-string.md) (SDK v2.6.10): The model string parser now recognizes the google-interactions provider, so you can select GeminiInteractions through a model string rather than importing and constructing the class yourself.
- [Refresh DeepSeek V4 thinking mode and defaults](https://www.agno.com/articles/refresh-deepseek-v4-thinking-mode-and-defaults.md) (SDK v2.6.10): Updated DeepSeek V4's thinking mode and default settings so agents on DeepSeek run against current, sensible defaults out of the box.
- [Register knowledge and managers in the registry](https://www.agno.com/articles/register-knowledge-and-managers-in-the-registry.md) (SDK v2.6.10): The registry now supports knowledge and managers alongside the components it already tracks, so you can register and reuse those pieces through the same mechanism rather than wiring them up by hand…
- [Run agents on four new model providers](https://www.agno.com/articles/run-agents-on-four-new-model-providers.md) (SDK v2.6.10): Agno adds first-party integrations for four more providers, widening the set of models you can run agents on without leaving the framework.
- [Run and monitor Parallel tasks, not just searches](https://www.agno.com/articles/run-and-monitor-parallel-tasks-not-just-searches.md) (SDK v2.6.11): The Parallel integration now reaches beyond web search. v2.6.11 added tools for Parallel's Task API and Monitor API, so an agent can kick off task executions and track their progress rather than only…
- [Search the web through You.com](https://www.agno.com/articles/search-the-web-through-you-com.md) (SDK v2.6.10): A new YouTools toolkit wires up the You.com Search API, so an agent can run web searches through You.com with no custom client to build.
- [Stream sub-agent events from context providers](https://www.agno.com/articles/stream-sub-agent-events-from-context-providers.md) (SDK v2.6.10): Context providers can now stream the events from their sub-agents, so a provider that runs its own agent surfaces that work as it happens instead of going quiet until the final result.

## 2026-05-21

- [Honor temperature=0 on Claude for deterministic output](https://www.agno.com/articles/honor-temperature-0-on-claude-for-deterministic-output.md) (SDK v2.6.9): Claude on Anthropic, AWS, and VertexAI used to silently drop an explicit 0 for temperature, top_p, or top_k, since a bare truthiness check treated 0.0 as unset and fell back to the API default near…
- [Make PgVector prefix matching actually match prefixes](https://www.agno.com/articles/make-pgvector-prefix-matching-actually-match-prefixes.md) (SDK v2.6.9): PgVector(prefix_match=True) used to be a silent no-op: it appended a * and then routed through websearch_to_tsquery, which ignores wildcards.
- [Read the full resolved-approval record in post-hooks](https://www.agno.com/articles/read-the-full-resolved-approval-record-in-post-hooks.md) (SDK v2.6.9): Post-hooks and observability integrations can now read the complete resolved approval record, including resolved_by and resolved_at, through run_response.metadata["approval"].
- [Stop server-side tool calls from breaking managed Gemini agents](https://www.agno.com/articles/stop-server-side-tool-calls-from-breaking-managed-gemini-agents.md) (SDK v2.6.9): On the agent path for Antigravity and Deep Research, the autonomous loop runs its tools inside Google's server-managed sandbox.

## 2026-05-20

- [Antigravity, two shapes](https://www.agno.com/articles/antigravity-two-shapes.md) (SDK v2.6.8): You can now give your agents a full code-running, web-browsing, file-editing sandbox without building or operating any of it.
- [Attribute Parallel MCP traffic to Agno](https://www.agno.com/articles/attribute-parallel-mcp-traffic-to-agno.md) (SDK v2.6.8): ParallelMCPBackend now sends a User-Agent: agno/<version> header on every request, so Parallel can attribute the traffic your agents generate.
- [Drive the Evals model dropdown from a single config](https://www.agno.com/articles/drive-the-evals-model-dropdown-from-a-single-config.md) (SDK v2.6.8): EvalsDomainConfig drops its unused available_models field, leaving the top-level AgentOSConfig.available_models as the only supported source for the model dropdown in the Evals UI.
- [Managed Deep Research and Antigravity](https://www.agno.com/articles/managed-deep-research-and-antigravity.md) (SDK v2.6.8): You can now run Google's two most capable managed agents, autonomous research and a code-running sandbox, without leaving the Gemini setup you already have.
- [Pick up the intended Chonkie version](https://www.agno.com/articles/pick-up-the-intended-chonkie-version.md) (SDK v2.6.8): Updated the Chonkie dependency pin as a follow-up to #7869, so installs resolve to the version Agno expects.
- [Point Gemini Interactions cookbooks at the current model](https://www.agno.com/articles/point-gemini-interactions-cookbooks-at-the-current-model.md) (SDK v2.6.8): Renamed gemini-3-flash-preview to gemini-3.5-flash across the Gemini Interactions cookbooks, so copied examples run against the current model ID instead of the preview name.

## 2026-05-15

- [Cut multi-turn token costs with Gemini's Interactions API](https://www.agno.com/articles/cut-multi-turn-token-costs-with-geminis-interactions-api.md) (SDK v2.6.7): A new GeminiInteractions model class builds on Google's stateful Interactions API, so agents can talk to the interactions endpoint directly instead of Gemini's generateContent.
- [Keep parent trace attribution correct when spans share a trace](https://www.agno.com/articles/keep-parent-trace-attribution-correct-when-spans-share-a-trace.md) (SDK v2.6.7): A child agent's spans no longer overwrite the parent trace's session_id, agent_id, or team_id when both share a trace_id.
- [Partition each user's data in a single AgentOS deployment](https://www.agno.com/articles/partition-each-users-data-in-a-single-agentos-deployment.md) (SDK v2.6.7): AgentOS now offers an opt-in per-user data isolation layer for authenticated endpoints, so one deployment keeps each user's data separated rather than pooling it together.
- [Restrict what your knowledge readers can fetch](https://www.agno.com/articles/restrict-what-your-knowledge-readers-can-fetch.md) (SDK v2.6.7): URL-fetching knowledge readers now take an allowed_hosts parameter, so a reader pulls only from hosts you trust and rejects everything else.
- [Resume workflow approvals cleanly in async code](https://www.agno.com/articles/resume-workflow-approvals-cleanly-in-async-code.md) (SDK v2.6.7): The workflow HITL continue path now calls the async acleanup_run when it runs in an async context, rather than the synchronous version.
- [Speed up Qdrant hybrid inserts by dropping a redundant encode](https://www.agno.com/articles/speed-up-qdrant-hybrid-inserts-by-dropping-a-redundant-encode.md) (SDK v2.6.7): Qdrant's async_insert no longer calls the sparse encoder twice. Hybrid inserts now encode once, cutting wasted compute on every write to a Qdrant collection.

## 2026-05-14

- [Back your agent's wiki with a Notion database](https://www.agno.com/articles/back-your-agents-wiki-with-a-notion-database.md) (SDK v2.6.6): WikiContextProvider now supports a NotionDatabaseBackend source, so you can point an agent's wiki at a Notion database and have it query that content directly rather than maintaining a separate store.
- [Carry dependencies and metadata through to continued runs](https://www.agno.com/articles/carry-dependencies-and-metadata-through-to-continued-runs.md) (SDK v2.6.6): The /continue endpoint now forwards dependencies and metadata through get_request_kwargs, so a resumed run sees the same context as the original call.
- [Catch duplicate tool names before they cause silent conflicts](https://www.agno.com/articles/catch-duplicate-tool-names-before-they-cause-silent-conflicts.md) (SDK v2.6.6): Registering two tools under the same name on an agent or team used to fail quietly, with one definition shadowing the other and no signal as to why a tool misbehaved.
- [Clear every pending approval without leaving Slack](https://www.agno.com/articles/clear-every-pending-approval-without-leaving-slack.md) (SDK v2.6.6): Reviewers no longer have to chase pending approvals one at a time. The Slack interface now supports multi-row approvals with all pause types covered, so a reviewer can resolve several pending…
- [Give teams the LearningMachine context they were missing](https://www.agno.com/articles/give-teams-the-learningmachine-context-they-were-missing.md) (SDK v2.6.6): LearningMachine now injects its context into the Team system prompt, not just the agent path. Teams get the same learned context that individual agents already received, so their behavior reflects it…
- [Reliably fetch the last run output, even with auto-generated agent IDs](https://www.agno.com/articles/reliably-fetch-the-last-run-output-even-with-auto-generated-agent-ids.md) (SDK v2.6.6): aget_last_run_output no longer returns None when agent.id is auto-generated during arun(). You get the run output back whether or not you set an explicit agent ID, so code that reads the last result…

## 2026-05-06

- [Give Slack agents file upload and download, off by default](https://www.agno.com/articles/give-slack-agents-file-upload-and-download-off-by-default.md) (SDK v2.6.5): File handling is powerful but not always wanted, so SlackContextProvider now puts it behind an enable_media_tools flag that defaults to False.
- [Ground agents in Gmail and Calendar without custom integrations](https://www.agno.com/articles/ground-agents-in-gmail-and-calendar-without-custom-integrations.md) (SDK v2.6.5): Email and calendar are two of the most-requested grounding sources, and wiring them up usually means custom API clients and token plumbing.
- [Recover gracefully when a conditional branch fails](https://www.agno.com/articles/recover-gracefully-when-a-conditional-branch-fails.md) (SDK v2.6.5): Conditional branches often wrap fragile work like external calls or tool execution, and until now a failure inside one would propagate unhandled and halt the run.
- [Restrict agent fetches to hosts you trust](https://www.agno.com/articles/restrict-agent-fetches-to-hosts-you-trust.md) (SDK v2.6.5): Fetch tools that follow links are an SSRF and data-exfiltration risk in production. LLMsTxtTools now takes an allowed_hosts parameter that closes that surface: an agent only fetches from hosts you…
- [Schedule agents on a cron with MongoDB](https://www.agno.com/articles/schedule-agents-on-a-cron-with-mongodb.md) (SDK v2.6.5): The AgentOS scheduler now supports Mongo and AsyncMongo as backing stores, so teams already running on Mongo can schedule recurring runs of agents, teams, and workflows without standing up a separate…
- [Search across images and text with Gemini file search](https://www.agno.com/articles/search-across-images-and-text-with-gemini-file-search.md) (SDK v2.6.5): Agno now supports multimodal inputs in the Gemini File Search API, so agents can index and semantically search across images alongside text rather than text alone.

## 2026-04-28

- [Bring your wiki into agent context](https://www.agno.com/articles/bring-your-wiki-into-agent-context.md) (SDK v2.6.4): Agno introduces WikiContextProvider, a context provider built specifically for wiki and knowledge-base content.
- [A clearer, more explicit SlackContextProvider](https://www.agno.com/articles/slackcontextprovider.md) (SDK v2.6.3): SlackContextProvider has been simplified to a single, self-documenting configuration surface. The for_bot_read(), for_assistant_search(), and for_write() factory methods have been removed in favor of…

## 2026-04-27

- [Default model IDs refreshed across providers](https://www.agno.com/articles/default-model-ids-refreshed-across-providers.md) (SDK v2.6.2): Agno has updated the default model id used by several model providers to newer, actively supported versions.
- [Project-aware context for repo-rooted agents](https://www.agno.com/articles/project-aware-context-for-repo-rooted-agents.md) (SDK v2.6.3): Agno introduces WorkspaceContextProvider, a context provider purpose-built for agents that operate inside a repository root.
- [Slack interface reliability and behavior improvements](https://www.agno.com/articles/slack-interface-reliability-and-behavior-improvements.md) (SDK v2.6.3): A round of fixes makes Slack-backed agents more predictable in production. The interface now gracefully falls back to public channels when the groups:read scope is missing, rather than failing…
- [Give agents safe, scoped access to the local workspace](https://www.agno.com/articles/workflow-hitl-guardrails-now-survive-deep-copies.md) (SDK v2.6.2): A new Workspace toolkit gives agents structured access to a configurable root directory, with operations grouped by capability and destructive actions gated by human-in-the-loop confirmation by…

## 2026-04-24

- [Lower cost and latency on Claude with multi-block prompt caching](https://www.agno.com/articles/lower-cost-and-latency-on-claude-with-multi-block-prompt-caching.md) (SDK v2.6.1): Agno now supports Anthropic's multi-block prompt caching for Claude models, giving teams granular control over what gets cached and for how long.
- [OpenAI agents now default to the Responses API](https://www.agno.com/articles/openai-agents-now-default-to-the-responses-api.md) (SDK v2.6.1): The openai: model prefix now resolves to OpenAIResponses rather than the legacy Chat Completions surface.
- [Web search and fetch through Parallel, no setup required](https://www.agno.com/articles/web-search-and-fetch-through-parallel-no-setup-required.md) (SDK v2.6.1): WebContextProvider now ships with a Parallel backend, giving agents access to high-quality web search and page fetch through Parallel's hosted research service.

## 2026-04-23

- [Add approval gates to multi-agent team workflows](https://www.agno.com/articles/add-approval-gates-to-multi-agent-team-workflows.md) (SDK v2.6.0): Teams now support approval flows through both the API and the AgentOS chat interface. Sensitive actions can be paused for explicit human sign-off before they execute, giving operators a clear control…
- [Bring human review into multi-agent teams](https://www.agno.com/articles/bring-human-review-into-multi-agent-teams.md) (SDK v2.6.0): Human-in-the-loop is now available for Teams, with full support in the AgentOS chat interface and a dedicated API layer.
- [Connect external knowledge sources to agents in a single line](https://www.agno.com/articles/connect-external-knowledge-sources-to-agents-in-a-single-line.md) (SDK v2.6.0): The new agno.context API lets agents reach into filesystems, web sources, SQL databases, Slack, Google Drive, and MCP servers as natural-language tools.
- [One call now returns all session types](https://www.agno.com/articles/one-call-now-returns-all-session-types.md) (SDK v2.6.0): The /sessions endpoint returns agent, team, and workflow sessions in a single response by default. This gives a complete view of session activity in one call, which is the most common use case for…
- [Provision agents, teams, and workflows on demand for multi-tenant systems](https://www.agno.com/articles/provision-agents-teams-and-workflows-on-demand-for-multi-tenant-systems.md) (SDK v2.6.0): AgentFactory, TeamFactory, and WorkflowFactory let you create agents, teams, and workflows dynamically at runtime instead of defining them statically at startup.
- [Recover long-running agent runs after interruptions](https://www.agno.com/articles/recover-long-running-agent-runs-after-interruptions.md) (SDK v2.6.0): Background runs streamed over Server-Sent Events can now reconnect and resume after a disconnection or page refresh. Operators rejoin the run exactly where they left off, with full context preserved.
- [Run agents from any framework on AgentOS](https://www.agno.com/articles/run-agents-from-any-framework-on-agentos.md) (SDK v2.6.0): AgentOS now runs agents built with the Claude Agent SDK, LangGraph, and DSPy alongside native Agno agents, all through a unified AgentProtocol interface.

## 2026-04-14

- [Authenticate MCP sessions correctly from the first request](https://www.agno.com/articles/authenticate-mcp-sessions-correctly-from-the-first-request.md) (SDK v2.5.17): We fixed an issue where headers supplied by header_provider were not being applied during MCP session initialization, only during subsequent requests.
- [Disable Claude file citations when they aren't needed](https://www.agno.com/articles/disable-claude-file-citations-when-they-arent-needed.md) (SDK v2.5.17): A new option lets you turn off file citations in Claude responses. This is useful when citations add noise to the output, for example in conversational flows, summarization tasks, or any context where…
- [Eliminate connection errors in concurrent workloads by isolating HTTP/2 clients](https://www.agno.com/articles/eliminate-connection-errors-in-concurrent-workloads-by-isolating-http-2-clients.md) (SDK v2.5.17): We fixed an issue where a shared HTTP/2 client was being injected across all model providers, causing connection conflicts and transient failures under concurrent load.
- [Ensure memory summarization runs reliably across all conversation shapes](https://www.agno.com/articles/ensure-memory-summarization-runs-reliably-across-all-conversation-shapes.md) (SDK v2.5.17): We fixed an issue where the memory pipeline gate check did not account for extra_messages, causing memory summarization to be skipped in runs where additional context messages were provided alongside…
- [Get a fully operational agent immediately after configuration without extra steps](https://www.agno.com/articles/get-a-fully-operational-agent-immediately-after-configuration-without-extra-steps.md) (SDK v2.5.17): We fixed an issue where knowledge databases were not being built live during configuration API calls, causing agents to operate without their knowledge base until a separate build step was triggered.
- [Keep custom database table names intact when reloading agent configurations](https://www.agno.com/articles/keep-custom-database-table-names-intact-when-reloading-agent-configurations.md) (SDK v2.5.17): We fixed an issue where custom db table names set on components were being overwritten with defaults when those components were loaded back from configuration.
- [Keep framework injected parameters out of user facing input schemas](https://www.agno.com/articles/keep-framework-injected-parameters-out-of-user-facing-input-schemas.md) (SDK v2.5.17): We fixed an issue where parameters automatically injected by the framework, such as agent, team, and run_context, were appearing in user_input_schema, presenting users with fields they should never…
- [Keep streaming responses clean when client connections are cancelled](https://www.agno.com/articles/keep-streaming-responses-clean-when-client-connections-are-cancelled.md) (SDK v2.5.17): We fixed an issue where cancellation of a client connection during streaming could surface as an unhandled error rather than being handled quietly.
- [Stop losing code blocks in structured outputs during JSON parsing](https://www.agno.com/articles/stop-losing-code-blocks-in-structured-outputs-during-json-parsing.md) (SDK v2.5.17): We fixed an issue where JSON cleaning was stripping or corrupting code blocks embedded in string values before the parse was even attempted.
- [Track nested workflow events accurately with correct identity and depth](https://www.agno.com/articles/track-nested-workflow-events-accurately-with-correct-identity-and-depth.md) (SDK v2.5.17): We fixed an issue where events emitted by inner workflows could lose their identity or be misattributed when bubbling up through outer workflows.
- [Work across multiple GitHub repositories without reconfiguring your agent](https://www.agno.com/articles/work-across-multiple-github-repositories-without-reconfiguring-your-agent.md) (SDK v2.5.17): GitHubConfig now accepts a repository override at the request level, allowing agents that work across multiple repositories to specify the target repo per call rather than being locked to a single…

## 2026-04-10

- [Connect agents directly to Salesforce CRM with SalesforceTools](https://www.agno.com/articles/connect-agents-directly-to-salesforce-crm-with-salesforcetools.md) (SDK v2.5.16): SalesforceTools gives agents native access to Salesforce CRM data, making it straightforward to build agents that query records, surface pipeline information, triage support cases, or answer questions…
- [Deploy Claude on Azure AI Foundry with a new dedicated model provider](https://www.agno.com/articles/deploy-claude-on-azure-ai-foundry-with-a-new-dedicated-model-provider.md) (SDK v2.5.16): A new Azure AI Foundry Claude model provider gives teams a first-class way to run Claude models through Microsoft's Azure AI infrastructure, with the same configuration patterns used across other Agno…
- [Get accurate knowledge retrieval when using separate agent and content databases](https://www.agno.com/articles/get-accurate-knowledge-retrieval-when-using-separate-agent-and-content-databases.md) (SDK v2.5.16): We fixed an issue where knowledge_table was being read from agent.db instead of contents_db, causing knowledge lookups to fail or return incorrect results when the two databases were configured…
- [Ingest LLM-friendly documentation from any site that supports the llms.txt standard](https://www.agno.com/articles/ingest-llm-friendly-documentation-from-any-site-that-supports-the-llms-txt-standard.md) (SDK v2.5.16): LLMsTxtTools and LLMsTxtReader add native support for the llms.txt standard — a Markdown-based file that websites publish at /llms.txt to provide LLMs with a concise, structured index of their…
- [Keep source data intact when loading team sessions](https://www.agno.com/articles/keep-source-data-intact-when-loading-team-sessions.md) (SDK v2.5.16): We fixed TeamSession.from_dict() so it no longer mutates the input mapping it receives. Previously, loading a team session from a dictionary could silently modify the original data structure, causing…
- [Pass images through workflow steps reliably regardless of how they're referenced](https://www.agno.com/articles/pass-images-through-workflow-steps-reliably-regardless-of-how-theyre-referenced.md) (SDK v2.5.16): We fixed an issue where workflow steps that included file path images were not being converted correctly, causing those images to be dropped or mishandled when passed between steps.
- [Run long OpenAI Responses API tasks in the background without holding an open connection](https://www.agno.com/articles/run-long-openai-responses-api-tasks-in-the-background-without-holding-an-open-connection.md) (SDK v2.5.16): OpenAIResponses now supports background mode for the OpenAI Responses API, allowing long-running agent tasks to execute asynchronously without holding an open connection.
- [See reasoning steps in real time in the AG-UI interface](https://www.agno.com/articles/see-reasoning-steps-in-real-time-in-the-ag-ui-interface.md) (SDK v2.5.16): We fixed two issues in the AG-UI interface: reasoning events are now correctly emitted as they occur so users can follow the model's thinking in real time, and input_content now stores the current…
- [Stream reasoning content as it arrives with OpenAI Responses](https://www.agno.com/articles/stream-reasoning-content-as-it-arrives-with-openai-responses.md) (SDK v2.5.16): We fixed handling of response.reasoning_summary_text.delta events in OpenAIResponses so that reasoning content is streamed incrementally as it is generated rather than being dropped or buffered.

## 2026-04-09

- [Compose complex pipelines with nested workflows as workflow steps](https://www.agno.com/articles/compose-complex-pipelines-with-nested-workflows-as-workflow-steps.md) (SDK v2.5.15): A Workflow can now be used directly as a step inside another workflow, enabling modular composition of reusable sub-pipelines.
- [Control session summary scope with last_n_runs and conversation_limit on SessionSummaryManager](https://www.agno.com/articles/control-session-summary-scope-with-last-n-runs-and-conversation-limit-on-sessionsummarymanager.md) (SDK v2.5.15): SessionSummaryManager now exposes last_n_runs and conversation_limit parameters, giving precise control over how much of the conversation history is fed into summary generation.
- [Eliminate duplicate messages from team conversation history](https://www.agno.com/articles/eliminate-duplicate-messages-from-team-conversation-history.md) (SDK v2.5.15): Resolved a bug where TeamSession.get_messages could return the same message more than once, causing downstream logic that relies on message history to process duplicates.
- [Enable full tracebacks in logs with AGNO_LOG_TRACEBACKS](https://www.agno.com/articles/enable-full-tracebacks-in-logs-with-agno-log-tracebacks.md) (SDK v2.5.15): A new AGNO_LOG_TRACEBACKS environment variable (opt-in) enables full Python tracebacks in log_error and log_warning calls.
- [Extend teams with Skills support](https://www.agno.com/articles/extend-teams-with-skills-support.md) (SDK v2.5.15): Skills—reusable, instruction-based capability modules—can now be attached to Teams directly via the skills parameter, giving the team leader access to domain expertise without delegating to a member…
- [Fetch pull requests reliably regardless of repository size](https://www.agno.com/articles/fetch-pull-requests-reliably-regardless-of-repository-size.md) (SDK v2.5.15): Resolved a crash in GitHubTools where get_pull_requests would raise an IndexError if the repository contained fewer pull requests than the specified limit.
- [Pause workflow steps after execution for human review with output review](https://www.agno.com/articles/pause-workflow-steps-after-execution-for-human-review-with-output-review.md) (SDK v2.5.15): Workflows can now pause after a step completes and wait for a human to inspect the output before it flows to the next step.
- [Reduce transient 400 errors in concurrent OpenAI and Azure OpenAI workloads](https://www.agno.com/articles/reduce-transient-400-errors-in-concurrent-openai-and-azure-openai-workloads.md) (SDK v2.5.15): Resolved an issue where a shared HTTP/2 client was being injected across concurrent OpenAI and Azure OpenAI requests, causing transient 400 errors under load.
- [Search Wikipedia reliably with disambiguation handling and configurable suggestions](https://www.agno.com/articles/search-wikipedia-reliably-with-disambiguation-handling-and-configurable-suggestions.md) (SDK v2.5.15): Resolved an unhandled DisambiguationError that caused WikipediaTools to crash when a search term matched multiple Wikipedia articles.
- [See audio token usage in run metrics for OpenAI, Perplexity, and LiteLLM](https://www.agno.com/articles/see-audio-token-usage-in-run-metrics-for-openai-perplexity-and-litellm.md) (SDK v2.5.15): audio_total_tokens is now correctly computed and included in run metrics for OpenAI, Perplexity, and LiteLLM.
- [Upload .msg, .xlsx, and .xls files to AgentOS without workarounds](https://www.agno.com/articles/upload-msg-xlsx-and-xls-files-to-agentos-without-workarounds.md) (SDK v2.5.15): Resolved an issue where .msg, .xlsx, and .xls files were not recognized on upload due to missing MIME type mappings. These file types now upload correctly without requiring manual workarounds.

## 2026-04-02

- [Authenticate with Azure Blob Storage using SAS tokens](https://www.agno.com/articles/authenticate-with-azure-blob-storage-using-sas-tokens.md) (SDK v2.5.14): AzureBlobConfig now supports Shared Access Signature (SAS) token authentication as an alternative to connection strings and service principal credentials.
- [Run Claude 4.6+ reliably across Anthropic, Bedrock, Vertex AI, and LiteLLM](https://www.agno.com/articles/fix-assistant-message-prefill-compatibility-for-claude-4-6-across-all-providers.md) (SDK v2.5.14): Claude 4.6 and later models do not support assistant message prefill, which previously caused silent failures or malformed requests when conversations ended with an assistant turn.
- [Keep agents running through provider failures with automatic fallback models](https://www.agno.com/articles/keep-agents-running-through-provider-failures-with-automatic-fallback-models.md) (SDK v2.5.14): Agents and teams can now be configured with fallback models that activate automatically when the primary model fails, whether from rate limits, outages, context window overflows, or other retryable…
- [Search across your Slack workspace from within an agent using SlackTools](https://www.agno.com/articles/search-across-your-slack-workspace-from-within-an-agent-using-slacktools.md) (SDK v2.5.14): SlackTools now includes a workspace search tool, letting agents query messages, files, and content across channels directly from a tool call.

## 2026-03-31

- [Improved reliability for large ChromaDB operations](https://www.agno.com/articles/improved-reliability-for-large-chromadb-operations.md) (SDK v2.5.13): We’ve made ChromaDB operations more reliable by automatically splitting large upsert and query requests into smaller batches at runtime.
- [Inspect instance scale at a glance with the new AgentOS /info endpoint](https://www.agno.com/articles/inspect-instance-scale-at-a-glance-with-the-new-agentos-info-endpoint.md) (SDK v2.5.13): A new /info API endpoint returns a lightweight count of agents, teams, and workflows registered in the AgentOS instance.
- [Respect chunk_size in default chunking strategies across reader classes](https://www.agno.com/articles/respect-chunk-size-in-default-chunking-strategies-across-reader-classes.md) (SDK v2.5.13): Reader classes now correctly propagate the chunk_size parameter to the default chunking strategy they apply when no explicit chunking configuration is provided.
- [Return richer session data from the AgentOS /sessions list API](https://www.agno.com/articles/return-richer-session-data-from-the-agentos-sessions-list-api.md) (SDK v2.5.13): The /sessions list endpoint now includes a significantly expanded set of fields per session, giving dashboards, monitoring tools, and integrations a more complete picture of each session without…
- [Strengthen agent evaluation with subset matching, argument validation, and improved tool call tracking in ReliabilityEval](https://www.agno.com/articles/strengthen-agent-evaluation-with-subset-matching-argument-validation-and-improved-tool-call-tracking-in-reliabilityeval.md) (SDK v2.5.13): ReliabilityEval has been extended with more precise evaluation capabilities: expected tool calls can now be matched as a subset of actual calls rather than requiring an exact full match, argument…
- [Surface member tool calls and handle long responses gracefully in the Slack interface](https://www.agno.com/articles/surface-member-tool-calls-and-handle-long-responses-gracefully-in-the-slack-interface.md) (SDK v2.5.13): Two improvements have been made to the Slack interface to give teams better visibility and more robust handling of long agent responses.

## 2026-03-30

- [Convert any document format directly from an agent with the new DoclingTools toolkit](https://www.agno.com/articles/convert-any-document-format-directly-from-an-agent-with-the-new-doclingtools-toolkit.md) (SDK v2.5.12): DoclingTools gives agents the ability to convert documents on demand using the Docling library — accepting PDFs, DOCX, PPTX, XLSX, HTML, images, audio, and video files as input and exporting to…
- [Restore full end-to-end reliability for Coda-integrated agents](https://www.agno.com/articles/fix-multiple-coda-integration-issues-across-tools-slack-team-streaming-and-learning.md) (SDK v2.5.12): Resolved a collection of bugs affecting agents deployed with Coda, including issues in CodingTools, Slack interface behavior, team streaming output, and the learning pipeline.
- [Keep Claude on track across multi-turn conversations with server tool results](https://www.agno.com/articles/fix-server-tool-blocks-being-dropped-from-claude-conversation-history.md) (SDK v2.5.12): Resolved an issue where server-side tool blocks in Claude conversations were not being preserved when building subsequent request messages.
- [Deliver long Slack responses without crashes or silent failures](https://www.agno.com/articles/fix-slack-streaming-path-crashing-on-messages-that-exceed-the-length-limit.md) (SDK v2.5.12): Resolved an unhandled msg_too_long error in the Slack streaming path that caused the agent to fail silently or crash when a streamed response exceeded Slack's message length limit.
- [Manage cron schedules directly from an agent with the new SchedulerTools toolkit](https://www.agno.com/articles/manage-cron-schedules-directly-from-an-agent-with-the-new-schedulertools-toolkit.md) (SDK v2.5.12): SchedulerTools Gives agents programmatic control over the AgentOS Scheduler, allowing them to create, list, update, enable, disable, trigger, and delete cron schedules as part of a run.

## 2026-03-26

- [Deduplicate hybrid search results in LanceDB](https://www.agno.com/articles/deduplicate-hybrid-search-results-in-lancedb.md) (SDK v2.5.11): Resolved an issue where LanceDB's search() could return the same document multiple times when hybrid search retrieved it via both vector similarity and full-text search.
- [Catch unsupported AWS Bedrock authentication for Claude before requests fail](https://www.agno.com/articles/fail-fast-for-unsupported-aws-bedrock-api-key-with-claude-on-aws-bedrock.md) (SDK v2.5.11): We added an early error when AWS_BEDROCK_API_KEY is set for Claude models on AWS Bedrock, which is not a supported authentication path, rather than failing silently later in the request lifecycle.
- [Give teams full visibility into async toolkit tools at the prompt level](https://www.agno.com/articles/fix-async-toolkit-tool-names-missing-from-team-system-message.md) (SDK v2.5.11): We resolved an issue where tools from async toolkits were not included in the tool name list injected into the team system message, leaving the team unaware of those tools at the prompt level.
- [Keep Azure OpenAI connections intact through agent and team setup](https://www.agno.com/articles/fix-azure-openai-client-references-lost-after-deepcopy.md) (SDK v2.5.11): We overrode deepcopy behavior on the Azure OpenAI model class to preserve live client references, preventing connection failures that occurred when the model object was copied during agent or team…
- [Cache responses reliably across a wider range of agent inputs](https://www.agno.com/articles/fix-cache-key-generation-for-non-serializable-types-in-response-caching.md) (SDK v2.5.11): We resolved a failure in cache key generation when the input contained types that are not directly JSON-serializable, ensuring caching works reliably across a broader range of agent inputs.
- [Get clean, deduplicated results from LanceDB hybrid search](https://www.agno.com/articles/fix-duplicate-results-in-lancedb-hybrid-search.md) (SDK v2.5.11): We resolved an additional case where hybrid search could surface the same document more than once when it matched across multiple search indices.
- [Preserve tool names throughout LiteLLM streaming responses](https://www.agno.com/articles/fix-litellm-overwriting-tool-names-with-empty-strings-during-streaming.md) (SDK v2.5.11): We resolved an issue where empty string values in streamed LiteLLM responses could overwrite previously accumulated tool names, resulting in tool calls with missing identifiers.
- [Restore correct output config, tool schemas, and streaming behavior for Claude](https://www.agno.com/articles/fix-several-claude-specific-issues-across-output-config-tool-schemas-and-streaming.md) (SDK v2.5.11): We fixed output_config not being applied correctly on Claude model wrappers, $defs being stripped from tool schemas, and file_ids and container information not being surfaced during streaming for…
- [Dispatch complete, accurate tool calls during streaming](https://www.agno.com/articles/fix-tool-call-streaming-using-assign-instead-of-append-in-parse-tool-calls.md) (SDK v2.5.11): We resolved a bug where streamed tool call data was overwriting accumulated state instead of appending to it, causing incomplete or incorrect tool calls to be dispatched.
- [New GoogleSlidesTools toolkit](https://www.agno.com/articles/new-googleslidestools-toolkit.md) (SDK v2.5.11): We’ve introduced GoogleSlidesTools to give agents full control over Google Slides. With it, you can create presentations, build out slides, and manage content end to end, all directly from your agent.
- [Search the web with recency and domain filtering using the new PerplexitySearch toolkit](https://www.agno.com/articles/search-the-web-with-recency-and-domain-filtering-using-the-new-perplexitysearch-toolkit.md) (SDK v2.5.11): A new PerplexitySearch toolkit gives agents access to the Perplexity Search API, returning ranked web results with titles, URLs, snippets, and publication dates in a single tool call.
- [Keep OpenRouter responses clean for non-reasoning models](https://www.agno.com/articles/skip-empty-reasoning-blocks-for-non-reasoning-models-on-openrouter.md) (SDK v2.5.11): We resolved an issue where empty reasoning blocks returned by OpenRouter for non-reasoning models were being processed unnecessarily, causing noise in parsed responses.
- [Swap models without rewriting tool definitions with cross-model tool call compatibility](https://www.agno.com/articles/swap-models-without-rewriting-tool-definitions-with-cross-model-tool-call-compatibility.md) (SDK v2.5.11): Tool call schemas are now normalized across model providers, so switching an agent from one model to another no longer requires adjusting how tools are defined or how their outputs are parsed.
- [Tailor document segmentation with custom prompt support on AgenticChunking](https://www.agno.com/articles/tailor-document-segmentation-with-custom-prompt-support-on-agenticchunking.md) (SDK v2.5.11): AgenticChunking now accepts a custom_prompt parameter, letting you override the default model-driven chunking instructions with domain-specific logic.
- [Update Seltz toolkit to SDK v0.2.0](https://www.agno.com/articles/update-seltz-toolkit-to-sdk-v0-2-0.md) (SDK v2.5.11): The Seltz toolkit has been updated to align with the breaking changes introduced in the Seltz SDK 0.2.0 release, replacing the previous 0.1.x integration.

## 2026-03-17

- [Upgrade to Mistral AI v2 SDK without changing your agent configuration](https://www.agno.com/articles/add-mistral-ai-v2-sdk-support-with-full-backward-compatibility.md) (SDK v2.5.10): The Mistral model provider now supports the mistralai v2 SDK while continuing to work with v1. Teams can upgrade their SDK dependency and take advantage of v2 improvements without any changes to their…
- [Cut search latency with native parallel AI search for Vertex AI](https://www.agno.com/articles/add-native-parallel-ai-search-for-vertex-ai-with-toolparallelaisearch.md) (SDK v2.5.10): AgentTools now includes ToolParallelAiSearch, a native integration with Vertex AI's Parallel AI Search that allows agents to issue multiple search queries concurrently and aggregate results.
- [Control Gemini request duration with a new timeout parameter](https://www.agno.com/articles/add-timeout-parameter-to-the-gemini-model-class.md) (SDK v2.5.10): The Gemini model class now accepts a timeout parameter, giving teams explicit control over how long a request is allowed to run before being cancelled.
- [Expand WhatsApp interface with media, interactive messages, team support, and encryption](https://www.agno.com/articles/expand-whatsapp-interface-with-media-interactive-messages-team-support-and-encryption.md) (SDK v2.5.10): The WhatsApp interface has been significantly extended in V2, adding support for rich media, interactive message types, teams, and workflows.
- [Expose agents, teams, and workflows as Telegram bots with the new Telegram interface and TelegramTools](https://www.agno.com/articles/expose-agents-teams-and-workflows-as-telegram-bots-with-the-new-telegram-interface-and-telegramtools.md) (SDK v2.5.10): The new Telegram interface mounts webhook endpoints directly on AgentOS, turning any agent, team, or workflow into a fully functional Telegram bot.
- [Fetch specific workflow versions and align run-level configuration with agents and teams](https://www.agno.com/articles/extend-workflow-api-with-version-fetching-and-run-level-parameter-parity.md) (SDK v2.5.10): The GET /workflows/{id} endpoint now accepts a version query parameter, allowing callers to fetch a specific version of a workflow rather than always receiving the latest.
- [Execute each tool call exactly once during streaming](https://www.agno.com/articles/fix-duplicate-tool-execution-in-streaming-mode-caused-by-shared-dict-references.md) (SDK v2.5.10): Resolved a bug in parse_tool_calls where shared dictionary references across parsed tool calls would cause the same tool to be executed multiple times during streaming.
- [Use MongoDB reliably in async agents and workflows](https://www.agno.com/articles/fix-incorrect-async-module-import-in-mongodb.md) (SDK v2.5.10): Resolved an incorrect import of the pymongo async modules that could cause runtime failures when using MongoDB with async agents or workflows.
- [Keep agents moving in parallel MCP tool calls without session conflicts](https://www.agno.com/articles/fix-race-condition-causing-duplicate-mcp-sessions-and-stuck-agents-in-parallel-tool-calls.md) (SDK v2.5.10): Resolved a race condition in MCPTools where parallel tool calls using a header_provider would each independently spin up their own MCP session instead of sharing one, leaving the agent in a stuck…
- [Get reliable structured output on all supported Claude models](https://www.agno.com/articles/fix-structured-output-detection-for-supported-claude-models.md) (SDK v2.5.10): Resolved an issue where structured output support was not correctly detected for certain Claude models, causing agents to fall back to less reliable output parsing strategies even when the model fully…
- [Get complete end-to-end agent trace visibility with MLflow](https://www.agno.com/articles/full-agent-trace-visibility-now-available-with-mlflow.md) (SDK v2.5.10): Production agent systems demand visibility. Agno now integrates with MLflow to deliver complete, end-to-end trace observability across every model call, tool invocation, and agent step—without custom…
- [Ingest any document format into your knowledge base with the Docling Reader](https://www.agno.com/articles/ingest-any-document-format-into-your-knowledge-base-with-the-docling-reader.md) (SDK v2.5.10): The DoclingReader provides a single, unified interface for processing the full range of document formats an AI agent encounters — PDFs, Word files, PowerPoint decks, Excel spreadsheets, images, and…

## 2026-03-10

- [Build smarter tool hooks with access to the full conversation history](https://www.agno.com/articles/access-full-message-history-inside-tool-hooks-for-richer-context-aware-logic.md) (SDK v2.5.9): Tool pre- and post-hooks, as well as agent-level tool_hooks, can now read the current run's complete message history via run_context.messages.
- [Complete multi-turn learning confirmations without losing context](https://www.agno.com/articles/auto-enable-chat-history-for-learningmode-propose-to-support-multi-turn-confirmation.md) (SDK v2.5.9): LearningMode.PROPOSE now automatically enables chat history for the session, ensuring that the multi-turn confirmation flow — where the agent proposes a learned fact and waits for user approval — has…
- [Followup suggestions in Agno: give users their next question](https://www.agno.com/articles/built-in-followup-suggestions-for-agents-and-teams.md) (SDK v2.5.9): Agno agents can suggest followup questions at the end of a response. Set followups=True on an Agent or Team and you get back a list of ready-to-run prompts on response.followups, built from the user's…
- [Automate more calendar workflows with expanded GoogleCalendarTools and service account auth](https://www.agno.com/articles/expand-calendar-automation-with-new-googlecalendartools-capabilities-and-service-account-auth.md) (SDK v2.5.9): GoogleCalendarTools has been extended with additional tools, a service account authentication path, and new cookbooks to help teams get started quickly.
- [Keep full conversation context across HITL multi-round runs](https://www.agno.com/articles/fix-add-history-to-context-support-in-hitl-multi-round-conversations.md) (SDK v2.5.9): Resolved a bug where add_history_to_context was not correctly applied during Human-in-the-Loop runs that involved multiple conversation rounds.
- [Connect to Siliconflow reliably with the corrected default base URL](https://www.agno.com/articles/fix-siliconflow-provider-base-url-to-use-correct-cn-domain.md) (SDK v2.5.9): Updated the default base_url for the Siliconflow model provider from .com to .cn to match Siliconflow's actual API endpoint.
- [Present datetime context in any format with datetime_format](https://www.agno.com/articles/format-datetime-context-to-match-your-locale-or-use-case-with-datetime-format.md) (SDK v2.5.9): A new datetime_format parameter on Agent and Team lets you control exactly how the current datetime is presented in the agent's context using any valid strftime format string.
- [Guide next steps with built-in followup suggestions for agents and teams](https://www.agno.com/articles/guide-next-steps-with-built-in-followup-suggestions-for-agents-and-teams.md) (SDK v2.5.9): Agents and teams can now automatically generate actionable followup prompts after each response by setting followups=True.
- [Keep tool schemas clean with accurate parameter descriptions](https://www.agno.com/articles/remove-spurious-none-prefix-from-tool-parameter-descriptions.md) (SDK v2.5.9): Fixed a formatting issue where tool parameter descriptions were incorrectly prefixed with (None) when no type annotation was present.

## 2026-03-06

- [Always receive generated media in run output, regardless of storage settings](https://www.agno.com/articles/always-include-generated-media-in-run-output-scrubbing-before-db-storage.md) (SDK v2.5.8): Images and audio generated during a run are now consistently included in run output regardless of the store_media setting.
- [Bring GitLab project and pipeline context into your agents in minutes](https://www.agno.com/articles/bring-gitlab-project-and-pipeline-context-into-your-agents-in-minutes.md) (SDK v2.5.8): Engineering and platform teams using GitLab can now connect agents directly to their repositories. GitlabTools brings read-focused GitLab access to Agno agents, covering projects, merge requests, and…
- [Automate the full Gmail workflow with new tools and service account authentication](https://www.agno.com/articles/expand-gmail-capabilities-with-new-tools-and-service-account-authentication.md) (SDK v2.5.8): GmailTools has been extended with a broader set of email management functions and a new service account authentication path.
- [Store structured agent state in MySQL without serialization errors](https://www.agno.com/articles/fix-json-serialization-for-mysql-database-connections.md) (SDK v2.5.8): A json_serializer is now passed during MySQL engine creation, ensuring that JSON fields are correctly serialized when reading from and writing to MySQL-backed databases.
- [Build on previous loop outputs with forward_iteration_output](https://www.agno.com/articles/fix-loop-iteration-input-chaining-and-add-opt-in-forwarding-via-forward-iteration-output.md) (SDK v2.5.8): Resolved a bug where each iteration of a Loop always received the original input rather than the output of the previous iteration, causing loops to repeat work instead of building on it.
- [Use external_execution and regular tools together in OpenAI Responses without conflicts](https://www.agno.com/articles/fix-tool-dispatch-when-mixing-external-execution-and-regular-tools-in-openai-responses.md) (SDK v2.5.8): Resolved a bug in OpenAIResponses where combining external_execution tools with standard tools caused incorrect dispatch behavior.
- [Simplify container deployments with environment variable support in serve()](https://www.agno.com/articles/simplify-container-deployments-with-environment-variable-support-in-serve.md) (SDK v2.5.8): serve() now reads AGENT_OS_HOST and AGENT_OS_PORT environment variables as fallbacks when explicit values are not passed.
- [Simplify debugging with human-readable agent and team identifiers](https://www.agno.com/articles/simplify-debugging-with-human-readable-agent-and-team-identifiers.md) (SDK v2.5.8): Agents and teams are now assigned human-readable IDs (e.g., brave-falcon-7x3k) instead of raw UUIDs.

## 2026-03-04

- [Improve session recall accuracy](https://www.agno.com/articles/improve-session-recall-accuracy.md) (SDK v2.5.7): The built-in session search tool has been upgraded from a single-pass lookup to a two-step process: the agent first calls search_past_sessions() to retrieve lightweight previews of recent sessions…

## 2026-03-02

- [Accelerate debugging with powerful, composable trace filtering](https://www.agno.com/articles/accelerate-debugging-with-powerful-composable-trace-filtering.md) (SDK v2.5.6): AgentOS now supports an advanced filtering DSL for traces, letting you construct precise, composable queries to isolate specific runs, models, components, or behaviors.
- [Broaden mobile capture by accepting native iOS image uploads](https://www.agno.com/articles/broaden-mobile-capture-by-accepting-native-ios-image-uploads.md) (SDK v2.5.6): File upload endpoints now accept image/heic and image/heif formats, removing the need to convert Apple-native image formats before ingestion.
- [Broader schema coverage with Literal type support](https://www.agno.com/articles/broader-schema-coverage-with-literal-type-support.md) (SDK v2.5.6): JSON schema generation now handles Literal types, ensuring that agents and tools using constrained value sets produce valid, complete schemas.
- [Cleaner Google tool imports with a unified sub-package](https://www.agno.com/articles/cleaner-google-tool-imports-with-a-unified-sub-package.md) (SDK v2.5.6): Google tools have been restructured into a dedicated agno.tools.google sub-package (e.g., from agno.tools.google import GmailTools).
- [Control in-flight agent runs with approval status tracking and admin enforcement](https://www.agno.com/articles/control-in-flight-agent-runs-with-approval-status-tracking-and-admin-enforcement.md) (SDK v2.5.6): A new approval status endpoint lets you query where a paused run stands in the approval process, and admin-gated enforcement ensures that only authorized users can continue execution.
- [Process file inputs directly with OpenAI Responses](https://www.agno.com/articles/process-file-inputs-directly-with-openai-responses.md) (SDK v2.5.6): OpenAIResponses now supports input_file, letting you pass files directly into OpenAI Responses API calls.
- [Reliable file search results with vector store polling fix](https://www.agno.com/articles/reliable-file-search-results-with-vector-store-polling-fix.md) (SDK v2.5.6): We resolved a race condition in OpenAI Responses where file_search could silently return empty results due to eventual consistency in OpenAI's vector store file listing API.
- [Render autonomous team tasks in real time with structured streaming events](https://www.agno.com/articles/render-autonomous-team-tasks-in-real-time-with-structured-streaming-events.md) (SDK v2.5.6): A new approval status endpoint lets you query where a paused run stands in the approval process, and admin-gated enforcement ensures that only authorized users can continue execution.
- [Securely ingest private GitHub content using GitHub App authentication](https://www.agno.com/articles/securely-ingest-private-github-content-using-github-app-authentication.md) (SDK v2.5.6): Knowledge sources now support GitHub App authentication (app_id, installation_id, private_key) in addition to personal access tokens.

## 2026-02-25

- [Expand creative automation with built-in text-to-image generation](https://www.agno.com/articles/expand-creative-automation-with-built-in-text-to-image-generation.md) (SDK v2.5.5): ModelsLabTools now supports text-to-image generation with PNG/JPG outputs, an image fetch endpoint, and sizing options.
- [Ship responsive Slack assistants with real-time streaming and per-workspace isolation](https://www.agno.com/articles/ship-responsive-slack-assistants-with-real-time-streaming-and-per-workspace-isolation.md) (SDK v2.5.5): Slack integrations now support real-time streaming with live progress cards, so end users see assistant activity as it happens rather than waiting for a final response.

## 2026-02-24

- [Enforce human checkpoints wherever you need them with Step-level HITL](https://www.agno.com/articles/enforce-human-checkpoints-wherever-you-need-them-with-step-level-hitl.md) (SDK v2.5.4): Workflows now support Human-in-the-Loop at the individual Step level, letting you pause execution to collect confirmation or user input before proceeding.
- [Update: PgVector now supports similarity_threshold!](https://www.agno.com/articles/filter-pgvector-results-by-minimum-similarity-score.md) (SDK v2.5.4): Stop returning low-quality matches. Set a quality floor on your vector search results so your agents only work with context that's actually relevant.
- [Keep your Seltz-powered search integrations current with SDK 0.1.3](https://www.agno.com/articles/seltztools-updated-to-seltz-sdk-0-1-3.md) (SDK v2.5.4): SeltzTools now uses Seltz SDK 0.1.3, incorporating the latest fixes and improvements from the upstream SDK.
- [Update: TeamMode.tasks now supports streaming events!](https://www.agno.com/articles/stream-real-time-events-during-autonomous-team-task-execution.md) (SDK v2.5.4): Stream real-time events during autonomous team task execution.
- [Track cost, tokens, and latency per model and per component](https://www.agno.com/articles/track-cost-tokens-and-latency-per-model-and-per-component.md) (SDK v2.5.4): The metrics system has been redesigned to provide granular, per-model and per-component tracking across the entire agent, team, and workflow lifecycle.

## 2026-02-19

- [Cleaner PDF text extraction reduces post-processing](https://www.agno.com/articles/cleaner-pdf-text-extraction-reduces-post-processing.md) (SDK v2.5.3): PDF extraction now includes a sanitize_content option (default: True) to normalize fragmented text by collapsing excessive whitespace.
- [Faster knowledge onboarding with native cloud and repository connectors](https://www.agno.com/articles/faster-knowledge-onboarding-with-native-cloud-and-repository-connectors.md) (SDK v2.5.3): We’ve added remote content sources including Amazon S3, Google Cloud Storage, Azure Blob, GitHub, and SharePoint, along with new APIs to list sources and browse files before ingestion.
- [Reliable access to workflow-generated files via API outputs](https://www.agno.com/articles/reliable-access-to-workflow-generated-files-via-api-outputs.md) (SDK v2.5.3): WorkflowRunOutput now exposes a files field and uses consistent JSON (de)serialization. This fixes prior serialization errors and makes file artifacts first-class, enabling teams to programmatically…

## 2026-02-17

- [Define tools, knowledge, and team members as callable factories for runtime flexibility](https://www.agno.com/articles/define-tools-knowledge-and-team-members-as-callable-factories-for-runtime-flexibility.md) (SDK v2.5): Tools, knowledge sources, and team members can now be defined as callable factories that are resolved at runtime.
- [Deploy agents on AWS with native EFS support](https://www.agno.com/articles/deploy-agents-on-aws-with-native-efs-support.md) (SDK v2.5): Agno infrastructure now supports AWS Elastic File System (EFS) for persistent, shared storage across agent deployments.
- [Enforce human approval gates in agentic workflows](https://www.agno.com/articles/enforce-human-approval-gates-in-agentic-workflows.md) (SDK v2.5): A new approval system lets you require human sign-off before agents execute sensitive actions. Using the @approval decorator alongside a HITL primitive (requires_confirmation, requires_user_input, or…
- [History messages are no longer stored by default](https://www.agno.com/articles/history-messages-are-no-longer-stored-by-default.md) (SDK v2.5): store_history_messages now defaults to False. Previously, agent run history messages were stored automatically.
- [Improved database integrity for sessions and component configuration](https://www.agno.com/articles/improved-database-integrity-for-sessions-and-component-configuration.md) (SDK v2.5): We updated the sessions, component configurations, and component links tables to use proper primary key constraints, replacing the previous unique constraint approach.
- [Give your teams persistent organizational memory with LearningMachine](https://www.agno.com/articles/learningmachine-now-available-for-teams-not-just-individual-agents.md) (SDK v2.5): Teams can accumulate and persist knowledge over time, improving their responses and decisions across sessions and runs.
- [Share vector databases across agents with isolated search results](https://www.agno.com/articles/share-vector-databases-across-agents-with-isolated-search-results.md) (SDK v2.5): A new isolate_vector_search option on the Knowledge class lets multiple agents or teams share the same vector database while keeping their search results isolated.

## 2026-02-12

- [Orchestrate multi-agent teams with four built-in execution modes](https://www.agno.com/articles/orchestrate-multi-agent-teams-with-four-built-in-execution-modes.md) (SDK v2.5): Teams now support four distinct execution modes (coordinate, route, broadcast, and tasks), giving you explicit control over how agents collaborate.
- [Richer human-in-the-loop controls for multi-agent teams](https://www.agno.com/articles/richer-human-in-the-loop-controls-for-multi-agent-teams.md) (SDK v2.5): The human-in-the-loop (HITL) system for teams has been significantly expanded. New run requirements support tool confirmation, user input collection, and external tool execution, giving you…
- [Schedule agents, teams, and workflows with cron-based automation](https://www.agno.com/articles/schedule-agents-teams-and-workflows-with-cron-based-automation.md) (SDK v2.5): Agno now includes a built-in scheduler for running agents, teams, and workflows on a recurring basis. Define cron schedules with support for retries, timeouts, and timezone configuration.

## 2026-02-03

- [Evaluate workflow logic with portable, serializable expressions](https://www.agno.com/articles/evaluate-workflow-logic-with-portable-serializable-expressions.md) (SDK v2.4.8): Workflow Condition, Loop, and Router steps now support CEL (Common Expression Language) evaluators. Expressions are defined as strings, making workflows fully serializable and easier to store, review…
- [Neosantara added as a supported model provider](https://www.agno.com/articles/neosantara-added-as-a-supported-model-provider.md) (SDK v2.4.8): Agno now supports Neosantara, an Indonesian LLM gateway that provides an OpenAI-compatible API. This allows teams to use Neosantara models without changing existing agent or workflow integrations.
- [Simpler, more flexible routing for complex workflows](https://www.agno.com/articles/simpler-more-flexible-routing-for-complex-workflows.md) (SDK v2.4.8): The Workflow Router step now supports returning the name of a step instead of the step object itself, and can route to a group of steps as a single choice.

## 2026-01-29

- [Enforce human approval for MCP tool calls](https://www.agno.com/articles/enforce-human-approval-for-mcp-tool-calls.md) (SDK v2.4.7): Human-in-the-loop confirmation now correctly applies to MCP Function tools via toolkit-level settings (e.g., requires_confirmation_tools).
- [Expand multimodal search with Cohere v4 embeddings on Bedrock](https://www.agno.com/articles/expand-multimodal-search-with-cohere-v4-embeddings-on-bedrock.md) (SDK v2.4.7): AwsBedrockEmbedder now supports Cohere Embed v4, including configurable output dimensions and multimodal (text + image) embeddings, with async variants.
- [Improve RAG answer quality with Bedrock-powered reranking](https://www.agno.com/articles/improve-rag-answer-quality-with-bedrock-powered-reranking.md) (SDK v2.4.7): Introduce smarter retrieval with the new AwsBedrockReranker, supporting Cohere Rerank 3.5 and Amazon Rerank 1.0.
- [Keep LanceDB integrations current with 0.26.0 support](https://www.agno.com/articles/keep-lancedb-integrations-current-with-0-26-0-support.md) (SDK v2.4.7): Agno now aligns with LanceDB’s latest API, replacing the deprecated table_names() with list_tables() and updating the minimum LanceDB version to 0.26.0.
- [Reduce workflow boilerplate with native if/else branching](https://www.agno.com/articles/reduce-workflow-boilerplate-with-native-if-else-branching.md) (SDK v2.4.7): Condition steps now support else_steps, allowing you to define a clear alternative path when a condition evaluates to false.

## 2026-01-28

- [Breaking change: website crawling uses per-page content hashes](https://www.agno.com/articles/breaking-change-website-crawling-uses-per-page-content-hashes.md) (SDK v2.4.6): We changed the WebsiteReader deduplication model to compute content hashes per page. This aligns skip_if_exists with page-level updates and ensures accurate re-crawls.
- [Consistent async file reads: empty files now return no documents](https://www.agno.com/articles/consistent-async-file-reads-empty-files-now-return-no-documents.md) (SDK v2.4.6): Async text_reader.aread() now returns an empty list ([]) for empty files, aligning behavior with the sync API. This removes special-case handling and simplifies downstream pipelines.
- [Learning agents work out of the box with default memory and clearer guidance](https://www.agno.com/articles/learning-agents-work-out-of-the-box-with-default-memory-and-clearer-guidance.md) (SDK v2.4.6): Learning is now simpler and more effective. When learning=True, user memory is enabled by default, and the LearnedKnowledgeStore captures organizational context (goals, constraints, policies) to guide…
- [Reliable website deduplication with per-page content hashing](https://www.agno.com/articles/reliable-website-deduplication-with-per-page-content-hashing.md) (SDK v2.4.6): WebsiteReader now computes a unique content hash per crawled URL, fixing skip_if_exists for multi-page crawls.

## 2026-01-27

- [Accelerate semantic search with native Seltz integration](https://www.agno.com/articles/accelerate-semantic-search-with-native-seltz-integration.md) (SDK v2.4.5): We’ve added a first-class SeltzTools toolkit that brings Seltz-powered semantic search directly into Agno.

## 2026-01-26

- [Accelerate image-rich apps with built-in Unsplash search and retrieval](https://www.agno.com/articles/accelerate-image-rich-apps-with-built-in-unsplash-search-and-retrieval.md) (SDK v2.4.4): We added UnsplashTools, a first-class toolkit for discovering and retrieving high-quality, royalty-free images directly in Agno.
- [Expand model choice with native Moonshot.ai provider](https://www.agno.com/articles/expand-model-choice-with-native-moonshot-ai-provider.md) (SDK v2.4.4): Agno now supports Moonshot.ai as a model provider with initial models and examples to help you get started quickly.

## 2026-01-23

- [Native Excel file ingestion with sheet-level controls and automatic routing](https://www.agno.com/articles/native-excel-file-ingestion-with-sheet-level-controls-and-automatic-routing.md) (SDK v2.4.3): We introduced a dedicated ExcelReader for .xls/.xlsx with sheet filtering, options to skip hidden sheets, and chunking controls. ReaderFactory now routes Excel files to ExcelReader automatically.

## 2026-01-22

- [Ensure reliable table creation with async database drivers](https://www.agno.com/articles/ensure-reliable-table-creation-with-async-database-drivers.md) (SDK v2.4.2): A fix restores reliable table creation across AsyncSQLiteDb, AsyncPostgresDb, AsyncMySQLDb, and FirestoreDb.
- [Simplify enterprise data ingestion with Azure Blob Storage](https://www.agno.com/articles/simplify-enterprise-data-ingestion-with-azure-blob-storage.md) (SDK v2.4.2): Knowledge now connects to private Azure Blob Storage as a first-class source — alongside SharePoint and GitHub — so Azure-centric organizations can centralize content without custom ETL.
- [Standardize model integrations with OpenAI Responses API compatibility](https://www.agno.com/articles/standardize-model-integrations-with-openai-responses-api-compatibility.md) (SDK v2.4.2): We introduced OpenAI Responses API–compatible clients, including a base OpenResponses and provider-specific clients for Ollama and OpenRouter.

## 2026-01-21

- [Accurate cost and usage reporting for Perplexity streaming](https://www.agno.com/articles/accurate-cost-and-usage-reporting-for-perplexity-streaming.md) (SDK v2.4.1): We corrected streaming token accounting for Perplexity by collecting usage only on the final chunk for providers that return cumulative metrics.
- [Bring Excel data into Knowledge with first-class .xlsx/.xls ingestion](https://www.agno.com/articles/bring-excel-data-into-knowledge-with-first-class-xlsx-xls-ingestion.md) (SDK v2.4.1): Knowledge now natively ingests Excel files by routing spreadsheets through the CSV reader. Each sheet is parsed into its own document with sheet-level metadata and normalized cell content.
- [Consistent error handling for async tools improves reliability](https://www.agno.com/articles/consistent-error-handling-for-async-tools-improves-reliability.md) (SDK v2.4.1): Async generator tools now capture and surface errors on the tool call —matching synchronous behavior — instead of re-raising exceptions.
- [Expand model choice with N1N, an OpenAI-compatible provider](https://www.agno.com/articles/expand-model-choice-with-n1n-an-openai-compatible-provider.md) (SDK v2.4.1): We’ve added n1n.ai as an OpenAI-compatible provider, giving teams more flexibility to optimize for cost, performance, and regional availability.
- [Securely ingest private GitHub and SharePoint content into Knowledge](https://www.agno.com/articles/securely-ingest-private-github-and-sharepoint-content-into-knowledge.md) (SDK v2.4.1): Knowledge can now ingest content from private GitHub repositories and SharePoint, via SDK and API. This enables organizations to consolidate code, docs, and operational knowledge from private systems…

## 2026-01-19

- [Centralize agent, team, and workflow configuration with AgentOS CRUD APIs](https://www.agno.com/articles/centralize-agent-team-and-workflow-configuration-with-agentos-crud-apis.md) (SDK v2.4.0): You can now persist and manage Agent, Team, and Workflow definitions in a database, with new AgentOS endpoints for programmatic create, read, update, and delete.
- [Clearer Knowledge APIs with insert/insert_many rename](https://www.agno.com/articles/clearer-knowledge-apis-with-insert-insert-many-rename.md) (SDK v2.4.0): Knowledge.add_content has been renamed to insert and insert_many for clarity and alignment with the new protocol direction.
- [Faster Gemini runs with direct GCS and URL file inputs](https://www.agno.com/articles/faster-gemini-runs-with-direct-gcs-and-url-file-inputs.md) (SDK v2.4.0): Gemini now accepts gs:// URIs and HTTPS URLs (including presigned URLs) directly, eliminating the need to download files before processing.
- [Migrate from DDG-specific search to the new generic WebSearchTools](https://www.agno.com/articles/migrate-from-ddg-specific-search-to-the-new-generic-websearchtools.md) (SDK v2.4.0): We replaced the DuckDuckGo-specific web search tool with a generic WebSearchTools interface. This standardization broadens provider choice and future-proofs search integrations.
- [Restored reliable Gemini file uploads](https://www.agno.com/articles/restored-reliable-gemini-file-uploads.md) (SDK v2.4.0): We resolved a 400 error caused by message formatting for file Part objects in Gemini (Vertex AI) uploads. Uploads now work as expected, unblocking multimodal use cases.
- [Simplified AgentOS configuration: use a single db parameter](https://www.agno.com/articles/simplified-agentos-configuration-use-a-single-db-parameter.md) (SDK v2.4.0): AgentOS now uses a unified db parameter and deprecates tracing_db. This reduces configuration complexity and clarifies data storage for both operational and tracing needs.
- [Streamlined tool and hook APIs; deprecated fields removed](https://www.agno.com/articles/streamlined-tool-and-hook-apis-deprecated-fields-removed.md) (SDK v2.4.0): We removed deprecated fields across tools/hooks and API parameters to simplify the surface area and reduce ambiguity. This change keeps the platform focused and easier to maintain at scale.
- [Unlock pluggable Knowledge backends with the new KnowledgeProtocol](https://www.agno.com/articles/unlock-pluggable-knowledge-backends-with-the-new-knowledgeprotocol.md) (SDK v2.4.0): We introduced KnowledgeProtocol, a unified interface that enables multiple Knowledge backends to work interchangeably with Agents and Teams.

## 2026-01-13

- [Per-request isolation delivers safer concurrency and simpler multi-tenant operations](https://www.agno.com/articles/per-request-isolation-delivers-safer-concurrency-and-simpler-multi-tenant-operations.md) (SDK v2.3.26): We’ve introduced request-scoped isolation for agents, teams, and workflows. Each incoming request now runs against a fresh copy of the component while expensive resources (database connections…

## 2026-01-12

- [Higher-quality code retrieval with AST-based chunking](https://www.agno.com/articles/higher-quality-code-retrieval-with-ast-based-chunking.md) (SDK v2.3.25): A new AST-based Code Chunker splits code into semantically meaningful units, preserving function and class boundaries across multiple languages and tokenizer options.
- [Reduce latency and cost by skipping retries on non-retryable LLM errors](https://www.agno.com/articles/reduce-latency-and-cost-by-skipping-retries-on-non-retryable-llm-errors.md) (SDK v2.3.25): We now classify common non-retryable conditions (e.g., 4xx responses, payload too large, context limit exceeded) and skip retries across both sync and async flows.
- [Unified learning across agent interactions to improve outcomes](https://www.agno.com/articles/unified-learning-across-agent-interactions-to-improve-outcomes.md) (SDK v2.3.25): We introduced a unified learning system that enables agents to learn from every interaction. Teams can choose learning types and plug in preferred storage backends, making continuous improvement a…

## 2026-01-08

- [Enable crawling in proxy-restricted environments](https://www.agno.com/articles/enable-crawling-in-proxy-restricted-environments.md) (SDK v2.3.24): Crawl4aiTools now supports proxy_config via BrowserConfig, allowing traffic to route through enterprise proxies and enabling browser-level network configuration.
- [Reduce risk with configurable filesystem isolation for tools](https://www.agno.com/articles/reduce-risk-with-configurable-filesystem-isolation-for-tools.md) (SDK v2.3.24): We introduced a restrict_to_base_dir parameter for PythonTools and MLXTranscribeTools, enabled by default.
- [Safer defaults: tools are restricted to their base directory by default](https://www.agno.com/articles/safer-defaults-tools-are-restricted-to-their-base-directory-by-default.md) (SDK v2.3.24): This release introduces a breaking change: PythonTools and MLXTranscribeTools now operate only within their defined base directory by default.

## 2026-01-07

- [Consistent usage metrics on every model response](https://www.agno.com/articles/consistent-usage-metrics-on-every-model-response.md) (SDK v2.3.23): Provider usage metrics (including token counts) are now propagated to the model response in both sync and async paths.
- [Non-blocking tool execution for async agents](https://www.agno.com/articles/non-blocking-tool-execution-for-async-agents.md) (SDK v2.3.23): Toolkit now supports async tool functions and automatically selects them when an agent runs in an async context.

## 2026-01-06

- [Broader native reasoning model coverage across OpenAI, Google, and DeepSeek](https://www.agno.com/articles/broader-native-reasoning-model-coverage-across-openai-google-and-deepseek.md) (SDK v2.3.22): Agno now supports native reasoning for OpenAI GPT-5.1/5.2, Google Gemini 3/3.5/deepthink, and DeepSeek r1/reasoner.
- [Per-run dynamic headers enable secure multi-tenant MCP integrations](https://www.agno.com/articles/per-run-dynamic-headers-enable-secure-multi-tenant-mcp-integrations.md) (SDK v2.3.22): MCPTools and MultiMCPTools now support a header_provider callback to generate request headers at run time.
- [Run and coordinate remote agents over A2A with the new A2AClient](https://www.agno.com/articles/run-and-coordinate-remote-agents-over-a2a-with-the-new-a2aclient.md) (SDK v2.3.22): You can now connect to and orchestrate remote agents via A2A using the new A2AClient, with cookbook examples to get started.
- [Simpler, streaming-by-default connections to external MCP servers](https://www.agno.com/articles/simpler-streaming-by-default-connections-to-external-mcp-servers.md) (SDK v2.3.22): When a URL is provided, MCPTools now default to StreamableHttp transport. This makes it easier to connect to external MCP servers and improves streaming behavior out of the box, reducing configuration…
- [Standardized Skills make agent capabilities reusable and easier to govern](https://www.agno.com/articles/standardized-skills-make-agent-capabilities-reusable-and-easier-to-govern.md) (SDK v2.3.22): We introduced a first-class Skills system, including a Skills class plus validation and loader utilities. Teams can now define, validate, and reuse skills across agents with a consistent interface.
- [Stronger JWT validation with audience checks](https://www.agno.com/articles/stronger-jwt-validation-with-audience-checks.md) (SDK v2.3.22): JWTMiddleware now supports a configurable audience parameter to validate the aud claim. This ensures tokens are intended for your services, reducing the risk of token replay or misrouting and…

## 2025-12-23

- [Full visibility and control of Agent-as-Judge evaluations in AgentOS](https://www.agno.com/articles/full-visibility-and-control-of-agent-as-judge-evaluations-in-agentos.md) (SDK v2.3.21): Agent-as-Judge evaluation runs are now returned on GET endpoints, making them fully visible and manageable in the AgentOS UI.

## 2025-12-22

- [Deeper observability with reasoning trace capture via LiteLLM](https://www.agno.com/articles/deeper-observability-with-reasoning-trace-capture-via-litellm.md) (SDK v2.3.20): Agno’s LiteLLM integration now extracts and surfaces reasoning_content for supported models, enabling richer, audit-ready reasoning traces.
- [Faster onboarding with a revamped getting started cookbook](https://www.agno.com/articles/faster-onboarding-with-a-revamped-getting-started-cookbook.md) (SDK v2.3.20): We overhauled the getting-started cookbook with structured examples, ready-to-use configs, and clear requirements.
- [Gain operational control with async run cancellation and pluggable managers](https://www.agno.com/articles/gain-operational-control-with-async-run-cancellation-and-pluggable-managers.md) (SDK v2.3.20): We introduced an async-capable cancellation manager with in-memory and Redis-backed options. This lets you reliably stop long-running or runaway work across distributed workers, improving cost control…

## 2025-12-20

- [Explicit service account authentication for Vertex AI improves governance and onboarding](https://www.agno.com/articles/explicit-service-account-authentication-for-vertex-ai-improves-governance-and-onboarding.md) (SDK v2.3.18): You can now pass Google OAuth2 service account credentials directly when configuring Vertex AI models.

## 2025-12-19

- [Automatic workflow event reconnection and replay for resilient real-time apps](https://www.agno.com/articles/automatic-workflow-event-reconnection-and-replay-for-resilient-real-time-apps.md) (SDK v2.3.17): Workflow event streams now support robust reconnection, catch-up, and replay. Clients automatically resume from the last known event after transient network issues, preventing gaps in dashboards…
- [Built-in cost visibility for OpenRouter usage in Metrics](https://www.agno.com/articles/built-in-cost-visibility-for-openrouter-usage-in-metrics.md) (SDK v2.3.15): We added a cost field to Metrics for OpenRouter-backed activity. This provides a reliable, standardized view of model spend without manual spreadsheets or custom aggregations, improving financial…
- [Hybrid search for ChromaDB improves retrieval accuracy and recall](https://www.agno.com/articles/hybrid-search-for-chromadb-improves-retrieval-accuracy-and-recall.md) (SDK v2.3.17): Hybrid search combines dense semantic similarity with keyword matching using reciprocal rank fusion (RRF) for Chroma-backed knowledge bases.
- [Operate AgentOS from anywhere with the new AgentOSClient](https://www.agno.com/articles/operate-agentos-from-anywhere-with-the-new-agentosclient.md) (SDK v2.3.17): AgentOSClient is a first-class client for connecting to and operating a remote AgentOS. It standardizes how you authenticate, manage agents/teams/workflows, and stream events, reducing integration…
- [Predictable auth precedence: JWT is now preferred over security key](https://www.agno.com/articles/predictable-auth-precedence-jwt-is-now-preferred-over-security-key.md) (SDK v2.3.15): When both JWT and security key authentication are enabled, JWT now takes precedence. This standardizes behavior, reduces ambiguity for clients, and aligns with common enterprise security practices.
- [Reliable, consistent reads across common data sources](https://www.agno.com/articles/reliable-consistent-reads-across-common-data-sources.md) (SDK v2.3.16): We resolved issues in read and async_read across multiple readers (CSV, field‑labeled CSV, JSON, Markdown, PDF, DOCX, PPTX, S3, Text, and Web Search).
- [Run agents, teams, and workflows remotely on AgentOS for scale and control](https://www.agno.com/articles/run-agents-teams-and-workflows-remotely-on-agentos-for-scale-and-control.md) (SDK v2.3.17): Introducing RemoteAgent, RemoteTeam, and RemoteWorkflow to execute orchestration on a remote AgentOS. This decouples runtime from application code so you can centralize governance and observability…
- [Semantic chunking now supports any embedder, including custom providers](https://www.agno.com/articles/semantic-chunking-now-supports-any-embedder-including-custom-providers.md) (SDK v2.3.17): SemanticChunking now works with all Agno embedders (e.g., Azure OpenAI, Mistral) and custom chonkie BaseEmbeddings via a wrapper, with new parameters for finer control.
- [Simplify multi-database upgrades with a single AgentOS migration endpoint](https://www.agno.com/articles/simplify-multi-database-upgrades-with-a-single-agentos-migration-endpoint.md) (SDK v2.3.15): AgentOS now exposes an API endpoint to migrate all managed databases in one operation. This reduces operational overhead in multi-tenant or multi-environment deployments and ensures consistent schema…

## 2025-12-18

- [Breaking: A2A endpoints moved to convention-based URLs for protocol alignment](https://www.agno.com/articles/breaking-a2a-endpoints-moved-to-convention-based-urls-for-protocol-alignment.md) (SDK v2.3.14): A2A protocol endpoints have been updated to follow standardized URL conventions, and related payloads were aligned to the protocol.
- [Breaking change: Update clients to always include JWT](https://www.agno.com/articles/breaking-change-update-clients-to-always-include-jwt.md) (SDK v2.3.14): JWTMiddleware now enforces token presence on every request. validate=False no longer permits requests without a token.
- [Enforced unique IDs across Agents, Teams, and Workflows to prevent collisions](https://www.agno.com/articles/enforced-unique-ids-across-agents-teams-and-workflows-to-prevent-collisions.md) (SDK v2.3.14): AgentOS now blocks initialization/resync if duplicate IDs are detected across Agents, Teams, or Workflows. This ensures unambiguous references and prevents hard-to-debug behavior at runtime.
- [Finer-grained Milvus queries for faster, more accurate retrieval](https://www.agno.com/articles/finer-grained-milvus-queries-for-faster-more-accurate-retrieval.md) (SDK v2.3.14): Milvus search and async_search now support radius, range_filter, and async search_parameters. These controls help teams tune recall vs. precision and reduce tail latency in high-throughput workloads.
- [Standardized A2A endpoints improve interoperability and client simplicity](https://www.agno.com/articles/standardized-a2a-endpoints-improve-interoperability-and-client-simplicity.md) (SDK v2.3.14): We introduced conventional A2A endpoints — including Agent Card retrieval — and aligned run endpoints and payloads to the updated protocol.
- [Stream reasoning in real time to speed iteration and oversight](https://www.agno.com/articles/stream-reasoning-in-real-time-to-speed-iteration-and-oversight.md) (SDK v2.3.14): You can now stream reasoning chunks whenever a reasoning model is used. A new ReasoningManager coordinates streaming and lifecycle, giving teams earlier visibility into model thinking, faster…
- [Use provider-native JSON schemas for structured outputs with zero translation](https://www.agno.com/articles/use-provider-native-json-schemas-for-structured-outputs-with-zero-translation.md) (SDK v2.3.14): output_schema now accepts provider-specific JSON schemas and passes them directly to model APIs (OpenAI, Claude, and OpenAI‑like).

## 2025-12-15

- [Gain fine-grained access control with built-in RBAC for Agents, Teams, and Workflows](https://www.agno.com/articles/gain-fine-grained-access-control-with-built-in-rbac-for-agents-teams-and-workflows.md) (SDK v2.3.13): We’ve added role-based access control (RBAC) to AgentOS via JWT middleware with per-endpoint authorization and per-resource scopes.

## 2025-12-13

- [Predictable costs and faster responses with cross-provider token counting and smart compression](https://www.agno.com/articles/predictable-costs-and-faster-responses-with-cross-provider-token-counting-and-smart-compression.md) (SDK v2.3.12): A new unified token counting utility provides consistent, accurate token estimates across OpenAI, Anthropic, AWS Bedrock, Google Gemini, and LiteLLM.

## 2025-12-11

- [Accelerate Shopify analytics with a ready-to-use toolkit](https://www.agno.com/articles/accelerate-shopify-analytics-with-a-ready-to-use-toolkit.md) (SDK v2.3.10): We introduced a Shopify toolkit that lets agents analyze store data such as sales, customers, and products without custom integration work.
- [Clearer run behavior: Streaming flags no longer persist](https://www.agno.com/articles/clearer-run-behavior-streaming-flags-no-longer-persist.md) (SDK v2.3.10): To make runs more predictable, the stream and stream_events flags no longer persist across run/arun calls.
- [Improve tracing and billing correlation with provider metadata in responses](https://www.agno.com/articles/improve-tracing-and-billing-correlation-with-provider-metadata-in-responses.md) (SDK v2.3.11): We now populate provider metadata for OpenAI Chat responses and surface it across key response and event objects.
- [Richer Gemini streaming with URL context and web search](https://www.agno.com/articles/richer-gemini-streaming-with-url-context-and-web-search.md) (SDK v2.3.10): Streaming experiences using Gemini now accept URL context and web_search_queries, enabling real-time retrieval and reasoning over live web content.

## 2025-12-09

- [Accelerate async MySQL workloads with first-class Async MySQLDb](https://www.agno.com/articles/accelerate-async-mysql-workloads-with-first-class-async-mysqldb.md) (SDK v2.3.9): We’ve added AsyncMySQLDb with native compatibility for the asyncmy driver, enabling fully asynchronous MySQL operations.
- [Operationalize model quality with Agent-as-Judge evaluations](https://www.agno.com/articles/operationalize-model-quality-with-agent-as-judge-evaluations.md) (SDK v2.3.9): A new built-in evaluation system lets you automate LLM quality checks with binary and numeric scoring, background execution, post-hooks, and customizable evaluator agents.
- [Simplify integrations with synchronous Knowledge operations](https://www.agno.com/articles/simplify-integrations-with-synchronous-knowledge-operations.md) (SDK v2.3.9): Knowledge add_content_ methods now support true synchronous execution. This removes the async-only limitation, making it straightforward to integrate content ingestion into synchronous services and…
- [Unlock OpenRouter reasoning messages for richer insight and control](https://www.agno.com/articles/unlock-openrouter-reasoning-messages-for-richer-insight-and-control.md) (SDK v2.3.9): Agno now supports reasoning messages from OpenRouter, enabling you to capture and act on models’ reasoning outputs where available.

## 2025-12-05

- [Faster setup with Memori SDK v3.0.5 and automatic conversation recording](https://www.agno.com/articles/faster-setup-with-memori-sdk-v3-0-5-and-automatic-conversation-recording.md) (SDK v2.3.8): Agno now ships with Memori SDK v3.0.5, enabling automatic recording of agent conversations without a separate tool.
- [Model-level retries improve reliability under provider rate limits](https://www.agno.com/articles/model-level-retries-improve-reliability-under-provider-rate-limits.md) (SDK v2.3.8): We’ve moved retry logic from Agents/Teams to the Model layer. When you set retries on a model, Agno now retries at the model execution level, which is more effective for handling provider throttling…
- [Streamlined Memori integration as MemoriTools is removed](https://www.agno.com/articles/streamlined-memori-integration-as-memoritools-is-removed.md) (SDK v2.3.8): MemoriTools has been removed in favor of Memori SDK v3’s built-in auto-recording. This consolidates functionality in the SDK, reduces integration complexity, and lowers maintenance overhead.

## 2025-12-04

- [Accelerate Spotify integrations with a ready-to-use toolkit and agent](https://www.agno.com/articles/accelerate-spotify-integrations-with-a-ready-to-use-toolkit-and-agent.md) (SDK v2.3.6): We’ve introduced a Spotify toolkit and example agent to manage and interact with Spotify, including library management.
- [Native Amazon Redshift toolkit simplifies data access and operations](https://www.agno.com/articles/native-amazon-redshift-toolkit-simplifies-data-access-and-operations.md) (SDK v2.3.7): We introduced RedshiftTools, giving agents first-class access to Amazon Redshift without custom glue. Teams can explore schemas, describe tables, inspect and run queries, and export data directly…
- [Run evaluations reliably on asynchronous databases](https://www.agno.com/articles/run-evaluations-reliably-on-asynchronous-databases.md) (SDK v2.3.7): AgentOS evaluation endpoints now work with asynchronous database backends. Teams using async DB classes can run evaluations without changing their stack, removing a key limitation for modern…
- [Streamline Human-In-The-Loop with a single, predictable requirement model](https://www.agno.com/articles/streamline-human-in-the-loop-with-a-single-predictable-requirement-model.md) (SDK v2.3.7): RunRequirement simplifies how agents request and manage human input. Requirements now surface directly in agent responses or as RunPaused events in streaming flows, providing a consistent pattern for…

## 2025-12-03

- [Built-in OpenTelemetry tracing delivers end-to-end visibility](https://www.agno.com/articles/built-in-opentelemetry-tracing-delivers-end-to-end-visibility.md) (SDK v2.3.5): We introduced native tracing with OpenTelemetry, including first-class spans and new endpoints to inspect traces.
- [DynamoDB Memory schema now requires GSI on created_at](https://www.agno.com/articles/dynamodb-memory-schema-now-requires-gsi-on-created-at.md) (SDK v2.3.5): To prevent CreateTable validation errors and ensure reliable, time-ordered queries, the DynamoDB schema for the user Memory table now requires a global secondary index (GSI) on created_at.
- [Non-blocking hooks accelerate AgentOS workflows](https://www.agno.com/articles/non-blocking-hooks-accelerate-agentos-workflows.md) (SDK v2.3.5): Agent and Team pre- and post-hooks now run as background tasks in AgentOS, so they no longer block the main operation.

## 2025-11-28

- [Accelerate Gemini 3 adoption with ready-to-run agents and configs](https://www.agno.com/articles/accelerate-gemini-3-adoption-with-ready-to-run-agents-and-configs.md) (SDK v2.3.4): We added a complete Gemini 3 demo, including example agents, configuration, and generated assets. This makes it faster to evaluate and roll out Gemini 3 within Agno by providing opinionated, runnable…
- [Capture model source citations in run data for better provenance](https://www.agno.com/articles/capture-model-source-citations-in-run-data-for-better-provenance.md) (SDK v2.3.4): Runs now support an optional citations field across single, team, and workflow executions. This lets you store and surface model-provided source citations directly in your run metadata, improving…

## 2025-11-27

- [Adapt outputs per run with runtime schema overrides](https://www.agno.com/articles/adapt-outputs-per-run-with-runtime-schema-overrides.md) (SDK v2.3.3): You can now override output_schema at runtime for both Agent and Team (streaming and non-streaming), with automatic restoration after the run.
- [Build retrieval-augmented agents with Gemini File Search](https://www.agno.com/articles/build-retrieval-augmented-agents-with-gemini-file-search.md) (SDK v2.3.3): Agno now offers full support for Google Gemini File Search, including store and document management, uploads/imports, metadata filters, citation extraction, and async APIs.
- [Faster Bedrock onboarding with optional API key authentication for Claude](https://www.agno.com/articles/faster-bedrock-onboarding-with-optional-api-key-authentication-for-claude.md) (SDK v2.3.3): We’ve added an optional API key path for AWS Bedrock Claude in addition to IAM. This reduces setup friction in environments where IAM is not feasible while preserving IAM as the default for…
- [Keep runs within context limits with automatic tool output compression](https://www.agno.com/articles/keep-runs-within-context-limits-with-automatic-tool-output-compression.md) (SDK v2.3.3): Automatically compress and summarize tool call results to keep agent context safely within model token windows.
- [Native async MongoDB support for higher throughput agents](https://www.agno.com/articles/native-async-mongodb-support-for-higher-throughput-agents.md) (SDK v2.3.3): MongoDB clients now support Motor and PyMongo async libraries with improved error handling and typing.
- [Reduce cost and drift with out-of-band memory optimization (beta)](https://www.agno.com/articles/reduce-cost-and-drift-with-out-of-band-memory-optimization-beta.md) (SDK v2.3.3): A new MemoryOptimizationStrategy framework and APIs allow you to summarize and optimize memories outside of agent runs.

## 2025-11-22

- [Topic-based memory retrieval in SQLite now returns correct results](https://www.agno.com/articles/topic-based-memory-retrieval-in-sqlite-now-returns-correct-results.md) (SDK v2.3.2): We fixed an issue where filtering memories by topic could return incorrect results when using SQLite or AsyncSQLite backends.

## 2025-11-21

- [Align with Nebius using the new default API endpoint](https://www.agno.com/articles/align-with-nebius-using-the-new-default-api-endpoint.md) (SDK v2.3.0): The default Nebius model endpoint is now api.tokenfactory.nebius.com. This aligns with Nebius updated platform and helps avoid legacy endpoints that may degrade or deprecate.
- [Cleaner streaming defaults across CLI and print helpers](https://www.agno.com/articles/cleaner-streaming-defaults-across-cli-and-print-helpers.md) (SDK v2.3.0): We removed the stream_events parameter from print_response/aprint_response and CLI. Streaming now works correctly by default, reducing configuration and edge cases.
- [Create production-grade sound effects with ModelsLab SFX support](https://www.agno.com/articles/create-production-grade-sound-effects-with-modelslab-sfx-support.md) (SDK v2.3.0): You can now generate WAV sound effects directly through ModelsLabTools. This adds an audio SFX modality to Agno, enabling teams to build sound-driven experiences — alerts, games, product interactions…
- [Explicit storage now required for knowledge filtering](https://www.agno.com/articles/explicit-storage-now-required-for-knowledge-filtering.md) (SDK v2.3.0): When using knowledge_filters, you must configure contents_db. This ensures deterministic, stateless filtering aligned with AgentOS and prevents silent mismatches.
- [Faster, validated image generation with Google Nano Banana](https://www.agno.com/articles/faster-validated-image-generation-with-google-nano-banana.md) (SDK v2.3.1): We introduced NanoBananaTools, a turnkey toolkit to generate images with Google’s Nano Banana model. It includes built-in parameter validation and a cookbook example, enabling faster adoption and…
- [Finalize AgentOS API names by removing deprecated parameters](https://www.agno.com/articles/finalize-agentos-api-names-by-removing-deprecated-parameters.md) (SDK v2.3.0): We removed deprecated AgentOS parameters to standardize on stable naming: os_id -> id, fastapi_app -> base_app, enable_mcp -> enable_mcp_server, replace_routes -> on_route_conflict.
- [Predictable, schema‑enforced Claude responses across sync, async, and streaming](https://www.agno.com/articles/predictable-schema-enforced-claude-responses-across-sync-async-and-streaming.md) (SDK v2.3.1): We added first-class support for Anthropic’s structured outputs, including schema enforcement, strict tool calling, and robust response parsing across synchronous, asynchronous, and streaming APIs.
- [Scale agent storage with Redis Cluster](https://www.agno.com/articles/scale-agent-storage-with-redis-cluster.md) (SDK v2.3.0): RedisDb now accepts RedisCluster clients, enabling high availability and horizontal scalability for agent state. Teams running clustered Redis can adopt Agno storage without architectural workarounds.
- [Simplified team configuration with clearer parameter naming](https://www.agno.com/articles/simplified-team-configuration-with-clearer-parameter-naming.md) (SDK v2.3.0): We renamed the Team parameter delegate_task_to_all_members to delegate_to_all_members. This clarifies intent and standardizes naming across the API.
- [Standardize web search with DuckDuckGoTools](https://www.agno.com/articles/standardize-web-search-with-duckduckgotools.md) (SDK v2.3.0): GoogleSearchTools has been removed. Please migrate to DuckDuckGoTools for web search capabilities. This streamlines support and ensures predictable results across environments.
- [Streamline upgrades with built‑in database migrations](https://www.agno.com/articles/streamline-upgrades-with-built-in-database-migrations.md) (SDK v2.3.0): We introduced a MigrationManager and the first migrations for sessions and memories tables. This provides a controlled, repeatable path for schema evolution, reducing upgrade risk and operational…
- [Unified message history APIs for simpler integrations](https://www.agno.com/articles/unified-message-history-apis-for-simpler-integrations.md) (SDK v2.3.0): We removed get_messages_for_session and get_messages_from_last_n_runs in favor of get_messages, get_session_messages, and get_chat_history. This unifies patterns and reduces mental overhead.
- [Unlock Gemini 3.0 Pro with thought signature support](https://www.agno.com/articles/unlock-gemini-3-0-pro-with-thought-signature-support.md) (SDK v2.3.0): Agno now supports the thought signatures required by Gemini 3.0 Pro, ensuring compatibility and unlocking the latest model capabilities without custom integration work.

## 2025-11-15

- [Reliable real-time event streaming in custom workflow steps](https://www.agno.com/articles/reliable-real-time-event-streaming-in-custom-workflow-steps.md) (SDK v2.2.13): We restored reliable live event streaming across Agents and Teams workflows, ensuring events from custom executor steps are delivered consistently.

## 2025-11-14

- [Achieve precise Knowledge retrieval with flexible metadata filters](https://www.agno.com/articles/achieve-precise-knowledge-retrieval-with-flexible-metadata-filters.md) (SDK v2.2.12): We introduced a metadata-based filter DSL for Knowledge searches, enabling precise, policy-aligned retrieval at scale.
- [Control Slack bot engagement with mention-only replies by default](https://www.agno.com/articles/control-slack-bot-engagement-with-mention-only-replies-by-default.md) (SDK v2.2.12): The Slack interface now replies only when mentioned, reducing channel noise and improving operator control by default. This change helps teams run bots in busy channels without overwhelming users.

## 2025-11-12

- [Deeper Gmail automation with label management](https://www.agno.com/articles/deeper-gmail-automation-with-label-management.md) (SDK v2.2.11): Agents can now list, create, apply/remove, and delete custom Gmail labels. This enables end-to-end email triage, routing, and compliance workflows without external glue code, speeding deployment and…
- [Faster, reliable web and PDF intelligence with built-in search and extraction](https://www.agno.com/articles/faster-reliable-web-and-pdf-intelligence-with-built-in-search-and-extraction.md) (SDK v2.2.11): We introduced ParallelTools with Search and Extract APIs to deliver LLM-ready excerpts and robust markdown from the open web and PDFs — including JS-heavy pages.
- [Lower token spend and reduce drift with Claude context management](https://www.agno.com/articles/lower-token-spend-and-reduce-drift-with-claude-context-management.md) (SDK v2.2.11): Opt-in support for Claude’s context editing helps automatically remove stale tool results from long conversations.
- [Seamless access to Anthropic beta capabilities via configuration](https://www.agno.com/articles/seamless-access-to-anthropic-beta-capabilities-via-configuration.md) (SDK v2.2.11): You can now enable Anthropic beta features by passing the betas parameter. When present, Agno automatically uses the appropriate beta client, making it easier to evaluate new capabilities without…

## 2025-11-08

- [Build custom workflow executors without event-wrapping overhead](https://www.agno.com/articles/build-custom-workflow-executors-without-event-wrapping-overhead.md) (SDK v2.2.10): Custom executors can now yield native objects — no need to wrap every output as an Agno event. This fix removes a key limitation, making it easier to integrate existing business logic and libraries…
- [Consistent state across all workflow steps, including parallel branches](https://www.agno.com/articles/consistent-state-across-all-workflow-steps-including-parallel-branches.md) (SDK v2.2.10): Run context now propagates automatically through every workflow step — including parallel branches. This delivers predictable, shared state across complex flows, reducing manual plumbing and avoiding…
- [Streamline run outputs with the new yield_run_output flag](https://www.agno.com/articles/streamline-run-outputs-with-the-new-yield-run-output-flag.md) (SDK v2.2.10): We introduced yield_run_output for Agent and Team runs, replacing yield_run_response (slated for deprecation).

## 2025-11-07

- [Guarantee schema‑true outputs across models to cut parsing errors](https://www.agno.com/articles/guarantee-schema-true-outputs-across-models-to-cut-parsing-errors.md) (SDK v2.2.9): We introduced a strict_output setting that enforces exact adherence to your output_schema by default across supported models.
- [Prevent configuration drift by scoping team model inheritance](https://www.agno.com/articles/prevent-configuration-drift-by-scoping-team-model-inheritance.md) (SDK v2.2.9): Child agents now inherit only the primary model from their parent. Auxiliary models for output, parsing, or reasoning are no longer inherited, ensuring predictable configurations and reducing hidden…
- [Simplify UI-to-agent orchestration with automatic session state mapping](https://www.agno.com/articles/simplify-ui-to-agent-orchestration-with-automatic-session-state-mapping.md) (SDK v2.2.9): AG-UI request state is now passed and mapped into session_state, ensuring agents receive the right UI context with no extra plumbing.
- [Unlock multi-tenant designs with multiple tables per database](https://www.agno.com/articles/unlock-multi-tenant-designs-with-multiple-tables-per-database.md) (SDK v2.2.9): AgentOS now supports multiple Agno tables of the same type within a single database. This enables clean tenant or namespace isolation without multiplying databases, lowering operational overhead while…

## 2025-11-05

- [Align ExaTools with Exa API v2 by removing deprecated “highlights”](https://www.agno.com/articles/align-exatools-with-exa-api-v2-by-removing-deprecated-highlights.md) (SDK v2.2.7): ExaTools no longer accepts or passes the removed highlights parameter, aligning with Exa API v2.0.0. This prevents runtime errors and ensures forward compatibility with the upstream service.
- [Breaking change: ExaTools no longer accepts “highlights”](https://www.agno.com/articles/breaking-change-exatools-no-longer-accepts-highlights.md) (SDK v2.2.7): To comply with Exa API v2.0.0, ExaTools has removed support for the highlights parameter. Calls that include it will fail. Update your integrations to avoid errors and maintain service compatibility.
- [High-throughput embeddings with vLLM, local or remote](https://www.agno.com/articles/high-throughput-embeddings-with-vllm-local-or-remote.md) (SDK v2.2.7): We added a vLLM embedder with batching support, enabling high-throughput, cost-controlled embeddings on your infrastructure or via remote endpoints.
- [Migration script provides a clear path to VectorDB v2](https://www.agno.com/articles/migration-script-provides-a-clear-path-to-vectordb-v2.md) (SDK v2.2.7): We introduced a migration script to move existing VectorDB data to the v2 format. This protects compatibility, unlocks improvements in the new version, and reduces risk during upgrades.
- [Namespaced MCP tools prevent collisions in multi-server environments](https://www.agno.com/articles/namespaced-mcp-tools-prevent-collisions-in-multi-server-environments.md) (SDK v2.2.7): MCPTools now supports a tool_name_prefix to avoid name collisions when sourcing tools from multiple MCP servers.
- [Redis-backed vector search streamlines Knowledge deployments](https://www.agno.com/articles/redis-backed-vector-search-streamlines-knowledge-deployments.md) (SDK v2.2.7): Knowledge now supports a Redis VectorDB backend. Teams standardized on Redis can consolidate infrastructure, simplify operations, and reduce latency by keeping vector search close to existing caches…
- [Simpler installs and broader deployment: VectorDb no longer requires Redis](https://www.agno.com/articles/simpler-installs-and-broader-deployment-vectordb-no-longer-requires-redis.md) (SDK v2.2.8): We removed Redis and redisvl as dependencies of the base VectorDb class. VectorDb can now be used without installing or configuring Redis, reducing setup time and expanding where Agno can run (for…
- [Unified RunContext streamlines orchestration across tools, hooks, and functions](https://www.agno.com/articles/unified-runcontext-streamlines-orchestration-across-tools-hooks-and-functions.md) (SDK v2.2.7): We introduced RunContext to carry session state, dependencies, metadata, and knowledge filters through every step of a run.

## 2025-11-01

- [Expanded FileTools enable precise edits and stronger guardrails](https://www.agno.com/articles/expanded-filetools-enable-precise-edits-and-stronger-guardrails.md) (SDK v2.2.6): FileTools now supports delete, chunked read, and partial replace operations, with new size limits and base_dir disclosure controls.
- [Faster, smarter chat experiences with conversational workflows](https://www.agno.com/articles/faster-smarter-chat-experiences-with-conversational-workflows.md) (SDK v2.2.6): We introduced WorkflowAgent, which powers chat-like workflows that decide when to answer from conversation history and when to execute workflow steps.
- [Native Notion toolkit accelerates content and knowledge automation](https://www.agno.com/articles/native-notion-toolkit-accelerates-content-and-knowledge-automation.md) (SDK v2.2.6): A new Notion toolkit and cookbook make it easy to connect agents to Notion. Teams can read, create, and update Notion pages and databases programmatically, reducing integration effort and speeding up…
- [Reduce runtime errors with schema-validated inputs for Agents and Teams](https://www.agno.com/articles/reduce-runtime-errors-with-schema-validated-inputs-for-agents-and-teams.md) (SDK v2.2.6): Agno now validates input schemas for Agents and Teams, enforcing strong contracts at the edge. This catches issues earlier, improves reliability in production, and shortens debugging cycles.

## 2025-10-30

- [Custom routers auto-refresh during AgentOS reprovision](https://www.agno.com/articles/custom-routers-auto-refresh-during-agentos-reprovision.md) (SDK v2.2.5): Updating AgentOS now automatically refreshes all API routers — including custom ones — so your routes stay aligned with the new OS state without manual intervention.
- [Simpler data hygiene with memory deletion enabled by default](https://www.agno.com/articles/simpler-data-hygiene-with-memory-deletion-enabled-by-default.md) (SDK v2.2.4): Agent-facing memory managers now include the deletion tool by default. This reduces setup friction and makes it easier to enforce data retention policies, remove outdated information, and keep…
- [Streamlined database configuration: AsyncPostgresDb now uses id (db_id deprecated)](https://www.agno.com/articles/streamlined-database-configuration-asyncpostgresdb-now-uses-id-db-id-deprecated.md) (SDK v2.2.4): AsyncPostgresDb now accepts id as the primary identifier, with db_id deprecated. This aligns naming with the broader platform and eliminates parsing edge cases, improving reliability and reducing…
- [Update running AgentOS instances during FastAPI lifecycle to reduce downtime](https://www.agno.com/articles/update-running-agentos-instances-during-fastapi-lifecycle-to-reduce-downtime.md) (SDK v2.2.4): You can now update a live AgentOS instance — adding Agents, Teams, and Workflows — inside FastAPI lifespan functions.

## 2025-10-29

- [Accelerate RAG pipelines with Tavily Extract and Reader](https://www.agno.com/articles/accelerate-rag-pipelines-with-tavily-extract-and-reader.md) (SDK v2.2.2): We’ve integrated TavilyReader for knowledge base ingestion and enhanced Tavily tools with URL content extraction (sync and async).
- [Async MongoDB enables non-blocking data flows across agents and workflows](https://www.agno.com/articles/async-mongodb-enables-non-blocking-data-flows-across-agents-and-workflows.md) (SDK v2.2.3): We introduced AsyncMongoDb to provide fully asynchronous MongoDB access end to end. Teams can process more requests concurrently, reduce I/O bottlenecks, and improve responsiveness in agent…
- [Cut inference latency and cost with built-in LLM response caching](https://www.agno.com/articles/cut-inference-latency-and-cost-with-built-in-llm-response-caching.md) (SDK v2.2.2): We’ve added native caching for model responses across sync, async, and streaming APIs. You can configure TTL and cache storage to accelerate repeated prompts and reduce token spend — without building…
- [Eliminate blocking with fully Async SQLite operations](https://www.agno.com/articles/eliminate-blocking-with-fully-async-sqlite-operations.md) (SDK v2.2.2): Introducing Async SqliteDb for end-to-end async database access. This enables higher throughput and lower tail latency in asyncio-based services by removing blocking I/O, while reducing boilerplate…
- [Expand document and spreadsheet automation with Claude native skills](https://www.agno.com/articles/expand-document-and-spreadsheet-automation-with-claude-native-skills.md) (SDK v2.2.2): Agents can now leverage Claude’s native skills for documents, spreadsheets, and presentations. This expands what agents can execute natively — from editing and analysis to content generation — without…
- [Maintain context across agents with shared sessions](https://www.agno.com/articles/maintain-context-across-agents-with-shared-sessions.md) (SDK v2.2.2): Agents and Teams can now share and reuse the same session to preserve context and history across handoffs.
- [Non-blocking session summaries for real-time systems](https://www.agno.com/articles/non-blocking-session-summaries-for-real-time-systems.md) (SDK v2.2.2): New async APIs (aget_session_summary()) let Agents and Teams return session summaries without blocking.

## 2025-10-24

- [Cap historical tool calls to control context size and spend](https://www.agno.com/articles/cap-historical-tool-calls-to-control-context-size-and-spend.md) (SDK v2.2.1): A new parameter, max_tool_calls_from_history, lets you cap how many historical tool call pairs are loaded into context.
- [Native PowerPoint ingestion streamlines knowledge workflows](https://www.agno.com/articles/native-powerpoint-ingestion-streamlines-knowledge-workflows.md) (SDK v2.2.1): You can now ingest Microsoft PowerPoint (.pptx) files natively with a dedicated PPTX reader. This removes the need for custom loaders, shortens integration time, and expands the types of enterprise…

## 2025-10-23

- [Agents now support media-only inputs](https://www.agno.com/articles/agents-now-support-media-only-inputs.md) (SDK v2.2.0): Agents can now process inputs that contain only media (no text). This unlocks use cases such as camera-to-answer, voice-only prompts, and file-first interactions, reducing friction in multimodal…
- [Better context sharing and delegation behavior in Teams](https://www.agno.com/articles/better-context-sharing-and-delegation-behavior-in-teams.md) (SDK v2.2.0): We added add_team_history_to_members to simplify sharing team history with members, improving context continuity in multi-agent work.
- [Clearer delegation contracts in Teams](https://www.agno.com/articles/clearer-delegation-contracts-in-teams.md) (SDK v2.2.0): We removed the default expected_output from Team delegate_task_to_member. Callers now explicitly define the expected output or handle its absence, improving predictability and intent in task…
- [Consistent async streaming for Workflows](https://www.agno.com/articles/consistent-async-streaming-for-workflows.md) (SDK v2.2.0): Workflow.arun now returns an AsyncIterator, aligning with Agent and Team. This change standardizes how you consume streaming results and reduces integration complexity across components.
- [Deprecated stream_intermediate_steps in favor of stream_events](https://www.agno.com/articles/deprecated-stream-intermediate-steps-in-favor-of-stream-events.md) (SDK v2.2.0): The stream_intermediate_steps setting is deprecated. Use stream_events for a single, consistent way to control streamed event emission across components, reducing configuration overhead and drift.
- [Explicit control over non-content events during streaming](https://www.agno.com/articles/explicit-control-over-non-content-events-during-streaming.md) (SDK v2.2.0): Non-RunContent events (e.g., operational or summary signals) now emit only when stream_events=True. This reduces noise in streaming responses and ensures UIs receive only the events they opt into.
- [Faster runs with concurrent memory updates and richer end-of-run events](https://www.agno.com/articles/faster-runs-with-concurrent-memory-updates-and-richer-end-of-run-events.md) (SDK v2.2.0): Memory updates now occur in a background thread, reducing latency for users and UIs. We’ve also added events for run content completion and session summaries to make end-of-stream handling more…
- [Manage session lifecycles programmatically with new AgentOS endpoints](https://www.agno.com/articles/manage-session-lifecycles-programmatically-with-new-agentos-endpoints.md) (SDK v2.2.0): We’ve added API endpoints to create empty sessions, fetch a run by ID, and update sessions. This gives teams precise control over session state, enabling cleaner orchestration, simpler retries, and…
- [One flag to control streaming events across Agents, Teams, and Workflows](https://www.agno.com/articles/one-flag-to-control-streaming-events-across-agents-teams-and-workflows.md) (SDK v2.2.0): A unified stream_events parameter now governs the emission of non-content events during streaming across Agent, Team, and Workflow APIs.
- [Reliable multimodal streaming for Teams](https://www.agno.com/articles/reliable-multimodal-streaming-for-teams.md) (SDK v2.2.0): Team streaming now respects input media. Images, video, audio, and files provided via run_input are properly processed during streaming, restoring expected multimodal behavior and improving parity…
- [Stream Workflow results without awaiting](https://www.agno.com/articles/stream-workflow-results-without-awaiting.md) (SDK v2.2.0): When stream=True, Workflow.arun now returns an AsyncIterator instead of requiring await. This clarifies intent and avoids mixed patterns in streaming code.

## 2025-10-21

- [Establish consistent multi‑agent behavior with an experimental Culture Manager and shared knowledge space](https://www.agno.com/articles/establish-consistent-multi-agent-behavior-with-an-experimental-culture-manager-and-shared-knowledge-space.md) (SDK v2.1.10): Define company-wide tone, domain norms, and editorial standards once and apply them across agents. The new, experimental Culture Manager centralizes “cultural context” so agents can think, write, and…
- [Unblock reuse and governance by decoupling knowledge from agents and teams](https://www.agno.com/articles/unblock-reuse-and-governance-by-decoupling-knowledge-from-agents-and-teams.md) (SDK v2.1.10): Knowledge can now be created and managed directly in AgentOS, independent of any specific Agent or Team.

## 2025-10-20

- [Message-level IDs improve traceability and storage workflows](https://www.agno.com/articles/message-level-ids-improve-traceability-and-storage-workflows.md) (SDK v2.1.9): We’ve added a stable id field to every Message and exposed it in RunOutput messages. This makes it straightforward to correlate logs, audit events, and persist messages across systems without custom…
- [Smarter workflow decisions with session-aware routing and conditions](https://www.agno.com/articles/smarter-workflow-decisions-with-session-aware-routing-and-conditions.md) (SDK v2.1.9): Workflow Condition (evaluator) and Router (selector) functions now receive session_state, enabling context-aware decisioning.

## 2025-10-17

- [Align to the new knowledge search route for reliability](https://www.agno.com/articles/align-to-the-new-knowledge-search-route-for-reliability.md) (SDK v2.1.8): To ensure consistent behavior across releases, the knowledge search endpoint has been renamed from search_vectors to search_knowledge. Update clients to avoid failures against the deprecated route.
- [Automate attendee communications in Google Calendar workflows](https://www.agno.com/articles/automate-attendee-communications-in-google-calendar-workflows.md) (SDK v2.1.8): The Google Calendar integration now supports attendee notifications when creating, updating, or deleting events.
- [Automate Jira time tracking with worklog support](https://www.agno.com/articles/automate-jira-time-tracking-with-worklog-support.md) (SDK v2.1.8): JiraTools now supports adding worklogs to issues. This enables automated time tracking and better reporting directly from your agents and workflows, reducing manual entry and improving governance.
- [Clearer control over tool-call message storage with paired scrubbing](https://www.agno.com/articles/clearer-control-over-tool-call-message-storage-with-paired-scrubbing.md) (SDK v2.1.6): We renamed the configuration flag from store_tool_results to store_tool_messages and aligned behavior so tool-call and tool-result messages are scrubbed together.
- [Improved reliability for async memory with Postgres and other async databases](https://www.agno.com/articles/improved-reliability-for-async-memory-with-postgres-and-other-async-databases.md) (SDK v2.1.7): We resolved errors that occurred when reading or updating user memories when using asynchronous databases such as Postgres.
- [Native SurrealDB support expands deployment options for agents, teams, and workflows](https://www.agno.com/articles/native-surrealdb-support-expands-deployment-options-for-agents-teams-and-workflows.md) (SDK v2.1.6): We added full SurrealDB support—including models, queries, metrics, and utilities—so you can run agents, teams, workflows, and memory on SurrealDB with first-class parity.
- [Simplify streaming post-processing with post-hooks](https://www.agno.com/articles/simplify-streaming-post-processing-with-post-hooks.md) (SDK v2.1.8): You can now attach post-hooks to streaming flows, making it easier to run instrumentation, filtering, logging, or persistence tasks as soon as a stream completes.
- [Standardized response_audio field type for cleaner integrations](https://www.agno.com/articles/standardized-response-audio-field-type-for-cleaner-integrations.md) (SDK v2.1.8): The response_audio field type has changed from Optional[List[dict]] to Optional[dict]. This brings consistency across Run, Team, and Workflow schemas and reduces edge cases in client code.
- [Update configuration to new tool message storage flag](https://www.agno.com/articles/update-configuration-to-new-tool-message-storage-flag.md) (SDK v2.1.6): The configuration flag store_tool_results has been renamed to store_tool_messages. This is a breaking change and requires updating any configs, environment variables, or automation that reference the…
- [Update required: Knowledge search API endpoint renamed](https://www.agno.com/articles/update-required-knowledge-search-api-endpoint-renamed.md) (SDK v2.1.8): We renamed the knowledge search endpoint from search_vectors to search_knowledge. This is a breaking change and aligns the API with terminology used across the platform.
- [Update required: response_audio schema standardized](https://www.agno.com/articles/update-required-response-audio-schema-standardized.md) (SDK v2.1.8): We corrected the response_audio field to be an optional single object (previously an optional list) across Run, Team, and Workflow schemas.

## 2025-10-15

- [Easier client parsing with streamlined media fields in GET runs](https://www.agno.com/articles/easier-client-parsing-with-streamlined-media-fields-in-get-runs.md) (SDK v2.1.5): GET runs responses now return media at the top level of array objects, making responses simpler to consume and reducing client-side parsing logic.
- [Harden Google Sheets integrations with service account authentication](https://www.agno.com/articles/harden-google-sheets-integrations-with-service-account-authentication.md) (SDK v2.1.5): GoogleSheetsTools now supports service account authentication. This enables secure, headless server-side access aligned with enterprise policies and CI/CD workflows — no manual OAuth flows or user…
- [Programmatic vector search in your knowledge bases](https://www.agno.com/articles/programmatic-vector-search-in-your-knowledge-bases.md) (SDK v2.1.5): AgentOS now exposes API endpoints for vector search across your knowledge bases. This makes it easy to build retrieval-augmented workflows and search experiences without custom indexing or ad-hoc…
- [Reduce noise and risk with AgentOS access logs off by default](https://www.agno.com/articles/reduce-noise-and-risk-with-agentos-access-logs-off-by-default.md) (SDK v2.1.5): AgentOS access logs are now disabled by default, resulting in leaner deployments and minimizing the risk of unintentionally logging sensitive data.
- [Scale non-blocking deployments with async Postgres across the platform](https://www.agno.com/articles/scale-non-blocking-deployments-with-async-postgres-across-the-platform.md) (SDK v2.1.5): We’ve added end-to-end async database support, including a complete async Postgres implementation, across Agents, Teams, AgentOS, Evals, and Knowledge.
- [Simpler MCP toolbox configuration without enforced URL suffixes](https://www.agno.com/articles/simpler-mcp-toolbox-configuration-without-enforced-url-suffixes.md) (SDK v2.1.5): We’ve removed the requirement to append “/mcp” to MCPToolbox database URLs. This simplifies configuration and aligns with common DSN formats.
- [Unlock advanced reasoning with Gemini 2.5+ and Claude models](https://www.agno.com/articles/unlock-advanced-reasoning-with-gemini-2-5-and-claude-models.md) (SDK v2.1.5): Agno now supports native reasoning/thinking modes for Gemini 2.5+, Anthropic Claude, and Vertex AI–hosted Claude.

## 2025-10-10

- [Built-in Google Drive toolkit streamlines file operations](https://www.agno.com/articles/built-in-google-drive-toolkit-streamlines-file-operations.md) (SDK v2.1.4): A new Google Drive toolkit allows agents to list, upload, and download files directly from Google Drive.
- [Lower latency and richer telemetry with real-time workflow event streaming](https://www.agno.com/articles/lower-latency-and-richer-telemetry-with-real-time-workflow-event-streaming.md) (SDK v2.1.4): Workflows now emit events immediately from parallel and custom function steps, with workflow context injected into each event.
- [Step-level workflow history enables more context-aware execution](https://www.agno.com/articles/step-level-workflow-history-enables-more-context-aware-execution.md) (SDK v2.1.4): Workflows can now persist and access history at the step level or across all steps. This makes agents and orchestrations more context-aware without manual state passing, reducing boilerplate and…

## 2025-10-08

- [Adopt MCP tools across Agents, Teams, and Workflows with better server lifecycle control](https://www.agno.com/articles/adopt-mcp-tools-across-agents-teams-and-workflows-with-better-server-lifecycle-control.md) (SDK v2.1.3): The AgentOS MCP server now runs cleanly in more environments and integrates with custom FastAPI base apps.
- [Increase reliability with OpenRouter failover and dynamic model routing](https://www.agno.com/articles/increase-reliability-with-openrouter-failover-and-dynamic-model-routing.md) (SDK v2.1.3): OpenRouter now supports fallback models for transparent failover. If a primary model is degraded or unavailable, requests automatically route to healthy alternatives — reducing errors and customer…
- [Use Claude on Vertex AI with multimodal, tool use, and prompt caching](https://www.agno.com/articles/use-claude-on-vertex-ai-with-multimodal-tool-use-and-prompt-caching.md) (SDK v2.1.3): We added a new Claude model integration backed by Vertex AI, enabling you to run Claude where your data and governance live on Google Cloud.

## 2025-10-07

- [Better default outcomes with Claude Sonnet 4.5 as the standard model](https://www.agno.com/articles/better-default-outcomes-with-claude-sonnet-4-5-as-the-standard-model.md) (SDK v2.1.2): Claude-based agents now default to Claude Sonnet 4.5, improving baseline reasoning quality and response consistency without any configuration changes.
- [Broaden media ingestion with expanded audio and file format support](https://www.agno.com/articles/broaden-media-ingestion-with-expanded-audio-and-file-format-support.md) (SDK v2.1.2): AgentOS routers now recognize more audio and file types out of the box. This increases ingestion success rates for multimodal workloads and reduces the need for pre-processing, accelerating…
- [Expand tool coverage by running local binaries as MCP servers](https://www.agno.com/articles/expand-tool-coverage-by-running-local-binaries-as-mcp-servers.md) (SDK v2.1.2): You can now run local binaries (for example, ./script) as MCP servers via MCPTools and MultiMCPTools. This makes it simple to operationalize existing scripts and CLIs as governed, observable tools…
- [Improve multi-tenant safety with user-scoped memory deletions](https://www.agno.com/articles/improve-multi-tenant-safety-with-user-scoped-memory-deletions.md) (SDK v2.1.2): All memory delete operations now honor user_id across every supported database backend. This ensures strict tenant isolation, reduces the risk of accidental cross-user deletions, and strengthens…
- [Standardize system to system integrations with an A2A JSON-RPC interface](https://www.agno.com/articles/standardize-system-to-system-integrations-with-an-a2a-json-rpc-interface.md) (SDK v2.1.2): Expose and run Agents, Teams, and Workflows over an A2A‑compatible JSON‑RPC interface with streaming responses and structured events.
- [Strengthen network security by upgrading h11 to a non-vulnerable version](https://www.agno.com/articles/strengthen-network-security-by-upgrading-h11-to-a-non-vulnerable-version.md) (SDK v2.1.2): We upgraded h11 from 0.14.0 to mitigate a known CVE. This reduces exposure to HTTP handling vulnerabilities and aligns deployments with current security best practices.

## 2025-10-04

- [MongoDB session serialization updated for higher reliability](https://www.agno.com/articles/mongodb-session-serialization-updated-for-higher-reliability.md) (SDK v2.1.1): We’ve updated the MongoDB session serialization format to improve consistency and resilience. Deployments that read existing session documents or rely on the previous format may encounter read/write…
- [Scale interface deployments with custom route prefixes](https://www.agno.com/articles/scale-interface-deployments-with-custom-route-prefixes.md) (SDK v2.1.1): You can now run multiple user interfaces on a single AgentOS instance by setting an optional route prefix for AGUI, WhatsApp, and Slack.

## 2025-10-01

- [Expand provider choice with Requesty LLM gateway](https://www.agno.com/articles/expand-provider-choice-with-requesty-llm-gateway.md) (SDK v2.1.0): Agno now supports Requesty, an affordable LLM gateway with advanced governance. Teams gain more control over costs and policy enforcement while keeping model choice flexible.
- [Increase safety and control with pre/post hooks and built-in guardrails](https://www.agno.com/articles/increase-safety-and-control-with-pre-post-hooks-and-built-in-guardrails.md) (SDK v2.1.0): You can now configure pre- and post-execution hooks for agents and enable built-in guardrails including prompt injection checks, PII detection, and OpenAI Moderation.
- [Scale retrieval with async batch embeddings across providers and vector stores](https://www.agno.com/articles/scale-retrieval-with-async-batch-embeddings-across-providers-and-vector-stores.md) (SDK v2.1.0): We added async batch embeddings for major providers and integrated them into most vector databases. This significantly improves throughput and reduces latency and cost for data ingestion, reindexing…
- [Simplify authentication and governance with pluggable middleware and built-in JWT](https://www.agno.com/articles/simplify-authentication-and-governance-with-pluggable-middleware-and-built-in-jwt.md) (SDK v2.1.0): AgentOS now supports user-supplied FastAPI-compatible middleware and includes a built-in JWT middleware for token validation and claims extraction.

## 2025-09-26

- [Enable better observability, routing, and control across model usage](https://www.agno.com/articles/enable-better-observability-routing-and-control-across-model-usage.md) (SDK v2.0.11): Agno’s LiteLLM model support now includes first-class metadata and additional fields. Teams can attach and propagate structured context alongside model calls, making it easier to track usage, apply…
- [Reduce configuration errors and avoid runtime conflicts](https://www.agno.com/articles/reduce-configuration-errors-and-avoid-runtime-conflicts.md) (SDK v2.0.11): AgentOS now more reliably discovers and registers MCP tools and databases during setup. The system also detects and rejects incompatible database instances that share the same identifiers, preventing…

## 2025-09-25

- [Explicit control over session state during AgentOS runs](https://www.agno.com/articles/explicit-control-over-session-state-during-agentos-runs.md) (SDK v2.0.10): Teams can now deterministically control whether an incoming run should override the session state already stored in the database.

## 2025-09-24

- [Accelerate database operations by writing multiple sessions and memories in a single call](https://www.agno.com/articles/accelerate-database-operations-by-writing-multiple-sessions-and-memories-in-a-single-call.md) (SDK v2.0.9): All Agno database implementations now support bulk writes, enabling multiple Sessions and Memories to be persisted with one operation.
- [Enable agents to run workflows directly through Slack.](https://www.agno.com/articles/enable-agents-to-run-workflows-directly-through-slack.md) (SDK v2.0.9): Agno now allows AgentOS workflows to integrate with Slack, expanding where agents can operate and interact with users.
- [Give agents direct access to Google’s MCP Toolbox for Databases to expand workflow capabilities](https://www.agno.com/articles/give-agents-direct-access-to-googles-mcp-toolbox-for-databases-to-expand-workflow-capabilities.md) (SDK v2.0.9): Agno introduces the MCP Toolbox for Databases, a new toolkit that allows agents to interact with structured data in Google’s MCP ecosystem.

## 2025-09-22

- [Control costs and prevent runaway tool calls in reasoning agents](https://www.agno.com/articles/control-costs-and-prevent-runaway-tool-calls-in-reasoning-agents.md) (SDK v2.0.8): Control costs and prevent runaway tool calls in reasoning agents
- [Increase workflow resilience with partial tool execution](https://www.agno.com/articles/increase-workflow-resilience-with-partial-tool-execution.md) (SDK v2.0.8): Agno now supports an allow_partial_failure option in MultiMCPTools, letting workflows continue even if some tools fail.
- [Integrate agents with CometAPI models for expanded capabilities](https://www.agno.com/articles/integrate-agents-with-cometapi-models-for-expanded-capabilities.md) (SDK v2.0.8): Agno now supports CometAPI as a model provider, giving teams more flexibility in how they build and deploy agentic workflows.
- [Simplify agent workflows by automatically passing dependencies to tools](https://www.agno.com/articles/simplify-agent-workflows-by-automatically-passing-dependencies-to-tools.md) (SDK v2.0.8): Agno now exposes tool dependencies as built-in arguments, making it easier to configure and run custom tools within agent workflows.
- [Streamline knowledge ingestion by adding multiple pieces of content at once](https://www.agno.com/articles/streamline-knowledge-ingestion-by-adding-multiple-pieces-of-content-at-once.md) (SDK v2.0.8): Agno now allows adding multiple text entries in a single call to Knowledge. Teams can populate knowledge bases faster, with fewer API calls and less orchestration overhead.

## 2025-09-18

- [Embed AgentOS in custom apps without CORS friction](https://www.agno.com/articles/embed-agentos-in-custom-apps-without-cors-friction.md) (SDK v2.0.7): AgentOS now works more reliably when embedded in custom applications. When teams provide their own app framework, Agno automatically ensures the AgentOS UI has the access it needs—without breaking…
- [Prevent unintended media from being sent to models](https://www.agno.com/articles/prevent-unintended-media-from-being-sent-to-models.md) (SDK v2.0.7): Agno now handles media routing more precisely when agents use tools that generate images or other media.
- [Run agents with local and self-hosted models using Llama CPP](https://www.agno.com/articles/run-agents-with-local-and-self-hosted-models-using-llama-cpp.md) (SDK v2.0.7): Agno now supports Llama CPP as a first-class model option, enabling teams to run agents on local or self-hosted LLMs.

## 2025-09-17

- [Build and route agents with a clearer, more extensible model abstraction](https://www.agno.com/articles/build-and-route-agents-with-a-clearer-more-extensible-model-abstraction.md) (SDK v2.0.6): We’ve added a first-class Nexus Model abstraction, giving teams a cleaner and more consistent way to define, route, and manage models across agentic systems.
- [Build more dependable agents with Gemini structured outputs](https://www.agno.com/articles/build-more-dependable-agents-with-gemini-structured-outputs.md) (SDK v2.0.6): Agno now handles Gemini schemas with nullable fields and complex definitions more reliably. This resolves issues that could previously cause structured outputs to fail or behave unpredictably.
- [Create PDFs, CSVs, JSON, and text files directly from agent workflows](https://www.agno.com/articles/create-pdfs-csvs-json-and-text-files-directly-from-agent-workflows.md) (SDK v2.0.6): Agno now supports file generation tools, enabling agents to produce structured file artifacts as part of normal execution.
- [Ensure accurate and private session history for every user](https://www.agno.com/articles/ensure-accurate-and-private-session-history-for-every-user.md) (SDK v2.0.6): Search results for session history are now correctly scoped to the current user. This change improves correctness, privacy, and trust—ensuring users only see and interact with their own session data.
- [Improve observability and auditability of agent workflows](https://www.agno.com/articles/improve-observability-and-auditability-of-agent-workflows.md) (SDK v2.0.6): Agno now persists the original workflow input alongside workflow run outputs. This gives teams full visibility into what triggered a run, making it easier to debug issues, audit behavior, and reason…
- [Upload and use PDFs and other files in chat without interruptions](https://www.agno.com/articles/upload-and-use-pdfs-and-other-files-in-chat-without-interruptions.md) (SDK v2.0.6): File uploads in the AgentOS chat experience are now more reliable, restoring support for common formats such as PDFs.

## 2025-09-16

- [Improve cost visibility and operational insight](https://www.agno.com/articles/improve-cost-visibility-and-operational-insight.md) (SDK v2.0.5): Agno now provides more accurate metric tracking for workflows using OpenAI Responses. This gives teams clearer visibility into usage, performance, and cost drivers when running agentic systems in…
- [Integrate AgentOS into existing APIs with more control](https://www.agno.com/articles/integrate-agentos-into-existing-apis-with-more-control.md) (SDK v2.0.5): AgentOS now offers expanded support for custom FastAPI applications. When routes overlap, AgentOS routes are applied by default, with the option to disable this behavior.
- [Lower risk and faster upgrades to Agno v2](https://www.agno.com/articles/lower-risk-and-faster-upgrades-to-agno-v2.md) (SDK v2.0.5): The v1 to v2 migration tooling has been updated to support metrics parsing and MongoDB migrations. This reduces manual work and uncertainty during upgrades, helping teams move to v2 with greater…
- [More predictable behavior for complex agent workflows](https://www.agno.com/articles/more-predictable-behavior-for-complex-agent-workflows.md) (SDK v2.0.5): AG-UI reliability has been improved to address issues such as duplicate events and multiple tool-calling inconsistencies.
- [More reliable interaction between UI and runtime](https://www.agno.com/articles/more-reliable-interaction-between-ui-and-runtime.md) (SDK v2.0.5): Ag-UI and AgentOS now integrate more cleanly, resolving issues that could cause inconsistent behavior between the interface and the underlying system.
- [Reduce friction and improve operational reliability](https://www.agno.com/articles/reduce-friction-and-improve-operational-reliability.md) (SDK v2.0.5): AgentOS now includes a default / route, eliminating unexpected 404 errors when accessing the service root.

## 2025-09-12

- [Enhance workflow reliability with structured input](https://www.agno.com/articles/enhance-workflow-reliability-with-structured-input.md) (SDK v2.0.4): Agno now supports TypedDict in addition to Pydantic for defining input schemas in agents, teams, and workflows.
- [Expand AI capabilities with a new model provider](https://www.agno.com/articles/expand-ai-capabilities-with-a-new-model-provider.md) (SDK v2.0.4): Agno introduces the SiliconFlow model class, enabling teams to integrate SiliconFlow models into their agentic workflows.
- [Run evaluation workflows at higher scale and speed](https://www.agno.com/articles/run-evaluation-workflows-at-higher-scale-and-speed.md) (SDK v2.0.4): All AgentOS evaluation workflows now support async tools through MCP, allowing parallel execution of tasks and faster throughput.
- [Visualize and interact with Vision AI workflows](https://www.agno.com/articles/visualize-and-interact-with-vision-ai-workflows.md) (SDK v2.0.4): Agno now includes a Streamlit application for Vision AI, providing an interactive interface to experiment with vision models and integrate them into workflows quickly.

## 2025-09-09

- [Access runs, summaries, and chat history more efficiently](https://www.agno.com/articles/access-runs-summaries-and-chat-history-more-efficiently.md) (SDK v2.0): Agno introduces new session convenience methods, providing streamlined access to runs, session summaries, and aggregated chat histories for easier workflow management and insights.
- [Comprehensive examples and clearer guidance for multi-agent systems](https://www.agno.com/articles/comprehensive-examples-and-clearer-guidance-for-multi-agent-systems.md) (SDK v2.0): Agno’s Cookbook documentation has been fully updated with more examples and structured guidance for building production-ready multi-agent workflows.
- [Create and track workflow events tailored to your business](https://www.agno.com/articles/create-and-track-workflow-events-tailored-to-your-business.md) (SDK v2.0): Agno now enables custom events in workflows, allowing teams to instrument workflows with domain-specific signals and outcomes.
- [Easier access to session, memory, and metric data](https://www.agno.com/articles/easier-access-to-session-memory-and-metric-data.md) (SDK v2.0): Sessions, memories, evals, and metrics are now managed through a simplified, unified storage system, reducing the complexity of tracking agent runs and workflow performance.
- [Introducing AgentOS: Simplify multi-agent management in production](https://www.agno.com/articles/introducing-agentos.md) (SDK v2.0): Agno introduces AgentOS, a production-ready API that consolidates agent, team, and workflow management into a single platform.
- [Migrate to AgentOS for a more unified workflow experience](https://www.agno.com/articles/migrate-to-agentos-for-a-more-unified-workflow-experience.md) (SDK v2.0): Playground, AGUIApp, SlackApi, WhatsappApi, and FastAPIApp have been replaced or integrated into AgentOS. Migration consolidates capabilities into a single platform, reducing operational overhead.
- [Run agents and workflows without session complexity](https://www.agno.com/articles/run-agents-and-workflows-without-session-complexity.md) (SDK v2.0): Agno now supports fully stateless agents, teams, and workflows, simplifying session management and making scaling more predictable.
- [Simplify deployment and reduce operational complexity](https://www.agno.com/articles/simplify-deployment-and-reduce-operational-complexity.md) (SDK v2.0.1): AgentOS no longer relies on the MCP dependency, making it easier to deploy and maintain across environments.
- [Standardized outputs and metadata for better observability](https://www.agno.com/articles/standardized-outputs-and-metadata-for-better-observability.md) (SDK v2.0): Run outputs and events are now standardized, with additional metadata for enhanced tracking, reporting, and debugging.
- [Stop agent or workflow runs safely and predictably](https://www.agno.com/articles/stop-agent-or-workflow-runs-safely-and-predictably.md) (SDK v2.0): Teams can now cancel agent, team, or workflow runs mid-execution while maintaining event integrity and state awareness. This feature improves operational control and reduces wasted compute.
- [Streamline knowledge management across agentic workflows](https://www.agno.com/articles/streamline-knowledge-management-across-agentic-workflows.md) (SDK v2.0): Agno’s unified knowledge system now supports multiple content types in a single structure. Teams can manage, access, and update knowledge consistently, enabling agents to deliver more accurate and…
