Tool developments
Material changes that may affect your AI stack
A maintained buyer watchlist of meaningful product, pricing, availability, privacy, security, policy, and company changes—not a comprehensive or real-time AI news feed.
Two different histories
Tool developments
Vendor and market changes that may alter a buy, try, wait, govern, or avoid decision.
Choose AI Stack updates
Our editorial checks, content changes, and recommendation history remain in the Update Log.
Current window
Current developments
Material records published within the maintained trailing 14-day window.
- Availability
DeepSeek will retire its legacy API model names on July 24
DeepSeek says the legacy deepseek-chat and deepseek-reasoner API model names will become inaccessible after July 24, 2026 at 15:59 UTC; both currently route to DeepSeek-V4-Flash modes.
Why it matters
Teams using DeepSeek in production should migrate explicit model configuration to deepseek-v4-flash or deepseek-v4-pro before the cutoff and verify thinking-mode behavior, compatibility, latency, and cost in their own integrations. Leaving legacy names in deployed clients creates a near-term continuity risk.
Affected tools
References
- DeepSeek: DeepSeek API change log: DeepSeek-V4
- Product
Linear Agent adds recurring Loops
Linear introduced Loops, recurring jobs for Linear Agent that can run on a schedule or in response to an event while using workspace and connected-tool context.
Why it matters
Teams evaluating Linear as an agent-operations layer should assess shared visibility, trigger controls, connected-tool permissions, and failure ownership before relying on Loops for recurring triage, investigation, or project-maintenance work.
Affected tools
References
- Linear: Introducing Loops
- Product
Cursor expands Slack agent planning and repository context
Cursor added pre-run plans, multi-repository environments, and broader channel and thread context to its Slack agent integration.
Why it matters
Teams evaluating Cursor for delegated work from Slack can inspect a plan before execution and coordinate changes that span repositories. Repository access and channel-context permissions still need an explicit governance review.
Affected tools
References
- Cursor: Improvements to Cursor in Slack
- Product
GitHub Copilot adds repository-level usage metrics
GitHub added enterprise and organization REST endpoints that report daily, per-repository pull request activity for Copilot coding agent and Copilot code review.
Why it matters
Platform and engineering leaders can identify which repositories are actually using Copilot agents and reviews, then target enablement and governance more precisely. Access still requires the Copilot usage metrics policy and an eligible owner, billing-manager, or custom-role permission.
Affected tools
References
- Product
Bolt launches agent-built interactive presentations
Bolt launched Bolt Slides, an open-source presentation builder that turns prompts or uploaded material into live, shareable decks with interactive data, prototypes, and 3D experiences. It is available to free and paid Bolt.new users and can also run with Claude Code, Codex, or Cursor.
Why it matters
Teams considering Bolt for rapid internal tools or prototypes can now test the same build workflow for presentations and interactive leave-behinds. Review public-link access, live-data connections, and audience controls before using it for sensitive or externally shared material.
Affected tools
References
- Product
ChatGPT desktop adds Work continuity and unified recents
OpenAI updated the macOS and Windows apps with a Chat and Work switcher, unified recent conversations, Project access, and cross-device continuation for cloud Work conversations.
Why it matters
Buyers comparing ChatGPT with dedicated agent workspaces should account for a more continuous desktop workflow: longer-running Work sessions can now move between web, mobile, and desktop while staying connected to Projects. Local conversations remain device-bound.
Affected tools
References
- Product
Figma preserves variables when code-backed screens return to design
Figma now binds colors, type, and spacing to existing file variables when teams bring code-backed screens onto the canvas from Figma Make, the Figma MCP server, or the Chrome extension, while importing more frames with auto layout.
Why it matters
Teams evaluating AI-assisted design-to-code workflows can expect less manual rebuilding when moving implemented screens back into design. Preserved variables and layout behavior make the round trip more compatible with governed design systems, but teams should still verify complex component and token mappings in their own files.
Affected tools
References
- Product
Google Search AI Mode adds connected app actions
Google began rolling out connected apps in AI Mode in the U.S., letting users link services such as Instacart, Canva, and YouTube Music to add items, find templates, or save playlists from Search.
Why it matters
People evaluating AI Mode as an action layer should account for a broader set of third-party app permissions and handoffs, not only search and personalized answers. Teams should verify which accounts are linked, what actions require confirmation, and whether the rollout supports their region and managed-account policies before depending on it.
Affected tools
References
- Product
Grok 4.5 expands into coding and knowledge-work surfaces
SpaceXAI released Grok 4.5 for coding, agentic tasks, and knowledge work, with availability through Grok Build, Cursor, and the SpaceXAI API.
Why it matters
Teams evaluating Grok beyond conversational use should retest real coding and agent workflows rather than carrying forward assumptions from earlier models. Buyers should compare API cost, tool permissions, repository access, and review controls before standardizing on it.
Affected tools
References
- SpaceXAI: Introducing Grok 4.5
- Product
Kimi releases K3 across chat, agent, and coding surfaces
Moonshot AI released Kimi K3 with native vision and a one-million-token context window across Kimi chat, agent, swarm, coding, and API surfaces.
Why it matters
Teams considering Kimi for long-context or multimodal work should retest their real documents, tool calls, and coding tasks on K3 rather than carrying forward older-model assumptions. Verify API availability, cost, and governance requirements before rollout.
Affected tools
References
- Moonshot AI: Kimi K3 overview
- Product
NotebookLM becomes Gemini Notebook and adds code execution
Google renamed NotebookLM to Gemini Notebook and announced code execution for deeper notebook analysis, with broader synchronization across the Gemini app and Google Search planned.
Why it matters
Research teams should expect the product to become more tightly connected to Google's broader Gemini environment rather than remain an isolated notebook tool. Code execution may reduce handoffs to separate analysis tools, but rollout timing and cross-product sync availability still need verification before standardizing workflows.
Affected tools
References
- Product
Notion Agent adds calendar actions
Notion added calendar tools that let its Agent inspect and manage schedules, send invitations, join calls, and schedule time from the desktop app.
Why it matters
The Agent can now act across another high-value workplace system, which increases its usefulness for coordination but also raises the importance of limiting calendar access and reviewing actions before invitations or schedule changes are sent.
Affected tools
References
- Notion: Calendar tools for Notion Agent
- Product
Elicit opens research workflows through API and MCP access
Elicit launched API and MCP access for Pro plans and above, exposing literature search, research reports, and systematic-review workflows to external tools and agents.
Why it matters
Research teams can incorporate Elicit into repeatable internal workflows instead of relying only on the hosted interface. Buyers should include API usage, external-agent permissions, and source-review controls in the pilot design.
Affected tools
References
- Product
Microsoft 365 Copilot adds governed agent and prompt publishing
Microsoft added administrator-reviewed Agent Builder submissions to the organization Agent Store and tenant-wide collections in Prompt Gallery.
Why it matters
Enterprise teams gain a clearer approval and distribution path for employee-built agents and shared prompts. Administrators still need ownership, connector-permission, and lifecycle rules before broad internal publication.
Affected tools
References
- Product
Zapier moves standalone Agents into AI by Zapier
Zapier made looping tool calls generally available in AI by Zapier and began migrating standalone Agents into AI steps inside the Zap editor, with automatic conversion of prompts, tools, and triggers.
Why it matters
Teams evaluating Zapier agents should plan around the core Zap editor rather than the standalone Agents product. The integrated model makes it easier to combine agentic reasoning with deterministic steps, branching, filters, and automation history, while Enterprise trial users face an August 15 migration deadline.
Affected tools
References
- Availability
Anthropic launches free Claude access for verified US K–12 educators
Anthropic introduced Claude for Teachers with free premium access for verified US K–12 educators, teaching-oriented skills, curriculum resources, and access to Claude Code and Cowork.
Why it matters
Eligible educators can pilot a broader Claude toolset without a paid seat. Schools should still verify eligibility, district approval, and student-data boundaries before using it in classroom or administrative workflows.
Affected tools
References
- Anthropic: Introducing Claude for Teachers
- Product
Figma adds AI credit usage exports for admins
Figma Organization and Enterprise admins can download a CSV of AI-credit beta usage, including member activity and feature-level consumption.
Why it matters
Teams piloting Figma AI can use the export to identify adoption and heavy usage before broader rollout. Admins should still confirm how credits, retention, and access policies map to their plan before using the report for governance decisions.
Affected tools
References
- Figma: Figma release notes
- Product
Figma adds organization-level AI credit usage exports
Figma added a downloadable CSV for Organization and Enterprise administrators to review AI credit usage in beta features.
Why it matters
Larger teams can measure adoption and identify heavy usage before AI-credit purchasing decisions. Because the reporting covers beta usage, buyers should confirm how it maps to future billing and enforcement.
Affected tools
References
- Figma: See AI credit usage in betas
- Product
Superhuman Auto Drafts now prepare replies with calendar and web context
Superhuman updated Auto Drafts so every message that needs a reply can arrive with a draft informed by inbox history, calendar availability, and web research, with recipient-specific style adaptation.
Why it matters
Teams evaluating AI email assistants can test a broader reply workflow instead of a narrow follow-up generator. Buyers should still review each draft and verify what inbox, calendar, and web context is used before relying on it for sensitive or time-critical communication.
Affected tools
References
- Superhuman: Introducing: Auto Drafts 2.0
- Product
Codex adds inline visualizations and stronger task controls on iOS
OpenAI added inline visualizations to Codex tasks on iOS and improved task creation, task links, approval-preset handling, tool-activity feedback, and file-opening feedback.
Why it matters
Teams evaluating mobile supervision for coding agents can review richer task output and manage work more reliably from iOS. Buyers should still test approval behavior and handoff clarity in their own repositories before relying on mobile controls for consequential changes.
Affected tools
References
- Product
ElevenLabs adds service accounts and workspace credit caps
ElevenLabs added service-account management, workspace-member listing, and monthly credit caps when inviting workspace members, alongside per-agent sentiment analysis.
Why it matters
Teams can separate automation identities from human accounts and constrain new members' consumption at onboarding. Administrators should still test whether the available controls match their least-privilege and cost-allocation requirements.
Affected tools
References
- ElevenLabs: ElevenLabs changelog — July 13, 2026
- Product
Perplexity expands Computer context, publishing, and admin controls
Perplexity introduced Brain for source-linked context, faster Computer models, website publishing, organization controls for public publishing, and model-level usage analytics.
Why it matters
The release broadens Perplexity from research into persistent context and publishable work. Organizations should decide whether public publishing is allowed and review what content enters Brain before enabling wider use.
Affected tools
References
- Product
Superset adds rich terminal input for coding-agent prompts
Superset added a multiline rich-input composer for terminal panes, with file mentions and a persisted global toggle, so prompts for CLI agents such as Claude Code, Codex, and OpenCode do not have to be typed into a raw terminal line.
Why it matters
For teams piloting Superset as a local agent workspace, this lowers prompt-entry friction and makes agent runs easier to prepare and review. It supports a Try posture for macOS agent-heavy workflows, but it does not resolve platform, remote-workspace, or enterprise-governance caveats.
Affected tools
References
- Product
Cursor adds durable side chats and conversation search
Cursor introduced side chats that persist with a task, local transcript search across conversations, new cloud-agent hooks, and multi-repository selection.
Why it matters
Developers can investigate a tangent without disrupting the main agent run and recover prior decisions more easily. Teams should still treat locally indexed transcripts as project data and define retention expectations for sensitive repositories.
Affected tools
References
- Product
Replit lets workspace editors answer routine Agent questions
Replit expanded collaborative Agent access so editors can answer routine Agent questions while reserving sensitive steps involving secrets and integrations for owners.
Why it matters
Teams can share more of the build loop without making every collaborator an owner. Buyers should test the sensitive-step boundary against their own repository, deployment, secret, and integration controls.
Affected tools
References
- Replit: Replit updates — July 10, 2026
- Availability
Zendesk opens an employee-service AI agents early access program
Zendesk announced an early access program for employee-service AI agents that replace the traditional help-center search entry point with a conversational agent connected to an organization's knowledge.
Why it matters
Organizations considering Zendesk beyond customer support can now test an internal employee-service use case. Because the capability is early access and relies on company knowledge, buyers should validate access controls, content scope, escalation behavior, and rollout readiness before treating it as a production help-desk replacement.
Affected tools
References
Retained history
Earlier developments
Older records are preserved for decision history and are not presented as newly current.
- Product
Figma Make adds GPT-5.6 across all plans
Figma made GPT-5.6 available in Figma Make for users on every plan.
Why it matters
Teams can test the newer generation model without changing plan tiers, which may affect prototype quality and iteration speed. Model availability alone does not remove the need to review AI-credit consumption and generated-code quality.
Affected tools
References
- Figma: GPT-5.6 is now in Figma Make
- Product
Fin adds Zapier actions through an MCP connector
Fin added a Zapier MCP connector that lets teams authorize selected Zapier actions for Fin to run during customer conversations and reuse in Workflows and Procedures without custom action code.
Why it matters
Support teams considering Fin for action-taking automation should pilot the connector with a narrow action allowlist and human review for high-impact changes. The broader integration reach reduces custom-build work, but it also increases the importance of permission design, auditability, and rollback planning.
Affected tools
References
- Intercom: Zapier MCP Connector
- Availability
GitHub Copilot adds three GPT-5.6 model options
GitHub began rolling out GPT-5.6 Sol, Terra, and Luna in Copilot with plan-specific availability, model-dependent billing, and administrator enablement required for organizations.
Why it matters
Teams receive new speed and capability choices but need to compare plan access, usage multipliers, and organization policy before standardizing on a model. The existing Try posture remains appropriate while those tradeoffs are measured.
Affected tools
References
- Availability
Microsoft 365 Copilot makes GPT-5.6 its preferred model
Microsoft 365 Copilot made GPT-5.6 the preferred model across Copilot Chat, Word, Excel, PowerPoint, and Cowork.
Why it matters
Existing Microsoft 365 buyers can test the newer model inside core productivity workflows without adopting a separate assistant. Teams should recheck output quality and governance in their own documents and connected data before expanding use.
Affected tools
References
- Product
OpenAI introduces ChatGPT Work and unifies desktop agent workflows
OpenAI announced ChatGPT Work for longer-running tasks across apps and files, with plugins, Sites, Scheduled Tasks, administration and spend controls, and a desktop experience that brings Chat, Work, and Codex together.
Why it matters
Buyers can evaluate research, knowledge-work, and coding agents in one product surface, with initial Work rollout concentrated in higher tiers. Organizations should still pilot connector permissions, action approvals, retention, and spend controls before broad deployment.
Affected tools
References
- Product
Wispr Flow reports lower latency and fixes an accuracy regression
Wispr Flow reported 99.9% dictation uptime over recent weeks, 30% lower latency since the start of 2026, and a fix for an Auto Cleanup setting that had become too aggressive for some users.
Why it matters
Teams evaluating voice as a primary input layer have a stronger current reliability signal, but vendor-reported uptime and latency do not replace a pilot in the team's own apps, accents, and editing workflows. Buyers should verify whether cleanup behavior remains predictable for specialized vocabulary before wider rollout.
Affected tools
References
- Wispr Flow: Reliability and accuracy: where things stand
- Product
GitHub Copilot adds managed settings and telemetry controls
GitHub added mobile-device-management deployment for Copilot settings in VS Code and the CLI, plus enterprise-managed controls for OpenTelemetry export.
Why it matters
Enterprise administrators can enforce more consistent client configuration and govern whether Copilot telemetry leaves managed environments. Teams should map the available settings to their endpoint and observability policies before rollout.
Affected tools
References
- Availability
Runway retires Gen-3 models and points workflows to newer replacements
Runway removed Gen-3 Alpha on July 8 and plans to remove Gen-3 Alpha Turbo after July 30, directing text and image generation, keyframes, and video editing to newer model and app replacements.
Why it matters
Teams with saved Gen-3 workflows should migrate and compare outputs before the remaining Turbo cutoff. Replacement behavior, credit use, and edit controls may differ, so production templates should not assume drop-in parity.
Affected tools
References
- Product
Superhuman launches Docs as an AI-native collaboration surface
Superhuman launched Superhuman Docs as the evolution of Coda, adding a rebuilt Docs AI experience, AI-generated interactive views, enterprise-scale databases in beta, and MCP access for connected assistants.
Why it matters
Teams evaluating Superhuman as more than an email and writing layer should reassess it as a broader collaboration suite. Existing Coda customers should also review the new AI usage controls, workspace integration, and migration implications before standardizing on the bundle.
Affected tools
References
- Pricing
GitHub adds per-user Copilot budgets within cost centers
GitHub Enterprise Cloud customers can assign per-user budgets within cost centers, including controls for metered Copilot usage and AI credits.
Why it matters
Organizations can cap individual consumption without relying only on a shared cost-center ceiling, improving pilot cost control. Finance and platform teams should still decide how exhausted budgets affect developer workflows and exceptions.
Affected tools
References
- Product
v0 adds grouped approvals and team deployment policies
v0 added grouped approvals for MCP and shell actions, richer MCP tool support, and paid-team policies that can block disallowed repositories or production deployments.
Why it matters
Teams can reduce repetitive approval prompts while setting clearer deployment boundaries for generated applications. Administrators should test policy coverage and avoid treating grouped approval as a substitute for reviewing high-impact actions.
Affected tools
References
- Product
Superset opens terminal-agent registration beyond built-ins
Superset added Bring Your Own Terminal Agents, letting users register custom CLI agents with their own name, icon, and launch command alongside the built-in agent roster.
Why it matters
This matters for engineering teams that do not want their agent workspace decision locked to a fixed set of vendors. It makes Superset more maintainable as the coding-agent market changes, while procurement still needs to review each registered agent's own data handling and repository permissions.
Affected tools
References
- Product
Warp launched Oz for cloud coding-agent orchestration
Warp announced Oz as a cloud-based platform for running, managing, and orchestrating coding agents at scale, including team-visible agent runs, CLI/API access, and cloud environments.
Why it matters
This is a material part of why Warp is evaluated as an agentic development environment rather than only a terminal. The record stays in source history, but it is outside the public 14-day tool-page window.
Affected tools
References