ChatGPT adds Zendesk and OneNote plugins for workspaces
ChatGPT and Codex have added beta Zendesk and OneNote plugins to the Plugin directory. The integrations let authorised workspace members work with support-ticket context or notes through supported actions, while keeping account connections and access permissions with each individual member within their established service permissions.
Google has launched agentic video understanding for Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite. The API feature lets models actively inspect video frames, audio and transcripts, aiming to cut token use and analysis cost while improving retrieval, counting and anomaly-detection tasks for developers working with long recordings.
OpenClaw adds macOS installer and RTX local-model onboarding
OpenClaw has added a guided macOS installer and automatic detection of existing AI access. On supported Windows NVIDIA RTX PCs, onboarding can now suggest and manage local 30B-class models, aiming to reduce setup friction while keeping device permissions visible and configurable. Users should still verify model sources, storage requirements and agent permissions before deployment.
NVIDIA introduces PAIR for local agent inference across PCs
NVIDIA has introduced PAIR, an open-source router that distributes local AI inference across compatible PCs. The IFA 2026 update also covers faster llama.cpp and vLLM inference, streamlined local-agent setup, and RTX Spark Windows PCs planned for October. Teams should test routing, device membership and data boundaries before relying on distributed local inference.
OpenAI has introduced GPT-6 Astra to a limited group of organisations in ChatGPT. The new model targets coding, research, computer use and longer multi-step work, while adding monitoring intended to flag possible instruction misunderstandings before wider availability. Administrators should validate workspace eligibility and retain human review for consequential actions.
Claude blueprint gives commerce teams a starting point for AI agents
Anthropic has released an open commerce-agent blueprint for teams building shopping and merchant assistants on Claude. It supplies reference implementations, integrations and guardrails for retail, travel, telecom and ticketing, while keeping checkout and consequential merchant changes under business control. The resources let organisations inspect practical designs before connecting live commercial data.
ChatGPT for Clinicians adds Healthcare Public Data search
Eligible US ChatGPT for Clinicians users can now add Healthcare Public Data, a read-only plugin that searches nine public health sources. The rollout expands research access without connecting patient charts, but organisations still need clear governance around prompts, availability, source validation and protected health information.
OpenClaw 2.0 rebuilds setup, browser experience and shared sessions
OpenClaw has released version 2.0, a major update that simplifies first-time setup, rebuilds its browser app and adds shared cloud sessions. The open-source personal AI runtime says the release also spans messaging, memory, skills, models, automations, plugins, security and a large collection of stability fixes.
xAI connects Grok Bot to X for social monitoring workflows
xAI has connected Grok Bot with X, letting paid users authorise Bots to search posts, read timelines, monitor mentions and collect current information. The integration creates an X developer account when required and supplies free API credits, turning a background AI colleague into a more capable social-workflow assistant for teams.
Anthropic extends Claude for Teachers to school districts
Anthropic has made Claude for Teachers available as a free Enterprise offering for eligible US schools and districts. The rollout adds centrally managed accounts, single sign-on, role controls, district privacy terms and new teaching resources, shifting the service from individual educator access toward organisation-wide deployment.
NVIDIA expands NVLink Fusion with NVHBM for custom AI infrastructure
NVIDIA has added NVHBM to its NVLink Fusion platform, putting the memory controller inside the HBM stack. The company says the design can improve bandwidth and power efficiency while giving cloud and AI-chip builders a more standard route to custom, rack-scale systems for enterprise AI teams.
Google previews Gemini Enterprise for financial services workflows
Google Cloud has previewed Gemini Enterprise for Financial Services, combining purpose-built agent skills, secure data connectors and a Financial Research agent. The platform is designed for regulated workflows, but its useful deployment will depend on each institution’s data entitlements, governance and review processes, including human validation of consequential outputs.
Google previews Gemini Enterprise for legal teams and law firms
Google Cloud has previewed Gemini Enterprise for Legal, a governed agent platform for law firms and legal departments. It combines legal skills, permission-aware connectors and partner agents for research, contract and compliance workflows, while keeping human legal judgment and existing data controls central across confidential, matter-specific work.
Claude Cowork adds a separate built-in browser for desktop agents
Anthropic is rolling out a built-in browser for Claude Cowork on desktop, letting users delegate website tasks without granting access to their personal tabs. The feature arrives for Pro, Max and Team users, with enterprise controls, but Anthropic says prompt-injection risks remain and high-impact tasks still need careful supervision.
Google Cloud adds flexible billing and spend controls for Gemini Enterprise agents
Google Cloud has introduced new Gemini Enterprise billing options and cost controls for agent workloads, including consumption pricing, pooled quotas, monthly caps and Flexible Savings Plans. The changes are designed to give organisations more predictable ways to scale AI agents while retaining centralised governance over spend.
OpenAI expands ChatGPT for Teachers to more US school districts
OpenAI is expanding ChatGPT for Teachers through 55 additional school systems, reaching more than 100,000 educators and staff. The announcement also introduces a 16-state data privacy agreement, aiming to make responsible district adoption, training, managed-workspace governance and local implementation planning easier to evaluate well.
Claude in Chrome reaches general availability with autonomous browser actions
Anthropic has made Claude in Chrome generally available across paid plans, adding autonomous browser actions guarded by prompt-injection defences, action checks and enterprise domain controls. The release turns the browser into a more capable workspace while keeping the desktop app necessary for local files and other applications.
Claude unifies memory across chat and Claude Cowork
Anthropic has made Claude memory work across chat and Claude Cowork, so context follows users between conversations and cloud tasks. People can inspect, edit or delete saved topics, while sensitive subjects remain excluded by default. The feature is available across consumer plans, with separate administrator controls for Team and Enterprise.
OpenAI adds an Admin plugin to ChatGPT Work and Codex
OpenAI has introduced an Admin plugin for ChatGPT Work and Codex, giving authorised workspace administrators a conversational way to inspect activity, manage membership and permissions, and handle supported usage actions. The release keeps existing role controls and approval requirements in place while reducing switching between consoles and reports.
NVIDIA puts Groq 3 LPX into production for agent inference
NVIDIA says its Groq 3 LPX inference system is now in full production alongside Vera Rubin NVL72. The company positions the LPU-based platform for faster token generation and long-context responsiveness in agentic applications, citing early deployments at Nebius and CoreWeave and a tighter integration with its wider AI-factory stack.
Grok Bot expands to SuperGrok Plus and Cursor plans
xAI has widened Grok Bot access beyond its initial beta. The always-on agent service is now included with SuperGrok Plus, Cursor Pro+ and Cursor Teams, giving more individual and team subscribers access to agents that can work across connected tools while requesting approval when needed.
Grok 4.6 becomes available through Google Enterprise Agent Platform
xAI has made Grok 4.6 available through Google Enterprise Agent Platform’s Model Garden, giving enterprise developers another route to use the model for long-running agent and interactive workloads. The announcement sets out a 500,000-token context window, configurable reasoning effort and listed token pricing.
OpenAI will expand ChatGPT Ads to 31 European countries next week, its largest rollout so far. The company says ads remain confined to Free and Go plans, separated from answers and unavailable to paid subscribers, while advertisers gain access through OpenAI, agencies and technology partners ahead of self-service tools.
Grok 4.6 becomes generally available through Amazon Bedrock
Grok 4.6 is now generally available through Amazon Bedrock, according to xAI. Supported AWS-region customers can access the flagship model through Bedrock with a 500,000-token context window, configurable reasoning effort and listed pricing of US$2 per million input tokens and US$6 per million output tokens.
OpenAI says it paused frontier training to strengthen cyber safeguards
OpenAI says it temporarily slowed frontier reinforcement-learning work while strengthening security, monitoring and alignment safeguards for models with advanced cyber capabilities. The company outlines workload and network isolation, continuous security testing, expanded chain-of-thought monitoring and a revised preparedness approach for higher-risk research.
OpenAI previews privacy-preserving safety checks for frontier APIs
OpenAI has previewed Private Safety Processing, a system intended to preserve Zero Data Retention for frontier-model API customers while identifying risk patterns across related interactions. The approach keeps customer content under customer-controlled encryption or infrastructure, with OpenAI receiving limited automated safety signals rather than readable prompts.
OpenAI and CodeAI launch education partnership around AI literacy
OpenAI and CodeAI have announced an education partnership to help students understand, question and create with AI. The programme combines an advisory council, classroom resources, Hour of AI activities, a national Builders Challenge and career exposure, alongside the newly introduced ChatGPT for Teens experience.
OpenAI launches a protected ChatGPT experience for teenagers
OpenAI has introduced ChatGPT for Teens, a learning-focused experience for users aged 13 to 17. It applies age-appropriate safeguards by default, adds parent controls and study features, and pairs the launch with an education partnership intended to help students use AI critically rather than treat it as an answer machine.
Grok 4.6 arrives in GitHub Copilot for everyday coding workflows
xAI has made Grok 4.6 available through GitHub Copilot, bringing its latest coding model to developers working in VS Code and across GitHub. Organisations may need to enable the model in Copilot settings, while developers can also access Grok 4.6 directly through xAI’s own console and API.
Cursor joins SpaceX, linking its AI coding product to a larger compute base
Cursor says it has completed its acquisition by SpaceX, following a partnership announced in April. The AI coding company expects access to SpaceX’s GPU fleet to support stronger and more economical models, while pointing to Grok 4.6 as an early example of what the combined organisations can build.
Google introduces Gemini 3.7 Flash with lower introductory pricing for coding agents
Google has introduced Gemini 3.7 Flash, positioning it as a stronger workhorse model for coding and agents. The company reports gains across software engineering, web development and knowledge work, and has set an introductory price through year-end that is half the original Gemini 3.6 Flash price per million tokens.
Claude explains how its upcoming text watermarking will work
Anthropic has outlined the technical approach behind text watermarks planned for future Claude models. The company says the system uses a version of SynthID-Text to help assess whether Claude contributed to longer passages, while preserving output quality and avoiding any user or organisation identifiers.
DeepSeek V4-Pro reaches general availability with new agent controls
DeepSeek has made V4-Pro generally available across its app, web service and API, adding native Responses API support and three thinking-effort settings. The company will also introduce peak and off-peak API pricing from 16 August, with off-peak rates set at half the peak price for scheduled workloads.
Mintlify lets coding agents create documentation accounts from the CLI
Mintlify has added a mint signup command that lets coding agents create an account, initialise a documentation project and deploy from the terminal. The release is designed to remove the browser-based sign-up step for agent-led documentation workflows while retaining email verification and locally stored authentication.
Grok Bot puts always-on cloud agents into xAI's beta
xAI has opened Grok Bot in early beta, offering cloud-based AI teammates that can work across apps, inboxes and websites while users are away. The service is initially limited to selected SuperGrok and Cursor subscribers, with enterprise access available through a waitlist. with a general release date yet to be announced.
Grok 4.6 targets longer-running AI agents and visual work
xAI has released Grok 4.6, positioning the model for longer-running agent tasks, coding and interactive visual projects. It is available through Grok Build, Cursor, the xAI API and selected partners, with input pricing from US$2 per million tokens and output pricing from US$6 per million tokens.
Claude outlines watermarks and provenance marks for AI content
Anthropic says Claude will add imperceptible watermarks to supported generated text and signed C2PA provenance metadata to selected image files. The approach is intended to help identify content Claude has processed, while recognising that marks cannot establish authorship or survive every form of editing unchanged.
Firebird has opened an NVIDIA-powered AI factory in Armenia that the companies describe as the CIS region’s largest. The project combines Dell systems, NVIDIA DSX infrastructure and planned Rubin and Blackwell GPU deployments, signalling a push to build AI capacity closer to regional research, enterprise and public-sector users.
GlideOS opens its app builder to external AI agents
GlideOS now exposes its app-building tools through an MCP server, allowing Claude, Cursor, ChatGPT and other compatible agents to create or edit Glide apps today. Connections inherit the user's organisation and plan access, while permissions and separate provider billing shape how teams control automated work.
Grok Imagine Image 2.0 brings precision editing to users
xAI has made Imagine Image 2.0 generally available as Grok's new Quality Mode, adding targeted edits, segmentation, background removal, multi-reference inputs and smart resizing. The release reaches Grok's web and mobile products now, with reusable creative templates included, while API access remains a future addition.
Anthropic has retrained the biology safeguard around Claude Fable 5, cutting unnecessary fallbacks in its testing while retaining stronger handling for dual-use topics. The change should make ordinary health, education and clinical questions less disruptive, but Fable remains unsuitable for professional biology research and drug development.
Friendli has added LG AI Research’s K-EXAONE 2.0 750B model to its serverless Model APIs. The 37-billion-active-parameter mixture-of-experts model offers a 262,144-token context window, reasoning and tool use across ten languages, giving developers access without operating the multi-node GPU setup documented for self-hosting.
Scenario rolls out model controls and new media tools
Scenario has released coordinated API, compute and web-app updates that tighten project-level model control while expanding video, 3D and workflow options. The 6 August package adds Grok and MiniMax video models, Meshy smart-topology assets, custom Cartwheel characters, richer Gaussian splats and more reliable moderation and automation controls.
SeedRealtime brings full-duplex audio and vision to live AI
ByteDance Seed has released SeedRealtime, a full-duplex model that continuously processes audio, video and text while deciding when to speak or act. The model is designed for fluid, interruption-aware interaction, with demonstrations spanning noisy group conversations, visual guidance and proactive assistance. Access and pricing details remain limited.
ThoughtSpot opens AgentSpot for governed business agents
ThoughtSpot has launched AgentSpot for creating agents and workflows from plain-language instructions while keeping them connected to governed business data. It combines model routing, identity controls and activity logs in one environment. Organisations can start with three custom agents at no charge, although broader pricing has not been disclosed.
Chatbase API now manages the full AI agent lifecycle
Chatbase has expanded its API so developers can create, configure, train, clone and delete AI agents programmatically. It covers data sources, model settings, widgets, access controls and retraining, making repeatable deployments easier. Teams should protect destructive endpoints and monitor training before exposing new agents in production.
Baseten has added NVIDIA Nemotron 3.5 ASR Streaming models to its model library, offering English and 40-locale multilingual transcription through NVIDIA NIM. Baseten reports support for 100 concurrent real-time WebSocket streams on one H100, with finalisation latency below 140 milliseconds in its test. Production results will vary by workload.
Qdrant 1.19 compresses vectors and unifies memory controls
Qdrant 1.19 introduces a four-bit Turbo4 vector datatype, a unified memory configuration and more precise tenant and text filtering. The release can substantially reduce storage, but compressed vectors trade some recall for efficiency. It also removes legacy query endpoints, so self-hosted users should plan staged upgrades and client compatibility checks.
Braintrust brings Cloudflare agent traces into AI evaluation
Braintrust now supports native tracing for agents running on Cloudflare, using OpenTelemetry data from the Agents SDK and related packages. Developers can inspect model calls, tools, sub-agents, tokens and Workers infrastructure in one trace, then turn production activity into evaluation datasets. Version and runtime prerequisites apply.
You.com Answer API pairs fast web answers with checked citations
You.com has added an Answer API that returns a web-grounded Markdown response with inline citations through one endpoint. It performs a single search, checks cited claims against source text and exposes the pages considered. At US$5 per 1,000 calls, it targets applications that need quick answers without building a retrieval pipeline.
FLUX.3 Video arrives with native sound and longer clips
Black Forest Labs has released FLUX.3 Video through its API and selected partners, combining video generation, native audio and editing controls in one model family. The launch supports clips of up to 20 seconds, multiple shots, keyframes and continuation from an existing video segment.
Harvey turns legal source material into review playbooks
Harvey has introduced a playbook builder that converts legal teams’ existing contracts, templates, marked-up documents and guidance into structured review rules. The builder identifies gaps, preserves source citations and lets lawyers refine the resulting playbook through a guided conversation before using it in Word or the web app.
Qodo brings organisation-wide code rules into Kiro
Qodo has released a Kiro Power that brings its code-review rules and cross-repository context into AWS’s agentic IDE. Developers can ask Kiro to review local changes, retrieve organisation standards, investigate the wider codebase and resolve findings before opening a pull request.
You.com APIs add keyless payments for autonomous agents
You.com now lets software agents pay per request for its Web Search and Finance Research APIs without an API key, account or prepaid balance. The endpoints advertise x402 and Machine Payments Protocol terms, allowing a funded wallet to settle a USDC payment in the same HTTP exchange as the data request.
GlideOS opens to the public with apps and AI workflows
Glide has publicly launched GlideOS with a new App Store, AI-generated workflows and connections to more than 1,000 business tools. Users can copy pre-built operational apps, describe automations in plain language and inspect each workflow’s triggers, steps and run history from a shared project workspace.
Databricks makes Unity AI Gateway generally available
Databricks has made Unity AI Gateway generally available, bringing access control, observability, cost attribution and runtime policy enforcement to agents, models, tools, skills and MCP servers. Smart Routing is also entering beta, with routing decisions based on quality, cost, performance, availability and budget.
OpenAI Details the Architecture Behind GPT-Live Voice
OpenAI has explained the low-latency architecture behind GPT-Live, the full-duplex voice system now powering newer ChatGPT Voice capabilities. Its design separates continuous audio from slower reasoning and tool work, while an upcoming API is intended to extend the same foundation to developers.
Databricks Makes Variant and Variant Shredding Generally Available
Databricks has moved its Variant data type and Variant Shredding optimisation into general availability. The release is designed to preserve flexible ingestion for semi-structured data while automatically improving query performance across Delta and Iceberg workloads, including data used by analytics and AI applications.
Databricks Completes Panther Deal to Expand Agentic Security
Databricks has completed its acquisition of Panther, bringing the security operations vendor’s detection engine, workflows and integrations into its Lakewatch security lakehouse strategy. The combined pitch centres on open telemetry storage and AI agents that can triage, investigate and refine detections across enterprise data.
Chatbase Adds SIP Trunking for Existing Business Numbers
Chatbase now lets organisations connect existing phone numbers to its AI voice agents through SIP trunks. The feature is designed for businesses that want to retain current numbers and telephony infrastructure while routing incoming calls to an agent configured with their own voice, knowledge and behaviour.