Gemini CLI v0.50.0: Tool Discovery via Tool Registry
Gemini CLI version 0.50.0 (July 8) introduces Tool Registry Discovery — automatic discovery and registration of external tools for AI agents.
60 noviniek
Gemini CLI version 0.50.0 (July 8) introduces Tool Registry Discovery — automatic discovery and registration of external tools for AI agents.
Perplexity expanded the Sonar API with access to GPT-5.6 and Grok 4.5 models and added a new finance_search tool for real-time financial data.
GitHub expanded Kimi K2.7 Code from Moonshot AI to Business and Enterprise plans on July 7, making it the first open-weight model in the Copilot model picker — 1T total parameters, 32B active, hosted on Azure with admin controls.
Starting July 9, 2026, ChatGPT users can no longer create new group chats — OpenAI is retiring the pilot feature to simplify the product and redirect development toward ChatGPT Work and agentic capabilities.
Google, alongside the Gemini Spark expansion to the AI Pro tier (July 24), rolled out key improvements: the agent is 50% faster via parallel source retrieval and can now edit shared Workspace files.
Anthropic launched Claude Reflect in beta on July 9, 2026 — a 'Spotify Wrapped' for AI usage showing monthly topic breakdowns, peak activity patterns, and optional wellbeing settings like break reminders and quiet hours.
GitHub made the Copilot agent-native desktop app available to all plans including Free on July 7, and rolled out GPT-5.6 Sol, Terra, and Luna in Copilot for paid plans on July 9.
xAI released Grok 4.5 — its first model purpose-built for coding and agentic tasks, co-trained with Cursor on real engineering tasks. Available via xAI API at $2/M input tokens and $6/M output tokens, with a 500K-token context window and configurable reasoning effort (low/medium/high).
New Claude Code version sets Claude Opus 5 as default model, enables nested subagents to depth 3, brings stricter sandbox network controls, and improved MCP error reporting.
OpenAI on July 8, 2026 launched GPT-Live, a new family of full-duplex voice models that listen and speak simultaneously, handle interruptions naturally, and can use web search and memory mid-conversation.
Anthropic expanded Claude Cowork with Microsoft 365 write access — Claude can now draft and send emails, manage calendar events, and create or update files in OneDrive and SharePoint on the user's behalf.
GitHub made agentic browser tools in GitHub Copilot generally available without experimental flags — an agent can now autonomously open a browser, navigate a running app, capture screenshots, and validate UI as part of a coding task.
Anthropic launched Claude for Government Desktop in public beta with FedRAMP High authorization, bringing Claude Code, Cowork, hash-chained audit logs, and SCIM-based budget controls to U.S. government agencies.
OpenAI released new Realtime API models on July 6—gpt-realtime-2.1 cuts p95 latency by 25% and the mini variant brings reasoning and tool use at 6× lower cost.
SpaceXAI added /goal to Grok Build for handing off large implementation tasks autonomously, and /voice (Ctrl+Space) for hands-free prompt dictation — the first native voice interface in Grok Build.
Anthropic made Claude Cowork available on web and mobile on July 7—tasks run autonomously in the cloud even when devices are offline, with status notifications.
GitHub Copilot agents in VS Code can now drive a real browser — navigating, clicking, typing, reading console output, and taking screenshots — enabled by default for all users as of July 1, 2026.
GitHub added AI credit pool support to enterprise cost centers on July 2, 2026, letting organizations cap monthly Copilot AI credits per cost center via REST API, with a UI management interface planned soon.
Cursor 3.10 (June 30, 2026) lets admins configure Team MCP servers once and distribute them across cloud agents, IDE, and CLI for all team members, while also adding org group support for marketplace access control.
DeepSeek's official V4 release is coming mid-July with a 1-million-token context window and the platform's first-ever peak/off-peak API pricing — calls during Beijing business hours will cost 2x the off-peak rate.
Codex 0.144.5 (July 16, 2026) expands dangerous shell command detection (rm and variants) and improves rejection informativeness — developers now receive a specific reason, not just a generic error message.
Perplexity AI introduces a hybrid orchestrator that automatically routes AI tasks between local device and cloud frontier models based on data sensitivity and compute requirements — no manual configuration needed.
GitHub Copilot Enterprise now streams agent session data (prompts, responses, tool calls) to external SIEM tools including Microsoft Purview. A REST API provides the last 48 hours of activity on demand.
Notion released version 3.6 on July 1, 2026, introducing External Agents support (Claude and Cursor as the first two), native Excel and PowerPoint file handling for Notion AI, speaker labels in meeting notes, and expanded MCP connections with stronger admin controls.
Anthropic released an expanded admin dashboard for Claude Enterprise: cost breakdowns by group and user, model-level entitlements, spend threshold alerts, and an Analytics API. Claude Code gets dedicated insight tabs updated daily.
GitHub Copilot Vision became generally available on July 1, 2026 for all Copilot plans including Free. Developers can attach images and PDFs directly to chat prompts so Copilot can reason about visual context alongside code.
Copilot CLI from July 1 automatically selects the best model per task (Auto mode, 10% discount). Enterprise customers now get real-time streaming of Copilot agent session data into external tools via a streaming endpoint or REST API.
Anthropic on July 22 updated the Claude Managed Agents API: configurable effort levels on model configuration, webhooks for environment and memory store lifecycle events, session seeding with initial events, and event deltas for subagent thread streaming.
xAI shipped a rapid series of Grok Build patches (0.2.66–0.2.73) over three days: sandbox profiles with kernel-deny lists, no-restart MCP server management, new agent dashboard with model visibility, and a fix for stdio hangs on Windows.
Claude Opus 4.8 and Haiku 4.5 are now GA in Microsoft Azure Foundry with native Azure authentication, consolidated billing through Microsoft Enterprise Agreements, and optional US data zone for compliance.
Anthropic released Claude Apps Gateway — a self-hosted control plane for enterprise Claude Code on Amazon Bedrock and Google Cloud, with SSO (OIDC), central policy enforcement, per-user spend limits, and OTLP telemetry.
xAI added 21 multilingual voices to Grok Voice on July 6 (25+ languages) and launched a no-code Voice Agent Builder for production voice agents with built-in telephony, MCP connectors, and observability.
The July 19, 2026 Claude Code update changes /verify and /code-review skill behavior — Claude no longer triggers them autonomously, only on direct invocation. Increases agentic pipeline predictability.
GitHub Copilot added Kimi K2.7 Code from Moonshot AI as the first open-weight model in the model picker, rolling out to Pro, Pro+, and Max plans with Business and Enterprise expansion planned.
Cursor updated CursorBench to v3.1 focused on codebase understanding, bug-finding, and code review across 36 models. Fable 5 Max leads at 72.9%, while Composer 2.5 offers the best value at $0.55 per task.
OpenAI opened a public beta for ChatGPT Sites — a feature that allows creating and publishing websites directly from a ChatGPT conversation.
Copilot Agent is now available inside JetBrains AI Assistant, expanding agentic coding to millions of IntelliJ-based IDE users. Enterprise admins can now set per-user AI credit budgets for cost centers. Claude Sonnet 5 is also GA in Copilot.
Cursor released its native iOS app for iPhone and iPad into public beta for all paid plans, letting users spawn and manage cloud agents, remote-control a desktop session, and code by voice from mobile.
Anthropic raised Claude API rate limits across the board — Sonnet and Haiku now match Opus at every tier — and collapsed usage tiers into three (Start, Build, Scale), powered by a new SpaceX compute deal.
GitHub Copilot now offers Claude Opus 4.8 in a new fast-mode preview that delivers significantly faster output token speeds while keeping the same intelligence as standard Opus. The variant is rolling out to Copilot users in preview, aimed at speeding up agentic and long-form coding tasks.
Vercel's AI Gateway adds beta support for realtime voice agents, text-to-speech, and speech-to-text, so developers can build conversational voice apps without juggling separate providers. Audio models route through AI SDK 7 with the same observability, BYOK and spend controls as text models, at no extra Gateway fee.
xAI's Grok audio family — realtime voice, TTS, and STT — is now routable through Vercel's AI Gateway with the same observability and spend controls as Grok's text models. Developers can compose voice agents with grok-voice-think-fast-1.0, grok-tts and grok-stt straight from AI SDK 7.
The desktop Copilot app can now connect your own API keys to OpenAI, Azure, Anthropic, Ollama or LM Studio. Keys are stored in the OS keychain and traffic routes through your own account or gateway.
The Copilot usage metrics API adds a new total_pull_requests_merged field in adoption phase reports. Enterprise admins now see daily merged PR counts by user adoption phase.
New `vercel metrics` command exposes Web Analytics directly via the CLI. Supports filters, dimensions, and is positioned for coding agents to answer questions about UTM campaign performance or mobile-vs-desktop conversion.
Concurrent build capacity for Pro teams jumped from 12 to 500. Enabled by default on Pro and Enterprise plans, billed on actual build minutes consumed.
Vercel auto-detects server.ts in the root or src/ and deploys it as a Node.js application without configuration. Supports native Node HTTP, Express, Koa, and NestJS.
Perplexity Deep Research upgrades to Claude Opus 4.6 for Max users. Finance equity pages gained analyst ratings, 52-week price targets, and AI-synthesized news. The Agent API adds GPT-5.6 Sol/Terra/Luna and Grok 4.5 as first-party options.
Microsoft on June 26 made MAI-Code-1-Flash, its fast 5B-parameter coding model, generally available for GitHub Copilot Business and Enterprise plans.
GitHub Desktop 3.6 brings native Git worktree support and Copilot SDK-powered commit messages and merge conflict resolution.
Cursor 3.9 (June 22) unifies management of plugins, skills, MCPs, subagents, rules, commands, and hooks in a new Customize hub including a leaderboard of popular plugins.
Copilot Code Review switched to grep/rg/glob/view tools and now costs ~20% less. Admins can set default review depth for the whole org. CLI adds queueing and debug-logs.
Qwen-AgentWorld is the first model trained to predict the next environment state (not the next action) across seven agent domains. Two MoE sizes, 256K context, Apache 2.0.
GitHub Copilot added an enterprise admin setting strictKnownMarketplaces that locks extension/plugin installs to approved marketplaces. It applies to both VS Code and the Copilot CLI, closing a supply-chain attack vector via third-party extensions.
Vercel added two new adapters to the AI SDK Harness — LangChain Deep Agents and OpenCode — letting you run either coding-agent runtime through one unified HarnessAgent API inside a Vercel Sandbox.
xAI added /goal to Grok Build — a long-running autonomous mode that plans, executes and verifies multi-step coding tasks with status, pause, resume and clear controls.
OpenAI detailed how it preserved private network boundaries while supporting MCP streaming, authentication and an inspectable client, so enterprises don't have to expose internal MCP servers to the public internet.
Copilot for Jira moves from public preview to GA, with streaming agent progress in Jira tickets, post-session steering that continues on the same draft PR, and simplified org/repo onboarding.
Zhipu AI released GLM-5.2 — an open-weight coding model with 1M context competing with Opus 4.8 and GPT-5.5. Available via the new GLM Coding Plan at ~10× lower price, with MIT-licensed weights.
Midjourney added automatic random styles to Draft Mode. Each batch now injects stylistic variations to broaden creative exploration.