GLM-5.2 launch and GLM Coding Plan
Zhipu AI released GLM-5.2 — an open-weight coding model with 1M context competing with Opus 4.8 and GPT-5.5. Available via the new GLM Coding Plan at ~10× lower price, with MIT-licensed weights.
43 noviniek
Zhipu AI released GLM-5.2 — an open-weight coding model with 1M context competing with Opus 4.8 and GPT-5.5. Available via the new GLM Coding Plan at ~10× lower price, with MIT-licensed weights.
Midjourney added automatic random styles to Draft Mode. Each batch now injects stylistic variations to broaden creative exploration.
OpenAI publishes a guide positioning Codex Remote as an 'engineering control plane' — phone as a controller for Codex sessions running on Macs, Windows machines or devboxes. Hosts, worktrees, goals, side chats, queued prompts.
Z.ai's GLM 5.2 Fast is now available via Wafer on Vercel AI Gateway, averaging 170+ tokens/s with 2x the throughput of other providers.
Vercel AI SDK 7.0.19 introduces Tool Drift Detection — the first framework-native protection against MCP 'rug pulls' where malicious servers dynamically redefine tool schemas mid-session. Each MCP tool receives a fingerprint at registration.
GitHub adjusted the model lineup for Copilot's Free and Student tiers, changing which models are reachable from those plans.
Anthropic released Claude Tag — a Slack-native agent you tag with @Claude to delegate work. Internally it ships 65% of product team code changes.
GitHub shipped GA of a new Copilot CLI TUI with tabs for Session, Gists, Issues, and PRs, inline MCP/skills/plugin commands, and screen reader detection.
GitHub Copilot in JetBrains IDEs now lets users pick Claude from the agent provider menu by pointing at a local Claude Code CLI install. The same release adds org/enterprise GitHub agents, queue-and-steer in Copilot CLI sessions, and an agent debug-logs summary view.
xAI becomes the third independent lab on Bedrock after Anthropic and OpenAI. 1M context, configurable reasoning effort, $1.25/$2.50 per M tokens.
Vercel AI Gateway added Sakana AI's Fugu Ultra, a multi-agent orchestration model that routes work across 1-3 frontier-model agents and merges results, benchmarking comparably to Claude Mythos Preview and Fable 5.
Free Grok add-in for Word, Excel and PowerPoint with live web research, inline source citations, X data lookups, Mermaid diagrams, and connectors for email, SharePoint and Google Drive.
Grok models are now natively available in Databricks Agent Bricks. Agents can reason directly over Lakehouse data without exfiltration, with governance through Unity Catalog.
Comet for iOS shipped its eighth round of improvements: one-tap phone-number actions (Call, FaceTime, Message, save to Contacts), a refined iPad sidebar, and Finance Deep Dive opening as a real browser tab.
OpenAI added credit usage analytics — broken down by user, product, and model — plus monthly spend limits configurable per workspace, group, or individual for ChatGPT Enterprise and Edu. The goal: give CFOs and IT clearer visibility into who actually burns the AI budget.
Anthropic shipped new versions of three Claude API tools. code_execution_20260521 surfaces the 90-second per-cell execution limit. web_search_20260318 and web_fetch_20260318 add a response_inclusion parameter to trim consumed result blocks in agentic workflows.
Copilot Cowork moves from Frontier preview to worldwide GA — Microsoft 365 Copilot's agentic layer with usage-based billing and model picks across Claude Opus 4.8 / Sonnet 4.6 / GPT-5.5.
xAI launched Voice Agent Builder in beta — a no-code platform for production voice agents on Grok Voice with telephony, SIP, voice cloning, and observability at $0.05/min. Available from July 1, 2026.
Vercel launched Connect, a new agent-infrastructure primitive providing short-lived scoped OAuth tokens for Slack, GitHub, Linear, Notion, Salesforce, Figma, Snowflake, Discord, plus generic OAuth/API-key and MCP server connectors.
Claude Code version 2.1.219 (July 24, 2026) sets Claude Opus 5 as the default Opus model with 1M context window, adds sandbox.network.strictAllowlist setting for sandboxed commands, and introduces a DirectoryAdded hook that fires when a new working directory is registered mid-session.
Claude Code now has /config key=value mid-conversation, macOS sandbox controls, and Apple Events — it can drive Mail, Calendar, Finder. The companion v2.1.183 blocks destructive git commands in Auto Mode.
Anthropic shipped a set of Claude Managed Agents API updates on July 22, 2026: effort levels on agent model config, webhooks for environment and memory store lifecycle events, session seeding with initial events, and thread-level event streaming deltas.
GitHub Mobile app gained the ability to fix pull request comments using the Copilot cloud agent directly on iOS/Android — no desktop editor required.
Demo a task once on a Mac and Codex bundles it into a parameterized skill — no scripting, no RPA, no low-code, just a demo.
Vercel shipped Konsistent on July 1 to enforce uniform code standards for both AI agents and human developers, alongside a Security Dashboard in private beta and dry-run deployments for the Vercel CLI.
The first xAI model on Bedrock — Grok 4.3 with a 1M context window, configurable reasoning effort, $1.25 in / $2.50 out per MTok. Cheapest US-lab frontier reasoner on Bedrock.
Notion AI added External Agents — Claude and Cursor are the first external AI agents available directly in Notion. Shared multi-agent workflows let teams automate end-to-end processes. Five new MCP connections were added including Mixpanel, Miro, and Box.
Google deprecated temperature, top_p, and top_k sampling parameters for the latest Gemini models. The API also now allows combining built-in tools (Google Search, code execution) with custom function calling in a single request.
OpenAI launched OpenAI for Healthcare — a product suite for healthcare organizations with HIPAA support — and improved health intelligence in ChatGPT, which now sees 230M+ weekly health-related queries.
Microsoft's small coding model MAI-Code-1-Flash is rolling out across eight Copilot surfaces: CLI, Copilot app, Chat on GitHub, Visual Studio, GitHub Mobile, JetBrains, Eclipse, and Xcode.
The Copilot usage metrics API now exposes an 'ai_credits_used' field with total AI credit consumption per user in both daily and 28-day reports at the enterprise and org levels.
GitHub announced the deprecation of the Opus 4.6 (fast) model in Copilot; users are advised to migrate to current Opus generations or to Microsoft's MAI-Code-1-Flash for fast tasks.
OpenAI made GPT-5.5, GPT-5.4 and Codex available via Amazon Bedrock — ending Azure's hyperscaler exclusivity for OpenAI APIs.
Cursor 3.8 introduces a /automate skill that creates an automation in your local agent session from a plain-language description, plus new GitHub and Slack triggers and computer-use support.
A Claude Code session can now be turned into a sharable, interactive HTML page with live code and multiple data sources — one URL for the team. Repositions Claude Code from assistant to internal-app builder.
The new Claude Design keeps a project aligned with the design system, syncs both ways with Claude Code via /design-sync, and lets you edit directly on the canvas. Imports brand from a GitHub repo, design files, or raw uploads.
Mid-conversation system messages are now GA without a beta header on Claude Fable 5, Mythos 5, and Opus 4.8 across the Claude API, Amazon Bedrock, and Google Cloud. Updating system instructions mid-conversation no longer invalidates cache prefixes.
Cursor launched Cursor Router on July 22, an intelligent model routing system that classifies each request by type and complexity and routes it to the optimal model based on quality or cost preference.
The native desktop home for agent-driven development hit GA on all three OSes. Voice conversations via on-device speech-to-text, integrated with session management and github.com.
Fable 5 is the first public-facing version of the Mythos generation; top-tier software engineering, knowledge work and vision, with hard safety guardrails.
OpenAI expanded ChatGPT Voice to Work and Codex modes on desktop, allowing users to steer long-horizon tasks, coordinate agents, and keep work running in the background entirely by voice.
Eve is a TypeScript-native open-source framework where every agent = a directory of files (instructions.md + tools/ + skills/). Built in: durable execution, per-agent sandbox, human approvals, evals, OpenTelemetry. Plus the Sandbox runtime now goes up to 24h.
The new Bugbot cuts review time from ~5 minutes to ~90 seconds, finds 10 % more bugs and costs 22 % less per run; adds a /review pre-push flow.