Herdr: Agent multiplexer that lives in your terminal
A new open-source tool called Herdr lets developers run and orchestrate multiple AI coding agents in parallel from the terminal; it trended on Hacker News.
A new open-source tool called Herdr lets developers run and orchestrate multiple AI coding agents in parallel from the terminal; it trended on Hacker News.
The desktop Copilot app can now connect your own API keys to OpenAI, Azure, Anthropic, Ollama or LM Studio. Keys are stored in the OS keychain and traffic routes through your own account or gateway.
The Copilot usage metrics API adds a new total_pull_requests_merged field in adoption phase reports. Enterprise admins now see daily merged PR counts by user adoption phase.
New `vercel metrics` command exposes Web Analytics directly via the CLI. Supports filters, dimensions, and is positioned for coding agents to answer questions about UTM campaign performance or mobile-vs-desktop conversion.
Concurrent build capacity for Pro teams jumped from 12 to 500. Enabled by default on Pro and Enterprise plans, billed on actual build minutes consumed.
Vercel auto-detects server.ts in the root or src/ and deploys it as a Node.js application without configuration. Supports native Node HTTP, Express, Koa, and NestJS.
Perplexity Deep Research upgrades to Claude Opus 4.6 for Max users. Finance equity pages gained analyst ratings, 52-week price targets, and AI-synthesized news. The Agent API adds GPT-5.6 Sol/Terra/Luna and Grok 4.5 as first-party options.
OpenAI is leaning toward delaying its IPO to 2027 according to Bloomberg, which would put it on public markets after rival Anthropic, who is targeting October 2026.
Analog giant ON Semiconductor is acquiring Synaptics in an all-stock deal, absorbing its Edge AI compute, HMI and wireless connectivity portfolio. The move targets a $30B TAM expansion by 2030.
SoftBank shares fell as much as 13% on June 26 after a NYT report indicated OpenAI is leaning toward delaying its IPO to 2027 — the steepest single-day drop since August 2024.
DeepSeek open-sourced DSpark, a speculative-decoding framework that speeds up V4 Flash generation by 60–85% without retraining the model. The codebase is MIT-licensed.
Bloomberg flags that combined AI revenue at Meta, Alphabet, and Microsoft has for the first time exceeded depreciation on AI infrastructure — a milestone toward positive ROI.
Founder Conno Christou fed PET and MRI scans into Claude, which assigned 90% probability to post-chemo thymic rebound — later confirmed by a specialist. The case fuels the LLM-in-medicine debate.
Faros AI's review of telemetry from 22,000 developers across 4,000 teams shows AI coding tools boost output but cause bugs and incidents to surge faster — dubbed Acceleration Whiplash.
Google is dual-sourcing its 10th-gen TPU Icefish — TSMC produces the compute die on 1.4nm and Samsung Foundry the I/O chiplet on 2nm. Samsung's biggest AI win to date.
Krea released the 12B-parameter weights of Krea 2 — Raw for fine-tuning and Turbo for local 2K image generation in roughly 2 seconds.
Mistral AI raised $830M in debt (not equity) at an €11.7B valuation. A new 10 MW inference-focused data center at Les Ulis opens in Q3 2026.
Microsoft on June 26 made MAI-Code-1-Flash, its fast 5B-parameter coding model, generally available for GitHub Copilot Business and Enterprise plans.
A community guide for a multi-node AMD Strix Halo APU cluster with vLLM and RDMA, running outside the Nvidia stack, hit Hacker News. Strix Halo with 128GB unified memory is becoming a serious target for local LLMs.
GitHub Desktop 3.6 brings native Git worktree support and Copilot SDK-powered commit messages and merge conflict resolution.
Open-source project Wayfinder Router directs queries between local and hosted LLMs via deterministic rules instead of LLM-based classifiers. Aimed at audit-friendly hybrid deployments.
Cursor 3.9 (June 22) unifies management of plugins, skills, MCPs, subagents, rules, commands, and hooks in a new Customize hub including a leaderboard of popular plugins.
Copilot Code Review switched to grep/rg/glob/view tools and now costs ~20% less. Admins can set default review depth for the whole org. CLI adds queueing and debug-logs.
Qwen-AgentWorld is the first model trained to predict the next environment state (not the next action) across seven agent domains. Two MoE sizes, 256K context, Apache 2.0.
OpenAI publicly launched three new GPT-5.6 models for all ChatGPT users and API developers on July 9, 2026. For the first time in AI history, three frontier labs simultaneously had publicly accessible models.
Amazon unveiled Alexa+ Agentic Ads at Cannes Lions 2026 — a new conversational ad format letting users see an ad and finish a purchase entirely within an Alexa conversation. Launch partners include Papa Johns for food orders and artists Beck, Jill Scott and Omar Courtz for concert tickets.
Commerce Secretary Howard Lutnick wrote Anthropic that 'appropriate safeguards are in place' to release the Claude Mythos 5 cyber model to more than 100 US institutions including Fortune 500 companies.
Mistral released OCR 4, a structure-aware document AI model with bounding boxes, block classification and confidence scores across 170 languages. Pricing is $4 per 1000 pages, available via Mistral API, Amazon SageMaker, and same-day on Microsoft Foundry.
A TechCrunch analysis argues OpenAI is now in the same regulatory limbo as Anthropic — the government is approving model releases 'customer by customer', and both labs face the same downside if the system slows down.
Google has quietly pushed Gemini 3.5 Pro general availability from June to July 2026, citing tester feedback on token efficiency and long-horizon task performance. The delay puts the model head-to-head with GPT-5.6 and Claude Opus 4.7 in mid-July.
Wealspring Asset and Shanghai Banxia, two well-known Chinese hedge fund managers, warn that global AI stocks have entered a 'super bubble' territory and the collapse may not be far off.
Bloomberg's deep dive shows the AI industry has pledged $275M toward the 2026 US midterms, having already spent $44.5M on federal primaries. Marc Andreessen and OpenAI's Greg Brockman lead via the Leading the Future super-PAC, while Anthropic-funded Public First Action takes the opposite side.
A retail-fueled unwind in AI chip stocks reverberated through semiconductor shares across Asia and the US, hammering leveraged ETFs and denting newly launched SpaceX funds.
KPMG's Q2 2026 AI Pulse Survey shows enterprise agent deployment holding above 50% as organizations move from experimentation to multi-agent orchestration with concrete ROI. Investment levels remain steady, but the 'tokenmaxxing' waterfall is over.
At its first-ever Cannes Lions appearance, OpenAI publicly branded advertising as a core part of its strategy, with 900M weekly ChatGPT users and an internal target of $2.5B in ad revenue in 2026.
Cory Doctorow's Pluralistic essay chronicles Meta's escalating lawsuits against whistleblowers leaking internal documents about its AI, content moderation, and operations. Doctorow argues the tactic produces a Streisand effect amplifying the very leaks Meta is trying to suppress.
IEEE Spectrum's in-depth piece maps how frontier models now prove non-trivial mathematical claims and the existential questions this raises about mathematical creativity and verification.
In-depth analysis argues open-weight models (Llama, Qwen, DeepSeek, GLM) are rapidly closing the gap with closed-source frontier — backed by benchmarks and enterprise deployment implications.
GitHub Copilot added an enterprise admin setting strictKnownMarketplaces that locks extension/plugin installs to approved marketplaces. It applies to both VS Code and the Copilot CLI, closing a supply-chain attack vector via third-party extensions.
Sebastian Raschka's ~6,300-word essay walks through his practical local coding agent stack as a substitute for Claude Code and Codex subscriptions. The post responds to growing reader demand for self-hosted alternatives after recent Sonnet 4 / Opus 4 retirements and the Opus 4.6 fast deprecation in Copilot.
Vercel added two new adapters to the AI SDK Harness — LangChain Deep Agents and OpenCode — letting you run either coding-agent runtime through one unified HarnessAgent API inside a Vercel Sandbox.
Simon Willison breaks down OpenAI's GPT-5.6 Sol announcement — pricing, positioning in the model family, and practical implications for developers. He highlights that 'limited preview' with US-government vetting is a novel distribution pattern for frontier models.
xAI added /goal to Grok Build — a long-running autonomous mode that plans, executes and verifies multi-step coding tasks with status, pause, resume and clear controls.
OpenAI detailed how it preserved private network boundaries while supporting MCP streaming, authentication and an inspectable client, so enterprises don't have to expose internal MCP servers to the public internet.
Simon Willison covers Fernando Irarrázaval's public challenge where ~2,000 participants made 6,000 prompt-injection attempts against a Claude Opus 4.6-based assistant, with zero successful secret extractions.
Apple is reportedly skipping the M6 Pro/Max/Ultra variants and fast-tracking the AI-focused M7 family with higher memory bandwidth and an upgraded Neural Engine for on-device AI.
Apple raised Mac prices 15-20% and iPad prices 15-25% on June 25, citing an unprecedented AI-driven surge in DRAM and NAND demand. The hike is the first major consumer-facing sign of AI build-out costs bleeding into mainstream electronics.
OpenAI is leaning toward delaying its IPO to 2027 per NYT/Bloomberg reporting — with Altman pushing for a $1T valuation. SoftBank, whose OpenAI stake is projected to reach ~$65B by October, tumbled 13%.
CNBC reports the 'tokenmaxxing' era — where developers were rewarded for burning unlimited tokens — is ending as enterprises impose spending caps and pivot to cheaper Chinese models. The shift threatens growth assumptions baked into OpenAI's and Anthropic's IPO valuations.
Spatial-AI startup General Intuition raised a $320M Series A led by Khosla Ventures at a $2.3B valuation, training world models on gameplay clips from Medal's 17M monthly users.
Trase, founded by Grant Verstandig, raised a $107M seed led by ARCH Venture Partners to build an agentic operating system for regulated industries like healthcare and defense. Total funding now $117.5M.
Google released DiffusionGemma, an experimental 26B MoE model (3.8B active) with a diffusion head that generates text in parallel blocks rather than token-by-token. NVIDIA optimized it for RTX GPUs and DGX Spark.
Scaled Cognition raised $100M Series A led by Khosla Ventures at a ~$750M valuation, with Genesys joining as both investor and customer. The company sells its Agentic Pretrained Transformer (APT) model and simulation platform for high-stakes enterprise deployments.
Professional investors increasingly view public anger at AI — over rising electricity bills, data-center pushback, and job-loss fears — as a material risk to the tech-led market rally.
French digital health insurer Alan closed a €480M Series G led by Prosus at a €5.5B valuation. Q1 2026 ARR exceeded €800M with 53% YoY growth. One of Europe's largest 2026 health-tech rounds.
Google launched a dedicated Google Finance Android app featuring real-time market data, live news and an AI-powered Key Moments feature that explains why stocks are moving. An iOS version is coming later.
GPU cloud Runpod closed a $100M growth round led by Summit Partners at a $1B valuation. ARR doubled in six months to $240M. The company turned down acquisition offers above $500M to stay independent.
NYC officials backed away from a promised June release of AI guidance for the 1.1M-student school system. Over half of City Council members called for a two-year AI moratorium.
Patronus AI raised a $50M Series B led by Greenfield Partners and unveiled Digital World Models — large-scale simulation environments to stress-test AI agents before production. Revenue grew 15x year over year.
Xiaomi introduced HarnessX, a framework that treats the AI agent harness as a composable object the model can autonomously rewrite during a task. Smaller models benefit the most.