Meta Rolls Out Meta AI in Threads DMs for All Users Globally
Meta rolled out Meta AI to Threads DMs globally on July 27, giving half a billion users private access to the AI chatbot — no more sharing conversations publicly in the feed.
14 noviniek
Meta rolled out Meta AI to Threads DMs globally on July 27, giving half a billion users private access to the AI chatbot — no more sharing conversations publicly in the feed.
Chinese AI lab Moonshot AI released full open weights for Kimi K3, a 2.8-trillion-parameter sparse MoE model. It is the largest open-weight model release in history, natively handling text, image, and video with a 1M token context window.
Inflect-Micro-v2 delivers complete text-to-speech synthesis in just 9.36M parameters -- dramatically smaller than standard TTS models -- enabling practical on-device voice generation without cloud dependency.
Anthropic published a detailed guide to context engineering for Claude 5-generation models, covering long contexts, instruction hierarchies, and agent memory management.
Moonshot AI will release open weights for the 2.8-trillion-parameter Kimi K3 on July 27, but independent testing revealed a 51% hallucination rate that the company omitted from its benchmark charts.
Google expanded access to Gemini Spark, its autonomous AI agent, from the exclusive AI Ultra tier ($100-200/month) to all US AI Pro subscribers ($20/month) on July 24, 2026.
OpenAI launched ChatGPT Work on July 9, 2026 — an agent within ChatGPT that takes over full projects across apps, works autonomously for hours, and returns finished outputs like spreadsheets, slides, documents, and web apps.
TracerML launched Echo, which dynamically selects and combines open-weight AI models to achieve results comparable to Anthropic Fable at roughly one-third of the cost. It earned 402 upvotes on Hacker News.
Black Forest Labs launched FLUX 3, a multimodal model trained end-to-end across image, video (up to 20 seconds), and audio. Unlike prior pipeline solutions assembling separate models, FLUX 3 is trained jointly across all modalities simultaneously. It launches in limited access.
The Cactus team post-trained Google Gemma 4 2B with a 68K-parameter probe that reads hidden states to predict p(wrong), achieving 0.814 AUROC vs. 0.549 for token entropy. Only 15-35% of queries need cloud escalation while matching frontier model performance.
Nonprofit open-source Git platform Codeberg announced two community-approved policies: banning use of user data for LLM training and discouraging hosting of primarily AI-generated (vibe-coded) projects, citing infrastructure load and the health of open-source collaboration.
Poolside released Laguna S 2.1, a 118B MoE open-weight coding model trained from scratch in 9 weeks, free on Hugging Face — beating DeepSeek-V4-Flash and rivals 10x its size on agentic coding benchmarks.
Nativ is a free, open-source macOS application built natively on Apple MLX for running frontier open-weight models locally on Apple Silicon with no accounts, subscriptions, or cloud dependencies.
Alibaba unveiled Qwen 3.8-Max-Preview at WAIC 2026 in Shanghai — a 2.4 trillion-parameter multimodal model. The company claims it ranks second only to Claude Fable 5, though no independent benchmarks have been published.