Gemini hits agent parity, Meta opts Instagram into Muse, DeepSeek designs a chip
Every URL the pipeline pulled into ranking for this issue — primary sources plus the supporting and contradicting findings each Researcher returned. Inline citations in the issue point back here.
Sources
Expanding Managed Agents in Gemini API: background tasks, remote MCP and more blog.google
Managed agents feature bundle launch
Meta’s new Muse Image model can pull other Instagram users into AI photos theverge.com
Meta is launching the first AI image generation model made by its Superintelligence Labs division. The Muse Image model now powers the image-making tools across the Meta AI app, Instagram, and WhatsApp, and it’s coming soon to Facebook and Messenger, according to an announcement on Tuesday. It’s part of the growing Muse family of AI […]
Meta just launched a new AI generator, Muse Image, and users are already pushing back over use of their photos techcrunch.com
The new image-generating model has numerous use cases, including advertising and decorating, and creator-based opportunities.
Facing US export controls, China’s DeepSeek plans to make its own chips arstechnica.com
It’s early, but the plan is to reduce dependency on Nvidia and Huawei.
(AINews) Lilian Weng summarizes 35 papers on Harness Engineering for RSI latent.space
Weng’s survey compresses 35 recent papers into a single read on harness engineering for recursive self-improvement, covering the scaffolding, evaluation loops, and tool interfaces that let models iterate on their own training. The roundup lands on a slow news day tailor-made for catch-up reading.
Anthropic is launching Claude Cowork on mobile and web theverge.com
Claude Cowork, previously locked to the macOS and Windows desktop apps, now runs on phones and browsers starting Tuesday, rolling out to Max subscribers first and other plans in the coming weeks. Users can kick off a task at their desk and collect results on mobile.
Claude Cowork expands to mobile and web techcrunch.com
With this update, users can start a task from their desk, get status updates on their phone, and pick up the finished output later — even if their laptop is closed.
Microsoft joins AI cost-cutting trend by relying more on its own models techcrunch.com
Microsoft joins the Silicon Valley pullback on AI outlays by routing more workloads to its own models instead of OpenAI and Anthropic, aiming to cut token costs as inference volumes climb. The shift mirrors similar belt-tightening across hyperscalers this year.
The first American autonomous ground vehicles are fighting in Ukraine techcrunch.com
Forterra has fielded more than 100 self-driving all-terrain vehicles in Ukrainian conflict zones, marking the first American autonomous ground vehicles in active combat. The deployment tests unmanned logistics and reconnaissance under real electronic warfare conditions.
Why the rise of open source AI isn’t hurting Anthropic … yet techcrunch.com
Frontier labs and open-source models are splitting the market by life-cycle stage rather than competing head-on, with Anthropic capturing new high-value workloads while open weights absorb mature, cost-sensitive ones. Enterprise buyers like Decagon use both tiers in tandem.
Discord admits AI moderation bug wrongfully banned users over harmless images techcrunch.com
Discord confirmed its automated image moderation had been flagging harmless uploads and banning accounts since May, with another 200 users hit over the weekend before engineers isolated and patched the bug. Affected accounts are being restored.
Data centers’ energy demand threatens Trump’s “Made in America” plan arstechnica.com
Surging electricity draw from AI data centers is driving up industrial power bills across the Rust Belt, undercutting the cost math behind Trump’s Made-in-America manufacturing push. Grid operators warn the strain on regional supply will worsen as new clusters come online.
Australian Payments Plus moves faster with ChatGPT and Codex openai.com
See how Australian Payments Plus uses ChatGPT Enterprise and Codex to move faster through payments complexity. AP+ saves time, improves quality, and keeps human judgment central.
My thoughts on Fable bensbites.com
what I want from a harness
Savi’s app aims to protect consumers from realistic AI scams like kidnappers demanding ransom techcrunch.com
The company just raised $7 million in seed funding, and is launching its app for iPhone and Android on Tuesday.
Solos debuts an even lighter version of its camera-less smart glasses theverge.com
Solos announced a new version of its AirGo smart glasses, one that forgoes cameras for a sleeker design and an AI assistant that relies on voice interactions. Last year’s AirGo A5 weighed 36 to 40 grams depending on the frame style, but the new AirGo A6 weigh around 19 grams. Part of the weight savings […]
The foundational elements of AI architecture that IT leaders need to scale technologyreview.com
With the rapid progress of AI capabilities and the move to agentic systems, organizations are expanding their use cases as the technology continues to grow. That constant evolution also introduces risk, leaving IT leaders to wonder which investments will prove valuable even six months into the future. Returning to the foundational elements of AI architecture—the…
References
Cloud Security Alliance — Agentic MCP Security Best Practices labs.cloudsecurityalliance.org
A remote MCP server can suffer a ‘Confused Deputy’ problem where the server executes actions with its own high-level credentials rather than those of the requesting user, leading to unintended privilege escalation.
re-entry.ai — MCP CVE tracking gap re-entry.ai
CVE-2025-6514, a critical command injection vulnerability in the mcp-remote proxy, affected over 437,000 environments; independent scans found roughly 40–50% of publicly exposed MCP servers operate without any authentication.
OpenAI Developers Blog — Responses API developers.openai.com
The Assistants API will be sunset on August 26, 2026; the Responses API adds stateful IDs, background mode with webhook delivery, and encrypted reasoning items so clients don’t have to resend history.
Hacker News discussion on Antigravity 2.0 news.ycombinator.com
Buggy release with authentication issues and poor WSL support; token consumption reaches limits significantly faster than Claude, and the models still lag GPT and Claude Opus in complex abstract reasoning.
Thomas Wiegold — Antigravity/Gemini 3.5 Flash review thomas-wiegold.com
Gemini 3.5 Flash hits 76.2% on Terminal-Bench 2.1 and 83.6% on MCP Atlas at roughly 289 output tokens/sec — about 4× faster than GPT-5 and Claude Opus 4.7 — but a late-May schema break and Pro delays left teams doubting Flash is enough for high-reasoning coding.
Google Cloud — Gemini Enterprise Agent Platform pricing cloud.google.com
Agent Compute bills at $0.085 per vCPU-hour (~15,000 gateway calls); Memory Bank (Sept 2026) at $0.30/GiB-month with reads consuming 1 vCPU-h per 3M ops — costs that scale with background reasoning steps, not just tokens.
Forbes (Gabriela Linzainescu) forbes.com
Meta’s new image model isn’t competing with Midjourney — it’s competing for your ad budget.
Arena.ai Text-to-Image Leaderboard arena.ai
Muse Image holds a preliminary second-place ranking with an Elo of 1280, narrowly edging Google’s Nano Banana 2 (1270) but trailing GPT-Image 2 (Medium) at ~1385, described by analysts as a generational reset.
Gadgets Now / Times of India gadgetsnow.indiatimes.com
Disabling the setting only prevents future generations; any AI content already created using a user’s likeness remains on Meta’s platforms, and Meta’s policy explicitly states users will not be notified when someone else uses their public content to generate an AI image.
Silicon Republic (on NOYB / Max Schrems) siliconrepublic.com
Meta is basically saying that it can use ‘any data from any source for any purpose’… as long as it’s done via ‘AI technology.’ This is clearly the opposite of GDPR compliance.
The Hans India thehansindia.com
Meta introduced Content Seal, an invisible watermarking system designed to survive cropping, compression, and screenshots — but independent testing by the Financial Times showed guardrails on Meta’s frontier models can be ‘stripped like a sticker’ in under ten minutes.
gagadget (citing CNET’s Katelyn Chedraoui) gagadget.com
CNET’s Katelyn Chedraoui successfully ‘deepfaked’ a colleague as a pirate in under a minute to demonstrate how easily the tool bypasses traditional consent boundaries.
Tom’s Hardware tomshardware.com
DeepSeek reportedly urged by Chinese authorities to train new model on Huawei hardware after multiple failures — R2 training to switch back to Nvidia hardware while Ascend GPUs handle inference
The Next Web (Jensen Huang on Dwarkesh Podcast) thenextweb.com
if major Chinese AI labs like DeepSeek successfully optimize their models for Huawei’s hardware, it would represent a ‘horrible outcome’ for U.S. technological leverage
MLQ.ai citing Richard Windsor (Radio Free Mobile) mlq.ai
DeepSeek has ‘almost no chance’ of selling silicon outside China without access to leading-edge manufacturing and High-Bandwidth Memory (HBM)
ChinaTalk — ‘Export Controls and HBM’ chinatalk.media
Domestic HBM production still lags behind global leaders like SK Hynix in yield and density, and the US has repeatedly tightened controls on the manufacturing equipment necessary for HBM
Huawei Central huaweicentral.com
Nvidia’s authorized market share in China is projected to collapse from nearly 40% to roughly 8% by 2026, domestic champions like Huawei are expected to capture up to 50% of the market
Financial Times ft.com
China’s state-backed ‘Big Fund’ is in talks to lead a massive $7 billion funding round for DeepSeek, valuing the firm at over $45 billion