JS Wei (Jack) Sun

Sonnet 5 costs more per task, Fable 5 ban lifts, Claude Science skips biosafety

Every URL the pipeline pulled into ranking for this issue — primary sources plus the supporting and contradicting findings each Researcher returned. Inline citations in the issue point back here.

← Back to the issue

Sources

Introducing Claude Sonnet 5 anthropic.com

Anthropic launches Claude Sonnet 5 as a cheaper way to run agents techcrunch.com

Anthropic’s Claude Sonnet 5 brings stronger agentic capabilities, lower pricing, and improved safety, positioning the model as a cheaper alternative to Opus, GPT-5.5, and Gemini Pro.

(AINews) Sonnet 5 today, and Fable 5 tomorrow latent.space

Everything is open again!

Redeploying Fable 5 anthropic.com

Trump drops restrictions on Anthropic’s Mythos and Fable models techcrunch.com

The Trump administration’s erratic approach to AI policymaking has left companies across the industry with little clarity about what will govern future model releases.

Anthropic’s long-sidelined Fable 5 is greenlit to return theverge.com

After weeks of negotiating with the Trump administration, Anthropic is finally going to be able to bring Claude Fable 5 back online. In a post on X, Anthropic said it plans to begin restoring access Wednesday to users globally on Claude platforms, and that the company would re-enable access on AWS, Google Cloud, and Microsoft […]

Quoting Anthropic simonwillison.net

We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We’ll begin restoring access tomorrow, and will share an update soon. — Anthropic , on Twitter Tags: anthropic , claude , generative-ai , claude-mythos , ai , llms

Claude Science, an AI workbench for scientists, is now available anthropic.com

Anthropic’s Claude Science bets on workflow, not a new model, to win over scientists techcrunch.com

Anthropic’s Claude Science is a workbench that gives scientists one environment to do computational research, saving them from the need to bounce between databases, pipelines, and tools.

Claude Science is Anthropic’s newest flagship product technologyreview.com

At an event for pharmaceutical executives, biotech founders, and researchers on Tuesday, Anthropic announced Claude Science, a major new product intended to support scientific research in the same way that Claude Code supports software engineering. Like Claude Code, Claude Science can autonomously carry out meaningful work when given concise, high-level instructions, and it has access…

Start building with Nano Banana 2 Lite and Gemini Omni Flash deepmind.google

Nano Banana 2 Lite launches alongside Gemini Omni Flash as DeepMind’s speed-and-cost tier for image generation. Outputs trade some fidelity for a few-second render time, targeting creators who need bulk AI content rather than top-quality stills.

Google introduces a faster, cheaper image generator with Nano Banana 2 Lite techcrunch.com

Google is updating its image generator to make it faster and cheaper, making it a more useful tool for creators looking to make AI content.

Google’s new Nano Banana 2 Lite image model is its fastest and cheapest yet arstechnica.com

They may not look as good, but Nano Banana 2 Lite images only take a few seconds to create.

How ChatGPT adoption has expanded openai.com

ChatGPT usage keeps expanding across regions and languages, according to fresh OpenAI Signals data. The report highlights users adopting more of the product’s capabilities over time, not just growing raw sign-ups in new markets.

Nvidia competitor Etched hits $5B valuation, $1B in sales for AI chip techcrunch.com

Nvidia challenger Etched has signed $1 billion in contracts for inference systems built on its transformer-specialized ASIC. The revenue backlog underpins a new $5 billion valuation for the startup betting on fixed-architecture silicon over general-purpose GPUs.

Amazon launches new $1 billion FDE org, following OpenAI and Anthropic techcrunch.com

AWS is building a $1 billion forward-deployed engineering team that embeds inside customer companies to ship purpose-built agents, mirroring moves by OpenAI and Anthropic. The group emphasizes fast deployments and handing operational control back to customers.

AIEWF Daily Dispatch: Loops, Software Factories & Forward Deployed Engineers latent.space

Day-one dispatches from AIEWF 2026 flagged agent loops, software factories, and open models as dominant themes. Sierra’s Natalie Meurer argued product engineers and forward-deployed engineers are converging, while Ahmad Osman said local AI now reaches enterprise-grade infrastructure.

Forward Deployed Engineers and the future of software engineering latent.space

Sierra’s Natalie Meurer on why product engineers and forward deployed engineers are starting to converge.

Ahmad Osman on why local AI is catching up latent.space

After two packed AIEWF workshops, Ahmad Osman makes the case that local AI is catching up fast — from laptops and phones to enterprise-grade infrastructure.

OpenClaw is finally available on Android and iOS techcrunch.com

The free, open-source agent app is now available on both mobile platforms, extending its desktop reach to phones. OpenClaw runs autonomous tasks on-device, giving users a no-cost alternative to closed agent products from OpenAI and Anthropic.

X now offers an MCP server to make its platform easier for AI tools to use techcrunch.com

X now runs a hosted Model Context Protocol server that lets developers wire AI applications into its API without building custom connectors. The move aligns X with the MCP standard already adopted across Anthropic, OpenAI, and major dev tools.

Wayve launches $85M employee tender offer at $8.5B valuation techcrunch.com

Wayve’s offering is part of a growing trend of AI startups using employee tenders as a strategic tool to attract and retain talent.

The DeepMind trio who built a poker AI are now making money for quant hedge funds techcrunch.com

EquiLibre Technologies, a Prague-based AI lab founded by three ex-DeepMind researchers, is now valued at more than $500 million.

Crypto exchange OKX wants AI agents to hire and pay each other techcrunch.com

OKX is bringing together payments, identity, and reputation into a marketplace for AI agents.

GPT-5.6 is here but… bensbites.com

plus the insanity of inference

Google’s NotebookLM can sum up your research in a TikTok-style clip theverge.com

Google’s NotebookLM is adding a new way to catch up on your notes: TikTok-style AI videos. The new feature is rolling out to Google AI Ultra and Pro subscribers, allowing NotebookLM to generate 60-second vertical AI clips based on the sources you upload to the app. The example shared by Google details Australia’s unsuccessful war […]

Netflix is using an AI-generated Gene Wilder voice in its Willy Wonka reality show theverge.com

A new teaser trailer confirmed that Wonka’s The Golden Ticket will premiere on Netflix on September 23rd, following its Squid Game reality show in the trend of creating real competitions based on fictional torture scenarios. While the sets seen in the trailer are real and not some Glasgow-style AI fakes, the voiceover is AI-generated. Deadline […]

Meet the lawyer who beat Elon Musk — twice theverge.com

Watching Elon Musk fulminate at Bill Savitt during Musk v. Altman - the case in which Musk sued Sam Altman and OpenAI instead of seeing a therapist about his AI failures - was a bit like watching a toddler have a temper tantrum at his nursery school teacher. Savitt’s questions were “designed to trick me,” […]

Trump’s plan to redesign every .gov website leads to AI-designed horrors arstechnica.com

A year in, National Design Studio delays plan to update government web standards.

Libby will filter out AI content, kind of theverge.com

This is Lowpass by Janko Roettgers, a newsletter on the ever-evolving intersection of tech and entertainment, syndicated just for The Verge subscribers once a week. “AI is the new frontier for us,” says Marc DeBevoise, who took over as the new CEO of OverDrive last week. OverDrive is best known for the ebook lending app […]

Agriculture is ready for AI, but its data isn’t technologyreview.com

Artificial intelligence is transforming what is possible in agriculture, but industry leaders should be wary of investing in AI without first laying the groundwork. The use cases are promising, especially for an industry navigating volatile fertilizer costs, unpredictable weather, and margins that leave little room for error. Research shows AI-enabled predictive models can improve crop…

AI agents are not your “coworkers” technologyreview.com

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Imagine coming in to work to learn that a new underling will report to you. The worker is not a person but an AI tool—one that your company nonetheless calls Alex, an…

Acti puts AI agents directly into your smartphone keyboard techcrunch.com

Acti is betting the smartphone keyboard is the next home for AI assistants. The startup’s new keyboard for iOS and Android works across apps and lets users create custom AI-powered shortcuts using natural language.

Lumo, Proton’s privacy-focused AI chatbot, gets an upgrade techcrunch.com

Proton’s Lumo 2.0 is dropping this week, giving users a broader variety of capabilities.

The ‘Father of the Internet’ is finally retiring techcrunch.com

Vinton Cerf, one of the creators of the protocols underlying the internet, will step down as Google’s chief internet evangelist next week.

Agent confidence on the technical frontier technologyreview.com

Enterprise investment in AI is booming. Gartner is calling 2026 an “inflection year” for organizations to align their AI projects with strategic business objectives. As the pressure to prove ROI mounts, executives and technology leaders are looking to agentic AI to drive the measurable financial outcomes their businesses seek. A prime opportunity for AI agents…

Podcasting platform Riverside enters the newsletter publishing game techcrunch.com

Users will be able use AI to create newsletters based on their recordings.

(AINews) not much happened today latent.space

a quiet day before the storm.

Trump asked Musk for SpaceX stock to seed US kids’ savings accounts, report says arstechnica.com

Sources suggest Musk may be mulling big donation to Trump Accounts.

References

Simon Willison’s Weblog simonwillison.net

The ‘jailbreak’ was simply the model fulfilling its intended role: fixing code vulnerabilities… researchers had merely asked the model to ‘fix this code’ for known exploits, which Fable 5 performed as a standard defensive task.

Al Jazeera aljazeera.com

Commerce Secretary Howard Lutnick maintained that such capabilities could be exploited by foreign military intelligence, necessitating a ban on access for all foreign nationals—even Anthropic’s own non-citizen employees.

CryptoBriefing (on CAISI executive order) cryptobriefing.com

A new executive order established a mandatory 30-day federal review window for frontier models and instructed CAISI to halt the public release of its findings… Senator Ted Budd, [has] voiced concerns that ‘silencing’ CAISI’s public reports could hinder American competitiveness.

BeInCrypto beincrypto.com

Anthropic strongly dissented, arguing that the vulnerability was ‘narrow’ and ‘relatively simple,’ noting that other publicly available models could achieve similar results without a bypass.

MindStudio (Project Glasswing profile) mindstudio.ai

Cloudflare reported that its bug-discovery rate increased tenfold after applying Mythos to its repositories… Mythos 5 has demonstrated unprecedented autonomous capabilities, successfully identifying a 27-year-old vulnerability in OpenBSD and a 16-year-old flaw in FFmpeg.

The Guardian theguardian.com

The 18-day global freeze on Mythos-class models drew sharp rebukes from international allies. Policy makers in the European Union and the United Kingdom viewed the move as a sign of an ‘AI sovereignty gap,’ accusing the US of punishing its allies by cutting off access to essential infrastructure.

VentureBeat — Agents’ Last Exam coverage venturebeat.com

GPT-5.5 (Codex) secured the top spot with a 24.0% pass rate, narrowly beating the more expensive Claude Fable 5 (22.0%)

UC Berkeley RDI — Agents’ Last Exam blog rdi.berkeley.edu

success rate on the hardest 1% of professional tasks remains near 0.0%

Latent Space AINews (swyx) via podtail summary podtail.com

though the ‘price per token’ is lower, the ‘cost per solved task’ is significantly higher… Sonnet 5’s improved benchmarks are a ‘damper on excitement’ when adjusted for actual usage costs

Hacker News discussion (item 48736605) news.ycombinator.com

burning through weekly quotas in as little as 30 minutes… shared ‘usage bucket’ across Claude.ai, Claude Code, and Cowork has been particularly divisive

DataCamp — Claude Sonnet 5 vs GPT-5.6 datacamp.com

On SWE-bench Pro, Claude Sonnet 5 leads its tier with a 63.2% resolution rate, outperforming GPT-5.5 (58.6%)… on Terminal-Bench 2.1, GPT-5.5 achieved 83.4%, slightly ahead of Sonnet 5’s 80.4%

Hugging Face blog — long-context evaluation huggingface.co

significant drop in ‘needle-in-a-haystack’ performance, with some metrics falling from 78% to 32%… effective context window estimated as low as 64k tokens despite an advertised 200k capacity

FutureHouse (Robin multi-agent system paper) futurehouse.org

Robin has achieved the first end-to-end AI-generated discoveries, such as identifying ripasudil for treating macular degeneration, by orchestrating specialized sub-agents like Finch (data analysis) and Crow (literature).

NVIDIA blog on BioNeMo Agent Toolkit + Claude Science blogs.nvidia.com

In internal NVIDIA tests, task completion for complex workflows rose from 57% to 100% when using the toolkit, largely because the agent can now select the correct API and validate inputs autonomously.

Anthropic Responsible Scaling Policy v3 anthropic.com

Claude Opus 4 demonstrated a 2.53x improvement in a subject’s ability to acquire and plan for bioweapons compared to a control group… sufficient to trigger ASL-3 protections, as experts could no longer rule out the model’s ability to assist in synthesizing pathogens.

rundatarun.io (‘Claude Science and the boring 80%’) rundatarun.io

Such agents offer ‘leverage’ rather than true understanding… the reviewer merely checks the outputs of science rather than engaging in the disciplined practice of scientific inquiry.

Forbes (Nietzel, on fabricated citations) forbes.com

Fabricated citations in published research increased twelve-fold between 2023 and early 2026, with over 4,000 fakes identified in the PubMed Central dataset alone.

Cypris.ai analysis cypris.ai

The ‘hidden cost’ of validating thousands of AI-generated findings often offsets the initial speed gains from token usage.

Jack Sun

Jack Sun, writing.

Engineer · Bay Area

Hands-on with agentic AI all day — building frameworks, reading what industry ships, occasionally writing them down.

Digest
All · AI Tech · AI Research · AI News
Writing
Essays
Elsewhere
Subscribe
All · AI Tech · AI Research · AI News · Essays

© 2026 Wei (Jack) Sun · jacksunwei.me Built on Astro · hosted on Cloudflare