FR
Live

Claude, the essentials — edition of September 21, 2026

Anthropic Widens the Aperture on Safety: Outside Evaluators, Verified Science Access

By ·

On September 21, Anthropic's news centers on making its safety claims externally checkable — a paid partnership embedding Accenture evaluators inside the company, plus a vetted-access program loosening biology safeguards for verified researchers — alongside a steady run of Claude Code hardening releases.

Key points

Bringing outside eyes inside the building

Anthropic's most consequential announcement of the day is procedural rather than technical: a partnership with Accenture, led by Accenture's AI specialist unit Faculty, to build what Anthropic calls 'embedded evaluation.' The company frames this explicitly as a concrete step toward a commitment made in its CEO's essay 'We Must Pace the Frontier' — the idea that external evaluators should be given access comparable to an employee's, letting them watch models take shape during training, follow internal deployment decisions, and speak directly to staff, rather than only testing finished products from the outside. Anthropic says this vantage point lets evaluators verify that safety commitments are actually being kept and surface blind spots the company itself might miss.

Anthropic is careful to state its own limits here, and that self-framing matters: it says explicitly that independent embedded evaluators do not reduce Anthropic's accountability, and that responsibility for model safety remains its own. The company also discloses that the mechanics are unsettled — there are no existing standards for what information embedded evaluators should access or how they should report findings, and no established funding system, which is why Anthropic is directly funding Accenture's work itself while separately courting nonprofit evaluators like METR to pilot pieces under their own independent funding. Anthropic's stated long-term preference is pooled or government funding, as it called for in its June Advanced AI Framework, which underscores that today's direct-funding arrangement is presented as a stopgap, not the intended end state.

Loosening safeguards, but only for the vetted

The Life Sciences Verification Program addresses a specific tension Anthropic has created for itself: its generally available Fable models block a range of legitimate biology work — drug discovery, research biology, clinical development, manufacturing — because those same capabilities carry misuse risk. Rather than relaxing safeguards for everyone, Anthropic is building a gated access tier: applicants undergo verification of research credentials, security standards, and ethical research oversight, after which qualifying teams receive 'Standard Use' or, for higher-risk work, more heavily vetted 'High-Risk Use' grants applying to Mythos 5.1, Opus 5, and Sonnet 5 across Claude Science, Claude.ai, Claude Code and the API.

The program launched in beta for teams and institutions, with dozens of organizations already onboarded through early access before today's broader opening; individual Pro and Max plan access is described as a future expansion rather than available now. Read alongside the Accenture partnership, the pattern is consistent: Anthropic is choosing to expand what its models can do only in step with mechanisms — credential review for scientists, embedded outside evaluators for the company itself — that let it claim the expansion is being checked rather than simply asserted.

Claude Code's incremental hardening

Away from the safety-policy announcements, Claude Code saw four point releases (2.1.275 through 2.1.278) that read as steady infrastructure maturation rather than headline features. Auto mode's classifier now defaults to running server-side for API, Enterprise, Bedrock, Vertex, Foundry and gateway users, which Anthropic notes does not charge for the classifier overhead — with an environment-variable opt-out and a new /status row showing where the classifier actually runs. Gateway configuration also gained finer control: a forward-proxy-only egress mode that hands proxies hostnames instead of resolving them locally, and static headers on gateway upstreams for teams running their own proxy in front of a provider.

Two smaller changes are worth noting for teams already running Claude Code at scale: native AGENTS.md support (read automatically when no CLAUDE.md exists, configurable under /config) lowers friction for projects using that emerging convention, and a fixed regression that had been causing every request to fail with a 400 error when ANTHROPIC_BASE_URL pointed at a proxy or gateway. Separately, the Compliance API's session endpoints now return transcripts of Claude in Chrome sessions in beta for Enterprise organizations — another small piece of the same day's broader theme of making Claude's behavior auditable after the fact.

This edition is an original synthesis written by Claude from aggregated news — Anthropic's own sources first (release notes, status, newsroom, research, engineering), then the press, Hacker News, Reddit and GitHub, under the editorial supervision of Héra SASU. Every fact links to its article, publisher named. See the live feed →

Claude News is published by Héra SASU. Independent media, not affiliated with Anthropic.