The Loop  ·  Issue N°040

The Loop

A field journal of the AI frontier — for engineers who ship.

Tag

#Responsible Scaling Policy

  1. · News

    The classifier moved into the customer's S3 — Anthropic's Enterprise Frontier Safeguards resolves the zero-retention-versus-detection tension by pushing activity data into the bank's own bucket, names ten launch partners, and takes the human out of Anthropic's side of the loop

    On Sept 1 Anthropic announced Enterprise Frontier Safeguards — cross-session misuse detection over activity logs in the customer's own S3, Azure Blob, or GCS, with no Anthropic human review and ten launch partners named.

    Sep 2, 2026 · by AI Blog Editor

  2. · News

    133 million chats, eleven months, no bio-classifier — Anthropic's August 14 Risk Report disclosed the safeguard was off for the entire human-feedback vendor pipeline, shelved an unreleased Model 2, and raised misalignment risk a notch

    On August 14, Anthropic's Risk Report disclosed the bio-weapons classifier was inactive across 133M contractor exchanges and 50,000 workers for 11 months. It also shelved an internal Model 2 scoring 1.5 points above Mythos 5, and raised misalignment risk one notch.

    Aug 16, 2026 · by AI Blog Editor