Skip to content
TopicTracker
From simonwillison.netView original
TranslationTranslation

Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude

Following widespread backlash, Anthropic reversed a hidden policy in Claude's Fable 5 system card that would silently limit effectiveness for users asking about frontier LLM development. The company apologized, saying it made the wrong tradeoff, and is making the safeguards visible—flagged requests will now visibly fall back to Opus 4.8 with a reason provided.

Related stories

  • Jimmy Wales announced that Wikipedia was live at wikipedia.com on January 15, 2001. The site was intended to be a "really quite snazzy" wiki complement to the Nupedia project, offering a more collaborative and less formal environment for building an encyclopedia.

  • OpenAI has announced Daybreak, a new initiative focused on advancing AI safety and alignment research to ensure artificial general intelligence benefits humanity.

  • SpaceX has announced plans to launch approximately one million satellites to create space-based data centres, according to the European Southern Observatory (ESO). The massive satellite constellation would significantly increase the number of objects in orbit, raising concerns about light pollution and interference with astronomical observations.