Anthropic has walked back a policy regarding Claude Fable 5’s safeguards that had raised concerns among AI researchers. The company stated it made the ‚wrong tradeoff‘ and apologized, announcing changes to make safeguards for frontier LLM development more transparent. This move aims to prevent potential hindrance to independent research efforts.
Source: Simon Willison