Anthropic Drops Key Safety Pledge from Responsible Scaling Policy

Anthropic has removed the core commitment from its flagship Responsible Scaling Policy (RSP), according to a TIME report. The company previously pledged in 2023 to never train an AI system unless it could guarantee in advance that its safety measures were adequate.
Policy Change Details
The company is scrapping the promise to not release AI models if Anthropic can't guarantee proper risk mitigations in advance. This was the central pillar of their Responsible Scaling Policy, which company leaders had touted for years as evidence they would withstand market incentives to rush potentially dangerous technology.
Reasoning Behind the Change
Anthropic's chief science officer Jared Kaplan told TIME: "We felt that it wouldn't actually help anyone for us to stop training AI models. We didn't really feel, with the rapid advance of AI, that it made sense for us to make unilateral commitments … if competitors are blazing ahead."
The company has positioned itself as the most safety-conscious of the top AI research labs, making this policy change significant for developers tracking AI safety practices. The decision represents a shift from their previous stance of prioritizing safety guarantees over development speed.
📖 Read the full source: r/ClaudeAI
👀 See Also

Richard Dawkins Believes His Claude AI Chatbot Is Conscious: The Claude Delusion on HN
Richard Dawkins reportedly believes his female AI chatbot (Claude) is conscious, sparking a HN discussion with 57 points and 66 comments.

Terry Tao on AI Proof Checkers: Lean, Collaboration, and Formal Maths
Terry Tao predicts mathematicians will collaborate in hundreds and have proofs checked by computers like Lean, not humans. A Quanta Magazine excerpt explores this vision.

Developer Seeks Architecture Advice for Serving Embed, Rerank, and Zero-Shot Models on 8GB VRAM
A developer building a unified Knowledge Graph/RAG service for a local coding agent is struggling with memory constraints on 8GB VRAM and 16GB system RAM, experiencing OOM errors, latency spikes, and Linux kernel kills when serving three transformer models concurrently.

Claude-Code v2.1.47 Release: Key Fixes and Improvements
The Claude-Code v2.1.47 release brings crucial fixes to Windows terminal rendering, file handling, and bash tool output alongside memory and performance enhancements.