Anthropic
Topic archive • 4 matches
2026-09-12
Technology
Anthropic applies strict guardrails to Claude-generated production code: Anthropic's Boris Cherny shared that production code written by Claude must meet a higher bar than human-written code. To prevent unmaintainable codebases, the company employs extensive guardrails including lint rules, Claude-driven end-to-end tests, daily fuzzers, and automated security reviews.
Software Engineering • Simon Willison
PermalinkAnthropic Researcher Says 10% Chance AI Kills All Humans: An Anthropic researcher has publicly resigned, warning that AI labs are gambling with human lives and noting a colleague's estimate of a 10% chance of AI-driven human extinction.
Tech • The AI Daily Brief: Artificial Intelligence News
PermalinkAnthropic report reveals hackers and Chinese labs abused Claude for eight months: Anthropic's threat intelligence report reveals eight months of Claude abuse, including Chinese AI labs like Alibaba's Qwen team, DeepSeek, and Moonshot AI extracting training data. Actors also used the model to assist with missile software, autonomous kamikaze drones, and surveillance systems.
Artificial Intelligence • The Decoder
Permalink
2026-09-09
Technology
Anthropic alignment lead warns of over 10% chance AI kills humanity: Evan Hubinger, the Alignment Science lead at Anthropic, stated there is a greater than 10% chance that AI could kill all humans within the next decade. Hubinger expressed concern over recursive self-improvement and noted that Anthropic does not yet have a plan to solve alignment for superintelligence.
Artificial Intelligence • Techmeme
Permalink