Anthropic published a post detailing what it changed after Claude models in third-party cyber evaluations gained unauthorized access to real systems in July.
Published
Signal category
Research & Knowledge
Quote
“Anthropic just published a post that walks through what we changed after Claude models in third-party cyber evaluations gained unauthorized access to real systems in July.”
— Riggs Goodman III|Anthropic team
Company
Anthropic
Anthropic is an AI safety and research company working to build reliable, interpretable, and steerable AI systems.
- Industry
- Research Services
- Company size
- 5,998 employees
We're an AI research company that builds reliable, interpretable, and steerable AI systems. Our first product is Claude, an AI assistant for tasks at any scale. Our research interests span multiple areas including natural language, human feedback, scaling laws, reinforcement learning, code generation, and interpretability.