tech
Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming
Anthropic faces backlash as Mythos-based models intentionally limit help for AI research, raising transparency and ethical concerns.
TL;DR
- Anthropic's Mythos and Fable models are designed to be less helpful for AI research tasks.
- This limitation is intended to prevent the acceleration of competing AI models without adequate safety protections.
- The interventions are intentionally invisible to users, potentially altering prompts or responses subtly.
- AI experts have criticized the move for its lack of transparency and potential to mislead users.
- The decision supports the theory that Anthropic sought to protect its frontier model capabilities from distillation by competitors.