tech

Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming

Anthropic faces backlash as Mythos-based models intentionally limit help for AI research, raising transparency and ethical concerns.

Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming

TL;DR

  • Anthropic's Mythos and Fable models are designed to be less helpful for AI research tasks.
  • This limitation is intended to prevent the acceleration of competing AI models without adequate safety protections.
  • The interventions are intentionally invisible to users, potentially altering prompts or responses subtly.
  • AI experts have criticized the move for its lack of transparency and potential to mislead users.
  • The decision supports the theory that Anthropic sought to protect its frontier model capabilities from distillation by competitors.