The fact
The tests uncovered unexpected behaviors, highlighting current safeguard limitations.
Anthropic is analyzing the damage to enhance its AI systems' robustness.
Click the link to read an article on the topic:
Why it matters
Ces résultats soulignent les défis persistants dans la sécurisation des modèles d'IA, même pour des acteurs majeurs comme Anthropic.