[research]By ByteBulletin Editor
Anthropic CEO commits to embedded evaluators to pace AI
Dario Amodei outlines a three-part strategy for slowing frontier development, starting with third-party safety auditors holding company badges and desk access.
[tag]
4 stories
Dario Amodei outlines a three-part strategy for slowing frontier development, starting with third-party safety auditors holding company badges and desk access.
The RLHF pioneer joins the Safety and Security Committee as OpenAI faces renewed questions over agent containment failures.
As Anthropic prepares for a blockbuster public debut, scrutiny intensifies on its Long-Term Benefit Trust, a non-equity holding body that controls the majority of the board and aims to balance commercial viability with long-term safety.
The project’s new policy on generative AI in contributions is a pragmatic middle ground that leaves quality and responsibility where they always were.