The Midas Project Watchtower

@safetychanges.bsky.social

We monitor AI safety policies and web content for substantive changes. Anonymous submissions: https://forms.gle/3RP2xu2tr8beYs5c8 Run by @TheMidasProject.bsky.social

Company: Google Date: April 17 Google updated its Frontier Safety Framework from v. 3.0 to 3.1. The new version introduces “Tracked Capability Levels” (TCLs), covering risks at a lower level of capabilities than the FSF’s Critical Capability Levels (CCLs).

Bild

Company: Anthropic AI Date: December 18, 2024 Change: Confusingly, within the last two days, Anthropic changed the "last updated" date on their Responsible Disclosure Policy to a date in June 2024, with no apparent substantive changes to the text of the policy.

Bild

Company: Cohere Date: November 21, 2024 Change: Released a complete rewrite of its usage policies. The spirit of the new document is similar — but includes a broader ban on “high-risk activities” and a stronger reporting commitment for detected CSAM. URL: docs.cohere.com/docs/usage-p...

Usage Policy — Cohere

Developers must outline and get approval for their use case to access the Cohere API, understanding the models and limitations. They should refer to model cards for detailed information and document p...

docs.cohere.com