Samidh

@samidh.bsky.social

Co-Founder at Zentropi (Trustworthy AI). Formerly Meta Civic Integrity Founder, Google X and Google Civic Innovation Lead, and Groq CPO.

Runway is doing some very cool stuff using CoPE/Zentropi: "We have designed and tested a new fast moderation system that will run synchronously once a user’s input clears moderation.... The synchronous moderation through Zentropi narrows the exposure window to less than half a second."

Runway News | Moderation in Real Time

Runway built a synchronous moderation system for real-time video generation to scan streaming frames for harmful content in under half a second — layered on Runway's existing safety defenses.

runway.com

The safety team at RunwayML has been doing some of the most innovative work in the industry. It's exciting to see how they use Zentropi to ensure the safety of their videos in real-time **as they are being generated**. Details on their blog: runway.com/news/safety/...

Runway News | Moderation in Real Time

Runway built a synchronous moderation system for real-time video generation to scan streaming frames for harmful content in under half a second — layered on Runway's existing safety defenses.

runway.com

At TrustCon this year I talked about a technique we’ve developed for automatically optimizing content-moderation policies, using an inversion of the binocular labeling approach Zentropi had already pioneered. Today we're shipping the tool that technique became. blog.zentropi.ai/optimizing-o...

Optimizing our Policy Optimizers

Today, we are releasing our next-generation policy refinement tools: policy-only correction, label-only correction, and auto-optimization.

blog.zentropi.ai

The divide in America used to be rural vs urban. Then red states vs blue states. But now I think the real emerging gap is between W-2 Americans and 1099-DIV Americans.

The $12.7B in damages Meta will be paying to the states for harming children is less than... 1. How much it paid ScaleAI for Alexandr Wang ($14.3B) 2. Its one-day market cap increase upon launch of Muse Spark ($94.9B) 3. How much it burned on the Metaverse/VR in 2025 ($19.1B)

Banning all teens from social media hasn't ever seemed like a great idea to me. A much healthier approach is for platforms to build smoother on-ramps for teens that are developmentally appropriate for each child's age. 🧵 [1/n]...

I continue to be impressed with Zentropi. My first thought is that they were a replcement for Perspective API that Jigsaw was ending. And yes, but so much more. So far they can answer virtually any question about a text of almost any size my work cares about that CAN be answered with a Yes or a No.

Samidh@samidh.bsky.social · 4mo ago

Today, in a single day, @dwillner.bsky.social and I had meetings with people in England, Turkey, NYC, SF, Argentina, and Australia. Very inspiring to see four continents of folks all using Zentropi and united in the earnest work of building a better internet.

I distinctly remember being at Meta in the wake of the Christchurch massacre, when horrific videos were circulating across Facebook without end. The technologies we had for being able to accurately classify videos just didn't exist. [1/n]

Today we're releasing Coop 1.0, the world's first free, open source content review & enforcement system any org can self-host and build on. For the first time, any org, whatever its size or budget, can review, act on, and report CSAM end to end, for free. roost.tools/blog/coop-1-...

Coop 1.0: World’s First Free, Open Source Child Safety Infrastructure for Every Platform

Robust Open Online Safety Tools or ROOST is a new non-profit entity designed to address the urgent need for accessible, high-quality safety tools in the rapidly evolving digital landscape.

roost.tools

Today, in a single day, @dwillner.bsky.social and I had meetings with people in England, Turkey, NYC, SF, Argentina, and Australia. Very inspiring to see four continents of folks all using Zentropi and united in the earnest work of building a better internet.

Very cool! @julietshen.bsky.social's independent tests show that our new model CoPE-B cooks :-) Direct link to her results: github.com/julietshen/c...

github.com

Juliet Shen@julietshen.online · 5mo ago

I am NOT an AI engineer or AI researcher, but I tried to do a little evaluation of CoPE-B vs CoPE-A vs gpt-oss-safeguard github.com/roostorg/mod... lmk what you think, and we'd love for more evaluations to be part of the ROOST Model Community! cc @samidh.bsky.social

The ROOST Model Community is growing! Today we welcome Zentropi's CoPE-B-A4B, a bring-your-own-policy model that's got the power of 25B but runs on only 4B active parameters. It's a fast, low-cost model that can be used on its own or as a first pass before larger models! roost.tools/blog/welcomi...

Welcoming Zentropi's CoPE-B-A4B to the ROOST Model Community

Robust Open Online Safety Tools or ROOST is a new non-profit entity designed to address the urgent need for accessible, high-quality safety tools in the rapidly evolving digital landscape.

roost.tools

a few use cases I can think of for CoPE-B (and BYOP models): if you're a platform that's ok with NSFW role play but not age play, you can create a custom CoPE model that looks just for that. Many free text classifiers come with baked-in ideas of morality that might not fit your community

Dave Willner@dwillner.bsky.social · 5mo ago

We're releasing CoPE-B in collaboration with @roost.tools and will support it through the ROOST Model Community. It's a great forum for advancing AI-powered trust & safety tooling and we're excited to be a part of it. We'll be active there if you're building with CoPE-B and have questions! 🧵 9/9

We are THRILLED to announce that @roost.tools is jointly releasing CoPE-B with the Zentropi team. We believe that everyone should have access to openly licensed models designed for safety use cases that can be tuned to their community's norms. Join us at RMC office hours next week to learn more!

Dave Willner@dwillner.bsky.social · 5mo ago

We're releasing CoPE-B in collaboration with @roost.tools and will support it through the ROOST Model Community. It's a great forum for advancing AI-powered trust & safety tooling and we're excited to be a part of it. We'll be active there if you're building with CoPE-B and have questions! 🧵 9/9

Super pumped to release CoPE-B, our latest policy-adaptive content classification model. It delivers frontier-level accuracy in a self-hostable package that's orders of magnitude cheaper to run-- opening up new possibilities in trustworthy platform design. Details: blog.zentropi.ai/meet-cope-b-...

Meet CoPE-B: Frontier-Quality Content Classification You Can Self-Host

TL;DR: * Today we're releasing CoPE-B, our next-gen small language model for policy-adaptive content classification * CoPE-B-A4B (text-only) is open weights under Apache 2.0 and free to use * CoPE...

blog.zentropi.ai

One of the things we've been thinking about a lot at Zentropi is: what happens when AI agents need to make judgment calls about content — not humans reviewing a queue, but agents acting autonomously?