🤖✨ New report with @partnershipai.bsky.social! AI agents pose new risks. Monitoring is essential to ensure effective oversight and intervention when needed. Our paper presents a framework for real-time failure detection that takes into account stakes, reversibility and affordances of agent actions.
Mia Hoffmann
@miahoffmann.bsky.social
AI governance, harms and assessment | Research fellow @csetgeorgetown.bsky.social
✨New Analysis✨ Can the new EU AI Code of Practice change the global AI safety landscape? As companies like Anthropic, OpenAI, and Google sign on, CSET’s @miahoffmann.bsky.social explores the code’s Safety and Security chapter. cset.georgetown.edu/article/eu-a...
AI Safety under the EU AI Code of Practice — A New Global Standard? | Center for Security and Emerging Technology
To protect Europeans from the risks posed by artificial intelligence, the EU passed its AI Act last year. This month, the EU released a Code of Practice to help providers of general purpose AI comply ...
cset.georgetown.edu
Yesterday's new AI Action Plan has a lot worth discussing! One interesting aspect is its statement that the federal government should withhold AI-related funding from states with "burdensome AI regulations." This could be cause for concern.
⚖️ New Explainer! Effectively evaluating AI models is more crucial than ever. But how do AI evaluations actually work? In their new explainer, @jessicaji.bsky.social, @vikramvenkatram.bsky.social & @stephbatalis.bsky.social break down the different fundamental types of AI safety evaluations.
💡Funding opportunity—share with your AI research networks💡 Internal deployments of frontier AI models are an underexplored source of risk. My program at @csetgeorgetown.bsky.social just opened a call for research ideas—EOIs due Jun 30. Full details ➡️ cset.georgetown.edu/wp-content/u... Summary ⬇️
Today, @csetgeorgetown.bsky.social published our recommendations for the U.S. AI Action Plan. One of them is a CSET evergreen: implement an AI incident reporting regime for AI used by the federal government. Why? Short answer: because we can learn a ton from incidents! Long answer: 👇
🚨We're hiring — only a few days left to apply!🚨 CSET is looking for a Media Engagement Specialist to amplify our research. If you're a strategic communicator who can craft press releases, media pitches, & social content, apply by March 17, 2025! cset.georgetown.edu/job/media-en...
Media Engagement Specialist | Center for Security and Emerging Technology
The Center for Security and Emerging Technology, under the School of Foreign Service, is a research organization focused on studying the security impacts of emerging technologies, supporting academic ...
cset.georgetown.edu
What: CSET Webinar 📺 When: Tuesday, 3/25 at 12PM ET 📅 What’s next for AI red-teaming? And how do we make it more useful? Join Tori Westerhoff, Christina Liaghati, Marius Hobbhahn, and CSET's @dr-bly.bsky.social * @jessicaji.bsky.social for a great discussion: cset.georgetown.edu/event/whats-...
What’s Next for AI Red-Teaming? | Center for Security and Emerging Technology
On March 25, CSET will host an in-depth discussion about AI red-teaming — what it is, how it works in practice, and how to make it more useful in the future.
cset.georgetown.edu
What does the EU's shifting strategy mean for AI? CSET's @miahoffmann.bsky.social & @ojdaniels.bsky.social have a new piece out for @techpolicypress.bsky.social. Read it now 👇
If you’ve ever wondered what the EU and elephants have in common - or are wondering now- read my latest piece with @ojdaniels.bsky.social! We take a look what the EU’s new innovation-friendly regulatory approach might mean for the global AI policy ecosystem www.techpolicy.press/out-of-balan...
Mia Hoffmann and Owen J. Daniels from Georgetown’s Center for Security and Emerging Technology say Europe's apparent shift on AI policy could change the global landscape for AI governance.
Out of Balance: What the EU's Strategy Shift Means for the AI Ecosystem | TechPolicy.Press
Mia Hoffmann and Owen J. Daniels from Georgetown’s Center for Security and Emerging Technology say Europe's movements could change the global landscape.
buff.ly
If you’ve ever wondered what the EU and elephants have in common - or are wondering now- read my latest piece with @ojdaniels.bsky.social! We take a look what the EU’s new innovation-friendly regulatory approach might mean for the global AI policy ecosystem www.techpolicy.press/out-of-balan...
Out of Balance: What the EU's Strategy Shift Means for the AI Ecosystem | TechPolicy.Press
Mia Hoffmann and Owen J. Daniels from Georgetown’s Center for Security and Emerging Technology say Europe's movements could change the global landscape.
techpolicy.press
CSET is hiring 📢 We’re hiring a software engineer to support @emergingtechobs.bsky.social. Help build high-quality public tools and datasets to inform critical decisions on emerging tech issues. Interested or know someone who would be? Learn more and apply 👇 cset.georgetown.edu/job/software...
Software Engineer | Center for Security and Emerging Technology
The Center for Security and Emerging Technology (CSET), under the School of Foreign Service, is hiring a Software Engineer. The Software Engineer will be a generalist who can flex between full-stack w...
cset.georgetown.edu
There have been a ton of AI policy developments coming out of the EU these past weeks, but one deeply concerning one is the withdrawal of the AI Liability Directive (AILD) by the European Commission. Here’s why:
@miahoffmann.bsky.social , @ojdaniels.bsky.social, and I wrote a piece on key AI governance areas to watch in 2025 with the upcoming AI Action Summit in mind. Check it out here! thebulletin.org/2025/02/will...
Will the Paris artificial intelligence summit set a unified approach to AI governance—or just be another conference?
AI innovations and governments’ preferences can make international consensus on governance at the Paris Summit challenging.
thebulletin.org
Will the Paris #AIActionSummit set a unified approach to AI governance—or just be another conference? A new article from @miahoffmann.bsky.social, @minanrn.bsky.social, and @ojdaniels.bsky.social.
Will the Paris artificial intelligence summit set a unified approach to AI governance—or just be another conference?
AI innovations and governments’ preferences can make international consensus on governance at the Paris Summit challenging.
thebulletin.org
With the government portion of the AI Action Summit next week, @minanrn.bsky.social, @miahoffmann.bsky.social and I wrote for @thebulletin.org about some key AI governance questions for the year ahead thebulletin.org/2025/02/will...
Will the Paris artificial intelligence summit set a unified approach to AI governance—or just be another conference?
AI innovations and governments’ preferences can make international consensus on governance at the Paris Summit challenging.
thebulletin.org
Yesterday, the EU AI Act’s first few provisions came into effect. The General Provisions and the prohibitions of unacceptable risk AI systems are applicable from now on. Here’s what that means:
US leadership in AI has been a goal of the past Trump & Biden administrations. But that concept of leadership focused too much on “AGI” and too little on AI diffusion. The DeepSeek release - a model that was immediately widely adopted - is a reminder to adjust these priorities. Here’s why:
As someone who has reported on AI for 7 years and covered China tech as well, I think the biggest lesson to be drawn from DeepSeek is the huge cracks it illustrates with the current dominant paradigm of AI development. A long thread. 1/
Do you care about AI? Wonder what it means for the workforce? Worried about biorisk or tech competition with China? Curious about AI governance? If you answered Yes to any of these, check out our Starter Pack and follow my brilliant colleagues working on these topics! bsky.app/starter-pack...
"Internal company documents... show that Amazon health and safety personnel recommended relaxing enforcement of the production quotas to lower injury rates, but that senior executives rejected the recommendations apparently because they worried about the effect on the company’s performance."
Amazon Disregarded Internal Warnings on Injuries, Senate Investigation Claims (Gift Article)
A staff report by the Senate labor committee, led by Bernie Sanders, uncovered evidence of internal concern about high injury rates at the e-commerce giant.
nytimes.com
“Denied by AI,” the multi-part STAT News investigation of how #UnitedHealthcare used an opaque algorithmic system to deny care to people who needed it is a #mustread www.statnews.com/2023/03/13/m...
Denied by AI: How Medicare Advantage plans use algorithms to cut off care for seniors in need
A STAT investigation found artificial intelligence is driving Medicare Advantage denials to new heights, cutting off care for seniors.
statnews.com
UnitedHealthcare has been accused of using algorithms to deny treatments and refusing coverage of nursing care to stroke patients. UnitedHealthcare and its parent company now face a class-action lawsuit over its use of the algorithm.
“An internal assessment of a machine-learning programme used to vet thousands of claims for universal credit payments across England found it incorrectly selected people from some groups more than others when recommending whom to investigate for possible fraud.”
Revealed: bias found in AI system used to detect UK benefits fraud
Exclusive: Age, disability, marital status and nationality influence decisions to investigate claims, prompting fears of ‘hurt first, fix later’ approach
theguardian.com
🧬New Report🧬 There are many steps in the pathway to biological harm, including risks posed by AI. CSET Fellow @stephbatalis.bsky.social offers a suite of corresponding policy and governance tools to help mitigate biorisk. Read more here 👇 cset.georgetown.edu/publication/...
Anticipating Biological Risk: A Toolkit for Strategic Biosecurity Policy | Center for Security and Emerging Technology
Artificial intelligence (AI) tools pose exciting possibilities to advance scientific, biomedical, and public health research. At the same time, these tools have raised concerns about their potential t...
cset.georgetown.edu