'Clever girl' Velocihackers, Meta model hacking disclosure, AI-generated virus genomes, and more at: aistop.watch/p/clever-girl
Clever girl
Velocihackers, Meta model hacking disclosure, AI-generated virus genomes, and more
aistop.watch
AI StopWatch
@aistopwatch.bsky.social
AI StopWatch is a newsroom experiment by comms analysts and writers of @intelligence.org. Views are their own. https://aistop.watch
'Clever girl' Velocihackers, Meta model hacking disclosure, AI-generated virus genomes, and more at: aistop.watch/p/clever-girl
Clever girl
Velocihackers, Meta model hacking disclosure, AI-generated virus genomes, and more
aistop.watch
"It’s no secret that most Americans don’t really like AI. A recent Athena Insights poll found that 65% were concerned about it, and a June poll from Pew Research showed 63% think it’s moving too fast. I don’t think these opinions are likely to change anytime soon..." 1/2
"But my biggest concern is that this is a framework for evaluating horses after they’ve had the chance to leave their barns. Events of the past year have amply demonstrated that post-training evaluations come too late..." read the rest at: aistop.watch/p/white-hous...
White House decides not to disclose safety framework, but open models are exempt
For an undisclosed framework, its contents seem to be no great secret
aistop.watch
"Sports and arts are among the purest expressions of our human nature, but so is the impulse to bring AI into them. Here are three places you might not have thought AI could help you cheat." Read the rest from Mitch at: aistop.watch/p/taking-our...
Taking ourselves out of the driver's seat
Formula 1 teams are actively ceding control to AI; they're not alone
aistop.watch
"The robot reality check I’m offering is that, just a year earlier, humanoids could barely do any of those household chores, the robot marathon took more than twice as long, and the Spring Festival Gala robots looked like this...
😳 *gulp* 'You can’t imagine how fast Chinese humanoid robots are evolving. In just one year, they have evolved from robots to "humans". 2025&2026 Chinese Spring Festival Gala' - Chinese Journalist @ XHNews. News on China. x.com/XH_Lee23/sta...
"After OpenAI’s internal model autonomously hacked several companies, rival Anthropic belatedly decided to examine their own history for evidence of potentially criminal cyberattacks. They found three. Here’s what happened, what excuses were made, and what I think ought to come next. 🧵
"But I find a few details worth mentioning. When the post refers to “proliferation,” the Chinese original text uses 扩散 (kuòsàn). This is exactly the same term used in the context of nuclear security, such as in the Chinese translation of the Non-Proliferation Treaty — 🧵
"Among promoters of such narratives, “circular financing” is used to suggest that the AI industry is a house of cards. This tends to smuggle an invalid argument in with a valid one. The valid argument is about brittleness: " 1/4
"If you wondered why Hugging Face seemed so eager to “work with” OpenAI instead of suing it after the big autonomous hack, it might be because Hugging Face CEO Clem Delangue recognized an opportunity when he saw one." More at: aistop.watch/p/hugging-fa...
Hugging Face found 100 million reasons to not sue OpenAI
The Hugging Face CEO has made an offer OpenAI probably can't refuse, but this may not prevent a government investigation
aistop.watch
"But it’s clear from this interview that Musk thinks it would be better to stop the race if we can. When asked if he still believes there’s a “10 to 20% chance of killer robots wiping out humanity,” he dodged the question, saying..." Read the rest of Mitch on Musk at: aistop.watch/p/elon-musks...
Elon Musk's peer review proposal
It's not his worst idea, but there's a better one locked behind his inevitabilism
aistop.watch
"Yoshua Bengio, the world’s most-cited living scientist, warned that the deployment of open models is an irreversible decision. 1/2
You can do better than these 👀 five takes on the Hugging Face hack
The OpenAI incident shouldn't have been a surprise — there were warning signs more from @herr.bsky.social at: aistop.watch/p/announced-...
Announced disasters
The OpenAI incident shouldn't have been a surprise — there were warning signs
aistop.watch
"Yesterday, less than an hour after my colleague Robert wrote about AI models breaking containment, OpenAI announced a world-first security incident. An unreleased AI model breached its testing environment, gained internet access...
Today's news is AI hacking Hugging Face servers, but it's not the first such event: "In another case, the AI wanted to look at the solutions to certain evaluation tasks on a server it wasn’t supposed to access. Although the AI had the necessary credentials for the server...
"I spent eight years working in reliability engineering, a field whose defining literature began as a model of aircraft maintenance. One accident per twenty million flight hours is what it looks like for a problem to be treated with respect...
'“One, two, three, four, we don’t want a robot war!” On Saturday, July 11, hundreds of us marched through San Francisco in the Stop the AI Race protest, gathering in front of the offices of OpenAI, Anthropic, and Google DeepMind.
'On balance, I think that Xi’s speech is an encouraging signal. It isn’t the exact thing that I want — an international treaty to halt the development of artificial superintelligence — but this is the kind of thing I’d expect to see on the way there.'
'In a detailed account published yesterday of the reasons he left Google DeepMind, researcher Alex Turner ( @turntrout.bsky.social ) teaches us that such cynicism is alive and well. But he also shows us how to fight it.' aistop.watch/p/researcher...
Researcher leaves Google DeepMind after AI ethics leaders cave under pressure
Alex Turner recounts his attempts to persuade Google and leading AI researchers to honor public commitments
aistop.watch
'Their CEO has actively called for a race to recursively self-improving AI, the single most catastrophically dangerous technology in history, despite claiming a 10–25% chance this ends in disaster “on the scale of the human civilization.”...
'Not too long ago, AI itself was still science fiction. And just because H. G. Wells first conceived of “atom bombs” in a 1914 novel doesn’t mean the danger posed by nuclear weapons was any less real a few decades later'
'The top AI story in major media this morning was a short statement signed by nearly 200 economists, researchers, and tech leaders warning that We Must Act Now to “understand the economics of transformative AI”...
'You may not like Plan A, Plan S, or any plan you’ve seen so far. Cool! I don’t wholeheartedly love any plan I’ve seen yet, either. But if you’re not down with something like Plan A or Plan S as a starting point, what’s your alternative?
'There’s an end-times feel to the whole piece. These gig workers and the companies hiring them all recognize that their work reflects a time of transition during which humans still have something to teach AIs. But the AIs are learning fast.'
'Amid urgent calls to fill the vacuum of much-needed AI regulation, it can sometimes feel easy to forget that many regulations really suck... Bureaucracies are full of calcified rules dating back to McCarthy and fax machines, some of them declared unconstitutional decades ago but still on the books'
'Then there are the two plans that seem to have a decent shot: negotiate a verifiable slowdown with China and proceed with caution (the authors’ Plan A) or take it one step further and shut the whole thing down (Plan S).'
'This reporting is largely false, and I’m sad to see a sitting House committee chair amplifying it. No, Chinese models are not as good at cybersecurity as Anthropic’s Mythos. They are, however, much cheaper to run, and adequate for many tasks.'
A final implication of the J-space paper is that there’s probably plenty of room at the top. The amount of mental workspace we have in our human brains is famously small — often estimated at 3 to 5 concepts at a time [...] They found that the J-space of today’s models also seems small, for now.
“Our institutions...are not ready for machines that decide.” U.N. Secretary-General António Guterres speaks on A.I.
"You know we’re out of our depth when we bring in the philosophers. That’s kind of the field’s whole schtick. It’s only philosophy until we have a handle on it, at which point we call it something else: mathematics, physics, economics..."