Another happy read that we should ponder and reflect on: Forethought's "AI-Enabled Coups". It's a careful, refreshingly unhysterical paper on how a small group (or literally one person) could use advanced AI to seize a state. Ofc: in personal capacity, not on behalf of Google
Andreas Kirsch
@blackhc.bsky.social
My opinions only here. 👨🔬 RS DeepMind Past: 👨🔬 R Midjourney 1y 🧑🎓 DPhil AIMS Uni of Oxford 4.5y 🧙♂️ RE DeepMind 1y 📺 SWE Google 3y 🎓 TUM 👤 @nwspk
The Atlantic @TheAtlantic says generative AI is "an engineering disaster." I had Claude fact-check all 23 checkable claims against primary sources: 7 check out · 7 need context · 9 don't hold Verdict: the economics hold up. The computer science doesn't (I checked it too) 🧵
I work at Google DeepMind. This won't make me popular. But it's all public reporting: 2014: DeepMind reportedly sold to Google on conditions: no military use, independent oversight 2026: a Pentagon contract for "any lawful government purpose" Not one safeguard survived intact
A serious essay by me with many personal thoughts: utaw.tech/news/trust-i... It's on us as Google DeepMind employees to demand real governance, and our union's (@utaw.tech) recognition push is the most realistic path to get there before it's too late
UTAW: Trust is not Governance
DeepMind has bet that a strong safety culture and good leadership built on trust are sufficient to withstand outside pressure. The bet has failed.
utaw.tech
Vibe-improved a small useful tool to render markdown & html directly from GitHub URLs, so you don't have to setup GitHub Pages etc E.g. `mdrenderer․github․io/?https꞉//github․com/mdrenderer/mdrenderer․github․com/blob/master/readme․md` (All thanks to Claude Code)
A while back, Andrej Karpathy said the app store will be replaced by generated, disposable software," and Amjad Masad predicted that the value of all application software will go to zero I think this "ephemeral software hypothesis" is wrong, though, and I want to explain why:
We launched CoverDrop 🎉 providing sources with a secure and anonymous way to talk to journalists. Having started five years ago as a PhD research project, this now ships within the Guardian app to millions of users—all of which provide cover traffic. Paper, code, and more info: www.coverdrop.org
CoverDrop: Blowing the Whistle Through A News App
coverdrop.org
This is going to be big news in my field. While we wait for the dataset, the stuff about post-processing makes interesting reading (if you're me)
can't wait til they actually upload the dataset to go with this one arxiv.org/abs/2506.08300
I'm late to review the "Illusion of Thinking" paper, so let me collect some of the best threads by and critical takes by @scaling01 in one place and sprinkle some of my own thoughts in as well. The paper is rather critical of reasoning LLMs (LRMs): x.com/MFarajtabar...
If the last time you tried to use an LLM for math was ~4 or 5 months ago it’s worth firing up Gemini 2.5 (which you can try for free) or ChatGPT o3 and getting a sense of how rapidly things have progressed.
I want to share my latest (very short) blog post: "Active Learning vs. Data Filtering: Selection vs. Rejection." What is the fundamental difference between active learning and data filtering? Well, obviously, the difference is that: 1/11
Hive (and all of its expansions) has been added to OpenSpiel! 🎉🤩🐝🐜🕷️🐞🦟🪲 From Gen42: "Hive is an award-winning board game with a difference. There is no board. The pieces are added to the playing area thus creating the board. As more and more pieces are added the game becomes a fight to ... 🧵1/5
📢📢 Junior researchers attending #ICLR2025, be sure to check out the mentoring chat sessions! More info here: blog.iclr.cc/2025/04/23/i... You can find all the sessions on the ICLR.cc schedule!
I want to share a blog post on our paper "All Models are Wrong, Some are Useful: Model Selection with Limited Labels" which we will present at AISTATS 2025 next week With @pokanovic.bsky.social, Jannes Kasper, @thoefler.bsky.social, @arkrause.bsky.social, and @nmervegurel.bsky.social
I want to reshare @brandfonbrener.bsky.social's @NeurIPSConf 2024 paper on CoLoR-Filter: A simple yet powerful method for selecting high-quality data for language model pre-training! With @hlzhang109.bsky.social @schwarzjn.bsky.social @shamkakade.bsky.social
I want to reshare @brandfonbrener.bsky.social's @NeurIPSConf 2024 paper on CoLoR-Filter: A simple yet powerful method for selecting high-quality data for language model pre-training! With @hlzhang109.bsky.social @schwarzjn.bsky.social @shamkakade.bsky.social
The Ukrainian government has a list of places where you can donate to the war effort here. I personally just donated $100: war.ukraine.ua/donate/ Slava Ukraini.
Donate to Ukraine’s defenders
The National Bank of Ukraine has decided to open a special fundraising account to support the Armed Forces of Ukraine.
war.ukraine.ua
I am quite excited that our brand-new module "P79: Cryptography and Protocol Engineering" has its first lecture today! @martin.kleppmann.com and I designed the course to bridge the gap between mathematical ideas and the challenge of implementing secure cryptography in the real world. @cst.cam.ac.uk
Check out MODEL SELECTOR, a framework for label-efficient selection of pretrained classifiers. We reduce the labeling cost by up to 94.15% to identify the best model.
Ever wondered why presenting more facts can sometimes *worsen* disagreements, even among rational people? 🤔 It turns out, Bayesian reasoning has some surprising answers - no cognitive biases needed! Let's explore this fascinating paradox quickly ☺️
TMLR is now on Bluesky: be sure to follow @tmlrorg.bsky.social!
🎉Announcing... the 2024 TMLR Outstanding Certifications! (aka, our "best paper" awards!) Are you bursting with anticipation to see what they are? Check out this blog post, and read down-thread!! 🎉🧵👇 1/n medium.com/@TmlrOrg/ann...
Ever wondered why presenting more facts can sometimes *worsen* disagreements, even among rational people? 🤔 It turns out, Bayesian reasoning has some surprising answers - no cognitive biases needed! Let's explore this fascinating paradox quickly ☺️
I didn't talk about it but I also made heavy use of Claude 3.5 and also o1 and Gemini when creating my lecture series on info theory and active learning in 3.5 weeks: bsky.app/profile/bla...
Andreas Kirsch (@blackhc.bsky.social)
The slides for my lectures on (Bayesian) Active Learning, Information Theory, and Uncertainty are online now 🥳 They cover quite a bit from basic information theory to some recent papers: blackhc.github.io/balitu/ and I'll try to add proper course notes over time 🤗
bsky.app
The slides for my lectures on (Bayesian) Active Learning, Information Theory, and Uncertainty are online now 🥳 They cover quite a bit from basic information theory to some recent papers: blackhc.github.io/balitu/ and I'll try to add proper course notes over time 🤗
Just 10 days after o1's public debut, we’re thrilled to unveil the open-source version of the technique behind its success: scaling test-time compute By giving models more "time to think," Llama 1B outperforms Llama 8B in math—beating a model 8x its size. The full recipe is open-source!
Thanks for following me here! 🫶 I went through my notifications to follow people if they are in ML research, doing PhDs, etc, to have a nice feed focused on ML. Apologies to anyone I have missed! You can unfollow and refollow me to give me a new notification (I suppose)! Plz update your profiles 🙏
Excited to be presenting my work, "Big batch Bayesian active learning by considering predictive probabilities" at the Bayesian Decision Making & Uncertainty (BDU) Workshop @neuripsconf.bsky.social, as both a lightning talk and a poster!https://openreview.net/pdf?id=VikX9euujU (1/3)
The NeurIPS Workshop on Bayesian Decision-making and Uncertainty has started - our first talk is by @mvdw.bsky.social! Join us at East Meeting Room 8, 15, or online!