Will Monroe

@futurulus.bsky.social

Research scientist at Duolingo (natural language processing and speech recognition). Mostly posts funny AI fails, sometimes also cool linguistics facts and rockets.

My computer just woke me up to tell me it's hungry. I'm not kidding 😂 During a long-running task, it noticed the battery was draining, set the volume to 100% using Computer Use, then opened Google Translate and hit "Listen" so I could hear it asking. Pardon my French, but what the fuck? 🤯

Bild

I’m increasingly of the opinion that when AI Overviews disappears we’re going to realize it was the funniest thing ever to happen on the internet Here for instance, it invents a totally new form of literary criticism, without being asked to! (rest of this is all about batteries)


Google
Sign in
Q good night, westley
All -
Images
Videos
Shopping
Short videos
• Al Overview
0-+2 :
Good work. Sleep well. I'll most likely kill you in the morning
• IMDb
Have you ever wondered what makes that line from The Princess Bride so perfect? It is a great contrast. The first part sounds like a kind boss.
The second part is a scary threat.
• Instagram • Ekazop
It is just like a battery. A battery has two ends: a positive end and a negative end. Let's look at how that works.
Parts of a Battery
• Cathode: This is the positive end. It pulls things in. Think of it like a magnet grabbing tiny metal bits.
• Anode: This is the negative end. It pushes things out. Think of it like a crowded room pushing people out the door.

i understand why everyone is doing it, but creating your own persistent agent right now feels like those maniacs who got their hands on x-rays in the 30s and started pointing that shit at everything just to see what would happen

To train better open models, we need predictable scaling. Delphi is Marin’s first step: we pretrained many small models with one recipe, then extrapolated 300× to predict a 25B-param / 600B-token run with just 0.2% error. Getting there took some work 🧵

Starting to feel like most humans aren’t equipped to deal with something that confidently lies and communicates those lies with a fluency almost no actual humans possess

Dr. Casey Fiesler@cfiesler.bsky.social · last yr.

This is fascinating: www.reddit.com/r/OpenAI/s/I... Someone “worked on a book with ChatGPT” for weeks and then sought help on Reddit when they couldn’t download the file. Redditors helped them realized ChatGPT had just been roleplaying/lying and there was no file/book…

Simple example for how errors can creep into our papers through LLM use: I had a statistic for 2020. I googled about the same statistic for 2023. AI overview tells me the statistic for 2023 and provides a link to support the claim. The link is to a 2023 article citing the 2020 statistic.

NYT: “For One Hilarious, Terrifying Day, Elon Musk’s Chatbot Lost Its Mind” That’s not what losing your mind looks like THIS is what losing your mind looks like

Bild