@patqdasilva.bsky.social

🚩 AI-assisted scientific writing faces increasing scrutiny Can human or LLM editing fix it? 👇 240k edits on scientific abstracts expose LLM writing flaws, and show that human editing fails to fix them cc Sanchaita Hazra, Doeun Lee, @shocheen.bsky.social, Bodhisattwa Prasad Majumder 🧵

Bild

Steering language models by directly intervening on internal activations is appealing–but does it generalize? We study 3 popular steering methods with 36 models from 14 families (1.5-70B), exposing brittle performance and fundamental flaws in underlying assumptions 🧵👇 (1/10)

Bild