Maria Valentini

@mvalentini.bsky.social

computer science/cognitive science PhD student @ CU Boulder • computational psycholinguistics, NLP for education, AI ethics

Excited to be presenting our new paper at #CogSci2026 this coming week, which asks: Does Contextual Informativeness Predict Preschoolers’ Word Learning from Stories? I will be in poster session 3 on Friday evening. Full paper will be published in proceedings at the conclusion of the conference! :)

image of poster

Excited to announce that the PolyGloss paper has been accepted to @aclmeeting.bsky.social! Previously, we trained models to help in endangered language documentation workflows by automatically predicting interlinear glosses. But real-world user studies revealed crucial issues...

Bild

I don’t really have the energy for politics right now. So I will observe without comment: Executive Order 14110 was revoked (Safe, Secure, and Trustworthy Development and Use of Artificial Intelligence)

1. Can you stop companies from training generative AI using your data? No, not currently. 2. Is this dataset meant for training generative AI? 🤷‍♀️ but more likely for research and statistical analysis. 3. Is it ok to duplicate and distribute people’s data without agency to opt out? I’d argue no.

Daniel van Strien@danielvanstrien.bsky.social · 2y ago

First dataset for the new @huggingface.bsky.social @bsky.app community organisation: one-million-bluesky-posts 🦋 📊 1M public posts from Bluesky's firehose API 🔍 Includes text, metadata, and language predictions 🔬 Perfect to experiment with using ML for Bluesky 🤗 huggingface.co/datasets/blu...

So many people, CS researchers included, think that you can explore how an LLM works by simply asking it to tell you what it is doing or "thinking". Here @jennhu.bsky.social provides an excellent illustration of how that approach fails even at the most basic level.

Jennifer Hu@jennhu.bsky.social · 3y ago

To researchers doing LLM evaluation: prompting is *not a substitute* for direct probability measurements. Check out the camera-ready version of our work, to appear at EMNLP 2023! (w/ @rplevy.bsky.social) Paper: arxiv.org/abs/2305.13264 Original thread: twitter.com/_jennhu/stat...