Nick Vincent

@nickmvincent.bsky.social

Studying people and computers (https://www.nickmvincent.com/) Blogging about data and steering AI (https://dataleverage.substack.com/)

More thoughts on why the surge of interest in AI evaluation, auditing, and benchmarking will create a fresh window to establish healthy data flow and data markets more broadly (largely leaning on old data intermediary / data guild style ideas): dataleverage.substack.com/p/the-ai-eva...

The AI "Evaluation Crisis" Is an Opportunity to Get Data Flow Right

Why the AI evaluation crisis could force a reckoning on dataset provenance, attribution, and consent.

dataleverage.substack.com

Adding this to my schtick-toolbelt of recurring topics to post about; I really think the class of people who aim to shape AI discourse *must* ask: Is a commitment to "augment without replacing" a commitment to intentionally build less capable models or something else?

Nick Vincent@nickmvincent.bsky.social · 3mo ago

One issue I think that technical AI pundits / the publicly-engaged professoriate / related groups should really push on is trying to get consensus on a technically precise definition of what it would mean to train a model that "augments without replacing"

One issue I think that technical AI pundits / the publicly-engaged professoriate / related groups should really push on is trying to get consensus on a technically precise definition of what it would mean to train a model that "augments without replacing"

Going to #CHI2026 and going to try and engage and boost CHI-related stuff here (and linkedin too, as it seems like there's a lot of activity there)! Hoping to support a positive social media experience around conference-ing, as this was a big part of my positive experience at my first CHI!

Excited to be heading to Barcelona for #CHI2026 to host our workshop PoliSim: LLM Agent Simulation for Policy! This year, we’ve seen incredible interest from researchers across HCI, NLP, CSS, and Policy. We accepted 25 outstanding papers, with 5 selected as Best Paper nominees.

Bild

New blog post: a more explicit vision of an “attestation-forward” data policy strategy for AI. With the right approach to attestations, I think we can get a win-win-win-win (for: consumers of AI products, auditors, AI developers worried about distillation, information quality):

Thread for some misc thoughts / pins from @atproto.science talks and sessions: - big gap in demos that give potential users (eg scientists who used to post on other platforms) a “wow” moment for features that at proto enables (I had this recently with semble + margin interop)

We have an exciting panel tomorrow @11am with @nickmvincent.bsky.social Laure Haak (@verime.coop) and Ellie DeSota (@metagov.bsky.social , SciOS)! The panel will explore how governance & sustainability challenges facing the broader atproto ecosystem are mirrored in its open science applications >

Can decentralists cooperate? Rethinking commons and collective action in the age of platforms and AI

Event: Can decentralists cooperate? Rethinking commons and collective action in the age of platforms and AI

atmo.rsvp

🚨Collective action strategies in the age of AI w/ Nick Vincent I spoke to @nickmvincent, AI researcher and author of the Data Leverage substack, about how AI systems are built on the collective output of humanity's digital labor and what we can do about it. FULL EPISODE⬇️

Pretty excited about trying move more of my info consumption through semble.so, margin.at, and www.graze.social Seems like there's some built in integration between @semble.so and margin -- very interested in figuring out a flow here (and further connect to local file workflows)

Semble — A social knowledge network for research

Follow your peers' research trails. Surface and discover new connections. Built on ATProto so you own your data.

semble.so

Writing a follow up post on data aspects of coding agents. One thing that's really under-discussed, IMO -- as far as I can tell, NO coding agent allows for consumer users to trigger server-side deletion of transcripts or even metadata. Anyone seen anything to the contrary?

Seems plausible that some motivation for labs to restrict usage of subscription auth tokens is the value of structured data from using the official app, but unfortunate that the current data control for agents is super limited (30 days or 5 yrs, no indiv deletions, etc.)