Sign in

Nick Vincent

@nickmvincent.bsky.social
408 followers 464 following 118 posts

Studying people and computers (www.nickmvincent.com) Blogging about data and steering AI (dataleverage.substack.com)

PostsRepliesMedia
Nick Vincent @nickmvincent.bsky.social · 24/06/2026
Sadly will not be at FAccT this year and will be missing everyone there, but two awesome students from the lab will be there presenting on Social Simulation (dl.acm.org/doi/10.1145/...) and Overreliance dl.acm.org/doi/10.1145/...
dl.acm.org
Mechanism Plausibility in Generative Agent-Based Modeling | Proceedings of the 2026 ACM Conference on Fairness, Accountability, and Transparency
050
Nick Vincent @nickmvincent.bsky.social · 07/05/2026
More thoughts on why the surge of interest in AI evaluation, auditing, and benchmarking will create a fresh window to establish healthy data flow and data markets more broadly (largely leaning on old data intermediary / data guild style ideas): dataleverage.substack.com/p/the-ai-eva...
dataleverage.substack.com
The AI "Evaluation Crisis" Is an Opportunity to Get Data Flow Right
Why the AI evaluation crisis could force a reckoning on dataset provenance, attribution, and consent.
110
Nick Vincent @nickmvincent.bsky.social · 02/05/2026
Adding this to my schtick-toolbelt of recurring topics to post about; I really think the class of people who aim to shape AI discourse *must* ask: Is a commitment to "augment without replacing" a commitment to intentionally build less capable models or something else?
220
Nick Vincent @nickmvincent.bsky.social · 01/05/2026
One issue I think that technical AI pundits / the publicly-engaged professoriate / related groups should really push on is trying to get consensus on a technically precise definition of what it would mean to train a model that "augments without replacing"
130
Nick Vincent @nickmvincent.bsky.social · 28/04/2026
Made some updates to this post, and added a more concise bullet summary! Also on leaflet: dataleverage.leaflet.pub/3mizn5hsjg5vo
dataleverage.leaflet.pub
Attestation across the AI Supply Chain - Data Leverage
A proposal for interoperable attestation objects that connect training data, evaluation labor, and AI-generated outputs across the AI supply chain.
241
Nick Vincent @nickmvincent.bsky.social · 15/04/2026
Enjoying the “Discussions” session (about online discussions — we’re not just having a “discussion”) in #CHI2026 So far — - conversational voice agents in video calls - visualizing comment threads programs.sigchi.org/chi/2026/pro...
programs.sigchi.org
Conference Programs
000
Reposted by Nick Vincent
Jenn Wortman Vaughan @jennwv.bsky.social · 13/04/2026
At #CHI2026? Check out @sachnishal.bsky.social's talk this Tuesday on how LLM-infused writing tools reshape journalists’ agency — their ability to exercise independent judgment in alignment with their values — in editorial decision making. programs.sigchi.org/chi/2026/pro...
0102
Reposted by Nick Vincent
Asia Biega @asiabiega.bsky.social · 13/04/2026
Could we govern data consent through citizen assemblies? What would that look like? We design the concept of a "consent assembly" in our new #CHI2026 paper with @linkyi.bsky.social, Paul Goelz, @robin.berjon.com 🧵 #AIGovernance #DataGovernance #AIPolicy #GDPR #AIAct #FAccT2026 #DigitalOmnibus
dl.acm.org
From Clicks to Consensus: Collective Consent Assemblies for Data Governance | Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems
192
Nick Vincent @nickmvincent.bsky.social · 13/04/2026
Going to #CHI2026 and going to try and engage and boost CHI-related stuff here (and linkedin too, as it seems like there's a lot of activity there)! Hoping to support a positive social media experience around conference-ing, as this was a big part of my positive experience at my first CHI!
0110
Reposted by Nick Vincent
Yuxuan Li @yuxuanli1225.bsky.social · 11/04/2026
Excited to be heading to Barcelona for #CHI2026 to host our workshop PoliSim: LLM Agent Simulation for Policy! This year, we’ve seen incredible interest from researchers across HCI, NLP, CSS, and Policy. We accepted 25 outstanding papers, with 5 selected as Best Paper nominees.
282
Nick Vincent @nickmvincent.bsky.social · 11/04/2026
New blog post: a more explicit vision of an “attestation-forward” data policy strategy for AI. With the right approach to attestations, I think we can get a win-win-win-win (for: consumers of AI products, auditors, AI developers worried about distillation, information quality):
210
Nick Vincent @nickmvincent.bsky.social · 06/04/2026
Some thoughts on new policy paper from OpenAI: dataleverage.substack.com/p/people-fir... Lots of good stuff, it's great to foster more discussion of these kinds of ideas, and I think this particular set of ideas has the potential to really enable data flow as a powerful accountability lever!
030
Nick Vincent @nickmvincent.bsky.social · 27/03/2026
Thread for some misc thoughts / pins from @atproto.science talks and sessions: - big gap in demos that give potential users (eg scientists who used to post on other platforms) a “wow” moment for features that at proto enables (I had this recently with semble + margin interop)
120
Nick Vincent @nickmvincent.bsky.social · 27/03/2026
At UBC today for @atproto.science workshop! Much to discuss.
150
Reposted by Nick Vincent
ATProto Science @atproto.science · 26/03/2026
We have an exciting panel tomorrow @11am with @nickmvincent.bsky.social Laure Haak (@verime.coop) and Ellie DeSota (@metagov.bsky.social , SciOS)! The panel will explore how governance & sustainability challenges facing the broader atproto ecosystem are mirrored in its open science applications >
atmo.rsvp
Can decentralists cooperate? Rethinking commons and collective action in the age of platforms and AI
Event: Can decentralists cooperate? Rethinking commons and collective action in the age of platforms and AI
2104
Nick Vincent @nickmvincent.bsky.social · 23/03/2026
Longer post(s) coming on these topics, but for those interested in "data counterfactuals" and various "data licenses" proposals I have made some big updates to two static site resources: datacounterfactuals.org and datalicenses.org
datacounterfactuals.org
Data Counterfactuals
An interactive explainer for a unifying frame across valuation, scaling, selection, poisoning, privacy, and collective action.
112
Nick Vincent @nickmvincent.bsky.social · 11/03/2026
I greatly enjoyed this conversation, and I think it might be interesting to a wide range of folks interested in AI, data, collective action, etc! Thanks for having me @tbsocialist.bsky.social (and structuring a really exciting conversation, and putting together such a nice final product!)
042
Reposted by Nick Vincent
Joshua Dávila @allmodelsarewrong.online · 11/03/2026
🚨Collective action strategies in the age of AI w/ Nick Vincent I spoke to @nickmvincent, AI researcher and author of the Data Leverage substack, about how AI systems are built on the collective output of humanity's digital labor and what we can do about it. FULL EPISODE⬇️
152
Reposted by Nick Vincent
Jason W. Burton @jasonburton.bsky.social · 11/03/2026
🧑‍💻 New paper at #chi2026 w @lorenzspreen.eurosky.social and @stefanherzog.bsky.social Are you worried about how social media algorithms affect people’s beliefs? We are, so we tested engagement-based ranking algorithms against alternatives in a pre-reg’d collaborative filtering experiment... 🧵
2199
Nick Vincent @nickmvincent.bsky.social · 09/03/2026
Two natural allies of a "Data Transparency" agenda: capabilities forecasters and social simulators --- dataleverage.substack.com/p/two-natura... (Short post reacting to recent discourse around AI forecasting + social simulation; both stand to independently benefit from "Clear Data Rules")
dataleverage.substack.com
010
Nick Vincent @nickmvincent.bsky.social · 05/03/2026
Pretty excited about trying move more of my info consumption through semble.so, margin.at, and www.graze.social Seems like there's some built in integration between @semble.so and margin -- very interested in figuring out a flow here (and further connect to local file workflows)
semble.so
Semble — A social knowledge network for research
Follow your peers' research trails. Surface and discover new connections. Built on ATProto so you own your data.
1123
Nick Vincent @nickmvincent.bsky.social · 03/03/2026
New post: dataleverage.substack.com/p/a-short-gu...
dataleverage.substack.com
A Short Guide to Data Strikes and Conscious Data Contribution in the Context of 2026 Frontier AI
Back to the basics of data leverage.
020
Nick Vincent @nickmvincent.bsky.social · 17/02/2026
New post: thoughts on the paradox of reuse in 2026 after a few months of coding agent discourse. 'The Paradox of Reuse in 2026: A Case of Quasi-Enclosure, or "Subsidized Club Goods that Sort of Look Like Public Goods"' dataleverage.substack.com/p/the-parado...
dataleverage.substack.com
The Paradox of Reuse in 2026: A Case of Quasi-Enclosure, or "Subsidized Club Goods that Sort of Look Like Public Goods"
How we can understand, and react to, the complicated impacts of AI systems on online communities and knowledge commons
040
Reposted by Nick Vincent
Ronen Tamari @ronentk.me · 13/02/2026
1/ Finally wrote up “The Story of Mendeley”! Most people know the tool, few know about its rise and fall. The Mendeley story provides important clues for how to build self-sustaining AND non-extractive knowledge commons, which is why I think it deserves more attention 🧵
cairos.leaflet.pub
What happened to Science Goodreads and how do we rebuild it? A 65 million dollar question (at least) - CAIROS Blog
The story of the rise and fall of Mendeley
67437
Nick Vincent @nickmvincent.bsky.social · 27/01/2026
OpenAI launching an overleaf competitor seems like it could be a big deal, and particularly interesting in wake of the NeurIPS hallucinations discourse (an important issue, but a lot of the back-and-forth I saw seemed to be missing a lot of important factors): openai.com/index/introd...
openai.com
Introducing Prism
Accelerating science writing and collaboration with AI.
071
Nick Vincent @nickmvincent.bsky.social · 12/01/2026
Here's the post: dataleverage.substack.com/p/the-coding...
dataleverage.substack.com
The Coding Agent Data Deal
On user data control, coding agents as retrievers, and the value of your coding transcripts
060
Nick Vincent @nickmvincent.bsky.social · 11/01/2026
Writing a follow up post on data aspects of coding agents. One thing that's really under-discussed, IMO -- as far as I can tell, NO coding agent allows for consumer users to trigger server-side deletion of transcripts or even metadata. Anyone seen anything to the contrary?
361
Nick Vincent @nickmvincent.bsky.social · 09/01/2026
Seems plausible that some motivation for labs to restrict usage of subscription auth tokens is the value of structured data from using the official app, but unfortunate that the current data control for agents is super limited (30 days or 5 yrs, no indiv deletions, etc.)
010
Nick Vincent @nickmvincent.bsky.social · 05/01/2026
Coding agents are (1) a big deal, (2) very relevant to data leverage, and (3) able to help build tools that support data leverage! dataleverage.substack.com/p/coding-age...
dataleverage.substack.com
Coding agents are (1) a big deal, (2) very relevant to data leverage, and (3) able to help build tools that support data leverage!
Sharing an early reaction to recent coding agent discourse and two relevant projects
030
Reposted by Nick Vincent
B! 🐝 Cavello (they/them) @b-cavello.bsky.social · 19/12/2025
A bunch of us are working to advance #PublicAI: AI that is publicly accountable, accessible, and sustainable. A lot of us are interested in local-first, community-governed, and more open models of what this technology could be. We welcome allies in the @publicai.network! publicai.network/whitepaper/
publicai.network
Public AI Network
A coalition to build public AI
1189
Reposted by Nick Vincent
Hanlin Li @hanlinliii.bsky.social · 06/12/2025
Happening now! Join us in Upper Level Room 4 for our workshop on Algorithmic Collective Action #NeurIPS2025 We will have stellar talks to kick off the day, followed by contributed talks and posters by authors before lunch break.
193
Reposted by Nick Vincent
B! 🐝 Cavello (they/them) @b-cavello.bsky.social · 04/12/2025
TODAY is the first-ever #NeurIPS position paper track! Come hear thoughtful arguments about “digital heroin,” the nature of innovation, protecting privacy, machine unlearning, & how we can do ML research better as a community. See you: ballroom 20AB from 10-11a & 3:30-4:30p! #NeurIPS2025 #NeurIPSSD
021
Nick Vincent @nickmvincent.bsky.social · 26/11/2025
Longer blog post: AI companies and data creators actually have aligned incentives re: establishing clearer "Data Rules" (norms, rules, contracts that control use of both "fresh" data and of model outputs). Good Data Rules can also support commons! dataleverage.substack.com/p/almost-eve...
dataleverage.substack.com
Almost Everybody -- Including Both Data Creators and AI Companies -- Stands to Benefit from Clearer "Data Rules".
In fact, anyone who doesn't think they will be a "big winner" long term benefits from clear rules, even if it means training data costs more in the short term.
020
Reposted by Nick Vincent
Xnet - Instituto para Digitalización Democrática @xnet-x.net · 28/10/2025
"There are many challenges to transforming the AI ecosystem and strong interests resisting change. But we know change is possible, and we believe we have more allies in this effort than it may seem. There is a rebel in every Death Star." 🗣️ @b-cavello.bsky.social in our #4DCongress
044
Nick Vincent @nickmvincent.bsky.social · 19/10/2025
Heading to AIES, excited to catch up with folks there!
021
Nick Vincent @nickmvincent.bsky.social · 14/10/2025
New blog (a recap post): "How collective bargaining for information, public AI, and HCI research all fit together." Connecting these ideas + a short summary of various recent posts (of which there are many, perhaps too many!). On Substack, but also posted to leaflet
130
Reposted by Nick Vincent
Ronen Tamari @ronentk.me · 10/10/2025
V interesting twist on MCP! “user data is often fragmented across services and locked into specific providers, reinforcing user lock-in” - enter Human Context Protocol (HCP): “user-owned repositories of preferences designed for active, reflective control and consent-based sharing.” 1/
1184
Nick Vincent @nickmvincent.bsky.social · 18/09/2025
Anyone compiling discussions/thoughts on emerging licensing schemes and preference signals? eg rslstandard.org and github.com/creativecomm... ? externalizing some notes here datalicenses.org, but want to find where these discussions are happening!
rslstandard.org
RSL: Really Simple Licensing
The open content licensing standard for the AI-first Internet
021
Nick Vincent @nickmvincent.bsky.social · 14/08/2025
Excited to be giving a talk on data leverage to the Singapore AI Safety Hub. Trying to capture updated thoughts from recent years, and have long wanted to better connect leverage/collective bargaining to the safety context.
010
Nick Vincent @nickmvincent.bsky.social · 14/08/2025
About a week away from the deadline to submit to the ✨ Workshop on Algorithmic Collective Action (ACA) ✨ acaworkshop.github.io at NeurIPS 2025!
acaworkshop.github.io
About the workshop – ACA@NeurIPS
020
Nick Vincent @nickmvincent.bsky.social · 08/08/2025
🧵In several recent posts, I speculated that eventually, dataset details may become an important quality signal for consumers choosing AI products. "This model is good for asking health questions, because 10,000 doctors attested to supporting training and/or eval". Etc.
131
Nick Vincent @nickmvincent.bsky.social · 17/07/2025
Around ICML with loose evening plans and an interest in "public AI", Canadian sovereign AI, or anything related? Swing by the Internet Archive Canada between 5p and 7p lu.ma/7rjoaxts
lu.ma
Oh Canada! An AI Happy Hour @ ICML 2025 · Luma
Whether you're Canadian or one of our friends from around the world, please join us for some drinks and conversation to chat about life, papers, AI, and...…
042
Nick Vincent @nickmvincent.bsky.social · 24/06/2025
[FAccT-related link round-up]: It was great to present on measuring Attentional Agency with Zachary Wojtowicz at FAccT. Here's our paper on ACM DL: dl.acm.org/doi/10.1145/... On Thurs Aditya Karan will present on collective action dl.acm.org/doi/10.1145/... at 10:57 (New Stage A)
dl.acm.org
Algorithmic Collective Action with Two Collectives | Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency
You will be notified whenever a record that you have chosen has been cited.
152
Nick Vincent @nickmvincent.bsky.social · 24/06/2025
“Attentional agency” — talk in new stage b at facct in the session right now!
010
Nick Vincent @nickmvincent.bsky.social · 21/06/2025
Off to FAccT; Excited to see faces old and new!
060
Nick Vincent @nickmvincent.bsky.social · 05/06/2025
Another blog post: a link roundup on AI's impact on jobs and power concentration, another proposal for Collective Bargaining for Information, and some additional thoughts on the topic: dataleverage.substack.com/p/on-ai-driv...
dataleverage.substack.com
On AI-driven Job Apocalypses and Collective Bargaining for Information
Reacting to a fresh wave of discussion about AI's impact on the economy and power concentration, and reiterating the potential role of collective bargaining.
020
Nick Vincent @nickmvincent.bsky.social · 27/05/2025
New data leverage post: "Google and TikTok rank bundles of information; ChatGPT ranks grains." dataleverage.substack.com/p/google-and... This will be post 1/3 in a series about viewing many AI products as all competing around the same task: ranking bundles or grains of records made by people.
dataleverage.substack.com
Google and TikTok rank bundles of information; ChatGPT ranks grains.
Google and others solve our attentional problem by ranking discrete bundles of information, whereas ChatGPT ranks more granular chunks. This lens can help us reason about AI policy.
131
Nick Vincent @nickmvincent.bsky.social · 02/05/2025
Sharing a new paper (led by Aditya Karan): there's growing interest in algorithmic collective action, when a "collective" acts through data to impact a recommender system, classifier, or other model. But... what happens if two collectives act at the same time?
121
Nick Vincent @nickmvincent.bsky.social · 03/04/2025
New early draft post: "Public AI, Data Appraisal, and Data Debates" "A consortium of Public AI labs can substantially improve data pricing, which may also help to concretize debates about the ethics and legality of training practices." dataleverage.substack.com/p/public-ai-...
dataleverage.substack.com
Public AI, Data Appraisal, and Data Debates
A consortium of Public AI labs can substantially improve data pricing, which may also help to concretize debates about the ethics and legality of training practices.
010
Reposted by Nick Vincent
Alek Tarkowski @alek.eurosky.social · 02/04/2025
“Algo decision making systems are “leviathans”, harmful not for their arbitrariness or opacity, but systemacity of decisions" - @christinalu.bsky.social on need for plural #AI model ontologies (sounds technical, but has big consequences for human #commons) www.combinationsmag.com/model-plural...
combinationsmag.com
Model Plurality
Current research in “plural alignment” concentrates on making AI models amenable to diverse human values. But plurality is not simply a safeguard against bias or an engine of efficiency: it’s a key in...
262