Archive
- Data labeling work should be dignified, not dismissed
- AI Dividends Without Taxing Compute or State Ownership: A Presumptive Commons-Rent Tax Based on Capability Measurement and Data Attribution
- The AI "Evaluation Crisis" Is an Opportunity to Get Data Flow Right
- Attestation across the AI Supply Chain
- "People First" Policy Ideas that Complement Each Other (through better data flow)
- AI is driving the cost of polish down; some musings on fancy versus terse artifacts
- Two natural allies of a "Data Transparency" agenda: capabilities forecasters and social simulators
- A Short Guide to Data Strikes and Conscious Data Contribution in the Context of 2026 Frontier AI
- The Paradox of Reuse in 2026: A Case of Quasi-Enclosure, or "Subsidized Club Goods that Sort of Look Like Public Goods"
- The Coding Agent Data Deal
- Coding agents are (1) a big deal, (2) very relevant to data leverage, and (3) able to help build tools that support data leverage!
- Almost Everybody -- Including Both Data Creators and AI Companies -- Stands to Benefit from Clearer "Data Rules".
- How collective bargaining for information, public AI, and HCI research all fit together
- Which datasets should we assume are "in all the AI models"?
- Algorithmic Collective Action With Two Collectives [crosspost]
- On AI-driven Job Apocalypses and Collective Bargaining for Information
- How do we know our AI output is good? Double checks, bar charts, vibes, and training data.
- Each Instance of "AI Utility" Stems from Some Human Act(s) of Information Recording and Ranking
- Google and TikTok rank bundles of information; ChatGPT ranks grains.
- [microblog] One book is worth "0.06%" benchmark points to AI; is "no different from noise". What gives?
- Public AI, Data Appraisal, and Data Debates
- Evaluation Data Leverage: Advances like "Deep Research" Highlight a Looming Opportunity for Bargaining Power
- Tipping Points for Content Ecosystems
- AI Labs Should Open Source Data Protection Technologies
- Live by the free-content-for-training sword, die by the free-content-for-training sword
- Selling AGI like AG1: Will Consumers Push Back Against Proprietary Blends of Herbs and of Data?
- Perplexity CEO's Interaction with Striking New York Times Workers Does Not Reflect Well on the AI Industry
- Is Zuckerberg right to say that your specific creative work has no value to AI?
- "Many Models" and "Track Changes" for AI: Some Thoughts on LLM Interfaces
- Building a Data Pipeworks for Democratic AI: From Human Knowledge to Records to AI Systems
- Will the New York Times Data Strike Have a Large Impact on ChatGPT?
- A Harbinger of the Future of Content? The New York Times Starts a Data Strike
- The WGA Strike is a Canary in the Coal Mine for AI Labor Concerns
- Reddit, StackOverflow, and Europe: All Trending Towards Data Dignity
- Data Leverage Recap: December 2022 - April 2023
- Bing Rewards for the AI Age
- Plural AI Data Alignment
- AI Technologies are System Maps, and You are a Cartographer
- AI Artist or AI Art Thief? Innovation, Public Mandates, and the Case for Talking in Terms of Leverage
- ChatGPT is Awesome and Scary: You Deserve Credit for the Good Parts (and Might Help Fix the Bad Parts)
- The Paradox of Reuse, Language Models Edition
- Don’t give OpenAI all the credit for GPT-3: You might have helped create the latest “astonishing” advance in AI too