The knowledge that makes AI possible came from all of us.Every person has the right to benefit.
AI Commons began in 2016 out of discussions among people and organizations working on AI as a common good. For most of that time we shaped initiatives, advocated, and convened towards democratization of access to intelligence. We are now building the neutral layer to make trustworthy AI possible at scale.
A single place to look
Every AI system that can write, reason, translate, or create learned to do so from the recorded work of human beings. There is no single place to see what that record contains, where it came from, or under what terms it was used. We're building one.
AI compressed humanity's memory and threw away the index.
Intelligence requires memory, and memory requires a record of where it came from. Civilization solved this a long time ago and solved it institutionally, with libraries, archives, citation, the patent record, version control. Each one governs what enters, what is kept, and how errors get corrected.
AI systems were trained on everything those institutions held and connected to none of them. The model keeps the content and discards the provenance. What it knows, it cannot trace. What goes wrong, no one can follow back to the source.
A common language for where it came from.
A shared record of the books, archives, datasets and models the field runs on, describing where each came from, what is known about its origins, how it can be used, and what derives from it. Built with the institutions that have kept this material available, and published as we go.
Held in the open rather than owned. Mirrored rather than centralized, because a record that lives in one place usually only answers to whoever owns that place.
A record that outlasts the moment.
Ten years from now, nobody will be asking whether AI was built from the world's knowledge.
The question will be whether anyone can still see how.
A common language makes that visible. A record held in common keeps it visible after the funding cycles, the acquisitions and the policy moments have all moved on. That's the whole of it: knowledge that stays legible to the people it came from, and a coordinating layer that treats everyone running on it the same.
The Global Data Pledge is where that commitment gets made, by governments, memory institutions, publishers and rightsholder collectives, each at the level their holdings allow.