// model development

All signals tagged with this topic

Mozilla's Data Collective aims to remake AI training through privacy-first sourcing

Mozilla is positioning itself as a counterweight to big tech incumbents' indiscriminate data scraping by creating a marketplace where creators and publishers can directly license content for AI training at fair rates. This challenges the current model where OpenAI, Meta, and others train on internet-scraped content first and negotiate licenses later—or not at all. The test is whether Mozilla can make this economically viable when the status quo lets trillion-dollar companies train for free.