Wirestock, the platform that has quietly become one of the largest marketplaces for AI training data, announced a $23 million Series A round led by Nava Ventures with participation from SBVP, the venture fund co-founded by Sheryl Sandberg. The funding will accelerate the company's transformation from a creator distribution tool into what it calls the world's leading multimodal data platform.
Founded in 2021, Wirestock originally helped photographers and illustrators distribute their work across multiple stock photography platforms simultaneously. But the explosive growth of generative AI in 2023 and 2024 created an unexpected opportunity: the company's network of creators represented exactly the kind of high-quality, rights-cleared visual data that AI companies desperately needed.
"We realized that our creator network was sitting on the most valuable asset in the AI economy -- legally sourced, diverse, high-quality multimodal data," said Wirestock CEO Marat Mukhamedyarov. "The pivot was obvious."
The Data Supply Chain Problem
The AI industry's appetite for training data has grown exponentially, but the supply of legally usable data has not kept pace. Major lawsuits against AI companies over unauthorized use of copyrighted material -- including ongoing cases involving OpenAI, Stability AI, and Midjourney -- have made rights-cleared data increasingly valuable.
Wirestock addresses this problem by working directly with creators who consent to having their work used for AI training, with compensation flowing back to them based on usage. The platform currently hosts more than 80 million assets from over 400,000 creators across photography, illustration, video, and audio.
"The era of scraping the internet for training data is ending," said Sarah Chen, partner at Nava Ventures. "Companies that can provide legally clean, high-quality data at scale are going to be essential infrastructure for the next generation of AI models."
Business Model and Growth
Wirestock operates on a dual revenue model. It continues to earn commissions from stock photography distribution for creators, while its newer AI data licensing business sells curated datasets to model developers. The company declined to disclose specific revenue figures but said AI data licensing now accounts for the majority of its revenue.
The $23 million round brings Wirestock's total funding to approximately $30 million. The company plans to use the capital to expand its data curation capabilities, invest in metadata annotation tools, and grow its enterprise sales team to serve the increasing number of AI labs seeking compliant training data.
The timing aligns with a broader shift in the AI data market. As regulatory frameworks tighten around training data usage -- the EU AI Act now requires transparency about training data sources -- demand for platforms that can guarantee provenance and consent is surging.
The Competitive Landscape
Wirestock competes with other AI data platforms including Scale AI, Defined.ai, and Appen, though its creator-first approach differentiates it from companies that rely primarily on crowdsourced annotation. Getty Images has also entered the space with its own AI training data licensing program, leveraging its massive archive of professional photography.
The key advantage Wirestock claims is diversity and freshness. Unlike static archives, its creator network continuously produces new content, ensuring that datasets remain current and representative of evolving visual styles and subjects.
What to Watch
As AI training enters a phase where data quality matters as much as quantity, companies like Wirestock are positioned at a critical juncture. The question is whether creator-sourced data platforms can scale fast enough to meet the growing needs of frontier model developers while maintaining the rights-cleared provenance that makes them valuable in the first place.
"The era of scraping the internet for training data is ending."— Sarah Chen, Partner, Nava Ventures