Tech Startup Aims to Help Media License Content for AI Training

Avail is launching Corpus to bridge the gap between smaller media rights holders and AI model developers. While major tech giants secure licensing deals with large studios, independent creators and mid-sized companies are often excluded. This new platform empowers these smaller entities to license their valuable content libraries directly to AI firms, ensuring they participate in the growing data economy rather than being left behind by industry consolidation. The product addresses specific, varied data needs across the AI sector. Different model developers require distinct types of content, such as STEM-focused material, rather than a one-size-fits-all solution. By curating libraries to match these specific priorities, Avail helps AI companies train more effective models while providing creators with a clear pathway to monetize their work. This targeted approach ensures that unique data assets find the appropriate technical applications, fostering a more nuanced and efficient data marketplace. This initiative is highly relevant to open data because it establishes a transparent, commercial framework for data licensing. It challenges the current trend where data extraction often bypasses creator compensation, advocating instead for a model where rights holders are cited and paid. By formalizing how creative data is sourced for AI training, Corpus promotes ethical data practices and economic fairness, demonstrating how structured data marketplaces can support equitable access and remuneration in the digital ecosystem.

Source: hollywoodreporter.com
Published on 2024-07-12