{"id":2564,"date":"2025-11-19T03:32:40","date_gmt":"2025-11-18T22:32:40","guid":{"rendered":"https:\/\/paknews.centers.pk\/litigating-and-monetizing-content-licensing-to-llms\/"},"modified":"2025-11-19T03:32:40","modified_gmt":"2025-11-18T22:32:40","slug":"litigating-and-monetizing-content-licensing-to-llms","status":"publish","type":"post","link":"https:\/\/paknews.centers.pk\/ur\/litigating-and-monetizing-content-licensing-to-llms\/","title":{"rendered":"Litigating and Monetizing Content Licensing to LLMs"},"content":{"rendered":"<p><br \/>\n<\/p>\n<div>\n<div class=\"article_single_featured_img\"><img decoding=\"async\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/ArticleImages\/335908-veritone-ORG.png\" alt=\"Article Featured Image\"\/><\/div>\n<p>Licensing data to LLMs is a potential revenue stream for streamers much like advertising on CTV platforms: it is an opportunity that didn\u2019t exist until recently that has the potential to deliver dividends for years to come. But as with CTV advertising, its viability and profitability won\u2019t happen overnight.<\/p>\n<p>For now, there are numerous rights lawsuits in the works, and most companies have not come around to respecting content copyright.<\/p>\n<p><em><img decoding=\"async\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/ArticleImages\/InlineImages\/335909-krefetz-licensing-opener-ORG.png\" alt=\"gen ai llm content provenance lawsuits\" width=\"800\" height=\"420\"\/><br \/>Image source: <a href=\"https:\/\/chatgptiseatingtheworld.com\/wp-content\/uploads\/2025\/08\/Map-of-Copyright-Litigation-v.-AI-companies-in-United-States-AUG-28-2025.pdf\" target=\"_blank\" rel=\"noopener\">ChatGPT Is Eating the World<\/a><\/em><\/p>\n<p>One of the companies that media analyst <a href=\"https:\/\/www.linkedin.com\/in\/laura-martin-cfa-cmt-0bb401a\/\" target=\"_blank\" rel=\"noopener\">Laura Martin<\/a> tracks in her work as a Managing Director at Needham &amp; Company is OTT subscription streaming service <a href=\"https:\/\/curiositystream.com\/\" target=\"_blank\" rel=\"noopener\">Curiosity Stream (CURI)<\/a>. CURI recently reported Q3 2025 revenue of $18.36 million, up 46% year-over-year. \u201cWhat we like most about CURIs 3Q25 was its 425% LLM licensing revenue growth,\u201d Martin says. \u201cToday, CURI is licensing 2 million hours of premium TV, film, and video content to 19 LLMs. In 2026, CURI believes it can double its content supply and grow its LLM licensees by ~50% to ~30 LLMs.\u201d<\/p>\n<p>Of the 8 deals they have had in place for at least a year, CURI expects revenue from these licensing deals to double in 2026. The company owns ~300,000 hours of their own content and does a revenue share for 50% of any LLM fees they get on the other 1.7 million hours of content distributed through their platform.<\/p>\n<p>\u201cThis has only happened a couple of times in my career in different companies when you achieve enough scale in one area, that a derivative service line can become potentially the main service line,\u201d says <a href=\"https:\/\/www.linkedin.com\/in\/ryansteelberg\/\" target=\"_blank\" rel=\"noopener\" title=\"Veritone CEO Ryan Steelberg\">Veritone CEO Ryan Steelberg<\/a>.<\/p>\n<p>Veritone started in 2012 and is now processing 250,000 hours of unstructured data and content per day. In Q3 2024 they entered into LLM licensing, facilitating deals for backend model training. \u201cWe represent a very large number of audio and video IP owners. The majority of the deals that we\u2019ve done to date are datasets that I had no relationship with before.\u201d<\/p>\n<h2><strong>Content at Scale<\/strong><\/h2>\n<p>In terms of training mediums, text has a much clearer provenance. \u201cThe through line is very clear,\u201d says <a href=\"https:\/\/www.linkedin.com\/in\/brianmcneill\" target=\"_blank\" rel=\"noopener\">Brian McNeill, COO\/CPO and Co-Founder, Stringr<\/a>. <a href=\"https:\/\/www.streamingmedia.com\/Articles\/ReadArticle.aspx?ArticleID=165141\" target=\"_blank\" rel=\"noopener\">Stringr<\/a> has been in business for 12 years and has a videographer marketplace of 160,000 videographers. They have 2.5 million video assets where you can buy stock or have videographers go out and shoot specific content. <span style=\"text-decoration: line-through;\"\/><\/p>\n<p>In addition to finished content, there is a large ratio of outs to good shots. \u201cAll of these news agencies have huge archives of images and video that are sitting there,\u201d he says, \u201cThere is a large amount that either goes to archives or is even deleted.\u201d<\/p>\n<p>Right now, 100,000+ minutes seems to be the minimum for training. McNeill says library content \u201cis only valuable when you\u2019re talking about the 2 million video assets we have or the tens of millions of assets that sit over at Reuters or AP. That&#8217;s when the conversation starts to get interesting for the model company, because there\u2019s enough there to actually train on,&#8221; says McNeill.<span style=\"text-decoration: line-through;\"\/><\/p>\n<p>Do they use existing metadata? No. \u201cMost of the models are using their own image detection algorithms to identify what is in there, as opposed to using the metadata that was provided by the by the contributor.\u201d<\/p>\n<p>News organizations have been approached by deal aggregators with a wide variety of terms. Currently one consultant in the field is recommending:<\/p>\n<ul>\n<li>$20 to $30 per hour (net to stations) of licensed content<\/li>\n<li>prohibitions against replication or generation of facsimiles of specific human beings or their voices, or anything we would consider a \u201cderivative work\u201d<\/li>\n<li>ability to withhold from\u00a0specific buyers<\/li>\n<li>no grant of copyright or any other exhibition rights, replication rights, etc.<\/li>\n<li>one-year term with right to pull out upon notice<\/li>\n<\/ul>\n<p>\u201cWhat these firms are looking for seems to shift every three to six months based on the next model they\u2019re training. For a while it was, we need faces straight on and then it was, we need shots taken in a particular style so we can replicate that style, whether it\u2019s various camera angles, lighting techniques,\u201d says McNeill.<\/p>\n<h2><strong>Depth<\/strong><\/h2>\n<p>There are different issues with using media content. Converting audio and video into tokens\u2014the distillation of content into a linguistical unit, which is mapped to a numeric value and used to train the next generation of models both has more complicated rights clearances as well as a much deeper level of data for model training.<\/p>\n<p>\u201cIt [was] a lot easier to tokenize text, which in effect is on a per-word basis, to train an LLM than it is to convert audio and video into tokens and then use that as training data,\u201d says Veritone\u2019s Steelberg. \u201cNow we\u2019re entering into the very messy world of all these other forms of unstructured data dominated by audio and video.&#8221;<\/p>\n<p>Steelberg notes that each frame of video could contain \u201cas many as 2,000 tokens. How do you describe centricity of a face on screen, or movement as that head turns into shadows? You\u2019re looking at the capability to generate description on anything and everything inside a frame: shadowing, lighting, movement, interaction between two objects, what they\u2019re saying, etc.\u201d<\/p>\n<p>Veritone has seen the most demand for licensing reality content. Sports is another key area. Veritone is now representing the BCAA to turn their audio and video into training data. They have found that gaining clearances for sports are more straightforward than for scripted shows. \u201cScripted content may not necessarily represent a true representation of the real world,\u201d says Steelberg.<\/p>\n<h2><strong>Training Costs<\/strong><\/h2>\n<p>ChatGPT sourced some data on <a href=\"https:\/\/blog.truegeometry.com\/api\/exploreHTML\/1c70c647ee23c2924c3d0c1121083d87.exploreHTML?utm_source=chatgpt.com\" target=\"_blank\" rel=\"noopener\">True Geometry<\/a> on training costs for the hyperscalers:<\/p>\n<ul>\n<li>Minimum: $70 million (a smaller, less capable model)<\/li>\n<li>Typical: $100 million-$200 million+ (for a model with comparable capabilities)<\/li>\n<li>State-of-the-art (e.g., GPT-4, Gemini): $200 million-$1 billion+ (and potentially much higher).<\/li>\n<\/ul>\n<p>A 2025 analysis of model-training cost growth (\u201c<a href=\"https:\/\/arxiv.org\/abs\/2405.21015\" target=\"_blank\" rel=\"noopener\">The rising costs of training frontier AI models<\/a>\u201d) shows: \u201cThe amortized cost to train the most compute-intensive models has grown at ~2.4\u00d7\/year since 2016 \u2026 If the trend continues, the largest training runs will cost more than a billion dollars by 2027.\u201d<\/p>\n<p>How much will media be worth in the world of LLM training? \u201cThe hyperscalers that are competing to build LLMs are spending $400 billion in CapEx this year, and $500 billion, which is half a trillion next year,\u201d says Needham &amp; Company\u2019s Martin. \u201cThe way to get competitive advantage over your competitors is to have data other people don\u2019t have.\u201d<\/p>\n<p>This aspect of the business will give a whole new revenue stream to media and entertainment. Who knew all that longtail content or video b-roll and outs would be valuable to more than the storage companies one day?<\/p>\n<div class=\"cta_magazine_subscription_wrap\">\n<div class=\"cta_magazine_subscription\">\n<div class=\"cta_magazine_subscription_img\">\n            <img decoding=\"async\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/SiteImages\/168775-2025-Cover-Images---April-2025-ORG.png\" alt=\"Streaming Covers\"\/><br \/>\n<!--            <img decoding=\"async\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/SiteImages\/167595-2025-Cover-Images-ORG.png\" alt=\"Streaming Covers\">-->\n        <\/div>\n<\/p><\/div>\n<\/div>\n<p><noscript>Please enable JavaScript to view the <a href=\"https:\/\/disqus.com\/?ref_noscript\" target=\"_blank\" rel=\"noopener\">comments powered by Disqus.<\/a><\/noscript><\/p>\n<p>&#13;<br \/>\n            Related Articles&#13;<br \/>\n            &#13;\n        <\/p>\n<section class=\"article_grid\">&#13;<br \/>\n    &#13;<\/p>\n<div class=\"article_grid_single\">\n<div>\n                <a id=\"MainContentPlaceHolder_ctl01_ctl00_rptArticles_lnkImageLink_0\" href=\"https:\/\/www.streamingmedia.com\/Articles\/News\/Online-Video-News\/Sneak-Preview-LLMs-on-Air---Gen-AI-Use-Cases-for-News-Sports-and-Entertainment-168036.aspx?utm_source=related_articles&amp;utm_medium=gutenberg&amp;utm_campaign=editors_selection\" target=\"_blank\" rel=\"noopener\"><img decoding=\"async\" class=\"lazy\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/ArticleImages\/167881-WED3-Sneak-Preview-ORG.png\"\/><\/a>\n            <\/div>\n<h3>&#13;<br \/>\n                <a id=\"MainContentPlaceHolder_ctl01_ctl00_rptArticles_lnkArticleTitle_0\" href=\"https:\/\/www.streamingmedia.com\/Articles\/News\/Online-Video-News\/Sneak-Preview-LLMs-on-Air---Gen-AI-Use-Cases-for-News-Sports-and-Entertainment-168036.aspx?utm_source=related_articles&amp;utm_medium=gutenberg&amp;utm_campaign=editors_selection\" target=\"_blank\" rel=\"noopener\">Sneak Preview: LLMs on Air &#8211; Gen AI Use Cases for News, Sports, and Entertainment<\/a><\/h3>\n<p>&#13;<br \/>\n                On Wednesday, February 26, leading industry expert Brian Ring will moderate the Streaming Media Connect panel &#8220;LLMs on Air: Gen AI Use Cases for News, Sports, and Entertainment.&#8221; Large language models (LLMs) are making inroads everywhere in the streaming world. As a text-generating subset of Gen AI that stands apart from audio and video creation but is nonetheless having a significant impact on how broadcasters deliver content and how viewers experience it, how will these LLMs get commoditized? This expert panel addresses a host of issues around LLMs and streaming, from aggressive data scraping to metadata-driven discovery and more. Confirmed panelists include experts from Sinclair, Play Anywhere, and evision.&#13;\n            <\/p>\n<p>&#13;<br \/>\n                &#13;<br \/>\n                <span class=\"article_date\">&#13;<br \/>\n                    14 Feb 2025<\/span>&#13;\n            <\/p>\n<\/p><\/div>\n<p>&#13;<br \/>\n    &#13;<\/p>\n<div class=\"article_grid_single\">\n<div>\n                <a id=\"MainContentPlaceHolder_ctl01_ctl00_rptArticles_lnkImageLink_1\" href=\"https:\/\/www.streamingmedia.com\/Articles\/Editorial\/Featured-Articles\/AI-and-Streaming-Media-165141.aspx?utm_source=related_articles&amp;utm_medium=gutenberg&amp;utm_campaign=editors_selection\" target=\"_blank\" rel=\"noopener\"><img decoding=\"async\" class=\"lazy\" src=\"https:\/\/dzceab466r34n.cloudfront.net\/Images\/ArticleImages\/165725-ai-feature-ORG.png\"\/><\/a>\n            <\/div>\n<h3>&#13;<br \/>\n                <a id=\"MainContentPlaceHolder_ctl01_ctl00_rptArticles_lnkArticleTitle_1\" href=\"https:\/\/www.streamingmedia.com\/Articles\/Editorial\/Featured-Articles\/AI-and-Streaming-Media-165141.aspx?utm_source=related_articles&amp;utm_medium=gutenberg&amp;utm_campaign=editors_selection\" target=\"_blank\" rel=\"noopener\">AI and Streaming Media<\/a><\/h3>\n<p>&#13;<br \/>\n                This article explores the current state of AI in the streaming encoding, delivery, playback, and monetization ecosystems. By understanding the\u00a0developments and considering key questions when evaluating AI-powered solutions, streaming professionals can make informed decisions about incorporating AI into their video processing pipelines and prepare for the future of AI-driven video technologies.&#13;\n            <\/p>\n<p>&#13;<br \/>\n                &#13;<br \/>\n                <span class=\"article_date\">&#13;<br \/>\n                    29 Jul 2024<\/span>&#13;\n            <\/p>\n<\/p><\/div>\n<p>&#13;<br \/>\n    &#13;<br \/>\n        &#13;<br \/>\n        <\/section>\n<\/p><\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/www.streamingmedia.com\/Articles\/ReadArticle.aspx?ArticleID=172469\" target=\"_blank\" rel=\"noopener\">Source link <\/a><\/p>","protected":false},"excerpt":{"rendered":"<p>Licensing data to LLMs is a potential revenue stream for streamers much like advertising on CTV platforms: it is an opportunity that didn\u2019t exist until recently that has the potential to deliver dividends for years to come. But as with CTV advertising, its viability and profitability won\u2019t happen overnight. For now, there are numerous rights [&hellip;]<\/p>","protected":false},"author":1,"featured_media":2566,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[58],"tags":[],"class_list":["post-2564","post","type-post","status-publish","format-standard","has-post-thumbnail","category-live-session"],"_links":{"self":[{"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/posts\/2564","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/comments?post=2564"}],"version-history":[{"count":0,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/posts\/2564\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/media\/2566"}],"wp:attachment":[{"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/media?parent=2564"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/categories?post=2564"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/paknews.centers.pk\/ur\/wp-json\/wp\/v2\/tags?post=2564"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}