AI

OpenAI says it’s building a tool to let content creators ‘opt out’ of AI training

Kommentar

OpenAI CEO Sam Altman speaks during the OpenAI DevDay event on November 06, 2023 in San Francisco, California.
Image Credits: Justin Sullivan / Getty Images

OpenAI says that it’s developing a tool to let creators better control how their content’s used in training generative AI.

The tool, called Media Manager, will allow creators and content owners to identify their works to OpenAI and specify how they want those works to be included or excluded from AI research and training.

The goal is to have the tool in place by 2025, OpenAI says, as the company works with “creators, content owners and regulators” toward a standard — perhaps through the industry steering committee it recently joined.

“This will require cutting-edge machine learning research to build a first-ever tool of its kind to help us identify copyrighted text, images, audio and video across multiple sources and reflect creator preferences,” OpenAI wrote in a blog post. “Over time, we plan to introduce additional choices and features.”

It’d seem Media Manager, whatever form it ultimately takes, is OpenAI’s response to growing criticism of its approach to developing AI, which relies heavily on scraping publicly available data from the web. Most recently, eight prominent U.S. newspapers including the Chicago Tribune sued OpenAI for IP infringement relating to the company’s use of generative AI, accusing OpenAI of pilfering articles for training generative AI models that it then commercialized without compensating — or crediting — the source publications.

Generative AI models including OpenAI’s — the sorts of models that can analyze and generate text, images, videos and more — are trained on an enormous number of examples usually sourced from public sites and data sets. OpenAI and other generative AI vendors argue that fair use, the legal doctrine that allows for the use of copyrighted works to make a secondary creation as long as it’s transformative, shields their practice of scraping public data and using it for model training. But not everyone agrees.

OpenAI, in fact, recently argued that it would be impossible to create useful AI models absent copyrighted material.

But in an effort to placate critics and defend itself against future lawsuits, OpenAI has taken steps to meet content creators in the middle.

OpenAI last year allowed artists to “opt out” of and remove their work from the data sets that the company uses to train its image-generating models. The company also lets website owners indicate via the robots.txt standard, which gives instructions about websites to web-crawling bots, whether content on their site can be scraped to train AI models. And OpenAI continues to ink licensing deals with large content owners, including news organizations, stock media libraries and Q&A sites like Stack Overflow.

Some content creators say OpenAI hasn’t gone far enough, however.

Artists have described OpenAI’s opt-out workflow for images, which requires submitting an individual copy of each image to be removed along with a description, as onerous. OpenAI reportedly pays relatively little to license content. And, as OpenAI itself acknowledges in the blog post Tuesday, the company’s current solutions don’t address scenarios in which creators’ works are quoted, remixed or reposted on platforms they don’t control.

Beyond OpenAI, a number of third parties are attempting to build universal provenance and opt-out tools for generative AI.

Startup Spawning AI, whose partners include Stability AI and Hugging Face, offers an app that identifies and tracks bots’ IP addresses to block scraping attempts, as well as a database where artists can register their works to disallow training by vendors who choose to respect the requests. Steg.AI and Imatag help creators establish ownership of their images by applying watermarks imperceptible to the human eye. And Nightshade, a project from the University of Chicago, “poisons” image data to render it useless or disruptive to AI model training.

More TechCrunch

In early 2018, VC Mike Moritz wrote in the FT that “Silicon Valley would be wise to follow China’s lead,” noting the pace of work at tech companies was “furious”…

This is how bad China’s startup scene looks now

Fei-Fei Li, the Stanford professor many deem the “Godmother of AI,” has raised $230 million for her new startup, World Labs, from backers including Andreessen Horowitz, NEA, and Radical Ventures.…

Fei-Fei Li’s World Labs comes out of stealth with $230M in funding

Bolt says it has settled its long-standing lawsuit with its investor Activant Capital. One-click payments startup Bolt is settling the suit by buying out the investor’s stake “after which Activant…

Fintech Bolt is buying out the investor suing over Ryan Breslow’s $30M loan

The rise of neobanks has been fascinating to witness, as a number of companies in recent years have grown from merely challenging traditional banks to being massive players in and…

Dave and Varo Bank execs are coming to TechCrunch Disrupt 2024

OpenAI released its new o1 models on Thursday, giving ChatGPT users their first chance to try AI models that pause to “think” before they answer. There’s been a lot of…

First impressions of OpenAI o1: An AI designed to overthink it

Featured Article

Investors rebel as TuSimple pivots from self-driving trucks to AI gaming

TuSimple, once a buzzy startup considered a leader in self-driving trucks, is trying to move its assets to China to fund a new AI-generated animation and video game business. The pivot has not only puzzled and enraged several shareholders, but also threatens to pull the company back into a legal…

Investors rebel as TuSimple pivots from self-driving trucks to AI gaming

Welcome to Startups Weekly — your weekly recap of everything you can’t miss from the world of startups. Want it in your inbox every Friday? Sign up here. This week…

Shrinking teams, warped views, and risk aversion in this week’s startup news

Silicon Valley startup accelerator Y Combinator will expand the number of cohorts it runs each year from two to four starting in 2025, Bloomberg reported Thursday, and TechCrunch confirmed today.…

Y Combinator expanding to four cohorts a year in 2025

Telegram has had a tough few weeks. The messaging app’s founder, Pavel Durov, was arrested in late August and later released on a €5 million bail in France, charged with…

Telegram CEO Durov’s arrest hasn’t dampened enthusiasm for its TON blockchain

Martin Casado, a general partner at Andreessen Horowitz, will tackle one of the most pressing issues facing today’s tech world — AI regulation — only at TechCrunch Disrupt 2024, taking…

A fireside chat with Andreessen Horowitz partner Martin Casado at TechCrunch Disrupt 2024

Christina Cacioppo, CEO and co-founder of Vanta, will be on the SaaS Stage at TechCrunch Disrupt 2024 to reveal how Vanta is redefining security and compliance automation and driving innovation…

Vanta’s Christina Cacioppo takes the stage at TechCrunch Disrupt 2024

On Thursday, cybersecurity giant Fortinet disclosed a breach involving customer data.  In a statement posted online, Fortinet said an individual intruder accessed “a limited number of files” stored on a…

Fortinet confirms customer data breach

Meta has confirmed that it’s restarting efforts to train its AI systems using public Facebook and Instagram posts from its U.K. userbase. The company claims it has “incorporated regulatory feedback” into a…

Meta reignites plans to train AI using UK users’ public Facebook and Instagram posts

Following the moves of other tech giants, Spotify announced on Friday it’s introducing in-app parental controls in the form of “managed accounts” for listeners under the age of 13. The…

Spotify begins piloting parent-managed accounts for kids on family plans

Uber users in Austin and Atlanta will be able to hail Waymo robotaxis through the app in early 2025 as part of a partnership between the two companies. 

Waymo robotaxis to become available on Uber in Austin, Atlanta in early 2025

There are plenty of calendar and scheduling apps that take care of your professional life and help you slot in meetings with your teammates and work collaborators. Howbout is all…

Howbout raises $8M from Goodwater to build a calendar that you can share with your friends

Delhivery claims Ecom Express has inaccurately represented Delhivery’s business metrics when drawing comparisons in its IPO filing. 

SoftBank-backed Delhivery contests metrics in rival Ecom Express’ IPO filing

It was a matter of time, but Apple is going to allow third-party app stores on the iPad starting next week, on September 16. This change will occur with the…

Alternative app stores will be allowed on Apple iPad in the EU from September 16

The U.K.’s antitrust regulator has delivered its provisional ruling in a longstanding battle to combine two of the country’s major telecommunication operators. The Competition and Markets Authority (CMA) says that…

Three and Vodafone’s $19B merger hits the skids as UK rules the deal would adversely impact customers and MVNOs

Late Thursday evening, Oprah Winfrey aired a special on AI, appropriately titled “AI and the Future of Us.” Guests included OpenAI CEO Sam Altman, tech influencer Marques Brownlee, and current…

Oprah just had an AI special with Sam Altman and Bill Gates — here are the highlights

Antonio Moraes, the grandson of a late prominent Brazilian billionaire, was never interested in joining the family-owned conglomerate of construction companies and a bank. Shortly after graduating from college, he…

XP Health grabs $33M to bring employees more affordable vision care

A crew of four private astronauts made history in the early hours of Thursday when they opened the hatch of their SpaceX Dragon capsule and conducted the first commercial spacewalk. …

Polaris Dawn astronauts perform historic private spacewalk while wearing SpaceX-made suits

Keith Rabois, managing director of Khosla Ventures, was having dinner with a “very successful CEO” in October 2018 when the CEO asked him a question: How many people does it…

Keith Rabois says Miami is still a great place for startups, even as a16z leaves

By making the AI info label harder to find, it might be easier for users to be deceived by content that was edited with AI, especially as editing tools become…

Meta is making its AI info label less visible on content edited or modified by AI tools

Cohost, a would-be X rival launched to the public in June 2022, is shutting down, the company announced via the social network’s staff account earlier this week. The service had…

Cohost, the X rival founded with an anti-Big Tech manifesto, is running out of money and will shut down

At the MTV Video Music Awards (VMAs) on Wednesday night, new technology allowed fans to shop their favorite artists’ styles as they appeared on the screen. Though the drama from…

Shopsense AI lets music fans buy dupes inspired by red-carpet looks at the VMAs

Featured Article

A comprehensive list of 2024 tech layoffs

A complete list of all the known layoffs in tech, from Big Tech to startups, broken down by month throughout 2024.

A comprehensive list of 2024 tech layoffs

Working away on his PhD in Munich only a few years ago, Stephan Herrmann (now a doctor) couldn’t have conceived of a time when his idea for a carbon-negative power…

This startup is making manure out of other biogas power plants and now has $62M to play with

ChatGPT, OpenAI’s text-generating AI chatbot, has taken the world by storm since its launch in November 2022. What started as a tool to hyper-charge productivity through writing essays and code…

ChatGPT: Everything you need to know about the AI-powered chatbot

Faraday Future is doling out big raises and bonuses to its CEO and its founder, despite having delivered just 13 cars in its 10-year history and recently laying off or…

Faraday Future gives CEO and founder raises and bonuses after delivering 13 cars