In Cringe Video, OpenAI CTO Says She Doesn’t Know Where Sora’s Training Data Came From

Maggie Harrison Dupré

15 March 2024 at 9:01 am·4-min read

Wondering what data OpenAI used to train its buzzy new text-to-video AI? The company's CTO is similarly unsure.

Mira Murati, OpenAI's longtime chief technology officer, sat down with The Wall Street Journal's Joanna Stern this week to discuss Sora, the company's forthcoming video-generating AI. About halfway through the 10-minute-long interview, Stern straightforwardly asked Murati where the new model's training data was gleaned from. But Murati, in the most cringe-inducing way possible, couldn't find an answer beyond vague corporate language.

"We used publicly available data and licensed data," Murati responded to the resoundingly simple question.

Stern pushed back with more specific source examples: "So, videos on YouTube?"

"I'm actually not sure about that," said Murati, before rebuffing further queries about whether videos shared to Instagram or Facebook were fed into model.

"You know, if they were publicly available — publicly available to use," the CTO answered, "but I'm not sure. I'm not confident about it."

Stern then inquired about OpenAI's data training partnership with the stock image company Shutterstock, asking if videos on the partnered platform were sucked into Sora's training material. And this time? Murati decided to shut down the line of questioning altogether.

"I'm just not going to go into detail about the data that was used," Murati continued. "But it was publicly available or licensed data."

So, in sum, Murati can't tell you exactly where the videos gobbled up by Sora first came from. But rest assured, the sourceless data was definitely, one hundred percent publicly available or licensed. Convincing stuff!

It's a bad look all around for OpenAI, which has drawn wide controversy — not to mention multiple copyright lawsuits, including one from The New York Times — for its data-scraping practices. After all, if the company's CTO can't firmly tell you where its buzziest new model's training data was sourced from, it doesn't exactly communicate a particular amount of care for the issue from OpenAI's higher-ups.

https://twitter.com/JoannaStern/status/1768306032466428291

After the interview, Murati reportedly confirmed to the WSJ that Shutterstock videos were indeed included in Sora's training set. But when you consider the vastness of video content across the web, any clips available to OpenAI through Shutterstock are likely only a small drop in the Sora training data pond.

Online, reactions to the clip were mixed, with many chalking Murati's close-lipped responses up to a possible lack of candidness.

"So when *the CTO* of OpenAI is asked if Sora was trained on YouTube videos, she says 'actually I'm not sure' and refuses to discuss all further questions about the training data," former LA Times tech columnist Brian Merchant wrote in an X-formerly-Twitter post. "Either a rather stunning level of ignorance of her own product, or a lie — pretty damning either way!"

"You're the CTO ma'am," added another netizen, "you should know."

Others, meanwhile, jumped to Murati's defense, arguing that if you've ever published anything to the internet, you should be perfectly fine with AI companies gobbling it up.

"Why does it matter? That is the question," said one X user. "I find it insane that people make things public to everyone in the world and then complain when someone uses that public thing. If you want to be private, then be private."

That latter argument, though, speaks to the bizarre new reality that internet users have now found themselves in. Historically, when someone told you to be careful of what you post online, the reasoning was something akin to "you might regret that later" — and not "a multibillion-dollar AI company might turn a profit by vacuuming that Facebook video of you and your family, or a goofy YouTube video you made with your friends, into a generative AI model."

Whether Murati was keeping things close to the vest to avoid more copyright litigation or simply just didn't know the answer, people have good reason to wonder where AI data — be it "publicly available and licensed" or not — is coming from. And moving forward, vague corporate mumbling probably isn't going to cut it.

More on OpenAI and its data: OpenAI Says It's Fine to Vacuum Up Everyone's Content and Charge for It Without Paying Them

Australian Associated Press
Husband found not guilty of 'brutal' wedding night rape
A man accused of a series of sexual assaults on his wedding night and honeymoon has been found not guilty on all charges in a Sydney court.
Cosmo
Rosalía goes braless and *almost* frees the nip in a lace naked dress
Rosalía stepped out wearing a breathtaking naked dress at the Prelude to the Olympics in Paris. The design was a nude coloured see-through lace gown by Dior.
HuffPost
Stephen Colbert Taunts Trump With Absolutely Brutal Reminder About Melania
The "Late Show" host mocked the former president over one curious claim.
Yahoo News Australia
Passengers slammed over 'disturbing' train act attracting $500 fine
Commuters were noticeably annoyed by the disturbance, one man told Yahoo, and were 'shifting away' from the men in question.
The Independent
Is Donald Trump good at golf? We asked a professional coach to analyze his swing
With Joe Biden calling Trump’s alleged golfing prowess into question, is the 45th president as good as he claims to be?
BuzzFeed
Kamala Harris' Press Release About Donald Trump's Fox News Appearance Is Going Viral
"Something about the question mark after 'old and quite weird' is taking me out."
Yahoo Sport Australia
Tennis world erupts over massive news about Novak Djokovic and Rafa Nadal at Olympics
Rafa Nadal has left the tennis world stunned. Find out more here.
NewsWire
Why Aussies being turned away from Bali
Hundreds of Aussie tourists are being denied entry into Indonesia’s island paradise for one reason.
Parade
Prince William Reportedly Removes Decades-Old Position From Royal Staff
The royal staff member reportedly let go is a relative of Queen Camilla.
NY Daily News
Harris campaign roasts Trump as ‘old and quite weird’ after Fox News insults
Republican presidential candidate Donald Trump called in to Fox News Thursday, where he told supporters that presumptive Democratic nominee Kamala Harris is a “radical left, not very smart person” who’s part of a massive conspiracy to weaponize the nation’s legal system against him. Harris’ campaign fired back mere minutes later with an email blasting the “78-year-old convicted criminal’s Fox ...
HuffPost
Jimmy Fallon Trolls Donald Trump With 3 Words, Over And Over Again
The "Tonight Show" host envisioned an exchange between the Republican presidential nominee and Elon Musk.
BuzzFeed
18 Famous "Childless Cat Ladies" And Their Thoughtful Reasons For Never Having Kids
Don't show this post to JD Vance.
Yahoo News Australia
Eerie dashcam footage captures ‘rare’ sight on Aussie highway
It's barely noticeable at first glance, but it could cost you your life – or at least a hefty repair bill – if you collide with it.
Evening Standard
FBI director suggests Donald Trump may not have been struck by bullet during assassination attempt at rally
FBI director Christopher Wray said investigators did not know whether Trump’s ear was grazed by a bullet or shrapnel
Hello!
Amanda Holden stuns in mini dress alongside lookalike daughters during Greek getaway
BGT judge Amanda Holden looked flawless as she holidayed with her mini-me daughters Lexi and Hollie. Take a look inside their lavish Greek getaway…
Parade
Nicole Scherzinger Sizzles in See-Thru Lace Dress With Risqué Chest Cutout in the French Riviera
The Pussycat Dolls singer showed off the racy look in spicy new social media snaps.
The Independent
Passenger refuses to let mother and child sit in her plane seat by providing controversial reason
‘As a very tall and big man, I have had this happen more than a few times,’ one commenter related to the Reddit post
The Independent
Wife was convicted of killing her husband in violent hammer attack. She was found dead hours before sentencing
Linda Kosuda-Bigazzi killed her husband with a hammer before hiding his body in the basement of their home and pocketing his paychecks for months
Yahoo Lifestyle
Kmart shoppers raving about $12 kitchen item with multiple uses: 'I have three'
The popular Kmart product has quickly become a household essential. Here's why.
Parade
Selma Blair Rocks Red Bikini by the Pool As She Sends Team USA a Message
The Summer 2024 Olympics officially kick off in Pairs on Friday, July 26.

In Cringe Video, OpenAI CTO Says She Doesn’t Know Where Sora’s Training Data Came From

Latest stories

Husband found not guilty of 'brutal' wedding night rape

Rosalía goes braless and almost frees the nip in a lace naked dress

Stephen Colbert Taunts Trump With Absolutely Brutal Reminder About Melania

Passengers slammed over 'disturbing' train act attracting $500 fine

Is Donald Trump good at golf? We asked a professional coach to analyze his swing

Kamala Harris' Press Release About Donald Trump's Fox News Appearance Is Going Viral

Tennis world erupts over massive news about Novak Djokovic and Rafa Nadal at Olympics

Why Aussies being turned away from Bali

Prince William Reportedly Removes Decades-Old Position From Royal Staff

Harris campaign roasts Trump as ‘old and quite weird’ after Fox News insults

Jimmy Fallon Trolls Donald Trump With 3 Words, Over And Over Again

18 Famous "Childless Cat Ladies" And Their Thoughtful Reasons For Never Having Kids

Eerie dashcam footage captures ‘rare’ sight on Aussie highway

FBI director suggests Donald Trump may not have been struck by bullet during assassination attempt at rally

Amanda Holden stuns in mini dress alongside lookalike daughters during Greek getaway

Nicole Scherzinger Sizzles in See-Thru Lace Dress With Risqué Chest Cutout in the French Riviera

Passenger refuses to let mother and child sit in her plane seat by providing controversial reason

Wife was convicted of killing her husband in violent hammer attack. She was found dead hours before sentencing

Kmart shoppers raving about $12 kitchen item with multiple uses: 'I have three'

Selma Blair Rocks Red Bikini by the Pool As She Sends Team USA a Message