Clicky

  • Login
  • Register
  • Submit Your Content
  • Contact Us
Thursday, August 22, 2024
World Tribune
No Result
View All Result
  • Home
  • News
  • Business
  • Technology
  • Sports
  • Health
  • Food
Submit
  • Home
  • News
  • Business
  • Technology
  • Sports
  • Health
  • Food
No Result
View All Result
World Tribune
No Result
View All Result

AI video startup Runway reportedly trained on ‘thousands’ of YouTube videos without permission

July 25, 2024
in Technology
Reading Time: 3 mins read
A A
AI video startup Runway reportedly trained on ‘thousands’ of YouTube videos without permission
0
SHARES
ShareShareShareShareShare

READ ALSO

Microsoft’s revised Recall AI feature will roll out to beta testers in October

Microsoft’s latest accessible controllers include the Xbox Adaptive Joystick

AI company Runway reportedly scraped “thousands” of YouTube videos and pirated versions of copyrighted movies without permission. 404 Media obtained alleged internal spreadsheets suggesting the AI video-generating startup trained its Gen-3 model using YouTube content from channels like Disney, Netflix, Pixar and popular media outlets.

An alleged former Runway employee told the publication the company used the spreadsheet to flag lists of videos it wanted in its database. It would then download them without detection using open-source proxy software to cover its tracks. One sheet lists simple keywords like astronaut, fairy and rainbow, with footnotes indicating whether the company had found corresponding high-quality videos to train on. For example, the term “superhero” includes a note reading, “Lots of movie clips.” (Indeed.)

Other notes show Runway flagged YouTube channels for Unreal Engine, filmmaker Josh Neuman and a Call of Duty fan page as good sources for “high movement” training videos.

“The channels in that spreadsheet were a company-wide effort to find good quality videos to build the model with,” the former employee told 404 Media. “This was then used as input to a massive web crawler which downloaded all the videos from all those channels, using proxies to avoid getting blocked by Google.”

AI video startup Runway reportedly trained on ‘thousands’ of YouTube videos without permission

Runway

A list of nearly 4,000 YouTube channels, compiled in one of the spreadsheets, flagged “recommended channels” from CBS New York, AMC Theaters, Pixar, Disney Plus, Disney CD and the Monterey Bay Aquarium. (Because no AI model is complete without otters.)

In addition, Runway reportedly compiled a separate list of videos from piracy sites. A spreadsheet titled “Non-YouTube Source” includes 14 links to sources like an unauthorized online archive of Studio Ghibli films, anime and movie piracy sites, a fan site displaying Xbox game videos and the animated streaming site kisscartoon.sh.

In what could be viewed as a damning confirmation that the company used the training data, 404 Media found that prompting the video generator with the names of popular YouTubers listed in the spreadsheet spit out results bearing an uncanny resemblance. Crucially, entering the same names in Runway’s older Gen-2 model — trained before the alleged data in the spreadsheets — generated “unrelated” results like generic men in suits. Additionally, after the publication contacted Runway asking about the YouTubers’ likenesses appearing in results, the AI tool stopped generating them altogether.

“I hope that by sharing this information, people will have a better understanding of the scale of these companies and what they’re doing to make ‘cool’ videos,” the former employee told 404 Media.

When contacted for comment, a YouTube representative pointed Engadget to an interview its CEO Neal Mohan gave to Bloomberg in April. In that interview, Mohan described training on its videos as a “clear violation” of its terms. “Our previous comments on this still stand,” YouTube spokesperson Jack Mason wrote to Engadget.

Runway did not respond to a request for commeInt by the time of publication.

At least some AI companies appear to be in a race to normalize their tools and establish market leadership before users — and courts — catch onto how their sausage was made. Training with permission through licensed deals is one thing, and that’s another tactic companies like OpenAI have recently adopted. But it’s a much sketchier (if not illegal) proposition to treat the entire internet — copyrighted material and all — as up for grabs in a breakneck race for profit and dominance.

404 Media’s excellent reporting is worth a read.

Credit: Source link

ShareTweetSendSharePin
Previous Post

Overwatch 2 may test a return to six-player teams

Next Post

Mets’ contract talks with Francisco Alvarez hit roadblock

Related Posts

Microsoft’s revised Recall AI feature will roll out to beta testers in October
Technology

Microsoft’s revised Recall AI feature will roll out to beta testers in October

August 22, 2024
Microsoft’s latest accessible controllers include the Xbox Adaptive Joystick
Technology

Microsoft’s latest accessible controllers include the Xbox Adaptive Joystick

August 21, 2024
Volkswagen’s long-awaited electric ID.Buzz pricing and range revealed
Technology

Volkswagen’s long-awaited electric ID.Buzz pricing and range revealed

August 21, 2024
Reanimal promises a ‘more terrifying journey’ than Little Nightmares
Technology

Reanimal promises a ‘more terrifying journey’ than Little Nightmares

August 21, 2024
Indiana Jones and the Great Circle has a Nazi-slapping mechanic
Technology

Indiana Jones and the Great Circle has a Nazi-slapping mechanic

August 21, 2024
What to expect at the iPhone 16 keynote
Technology

What to expect at the iPhone 16 keynote

August 20, 2024
Next Post
Mets’ contract talks with Francisco Alvarez hit roadblock

Mets' contract talks with Francisco Alvarez hit roadblock

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

What's New Here!

Korean research utilises LLM to predict dementia risk

Korean research utilises LLM to predict dementia risk

August 6, 2024
What doomed Sha’Carri Richardson in the 100m final at Olympics

What doomed Sha’Carri Richardson in the 100m final at Olympics

August 4, 2024
Disney’s media assets are generating more excitement than parks

Disney’s media assets are generating more excitement than parks

August 7, 2024
MLB Hall of Famer Dennis Eckersley’s daughter, Alexandra, facing trial over child abandonment case

MLB Hall of Famer Dennis Eckersley’s daughter, Alexandra, facing trial over child abandonment case

July 26, 2024
Why this Mets’ road trip is especially brutal

Why this Mets’ road trip is especially brutal

August 2, 2024
Yankees bolster bullpen by acquiring Mark Leiter Jr. in Cubs trade

Yankees bolster bullpen by acquiring Mark Leiter Jr. in Cubs trade

July 30, 2024
Pete Alonso slugs two home runs as Mets pummel Rockies to win series

Pete Alonso slugs two home runs as Mets pummel Rockies to win series

August 9, 2024

About

World Tribune is an online news portal that shares the latest news on world, business, health, tech, sports, and related topics.

Follow us

Recent Posts

  • Star fund manager takes leave amid accusations of cherry picking
  • FTX Sam Bankman-Fried former partner Ryan Salame seeks to void guilty plea
  • Noah Lyles gushes over ‘fighter’ girlfriend Junelle Bromfield
  • Microsoft’s revised Recall AI feature will roll out to beta testers in October

Newslatter

Loading
  • Submit Your Content
  • Contact Us
  • Privacy Policy
  • Terms of Use
  • DMCA

© 2024 World Tribune - All Rights Reserved!

No Result
View All Result
  • Home
  • News
  • Business
  • Technology
  • Sports
  • Health
  • Food

© 2024 World Tribune - All Rights Reserved!

Welcome Back!

Login to your account below

Forgotten Password? Sign Up

Create New Account!

Fill the forms below to register

All fields are required. Log In

Retrieve your password

Please enter your username or email address to reset your password.

Log In