Tech
tech
Jon Keegan

Meta used a pirated library of millions of books and papers to train its AI because they thought everybody was doing it

In January, we learned from internal Meta communications revealed in a copyright lawsuit that the company downloaded LibGen, a massive collection of pirated, copyrighted works including millions of books and academic papers, to train its Llama AI model. This legally dubious move was approved by “MZ.”

More details are emerging surrounding this consequential decision as the lawsuit plays out. New court filings detail the internal deliberations within Meta involving researchers who knew using pirated works was a big no-no, but they did it anyway, as they suspected their competitors were using the archive, too. Meta employees wrote:

“everyone is using lib-gen (startups, but also google, openAI)”

“And I’m pretty sure other folks have no issues taking all of libgen 😊”

The Atlantic took a deeper look at what exactly is in this dataset. Using a “snapshot” of the archive (just a list of what is in there, not the works themselves), they created a search tool you can use to find exactly what works were in the archive. Authors who found that their works were in the dataset have taken to social media to express their outrage.

More details are emerging surrounding this consequential decision as the lawsuit plays out. New court filings detail the internal deliberations within Meta involving researchers who knew using pirated works was a big no-no, but they did it anyway, as they suspected their competitors were using the archive, too. Meta employees wrote:

“everyone is using lib-gen (startups, but also google, openAI)”

“And I’m pretty sure other folks have no issues taking all of libgen 😊”

The Atlantic took a deeper look at what exactly is in this dataset. Using a “snapshot” of the archive (just a list of what is in there, not the works themselves), they created a search tool you can use to find exactly what works were in the archive. Authors who found that their works were in the dataset have taken to social media to express their outrage.

More Tech

See all Tech
tech

Getty Images suffers partial defeat in UK lawsuit against Stability AI

Stability AI, the creator of the image generation tool Stable Diffusion largely defended itself from a copyright violation lawsuit filed by Getty Images , which claimed the company illegally trained its AI models on Getty’s image library.

Lacking strong enough evidence, Getty dropped the part of the case alleging illegal training mid-trial, according to Reuters reporting.

Responding to the decision, Getty said in a press release:

Today’s ruling confirms that Stable Diffusion’s inclusion of Getty Images’ trademarks in AI‑generated outputs infringed those trademarks. ...The ruling delivered another key finding; that, wherever the training and development did take place, Getty Images' copyright‑protected works were used to train Stable Diffusion.

Stability AI still faces a lawsuit from Getty in US courts, which is still ongoing.

A number of high-profile copyright cases are still working their way through the courts, as copyright holders seek to win strong protections for their works which were used to train AI models from a number of big tech companies.

Responding to the decision, Getty said in a press release:

Today’s ruling confirms that Stable Diffusion’s inclusion of Getty Images’ trademarks in AI‑generated outputs infringed those trademarks. ...The ruling delivered another key finding; that, wherever the training and development did take place, Getty Images' copyright‑protected works were used to train Stable Diffusion.

Stability AI still faces a lawsuit from Getty in US courts, which is still ongoing.

A number of high-profile copyright cases are still working their way through the courts, as copyright holders seek to win strong protections for their works which were used to train AI models from a number of big tech companies.

tech

Norway’s wealth fund, Tesla’s sixth-largest institutional investor, votes against Musk’s pay package

Norway’s Norges Bank Investment Management, the world’s largest sovereign wealth fund, said Tuesday that it voted against Tesla CEO Elon Musk’s $1 trillion pay package, ahead of the EV company’s annual shareholder meeting Thursday. The fund, which has a 1.2% stake in Tesla, is the company’s sixth-largest institutional investor, according to FactSet, and the first major investor to disclose how it voted on the matter.

Tesla is down nearly 3% premarket, amid a wider pullback in equities that’s most pronounced in AI-related stocks.

“While we appreciate the significant value created under Mr. Musk’s visionary role, we are concerned about the total size of the award, dilution, and lack of mitigation of key person risk- consistent with our views on executive compensation,” NBIM said in a statement.

Tesla’s board considers Musk’s mammoth, performance-based pay package necessary to retain Musk. For what it’s worth, prediction markets are quite certain investors will pass the proposition.

Tesla is down nearly 3% premarket, amid a wider pullback in equities that’s most pronounced in AI-related stocks.

“While we appreciate the significant value created under Mr. Musk’s visionary role, we are concerned about the total size of the award, dilution, and lack of mitigation of key person risk- consistent with our views on executive compensation,” NBIM said in a statement.

Tesla’s board considers Musk’s mammoth, performance-based pay package necessary to retain Musk. For what it’s worth, prediction markets are quite certain investors will pass the proposition.

tech

Waymo to expand robotaxi service to Detroit, Las Vegas, and San Diego

Google’s Waymo robotaxi service is expanding to three new cities — Detroit, Las Vegas, and San Diego — where it has previously tested its driverless vehicles. Waymo plans to bring its Jaguar I-Pace and Zeekr RT vehicles to those three markets this week, but they won’t be immediately available to the public.

Currently Waymo is available in five US cities: Atlanta, Austin, Los Angeles, Phoenix, and San Francisco.

Tesla is currently testing in Las Vegas, while Amazon’s Zoox has limited service in the city.

Currently Waymo is available in five US cities: Atlanta, Austin, Los Angeles, Phoenix, and San Francisco.

Tesla is currently testing in Las Vegas, while Amazon’s Zoox has limited service in the city.

Latest Stories

Sherwood Media, LLC produces fresh and unique perspectives on topical financial news and is a fully owned subsidiary of Robinhood Markets, Inc., and any views expressed here do not necessarily reflect the views of any other Robinhood affiliate, including Robinhood Markets, Inc., Robinhood Financial LLC, Robinhood Securities, LLC, Robinhood Crypto, LLC, or Robinhood Money, LLC.