Kate Knibbs reports in Wired magazine:

Against the company’s wishes, a court unredacted information alleging that Meta used Library Genesis (LibGen), a notorious so-called shadow library of pirated books that originated in Russia, to help train its generative AI language models. […] In his order, Chhabria referenced an internal quote from a Meta employee, included in the documents, in which they speculated, “If there is media coverage suggesting we have used a dataset we know to be pirated, such as LibGen, this may undermine our negotiating position with regulators on these issues.” […] These newly unredacted documents reveal exchanges between Meta employees unearthed in the discovery process, like a Meta engineer telling a colleague that they hesitated to access LibGen data because “torrenting from a [Meta-owned] corporate laptop doesn’t feel right 😃”. They also allege that internal discussions about using LibGen data were escalated to Meta CEO Mark Zuckerberg (referred to as “MZ” in the memo handed over during discovery) and that Meta’s AI team was “approved to use” the pirated material.

You are viewing a single thread.
View all comments View context
6 points

The pivot-to-ai writeup is out, they did seed! I assume it’s documented then.

Multinational corporations can act ethically after all.

permalink
report
parent
reply
5 points

Multinational corporations can act ethically after all.

I wouldn’t go that far

permalink
report
parent
reply
3 points

They can, they just choose deliberately not to most of the time.

In total honesty though, Meta had actually done some good things for Open Source. Sure, this is probably it of their own interest and neither outweighs nor make up for all the bad. But they can, and sometimes do.

permalink
report
parent
reply
3 points

It’s clear that they didn’t stop uploads of the torrents. It hasn’t been established in the documents we’ve seen so far that they actually had downloaders in turn. But they did clearly make the works available for upload.

permalink
report
parent
reply

TechTakes

!techtakes@awful.systems

Create post

Big brain tech dude got yet another clueless take over at HackerNews etc? Here’s the place to vent. Orange site, VC foolishness, all welcome.

This is not debate club. Unless it’s amusing debate.

For actually-good tech, you want our NotAwfulTech community

Community stats

  • 1.5K

    Monthly active users

  • 543

    Posts

  • 12K

    Comments

Community moderators