← Back to articles
Meta accused of using pirated data to develop Llama

Meta accused of using pirated data to develop Llama

Meta is facing new controversy regarding the development of its open-source AI, Llama. Recent leaks suggest the company may have exploited data from

By Veille IA Gennn··1 min read
🎧 Écouter le résumé
0:00 / 0:00

Meta is facing new controversy regarding the development of its open-source AI, Llama. Recent leaks suggest the company may have exploited data from LibGen (Library Genesis), a controversial platform known for hosting pirated content.

A controversy reaching the top

According to the revealed information, Mark Zuckerberg personally approved the use of this database, despite internal concerns. Reports indicate that Meta was aware of the risks associated with using potentially illegal data but proceeded nonetheless.

Meta's stance: no copyright infringement

In response to the accusations, Meta claims there are no legal issues in training Llama, citing the principle of Fair Use. According to the company, the origin of the data is irrelevant as long as its use complies with this legal framework. Meta also insists that no direct evidence has shown that LibGen was specifically used to train Llama.

A major issue for the future of AI

This case highlights the growing challenges related to ethics and legality in AI development. If the accusations are confirmed, Meta could face sanctions and see its image tarnished, in a context where AI regulation is becoming increasingly strict.