概要: 裁判記録によると、Metaの従業員らが、ライセンス取得に伴うコストと速度への懸念を理由に、LLaMA 3の学習用書籍の海賊版作成について協議していたことが明らかになった。内部メッセージによると、Metaはマーク・ザッカーバーグの承認を得て、750万冊以上の海賊版書籍を保管するLibGenにアクセスしていたことが示唆されている。従業員らは、データセットの出所を隠すための措置を講じていたとされている。OpenAIもLibGenの使用に関与していたとされている。
Editor Notes: Please refer to these two legal filings for more information; the incident date of 02/28/2023 is drawn from (2): (1) Case 3:23-cv-03417-VC, Document 417-6, filed 02/05/2025, Exhibit C, https://storage.courtlistener.com/recap/gov.uscourts.cand.415175/gov.uscourts.cand.415175.449.4.pdf; and (2) Case 3:23-cv-03417-VC, Document 449-4, filed 02/20/2025, Woodhouse Exhibit 4, Exhibit C, https://storage.courtlistener.com/recap/gov.uscourts.cand.415175/gov.uscourts.cand.415175.449.4.pdf. See also Incidents 995 and especially 996 for similarly related cases.
推定: OpenAI , Meta , OpenAI models , Llama 3 , Library Genesis (LibGen) , GPT-4 と BitTorrentが開発し提供したAIシステムで、Writers , publishers , Journalists , Authors と Academic researchersに影響を与えた
インシデントのステータス
Risk Subdomain
A further 23 subdomains create an accessible and understandable classification of hazards and harms associated with AI
2.1. Compromise of privacy by obtaining, leaking or correctly inferring sensitive information
Risk Domain
The Domain Taxonomy of AI Risks classifies risks into seven AI risk domains: (1) Discrimination & toxicity, (2) Privacy & security, (3) Misinformation, (4) Malicious actors & misuse, (5) Human-computer interaction, (6) Socioeconomic & environmental harms, and (7) AI system safety, failures & limitations.
- Privacy & Security
Entity
Which, if any, entity is presented as the main cause of the risk
Human
Timing
The stage in the AI lifecycle at which the risk is presented as occurring
Pre-deployment
Intent
Whether the risk is presented as occurring as an expected or unexpected outcome from pursuing a goal
Intentional