Hehe, that analogy is hilarious knowing that it expanded its activities into a new space voluntarily, and arguably without much due diligence. That's like saying: "In light of crackdowns on extortion, Mafia is in a fight to save itself."
@alexjc
-

Meta LLM Lawsuit Reaches Summary Judgment on Fair Use
By
–
The LLM lawsuit against Meta will reach 'summary judgement' on the Fair Use question soon. When the CEO says something like this at such an important phase, it's not a coincidence. This is the outcome that Meta has been lobbying, litigating, pressuring for — with all its power.
-

Judge Criticizes AI Copyright Lawsuit Team Competence
By
–
The Judge also said Saveri & his team seem "unwilling or unable to litigate properly." UNABLE means incompetent.
UNWILLING means conflicted. The judge is spelling out why this is going to be a complete disaster across the board—they don't even have Copyright lawsuit experience. -
Perplexity Feed Strategy Questioned: Profitability and Business Focus Concerns
By
–
Predicting this one to fail, because people mostly want humanity in their feed.
— Alex J. Champandard 🌱 (@alexjc) 14 septembre 2024
Also, it casts doubts about Perplexity's focus on their core business and weird incentives that lead to what seems like a sidequest. The core search not profitable so they must have a "feed" for ads? https://t.co/Fajfs8a28lPredicting this one to fail, because people mostly want humanity in their feed. Also, it casts doubts about Perplexity's focus on their core business and weird incentives that lead to what seems like a sidequest. The core search not profitable so they must have a "feed" for ads?
-

Staged StopAI Movement Undermines Legitimate AI Safety Concerns
By
–
People are obviously seeing how fake the staged StopAI movement is, which undermines legitimate concerns and solutions. The same questions should be asked of the XR movement too, likewise burying important points in the climate debate.
-
Web-Scale Data: Legal Risks and Compliance Requirements
By
–
Voilà! Overall, dealing with web-scale data is inherently risky and requires due care and good faith to reduce legal risk. In lawsuits, the burden of proof quickly shifts onto the defendants (because it's an affirmative defense) and they need to provide evidence of compliance.
-
Fair Use Licensing for AI Datasets: Legal Compliance Strategy
By
–
Use an appropriate license compatible with Fair Use. LAION's creation of the dataset required one or more Copyright Exceptions with strict terms. Picking a license that respects the terms of these exceptions would reduce legal risk significantly, and show good faith.
-
Partner with IWF upfront to prevent illegal content in AI datasets
By
–
Partner with IWF upfront, not when problem arise! LAION knew about the risks of illegal content (e.g. child abuse) being in their dataset. It would have been sensible to partner with professional organizations like IWF to check files *before* publishing the dataset.
-
LAION criticized for partnering with questionable data hosting site
By
–
Don't partner with dubious websites to host files! LAION shouldn't have requested TheEye, a questionable "archival" website, to host their data in order for them to link to it! Maybe instead get approval from platforms like HF so they share legal responsibility for hosting.
-
User Generated Content Exclusion from AI Training Datasets
By
–
Exclude social media and UGC websites. Popular websites that host user generated content or personal profiles could have been excluded from the dataset. Platforms don't have full rights to their users private information, and avoiding that kind of content reduces legal risk.