
Chamath Palihapitiya
@chamath · 28. Dez. 2023
Say what you will about the NYT v OAI lawsuit but one thing is clear:
Even non-profit crawlers have become biased left over the years. See below: Common Crawl is mostly a servile dump of left-biased content. And it’s the largest weight.
There needs to be a more balanced, non
Cecilia Ziniti@CeciliaZin· 27. Dez. 20231/ First, the complaint clearly lays out the claim of copyright infringement, highlighting the 'access & substantial similarity' between NYT's articles and ChatGPT's outputs. Key fact: NYT is the single biggest proprietary data set in Common Crawl used to train GPT.

Elon Musk
@elonmusk
True
08:07 · 29. Dezember 2023 · 80.948 Aufrufe
51
24
746