
Chamath Palihapitiya
@chamath · Dec 28, 2023
Say what you will about the NYT v OAI lawsuit but one thing is clear:
Even non-profit crawlers have become biased left over the years. See below: Common Crawl is mostly a servile dump of left-biased content. And it’s the largest weight.
There needs to be a more balanced, non
Cecilia Ziniti@CeciliaZin· Dec 27, 20231/ First, the complaint clearly lays out the claim of copyright infringement, highlighting the 'access & substantial similarity' between NYT's articles and ChatGPT's outputs. Key fact: NYT is the single biggest proprietary data set in Common Crawl used to train GPT.

Elon Musk
@elonmusk
True
08:07 AM · December 29, 2023 · 80.9K views
51
24
746