Replying to @⁨im_fine_sandy@nord.pub⁩

They actually did use old reddit, as confirmed by reddit itself, as confirmed by massive drop after the change. It was simply far easier to scrape old reddit (pre-rendered instead of relying on JavaScript and dynamic content load)

engadget.com/…/old-reddit-could-be-the-next-casua…

EngadgetOld Reddit could be the next casualty of Reddit's war on AI scraping - EngadgetThe company is also planning to retire its public API.
Edited ⁨⁨Aug⁩ ⁨22⁩, ⁨2026⁩, ⁨20:21⁩⁩en

Replying to an earlier post

You seem to have misunderstood, and the article you linked doesn’t contradict what I’ve said.

Wealthy, sophisticated LLM developers would not have been scraping old reddit. They would pay reddit for API access.

Shutting down old reddit closes the door on back yard developers, not OpenAI.

Additionally, when you run a prompt in a chatbot, it doesn’t scurry away and scrape old reddit and then formulate an answer. The scraping of content is going on while the model is being developed.

Replying to an earlier post

As if wealthy people pay for stuff they could have for free. Literally everybody had their copyright infringed from the training of GenAI. Authors, Journalists, Artists, Musicians, Regisseurs, the list is endless. Unless you sue, you won’t see any compensation from them. By now you should know that companies only play fair if they’re forced to, otherwise literally anything goes.

It’s quite naive to believe they would pay for API access, tbh.