Replying to @⁨artyom@piefed.social⁩

The scraping wouldn’t be a problem if Reddit simply provided an RSS feed or other data-efficient API. The “ramfucking” is caused by the attempt to block bots; it is entirely self-inflicted.

Remember, it’s all our content to begin with and Reddit does not have any right to try to lock it up for itself.

That doesn’t mean I like all the AI bullshit going on, BTW. But the problem is the generation of the slop, not the data accessibility.

Replying to @⁨artyom@piefed.social⁩

Okay, if efficient APIs existed and they weren’t incompetently failing to use them, it wouldn’t be a problem. Happy now?

(I should’ve addressed that in my previous comment, as I was aware of how one of the Lemmy instances was taken down by scrapers the other day despite the fact that they could easily get all the content simply by consuming ActivityPub directly. But I was naively hoping it wouldn’t be necessary because, as you can see from this text, it would’ve cluttered up my writing with double the words.)

Replying to an earlier post

I think it would be a problem because the scrapers are hammering all types of websites from small forums to reddit with tens of thousands of unique ip addresses at a time. Websites that have neither the money, hardware, or protection had to figure out solutions really quick or suffer what is essentially a constant ddos attack. This is the reality of the web now, it’s just an incredibly hostile place.

Replying to @⁨grue@lemmy.world⁩

The scraping wouldn’t be a problem if Reddit simply provided an RSS feed or other data-efficient API

Reddit does provide RSS feeds, e.g.: www.reddit.com/r/SonicTheHedgehog/.rss

Frankly, I’m surprised they still offer RSS feeds. They’ve been slowly but surely killing off all ways of accessing their content for years. One day they’ll disable them, but for now they still work.

Replying to @⁨Dudewitbow@lemmy.zip⁩

That’s just an excuse by reddit. Moving to the new style allows them to choke down and control the way that posts and replies are displayed and nested. This is good for them, because it allows them to offer white glove PR services to paying customers. It also obfuscates useful user supplied content so that it can be sold wholesale to anyone who has the money to buy it. That’s more important to them than offering a good user experience and useful website to the proles.

Anyone still posting on reddit (who isn’t a bot) is working for free for an unscrupulous company.

Replying to @⁨voluble@lemmy.ca⁩

Anyone still posting on reddit (who isn’t a bot) is working for free for an unscrupulous company.

Reddit was founded during the “crowdsourcing” craze. People realized that they could launch websites where all the content was created by the users, and it would snowball into daily views.

My point is that posting and commenting on Reddit has always been doing free work for an unscrupulous company. You could argue that it wasn’t unscrupulous before it went public, but even that is debatable.

Replying to an earlier post

I am so glad I’m off reddit. Even though I could use their help on a great number of subjects. Neverheless, because Isreal I can’t use reddit apparently. Not permabanned yet but they are on my shit with fake violations like right away now. I abandon them after a 2nd violation, go through one every 3 to 6 months when using it.

Every single company that goes public gets worse, reddit will be no exception. Those craven amoral cockscum are not to be trusted, nor patronized with our words they can use for their own benefit.

Replying to @⁨lemmydividebyzero@reddthat.com⁩

I now browse Wikipedia. Please don’t screw me over Wikipedia, I donated five bucks to one of your nags once.

For anyone who doesn’t know, you can download Wikipedia and host it yourself! I got the top 50k version (~7G) on my RPI3 and now no matter what fuckery they pull or the government pulls, I’ve got a pretty decent source of general information.

Replying to @⁨HAL_9_TRILLION@lemmy.world⁩

Yeah, I did that about 20 months ago (wonder why?) and the problem is, there it sits but I forget where on my LAN I stored it, how to access it - I suppose, were I absolutely desperate, I could launch a several hours research project and find it, maybe even figure out how to access (I believe I stored README with it about how to do all that), but… why? And, even worse, one why? might be to see how articles have drifted over time, which I guarantee they have and continue to do, Wikipedia is actually quite fluid, but does knowing how current articles compare to the ones you archived in 2024 do anything of greater value for you than the negative value of being pissed off at how the world is lying to itself (same as it ever has…)?

Replying to @⁨MangoCats@feddit.it⁩

I have my Pi-Hole doing DNS for me so I just set up an easy to remember redirect, wikipedia.local. I did it more for independence. I got Home Assistant and Voice PE so I could have a local smart speaker to get big tech out of my house. I don’t do anything on the cloud, so it stands to reason that if I want to make sure I always have access to rudimentary information, having my own version of WP is nice. I also just think it’s kind of cool and it never begs me for money.

Replying to @⁨HAL_9_TRILLION@lemmy.world⁩

Could you share what you did to achieve this? I’ve been planning on doing just that, and have it auto-update every week or so (keeping the previous versions archived, of course) by using kiwix-serve for a static ‘.zim’ file and maybe a cron job for the auto-update. But if you have a better solution, I’d love to know. The deployment I am planning is kind of convoluted to be honest.

Replying to @⁨jjlinux@lemmy.zip⁩

No, that’s exactly what I did, I’m running kiwix-serve, but I’m not going to bother updating it because I’m really worried about information degrading now that fascists are basically calling the shots on everything (and WP’s jackboot co-founder has a hard on for it). If I feel enough time has gone by to warrant an update I’ll just do it manually.

Replying to @⁨MrOtingocni@lemmy.world⁩

You run a server on a machine inside your house, it can be any computer on your local LAN/wifi, but it’s obviously best if it’s a machine that’s always on. I use a Raspberry Pi 3B+ (these can be had for about $50) that I have plugged into my wifi router and it’s running a little program called Kiwix-Server (free and open source). You download the WP file (it’s a huge single file with a .zim extension) and point the server to it and boom.

Replying to @⁨rumba@lemmy.zip⁩

My recommendations have been mostly fine and just keep the channels i regularly watch at the forefront. I’ve had my account for decades now so my subscribed feed is too much of a mess to wrangle now.

That saying, I do really miss the custom folders you could once make on YouTube. Back then I would categorise certain favourite YouTubers together and mostly use that. Using something like that again would be nice.

Replying to @⁨MrScottyTay@sh.itjust.works⁩

They also have this feature, you have subscription groups!

As for recommendations, it does one big sync once then anytime you open the app it seemed to sync everything (ime at least)

That alongside being able to sync a video to my PC at the exact timestamp is very nice for living room experiences. Bonus you can have peer tube sources show up next to your yt videos

Replying to @⁨Gsus4@mander.xyz⁩

Yeah that’s the killer. Just not enough content yet. I think it’s cause video hosting is expensive. My hope is that individual youtubers start hosting their own peertube instances. But they’d never do that because they’d lose money.

If peertube could get functionality which would allow creators to hide videos behind a subcription which could be paid in fiat/crypto that would be a game changer. Easy to donate to your favourite creators while keeping federation.

Maybe one day.

Defo would recommend changing your email though. Takes a while to do all your accounts but once it’s done you’re free! Feels good.

Replying to @⁨lemmydividebyzero@reddthat.com⁩

If you consider scraping a threat then yes plain HTML might as well be giving up. The advantage of new reddit for that is quite clear: they can collect a bunch of data about your browser before deciding if you are a bot and if the rest of the page should load. The embedded recaptcha call in the screenshots is a pretty good hint. I suspect blocking trackers on new reddit will break as soon as the scrapers move over.

As for why they don’t just kill old reddit: a significant chunk of their active posters use it and are attached to it. So if they kill it entirely they will lose content. Posters are of course logged in so this change is less likely to affect them.

Replying to @⁨PattyMcB@lemmy.world⁩

They are in the business of attracting new users into their new algorithmic engagement hellhole now. Show any propensity of interest, and you will get sidetracked to the most godawful side-communities that seem to have emerged to engage as many victims as possible. They do not want to focus on their old users as anything less than the content they already made that makes reddit show up as free advertisement to their new base in search engines. It’s all a game of “it’s the algorithm’s fault so you can’t blame us” now.

Replying to @⁨lemmydividebyzero@reddthat.com⁩

Yesterday, I visited new Reddit. No VPN, Chrome in incognito mode, so no extensions, clickedon a link directly from Google search. Got a message that my access was blocked for security reasons. Copied the link to Firefox (with uBO and a few privacy-centric extensions), changed it to old reddit (where I was already logged in), and it worked just fine. I found out that when Reddit kills old reddit, I won’t even have the choice to switch to the new one (not that I ever would) because I’d be blocked anyway.