this post was submitted on 29 Aug 2026
242 points (97.3% liked)

Technology

87649 readers
3291 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
top 50 comments
sorted by: hot top controversial new old
[–] MonkderVierte@lemmy.zip 4 points 26 minutes ago

This is because appending site: reddit.com to a search query is basically a surefire way to find results written by genuine humans.

The article is from 2026, not 2016? That bot-ridden Reddit? Am i in the wrong film?

[–] DeadSquirrel@lemmy.blahaj.zone 2 points 53 minutes ago (1 children)

There's actually a very big algorithmic difference in how the content is shown between old and new reddit.

I think that's also why they keep trying to make old reddit unusable little by little.

Moreover, without RES, you can't tag users, which has become essential nowadays (especially since users can now hide their comment history). It's already known by now, but it's super weird seeing people from one country pretending to be from another one.

[–] MonkderVierte@lemmy.zip 2 points 29 minutes ago (1 children)

tag users […] it's super weird seeing people from one country pretending to be from another one.

I see issues.

Implementation A: automatic by IP)

  • they moved
  • they use a VPN

Implementation B: field in profile settings)

  • can write whatever the fuck they want
  • can probably change it
[–] DeadSquirrel@lemmy.blahaj.zone 3 points 6 minutes ago* (last edited 32 seconds ago) (1 children)

In this case, tagging is something done manually and locally by the user of the extension (aka me), by going to the various subs (they're divided politically) and verifying what people are saying (non-English, local memes, culture, local news, etc.). Tags can be modified if I find out the user was badly categorized.

I've also tagged different users that were involved in viral incidents (unknown details beyond that), where it seems a whole lot of people decided to pile up on a topic.

It's interesting to see how much manipulation is going on in reddit. I've stopped interacting over there because of this. You just never know who is real, or a bot, or someone paid to post propaganda.

Edit: something I forgot to mention, and it's relevant to this blocking of non-logged in users, is that the RES tags are persistent and don't depend on an account (on my end). So I could (in the past) delete my reddit account, and still be able to see and apply tags. But now I can't do it unless I'm logged in.

[–] MonkderVierte@lemmy.zip 1 points 4 minutes ago

Ahh, so jt's you doing the tagging of others, only visible to you.

[–] lechekaflan@lemmy.world 12 points 3 hours ago

Fuck you spez.

[–] PattyMcB@lemmy.world 16 points 6 hours ago (3 children)

Reddit seems to be in the business of extracting as much value as it can from said forums without completely destroying them.

I call bs. It's been completely destroyed for a while.

[–] fodor@lemmy.zip 3 points 1 hour ago

I agree with you. But what's important is if they can sell anything. They destroyed their own product but they're still pretending it has value, and maybe they can fool some investors into paying for script and AI slop.

[–] rumba@lemmy.zip 1 points 46 minutes ago

While I agree, More celebs than ever have been asking for comments and ideas on their reddit pages.

[–] JackbyDev@programming.dev 6 points 4 hours ago

I think they're more upset AI companies scraped 'em and they didn't get paid.

[–] Treczoks@lemmy.world 1 points 3 hours ago (1 children)

Apart from all that, what would prevent an AI knowledge thieving tool from using an actual account on Reddit?

[–] lemmydividebyzero@reddthat.com 1 points 2 hours ago

One would need more than 1 account...

[–] carpelbridgesyndrome@sh.itjust.works 26 points 9 hours ago (1 children)

If you consider scraping a threat then yes plain HTML might as well be giving up. The advantage of new reddit for that is quite clear: they can collect a bunch of data about your browser before deciding if you are a bot and if the rest of the page should load. The embedded recaptcha call in the screenshots is a pretty good hint. I suspect blocking trackers on new reddit will break as soon as the scrapers move over.

As for why they don't just kill old reddit: a significant chunk of their active posters use it and are attached to it. So if they kill it entirely they will lose content. Posters are of course logged in so this change is less likely to affect them.

[–] heartSagan5@lemmy.zip 6 points 6 hours ago

if you are a bot

Or they’re bouncing banned people.

[–] PurpleFanatic@quokk.au 24 points 10 hours ago (1 children)

It's shitty moves like this that have me feeling deeply grateful about the fediverse. It's NOT without its myriad of problems, but how lucky are we? We're insulated from all this bullshit.

Mastodon is every bit as good (and better) as it was when I started using it in 2018. Can the same be said for Reddit, Instagram, Facebook or YouTube? Absolutely the fuck not.

[–] Gsus4@mander.xyz 6 points 7 hours ago (2 children)

I'll admit I'm having trouble moving from yt to peertube :/ and still use the old gmail accounts

[–] Truscape@lemmy.blahaj.zone 2 points 4 hours ago (1 children)

You can use something like GrayJay to watch YT and other sources (and have offline playlists and subscriptions) without a google account whatsoever. That helped me make the jump to fully degoogle.

[–] MrScottyTay@sh.itjust.works 2 points 4 hours ago (2 children)

The lack of a "cross-platform" (for lack of a better term) account is why I don't use things like GrayJay because I watch on different devices and want to ensure they all have similar recommendations and a watch history.

I use SmartTube next on tv like 80% of the time I watch YouTube.

[–] rumba@lemmy.zip 1 points 44 minutes ago (1 children)

Recommendations are their method of control. Eshew the algorithm, curate your own choices of who you watch. also drastically reduces your exposure to slop.

[–] MrScottyTay@sh.itjust.works 1 points 40 minutes ago

My recommendations have been mostly fine and just keep the channels i regularly watch at the forefront. I've had my account for decades now so my subscribed feed is too much of a mess to wrangle now.

That saying, I do really miss the custom folders you could once make on YouTube. Back then I would categorise certain favourite YouTubers together and mostly use that. Using something like that again would be nice.

[–] Truscape@lemmy.blahaj.zone 2 points 2 hours ago* (last edited 1 hour ago) (1 children)

Grayjay allows you to sync between devices if desired, including platforms. I have a desktop, laptop, and a phone, and I can sync everything locally by just pairing the devices and having them at least 2 active for the transfer.

[–] MrScottyTay@sh.itjust.works 1 points 1 hour ago

So another has to be active at the same time as accessing another? I can't always guarantee that so that's a bit too much of a faff around and having to do that multiple times a day would be annoying. I'm glad it's there for those if works for, if it works that way though still think it's not right for me sadly.

[–] rumba@lemmy.zip 6 points 6 hours ago (1 children)

PT, nebula, loops, odysee, and floatplane together can't fill the yt content gap.

You can start with moving to newpipe/grayjay and curate your own content which will lower your surface area.

At some point YT will manage to widevine and we'll be torrenting the best of that shit.

[–] Truscape@lemmy.blahaj.zone 4 points 4 hours ago

Ahead of the curve, already downloaded all of the greats onto my selfhosted storage.

[–] redditStinksSuperBad@lemmy.world 13 points 9 hours ago (1 children)
[–] Gsus4@mander.xyz 6 points 8 hours ago* (last edited 8 hours ago)

a corpse of a platform

[–] HAL_9_TRILLION@lemmy.world 16 points 10 hours ago (5 children)

I now browse Wikipedia. Please don’t screw me over Wikipedia, I donated five bucks to one of your nags once.

For anyone who doesn't know, you can download Wikipedia and host it yourself! I got the top 50k version (~7G) on my RPI3 and now no matter what fuckery they pull or the government pulls, I've got a pretty decent source of general information.

[–] jjlinux@lemmy.zip 5 points 9 hours ago* (last edited 9 hours ago) (1 children)

Could you share what you did to achieve this? I've been planning on doing just that, and have it auto-update every week or so (keeping the previous versions archived, of course) by using kiwix-serve for a static '.zim' file and maybe a cron job for the auto-update. But if you have a better solution, I'd love to know. The deployment I am planning is kind of convoluted to be honest.

[–] HAL_9_TRILLION@lemmy.world 8 points 8 hours ago* (last edited 8 hours ago) (1 children)

No, that's exactly what I did, I'm running kiwix-serve, but I'm not going to bother updating it because I'm really worried about information degrading now that fascists are basically calling the shots on everything (and WP's jackboot co-founder has a hard on for it). If I feel enough time has gone by to warrant an update I'll just do it manually.

[–] Truscape@lemmy.blahaj.zone 1 points 4 hours ago (1 children)

What year was your cutoff?

[–] HAL_9_TRILLION@lemmy.world 1 points 4 hours ago

About 6 months ago.

[–] MrOtingocni@lemmy.world 2 points 7 hours ago (1 children)
[–] HAL_9_TRILLION@lemmy.world 1 points 4 hours ago

You run a server on a machine inside your house, it can be any computer on your local LAN/wifi, but it's obviously best if it's a machine that's always on. I use a Raspberry Pi 3B+ (these can be had for about $50) that I have plugged into my wifi router and it's running a little program called Kiwix-Server (free and open source). You download the WP file (it's a huge single file with a .zim extension) and point the server to it and boom.

load more comments (3 replies)
[–] grue@lemmy.world 201 points 15 hours ago (3 children)

The entire fucking point of the Web was to make information as easily-accessible as possible, structured and semantically tagged, and consumable by humans and further machine transformation alike. "Scraping" is facilitated by design!

Using Javascript to deliberately break that is evil and every programmer who participates it is a piece of shit. No exceptions.

[–] PurpleFanatic@quokk.au 13 points 10 hours ago* (last edited 10 hours ago) (1 children)

Its amazing to me how consistently the shitty behaviours of these billionaire techbro oligarchs impact disabled or marginalised people... even when the point isnt to directly shit on them. Its fucking vile.

I honestly think many (too many, but certainly not all! I am one) programmers are some of the immoral, ethically spurious people around in the 21st century.

[–] voytrekk@sopuli.xyz 6 points 10 hours ago

That is why they want AI to replace programmers. AI morals are programmed, so they can be designed to do shitty things that a normal person would refuse.

[–] artyom@piefed.social 47 points 14 hours ago* (last edited 14 hours ago) (1 children)

That's true but they probably didn't account for AI data scrapers ramfucking your server so they could steal all the value you assembled for general consumption and serve it themselves for profit.

[–] grue@lemmy.world 51 points 14 hours ago (8 children)

The scraping wouldn't be a problem if Reddit simply provided an RSS feed or other data-efficient API. The "ramfucking" is caused by the attempt to block bots; it is entirely self-inflicted.

Remember, it's all our content to begin with and Reddit does not have any right to try to lock it up for itself.


That doesn't mean I like all the AI bullshit going on, BTW. But the problem is the generation of the slop, not the data accessibility.

[–] missingno@fedia.io 30 points 12 hours ago

Well they had an API, but...

[–] artyom@piefed.social 33 points 13 hours ago (1 children)

The scraping wouldn't be a problem if Reddit simply provided an RSS feed or other data-efficient API

That's simply not true. These bots are essentially DDOSing the entire internet, API or not.

[–] grue@lemmy.world 16 points 12 hours ago

Okay, if efficient APIs existed and they weren't incompetently failing to use them, it wouldn't be a problem. Happy now?

(I should've addressed that in my previous comment, as I was aware of how one of the Lemmy instances was taken down by scrapers the other day despite the fact that they could easily get all the content simply by consuming ActivityPub directly. But I was naively hoping it wouldn't be necessary because, as you can see from this text, it would've cluttered up my writing with double the words.)

load more comments (6 replies)
load more comments (1 replies)
[–] turdburglar@piefed.social 10 points 9 hours ago (1 children)

funny, that comes just on the heels of my deciding that reddit is unsafe.

[–] PattyMcB@lemmy.world 2 points 6 hours ago (1 children)

You're just now realizing it?

[–] turdburglar@piefed.social 2 points 3 hours ago

not really but it make for good joke pacing.

i’ve been here for a while. and not there for a while.

[–] tal@lemmy.today 2 points 6 hours ago* (last edited 6 hours ago)

Whether or not it shows it requires a login for old.reddit.com depends on IP. I see this on a few networks, not on others.

[–] Steve@startrek.website 7 points 10 hours ago (1 children)
load more comments (1 replies)
load more comments
view more: next ›