literature.cafe
  • Communities
  • Create Post
  • Create Community
  • heart
    Support Lemmy
  • search
    Search
  • Login
  • Sign Up
irelephant [he/him]🍭@lemm.ee to TechTakes@awful.systemsEnglish · 2 days ago

Ai scraping is an effective DDoS on the entire interent

pod.geraspora.de

external-link
message-square
24
link
fedilink
92
external-link

Ai scraping is an effective DDoS on the entire interent

pod.geraspora.de

irelephant [he/him]🍭@lemm.ee to TechTakes@awful.systemsEnglish · 2 days ago
message-square
24
link
fedilink
Excerpt from a message I just posted in a #diaspora team internal f...
pod.geraspora.de
external-link
Excerpt from a message I just posted in a #diaspora team internal forum category. The context here is that I recently get pinged by slowness/load spikes on the diaspora* project web infrastructure (Discourse, Wiki, the project website, ...), and looking at the traffic logs makes me impressively angry. In the last 60 days, the diaspora* web assets received 11.3 million requests. That equals to 2.19 req/s - which honestly isn't that much. I mean, it's more than your average personal blog, but nothing that my infrastructure shouldn't be able to handle. However, here's what's grinding my fucking gears. Looking at the top user agent statistics, there are the leaders: 2.78 million requests - or 24.6% of all traffic - is coming from Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.2; +https://openai.com/gptbot). 1.69 million reuqests - 14.9% - Mozilla/5.0 (Macintosh; Intel Mac OS X 10_10_1) AppleWebKit/600.2.5 (KHTML, like Gecko) Version/8.0.2 Safari/600.2.5 (Amazonb...
  • BlueMonday1984@awful.systems
    link
    fedilink
    English
    arrow-up
    7
    ·
    2 days ago

    That opens you up to getting accused of click fraud, as AdNauseam found out the hard way but its worth it if you can squeeze some cash out of them before that happens.

    • Sas [she/her]@beehaw.org
      link
      fedilink
      English
      arrow-up
      6
      ·
      2 days ago

      I mean, scraping bots would obviously obey robots.txt so those scraping - bots, i mean users can’t be bots

TechTakes@awful.systems

techtakes@awful.systems

Subscribe from Remote Instance

Create a post
You are not logged in. However you can subscribe from another Fediverse account, for example Lemmy or Mastodon. To do this, paste the following into the search field of your instance: !techtakes@awful.systems

Big brain tech dude got yet another clueless take over at HackerNews etc? Here’s the place to vent. Orange site, VC foolishness, all welcome.

This is not debate club. Unless it’s amusing debate.

For actually-good tech, you want our NotAwfulTech community

Visibility: Public
globe

This community can be federated to other instances and be posted/commented in by their users.

  • 176 users / day
  • 1.13K users / week
  • 2.37K users / month
  • 5.04K users / 6 months
  • 5 local subscribers
  • 1.86K subscribers
  • 866 Posts
  • 24.2K Comments
  • Modlog
  • mods:
  • David Gerard@awful.systems
  • BE: 0.19.11
  • Modlog
  • Legal
  • Instances
  • Docs
  • Code
  • join-lemmy.org