this post was submitted on 21 Mar 2025
1447 points (99.5% liked)

Technology

67536 readers
4504 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] lily33@lemm.ee 12 points 1 week ago (3 children)

while allowing legitimate users and verified crawlers to browse normally.

What is a "verified crawler" though? What I worry about is, is it only big companies like Google that are allowed to have them now?

[–] wingiee@lemm.ee 20 points 1 week ago (1 children)

I assume a crawler which adheres to robots.txt

[–] lily33@lemm.ee 5 points 6 days ago (1 children)

I would love to think so. But the word "verified" suggests more.

[–] killeronthecorner@lemmy.world 2 points 6 days ago

IP verification is a not uncommon method for commercial crawlers

Cloudflare isn't the best at blocking things. As long as your crawler isn't horribly misconfigured you shouldn't have much issues.