home.social

#crawlers — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #crawlers, aggregated by home.social.

  1. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  2. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  3. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  4. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  5. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  6. Thanks to content scraper from companies and protection solutions from companies like for delivering 56 kbps era response times in the age of fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  7. Thanks to content scraper #crawlers from #AI companies and #bot protection solutions from companies like #Cloudflare for delivering 56 kbps #dialup era #website response times in the age of #gigabit fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  8. Thanks to content scraper #crawlers from #AI companies and #bot protection solutions from companies like #Cloudflare for delivering 56 kbps #dialup era #website response times in the age of #gigabit fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  9. Thanks to content scraper #crawlers from #AI companies and #bot protection solutions from companies like #Cloudflare for delivering 56 kbps #dialup era #website response times in the age of #gigabit fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  10. Thanks to content scraper #crawlers from #AI companies and #bot protection solutions from companies like #Cloudflare for delivering 56 kbps #dialup era #website response times in the age of #gigabit fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  11. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  12. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  13. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  14. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  15. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  16. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  17. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  18. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  19. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  20. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  21. ICYMI: Microsoft Clarity now flags robots.txt violations inside Bot Analytics: Microsoft Clarity now surfaces robots.txt violations in Bot Analytics, showing publishers which AI crawlers break access rules and what content they target. ppc.land/microsoft-clarity-now #MicrosoftClarity #BotAnalytics #SEO #WebAnalytics #Crawlers

  22. ICYMI: Microsoft Clarity now flags robots.txt violations inside Bot Analytics: Microsoft Clarity now surfaces robots.txt violations in Bot Analytics, showing publishers which AI crawlers break access rules and what content they target. ppc.land/microsoft-clarity-now #MicrosoftClarity #BotAnalytics #SEO #WebAnalytics #Crawlers

  23. ICYMI: Microsoft Clarity now flags robots.txt violations inside Bot Analytics: Microsoft Clarity now surfaces robots.txt violations in Bot Analytics, showing publishers which AI crawlers break access rules and what content they target. ppc.land/microsoft-clarity-now #MicrosoftClarity #BotAnalytics #SEO #WebAnalytics #Crawlers

  24. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation

  25. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation

  26. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation

  27. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation

  28. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation