home.social

#crawlers — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #crawlers, aggregated by home.social.

fetched live
  1. Deadly spiders invade an apartment building in the Crawlers trailer. Watch it here bit.ly/4qJZdV7

    #Crawlers #film

  2. ICYMI: US sends 53.5% of global bot traffic, Decodo analysis finds: Iran runs 81.4% bots at home while retail absorbs 13% of automated requests, a split that now decides which crawlers reach product pages and which get shut out. ppc.land/us-sends-53-5-of-glob #BotTraffic #DigitalMarketing #Crawlers #SEO #Automation

  3. September 15, 2026, #Cloudflare will set updated defaults for new domains: #bots classified as #Training or #Agent will be #blocked on pages that display ads, and Search will remain allowed. #Crawlers that combine Search&Training will also be blocked. #AI

    developers.cloudflare.com/bots

  4. “AI webpage #crawlers are now being served their very own #ads that ordinary visitors never see, with one firm's boss openly describing a strategy to influence what #chatbots say about #brands. The next time you ask Claude about where to bank, its answer may have been influenced by this #BotTargeted content.
    We only have one documented example so far, and that's Time serving #AIOnly ads to selected AI crawlers, as spotted by Germany based freelance software developer #VincentSchmalbach.

    In a Wednesday blog post, Schmalbach detailed how he found #sponsored content embedded in #markdown versions of some Time pages that are served to AI crawlers but not ordinary browsers. Those pages also contained #advertisings tags from #AdTech vendor Mobian ahead of extensive #FAQs for online-only bank Ally. The FAQs include brand "facts," such as the number of fee-free ATMs associated with the branchless bank, alongside claims that Ally is "the only bank built for life today," putting it in "a category of one."

    AI poisoning with advertising.

    #AI / #advert <theregister.com/ai-and-ml/2026>

  5. EFF Joins 18 Civil Rights Organizations Calling on Governor Hochul to Reject the Stealth Crawler Prohibition Act www.eff.org/deeplinks/2026… #journalism #crawlers #privacy

  6. The effect of blocking AI crawlers on Blender's infrastructure. This is the CPU usage graph.

    So many AI bots are crawling Blender's infra all the time. Just do a `git clone` and investigate local files. It's way faster (also for the AI users themselves) and doesn't block actual Blender development (it got that bad).

  7. RE: social.edu.nl/@wlaatje/1169708

    "The web is full of independent archives, hobby databases, local news sites, forums, reference works. Decades of accumulated human effort, running on old code, maintained by small teams or single individuals, quietly holding up far more of our shared knowledge than anyone acknowledges."

    And proponents of so-called 'AI' are speedrunning their destruction.

    #noAI #scrapers #scraping #crawlers #AI #genAI

  8. Thanks to content scraper from companies and protection solutions from companies like for delivering 56 kbps era response times in the age of fibre internet connections! How nice of them to constantly satisfy our nostalgia!

  9. ICYMI: AI crawlers hit sites 50,000 times per human visit, Cloudflare data shows: AI crawlers reach up to 50,000 visits per referred reader, Cloudflare data shows, as a Munich court strips Google of its liability shield over AI Overviews. ppc.land/ai-crawlers-hit-sites #AI #Crawlers #DataAnalysis #Cloudflare #Google

  10. #Cloudflare gives #AI #crawlers a September deadline: pay publishers or get blocked - thenextweb.com/news/cloudflare "From 15 September, Cloudflare will block crawlers that harvest content for AI training from any page carrying ads, unless the owner opts in, and pay publishers when their work shapes an AI answer. "

  11. ICYMI: Microsoft Clarity now flags robots.txt violations inside Bot Analytics: Microsoft Clarity now surfaces robots.txt violations in Bot Analytics, showing publishers which AI crawlers break access rules and what content they target. ppc.land/microsoft-clarity-now #MicrosoftClarity #BotAnalytics #SEO #WebAnalytics #Crawlers

  12. ICYMI: New York passes bill forcing AI crawlers to identify themselves to news sites: New York's Assembly passed A11292 on June 5, 2026, requiring AI crawlers to disclose identity and purpose to news publishers or face $15,000-per-day penalties. ppc.land/new-york-passes-bill- #NewYork #AI #Crawlers #News #Legislation