home.social

#searchengines — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #searchengines, aggregated by home.social.

  1. FYI: Google loses fight to strip 90%+ of unique queries from EU rivals' data: Rival engines and AI chatbots get daily records, NUTS 3 location and 5-year access at cost-based prices. What does the full text show about Alphabet's campaign? ppc.land/google-loses-fight-to #Google #EURegulations #DataPrivacy #AIChatbots #SearchEngines

  2. FYI: Google loses fight to strip 90%+ of unique queries from EU rivals' data: Rival engines and AI chatbots get daily records, NUTS 3 location and 5-year access at cost-based prices. What does the full text show about Alphabet's campaign? ppc.land/google-loses-fight-to #Google #EURegulations #DataPrivacy #AIChatbots #SearchEngines

  3. FYI: Google loses fight to strip 90%+ of unique queries from EU rivals' data: Rival engines and AI chatbots get daily records, NUTS 3 location and 5-year access at cost-based prices. What does the full text show about Alphabet's campaign? ppc.land/google-loses-fight-to #Google #EURegulations #DataPrivacy #AIChatbots #SearchEngines

  4. FYI: Searches with words used by under 50 people won't reach Google rivals: Rival engines and AI chatbots only get records where 1,000 signed-in users share region, device and language. First data deliveries could come in early 2027. ppc.land/searches-with-words-u #Google #AI #searchengines #digitalmarketing #SEO

  5. FYI: Searches with words used by under 50 people won't reach Google rivals: Rival engines and AI chatbots only get records where 1,000 signed-in users share region, device and language. First data deliveries could come in early 2027. ppc.land/searches-with-words-u #Google #AI #searchengines #digitalmarketing #SEO

  6. FYI: Searches with words used by under 50 people won't reach Google rivals: Rival engines and AI chatbots only get records where 1,000 signed-in users share region, device and language. First data deliveries could come in early 2027. ppc.land/searches-with-words-u #Google #AI #searchengines #digitalmarketing #SEO

  7. FYI: Searches with words used by under 50 people won't reach Google rivals: Rival engines and AI chatbots only get records where 1,000 signed-in users share region, device and language. First data deliveries could come in early 2027. ppc.land/searches-with-words-u #Google #AI #searchengines #digitalmarketing #SEO

  8. FYI: Searches with words used by under 50 people won't reach Google rivals: Rival engines and AI chatbots only get records where 1,000 signed-in users share region, device and language. First data deliveries could come in early 2027. ppc.land/searches-with-words-u #Google #AI #searchengines #digitalmarketing #SEO

  9. Glenn M. brought Archive of Archives to my attention. (Thank you sir! Good to hear from you.) Based on the GitHub I would guess it launched around March. From the home page: “Archive of Archives is a curated collection of digital archives developed by SHAPE. All are welcome to contribute to the project. Archive submission guidelines can be found on GitHub.”

    https://rbfirehose.com/2026/08/17/archive-of-archives/
  10. I posted this about a year ago. Might be a good time to repost it.

    "Search With Stateful Chat" patent (Cf. patents.google.com/patent/US20 ) - appears to describe the Gemini app for smartphones.

    "Method for Text Ranking with Pairwise Ranking Prompting" (Cf. patents.google.com/patent/US20 ) - documents an experimental process described in this research paper titled "Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting" (Cf. arxiv.org/pdf/2306.17563 ). There is no indication this was introduced into a live agentic system like Gemini.

    "User Embedding Models for Personalization of Sequence Processing Models" (Cf. patents.google.com/patent/WO20 ) - documents an experimental process for improving recommender (sub-)systems (like movie searches) that incorporate large language models. The process is described in this research paper titled "User Embedding Model for Personalized Language Prompting" (Cf. arxiv.org/pdf/2401.04858 ).

    "Systems and methods for prompt-based query generation for diverse retrieval" (Cf. patents.google.com/patent/WO20 ) - updates a 2022 patent for a process named PROMPTAGATOR that generates queries more efficiently based on a small number of examples, as described in this research paper titled "Promptagator - Few-shot Dense Retrieval from 8 Examples" (Cf. arxiv.org/pdf/2209.11755 ). This could be used to generate query fan-outs (but query fan-out has been used in multiple systems at least since the 1990s, so there are many implementations).

    "Instruction Fine-Tuning Machine-Learned Models Using Intermediate Reasoning Steps" (Cf. patents.google.com/patent/US20 ) - documents an older method for fine-tuning instructions submitted to LLMs, as described in this 2022 research paper titled "Scaling Instruction-Finetuned Language Models" (Cf. jmlr.org/papers/volume25/23-08 ). The work has been superseded by this paper titled "Mixture-of-Experts Meets Instruction Tuning: A Winning Combination for Large Language Models" (Cf. arxiv.org/pdf/2305.14705 ).

    This is the AI Overviews patent, titled "Generative summaries for search results" (Cf. patents.google.com/patent/US11 )

    #google #aioverviews #aimode #machinelearning #search #searchengines #generativesearch #seo #searchengineoptimization #webmarketing #digitalmarketing #ai #patents

    seo-theory.com/how-to-read-pat

  11. Spotted in my RSS feeds: GovAuctions. From the About page: “GovAuctions lets you search every government surplus auction at once. You can filter by location, category, and price, save items to a watchlist, and get alerts when new auctions match what you’re looking for.”

    https://rbfirehose.com/2026/05/05/govauctions-auction-metasearch/
  12. #question of the day (#qotd?):

    Most #SearchEngines have paid #APIs one can use to retrieve results but there are also #metasearch engines which seem to pull these results without paying for them (e.g., on the small scale, independent #seeks instances). I'm curious about the #legal issues surrounding metasearch. Who pays for the #API? And who uses other methods and why?

    #search