home.social

#searchengines — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #searchengines, aggregated by home.social.

  1. I'm only trying to influence you to read this thing I wrote. I spent nearly fifteen minutes on it! So, c'mon!

    Check out a random #blogpost called "Worth It"

    8clicks.8r4d.com/2026/09/worth

    #Blog #SocialMedia #BloggingForFun #SearchEngines

  2. Ars Technica: Remembering the pre-Google web, when search was an experiment. “In the mid-’90s, the web was exploding, but finding anything of actual value on it felt like an elaborate negotiation with whatever proto-search engine happened to be standing closest to the door. Unlike now, when Google is widely seen as both portal and gatekeeper, sites like AltaVista, Lycos, Excite, HotBot, and […]

    https://rbfirehose.com/2026/08/08/ars-technica-remembering-the-pre-google-web-when-search-was-an-experiment/
  3. With the way many SEO agencies appear to be misrepresenting what AI Overviews are, how they are produced, it seems like it's only a matter of time before they convince some client to do something really stupid. AI Overviews start with what is in the search results.

    They are summaries of the information the search engine has indexed. Yes, sometimes the models hallucinate and make up stuff. But most hallucinations come from conflating opposing or contradictory information and opinions found in the search results.

    Parody and satire in the sources have led to many a fascinating and sometimes humorous AI generation. One can argue that it's on the search engines to figure out which sources to trust. But they DO make that attempt (I'm not saying it's enough).

    And that leads to yet more misinformation coming out of the SEO world. Simply flooding the index with listicles isn't enough to ensure that an AI summary will say what you want it to say.

    What we've been telling our premium newsletter subscribers for a long time now is that they need mentions on highly reputable sites (and I'm NOT talking about "domain authority" and all that nonsense). But they also have to publish reliable, authoritative information (not simply rehashes of whatever their freelancers or AI tools find on the Web).

    We now see a lot of SEO agencies offering that same advice. And that's a good thing. But the disconnect between these two positions is making it hard for people to understand what is happening in the search results.

    The AI summaries ALWAYS start with what is in the search results. You don't have control over what they will say. And every search produces a relatively unique experience for the user based on their search context. You can't control that. You can't influence it.

    At best, you can target a range of user search contexts to which you want your content to be relevant.

    #seo #ai #bing #google #search #searchengines #searchengineoptimization #content #aioverviews #aisummaries #marketing

  4. I posted this about a year ago. Might be a good time to repost it.

    "Search With Stateful Chat" patent (Cf. patents.google.com/patent/US20 ) - appears to describe the Gemini app for smartphones.

    "Method for Text Ranking with Pairwise Ranking Prompting" (Cf. patents.google.com/patent/US20 ) - documents an experimental process described in this research paper titled "Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting" (Cf. arxiv.org/pdf/2306.17563 ). There is no indication this was introduced into a live agentic system like Gemini.

    "User Embedding Models for Personalization of Sequence Processing Models" (Cf. patents.google.com/patent/WO20 ) - documents an experimental process for improving recommender (sub-)systems (like movie searches) that incorporate large language models. The process is described in this research paper titled "User Embedding Model for Personalized Language Prompting" (Cf. arxiv.org/pdf/2401.04858 ).

    "Systems and methods for prompt-based query generation for diverse retrieval" (Cf. patents.google.com/patent/WO20 ) - updates a 2022 patent for a process named PROMPTAGATOR that generates queries more efficiently based on a small number of examples, as described in this research paper titled "Promptagator - Few-shot Dense Retrieval from 8 Examples" (Cf. arxiv.org/pdf/2209.11755 ). This could be used to generate query fan-outs (but query fan-out has been used in multiple systems at least since the 1990s, so there are many implementations).

    "Instruction Fine-Tuning Machine-Learned Models Using Intermediate Reasoning Steps" (Cf. patents.google.com/patent/US20 ) - documents an older method for fine-tuning instructions submitted to LLMs, as described in this 2022 research paper titled "Scaling Instruction-Finetuned Language Models" (Cf. jmlr.org/papers/volume25/23-08 ). The work has been superseded by this paper titled "Mixture-of-Experts Meets Instruction Tuning: A Winning Combination for Large Language Models" (Cf. arxiv.org/pdf/2305.14705 ).

    This is the AI Overviews patent, titled "Generative summaries for search results" (Cf. patents.google.com/patent/US11 )

    #google #aioverviews #aimode #machinelearning #search #searchengines #generativesearch #seo #searchengineoptimization #webmarketing #digitalmarketing #ai #patents

    seo-theory.com/how-to-read-pat

  5. #поисковики #SearchEngines #FAIL #music #EBM #DarkElectro #Industrial

    Ищу пестню по фразе¹. «when i was walking down along the street i found one road where we shall meet". Да, не расслышал, должно быть "should". Но почему нашел только #Яша, причем только в виде «быстрого ответа» Алисы и одной (!) ссылки на #VK?

    Остальные — даже по полной корректной (!) фразе выдают максимум сслыки на «помогите с английским»? Что за нахуй, пардон майн френч?

    P.S. Нуичо что это Dance My Darling из этой страны — типа санкции и знать необязательно?

    В качестве пруфа прикладываю трек в версии ремикса от Holocoder, #LastFM считает, что его слушает 1 заблудший (?) человек, но на рутрахере раздача на 1,5 гига.

    ¹ Потому что чебурашки #Hoco играют с карточки, с приложением не коннектятся, так что тупо «режим радио» (что меня вполне устраивает), но иногда хочется найти, «а что это такое я лайкнул».

  6. How to add Google News, Google Books, and Google Scholar search directly in your browser, to search those sites with fewer steps, and less chance of being distracted.

    One of my hobbies is contributing to Wikipedia, and more so, creating new Wikipedia articles. Citations of reliable sources are key to both good contributions, and new articles, especially in making them stick (not get reverted/deleted).

    I use Google News to search for citations, and depending on the topic, sometimes Google Books and Google Scholar.

    Unfortunately, to search Google News, you have to first go to Google News (news.google .com - deliberately unlinked here), upon which you are immediately shown distracting (if not dire) news photos, headlines etc. which present a non-trivial challenge to staying focused.

    If only there was a way to directly search Google News from your browser search box / address bar (like you can search Wikipedia from your browser).

    Unfortunately, there is no site-specific search option for Google News (like there is for YouTube, e.g. see https://support.mozilla.org/en-US/kb/add-or-remove-search-engine-firefox#w_add-search-engines) that you can select an add to your browser in one click (enabled by OpenSearch support, link in footer).

    This is another technology I use for category 2 (defending focus) that I mentioned in my previous post on focus.

    Steps to add a Google News search option to Firefox:

    1. open Firefox Preferences ("Firefox" menu, "Preferences" item)
    2. select "🔍 Search" from the left column
    3. scroll down to "Search Shortcuts"
    4. click the "Add" button under the list of search engines which opens a dialog
    5. enter "Google News" into the Search engine name field (without quotes)
    6. enter "https://news.google.com/search?q=%s" into the URL field (without quotes)
    7. enter "gn" into the Keyword field (without quotes)

    The dialog should look like this:


    8. click (Add Engine)

    Now you can go to your address bar, type in "gn " (without the quotes), your news search term or phrase, e.g. "Broken Arrow Skyrace", and press return to directly see search results.

    Similarly for Google Books, follow the same steps except:
    5. enter "Google Books" into the Search engine name field (without quotes)
    6. enter "https://books.google.com/books?q=%s" into the URL field (without quotes)
    7. enter "gb" into the Keyword field (without quotes)

    And similarly for Google Scholar:
    5. enter "Google Scholar" into the Search engine name field (without quotes)
    6. enter "https://scholar.google.com/scholar?q=%s" into the URL field (without quotes)
    7. enter "gs" into the Keyword field (without quotes)

    Thanks to folks in the #indieweb informal chat who reminded me (when I complained about the distracting Google News home page) of this way to add new site-specific search capabilities to the browser even for sites (or subsites) without explicit OpenSearch support.

    Looking forward to using these capabilities to more quickly find citations for updating and creating new Wikipedia articles.

    Previously:
    * https://tantek.com/2026/158/t2/three-insights-improving-focus
    * https://tantek.com/2024/287/t2/setup-search-shortcuts-firefox

    OpenSearch FYI:
    * https://developer.mozilla.org/en-US/docs/Web/XML/Guides/OpenSearch

    #focus #Wikipedia #Firefox #GoogleNews #GoogleBooks #GoogleScholar #search #siteSearch #AddSearchEngine #searchEngine #searchEngines #OpenSearch #webSearch #SearchShortcuts #browserTip #FirefoxTip #searchTip

  7. Search Engine Journal: US Publishers Demand Common Crawl Stop Scraping Their Content. “Digital Content Next, a trade body representing US digital publishers, has sent a cease and desist letter to the Common Crawl Foundation. The letter demands Common Crawl stop collecting publisher content and remove material already in its datasets.”

    https://rbfirehose.com/2026/06/11/search-engine-journal-us-publishers-demand-common-crawl-stop-scraping-their-content/
  8. From answer engines to learning engines — Why fast answers are like fast food

    People crave fast answers. But the purpose of information systems is to help people gain knowledge. So we should seek better questions.

    duncanstephen.net/from-answer-

  9. Analytics India: Perplexity Announces Search API . “AI startup Perplexity, on September 25, launched the ‘Perplexity Search API’, providing developers access to the infrastructure that enables Perplexity’s services and an index that covers ‘hundreds of billions’ of webpages, the company announced.”

    https://rbfirehose.com/2025/09/27/analytics-india-perplexity-announces-search-api/

  10. Search Engine Land: Google still leads, but Gen Z and AI are reshaping search behavior: Survey. “Searching is no longer synonymous with ‘Googling.’ While traditional search engines still dominate for information retrieval, Americans are also turning to social media, AI tools, and ecommerce platforms, depending on what they’re searching for and who they are.”

    https://rbfirehose.com/2025/08/05/google-still-leads-but-gen-z-and-ai-are-reshaping-search-behavior-survey-search-engine-land/

  11. Focus on #TargetGroups 🎯 Every successful #onlinemarketingstrategy begins with analyzing the target groups in order to optimally align #content, #channels and approach to them.

    #Measurability as an Advantage 📊 #KPIs and #analysis tools allow the success of each measure to be precisely evaluated and continuously optimized across funnels.

    Choosing the Right #Channels 🌐 Whether #searchengines, #socialmedia or alternative #onlinemarketing.

    👉 OnlineMarketingStrategy.EU
    👉 @marketing #news & #strategy

  12. Mashable: Perplexity adds a Max tier just as expensive as its rivals . “Perplexity Max costs $200 per month or $2,000 a year; the tier down, Perplexity Pro, costs $20 per month or $200 per year. It’s the same monthly price as the top tier of OpenAI’s offering, ChatGPT Pro; in comparison, the most expensive tier of Google’s AI, Gemini, is Google AI Ultra, which costs $249.99 per month.”

    https://rbfirehose.com/2025/07/09/mashable-perplexity-adds-a-max-tier-just-as-expensive-as-its-rivals/

  13. 🚀 Breaking news: Search engine crawlers are taking longer coffee breaks than your average office worker! 🐢 Apparently, shrinking memory means faster crawling... except when it doesn't, because the last 0.1% is on vacation for a week! 📅 Clearly, Pareto forgot to distribute common sense! 😂
    marginalia.nu/log/a_117_crawl_ #SearchEngines #CoffeeBreaks #TechHumor #CrawlingChallenges #ParetoPrinciple #MemoryManagement #HackerNews #ngated

  14. As #socialmedia and #searchengines become more and more unreliable at providing #news and #web #content , don't #forget you can curate your #internet experience through #rss and #rssfeedreader .

    If you are not familiar with #rssfeeds check out The RSS Review to learn more!

    the-rss-review.surge.sh/

    Note: This is just a hobby project of mine, but hope it helps.

    #therssreview

  15. Got a minute? #SearchClub would love to get your feedback on some potential "easy starter projects" to help reclaim search for the searcher - and any other ideas you'd like to share.

    Check out the ChaosPad, and let us know what you think: pads.ccc.de/rNBu09Mr2M. Perhaps you'd like to get involved? Even better! :blobfoxhyper2:

    #38C3 #CCC #Search #Discovery #SearchEngines #LLM #LLMs #GenAI #MushroomForaging #PizzaGlue #Mwmbl #SearXNG #Blocklists #uBlock #uBlockOrigin #uBlacklist

  16. Got a minute? #SearchClub would love to get your feedback on some potential "easy starter projects" to help reclaim search for the searcher - and any other ideas you'd like to share.

    Check out the ChaosPad, and let us know what you think: pads.ccc.de/rNBu09Mr2M. Perhaps you'd like to get involved? Even better! :blobfoxhyper2:

    #38C3 #CCC #Search #Discovery #SearchEngines #LLM #LLMs #GenAI #MushroomForaging #PizzaGlue #Mwmbl #SearXNG #Blocklists #uBlock #uBlockOrigin #uBlacklist

  17. Got a minute? #SearchClub would love to get your feedback on some potential "easy starter projects" to help reclaim search for the searcher - and any other ideas you'd like to share.

    Check out the ChaosPad, and let us know what you think: pads.ccc.de/rNBu09Mr2M. Perhaps you'd like to get involved? Even better! :blobfoxhyper2:

    #38C3 #CCC #Search #Discovery #SearchEngines #LLM #LLMs #GenAI #MushroomForaging #PizzaGlue #Mwmbl #SearXNG #Blocklists #uBlock #uBlockOrigin #uBlacklist

  18. Got a minute? #SearchClub would love to get your feedback on some potential "easy starter projects" to help reclaim search for the searcher - and any other ideas you'd like to share.

    Check out the ChaosPad, and let us know what you think: pads.ccc.de/rNBu09Mr2M. Perhaps you'd like to get involved? Even better! :blobfoxhyper2:

    #38C3 #CCC #Search #Discovery #SearchEngines #LLM #LLMs #GenAI #MushroomForaging #PizzaGlue #Mwmbl #SearXNG #Blocklists #uBlock #uBlockOrigin #uBlacklist

  19. Got a minute? #SearchClub would love to get your feedback on some potential "easy starter projects" to help reclaim search for the searcher - and any other ideas you'd like to share.

    Check out the ChaosPad, and let us know what you think: pads.ccc.de/rNBu09Mr2M. Perhaps you'd like to get involved? Even better! :blobfoxhyper2:

    #38C3 #CCC #Search #Discovery #SearchEngines #LLM #LLMs #GenAI #MushroomForaging #PizzaGlue #Mwmbl #SearXNG #Blocklists #uBlock #uBlockOrigin #uBlacklist

  20. Judge rules against users suing Google and Apple over “annoying” search results - Enlarge (credit: SOPA Images / Contributor | LightRocket)

    Whil... - arstechnica.com/?p=2001666 #onlineadvertising #searchengines #antitrustlaw #onlinesearch #sundarpichai #defaultdeal #shermanact #antitrust #timcook #policy #google #safari #apple

  21. How can advocacy for a search engine monopoly be reduced (by practical editorial changes) in the English-language #Wikipedia?

    This is a fundamental meta-academic question (reviews of knowledge) for which practical participation - with evidence and arguments - would be better than postmodernist gobbledegook or neoliberal empty rhetoric.

    @academicchatter

    en.wikipedia.org/wiki/Wikipedi

    #Monopolies #SearchEngines #SelectionBias #OpenScience

  22. Finally finished modifying the way OddNugget.com handles its index to cram millions of pages into (very little) active memory!

    Also showing the actual match snippet now and only matching EXACTLY what was searched for instead of getting fancy with it. More to come!

    I've gone from -_- to :) and 0.0 in short order.

    #SearchEngines #searchengine #search #gopher #gopherspace #internet #protocol