home.social

#firecrawl — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #firecrawl, aggregated by home.social.

fetched live
  1. أصبح مكون Firecrawl الإضافي متاحًا الآن لمنصة Codex، مما يوفر قدرات متقدمة في البحث على الويب واستخراج البيانات من المواقع والتفاعل معها. يتميز هذا المكون بتحقيق دقة عالية تصل إلى 94.7 بالمائة عند اختباره على معيار SimpleQA. يمكن للمستخدمين المهتمين تثبيت Firecrawl بسهولة ومباشرة من سوق مكونات OpenAI الإضافية، للاستفادة من ميزاته المحسنة في معالجة المعلومات والوصول إليها.

    #Firecrawl #Codex #OpenAI

  2. J’ai publié la suite de mon setup opencode + Ollama sur RTX 3090.

    Cette fois, je remplace SearXNG par Firecrawl pour donner à opencode un accès web plus exploitable : recherche, scrape, markdown propre, MCP, self-hosted Docker, limites et points de sécurité.

    Le but n’est pas de “scraper Internet”, mais de fournir au modèle local un contexte plus propre quand il doit lire de la documentation vivante.

    cryptolab.re/posts/2026/openco

    #opencode #Firecrawl #Ollama #LLM #SelfHosting #DevOps

  3. Как мы автоматизировали мониторинг цен конкурентов: мультиагентная система на CrewAI + n8n + Firecrawl

    0. TL;DR для тех, кто спешит Статья о том, как собрать из подручных open-source инструментов систему, которая ежедневно: — Сканирует цены и отзывы у конкурентов — Анализирует их ИИ‑агентами — Присылает готовый отчёт в Telegram Стек: n8n (оркестрация) → Firecrawl (парсинг) → CrewAI (анализ) → Telegram (доставка) 1. Проблема: ручной мониторинг — это боль Представьте: вы продаёте электронику. У вас 15 конкурентов на Ozon, 8 — на Wildberries, плюс 3 собственных сайта. Каждое утро менеджер открывает 26 вкладок, сверяет цены, записывает в Excel. Занимает 45 минут. Человек ошибается, пропускает, уходит в отпуск. Мы решили: пусть роботы следят за роботами (ценами).

    habr.com/ru/articles/1048110/

    #n8n #firecrawl #llm #agents #python #парсинг #мониторинг_цен_конкурентов #ecommerce #aiбезопасность #prompt_injection

  4. 🔍 #Firecrawl /monitor: keep your #AI #agent in sync with the web.
    Point it at a URL, describe changes in plain English, get a webhook or email when something meaningful changes ☝️ #webscraping #devtools

    docs.firecrawl.dev/features/mo

  5. 🔍 #Firecrawl /monitor: keep your #AI #agent in sync with the web.
    Point it at a URL, describe changes in plain English, get a webhook or email when something meaningful changes ☝️ #webscraping #devtools

    docs.firecrawl.dev/features/mo

  6. Как научить Claude Code работать с вебом и не сжигать на этом лимиты

    Попросить LLM-агента типа Claude Code "сходи в интернет и собери мне данные" - это как играть в казино. Иногда везет, и ты получаешь то что искал. А иногда сжигаешь половину дневного лимита на двух сайтах, упираешься в антибот защиту и в итоге получаешь кашу из тегов вперемешку с куском нужного контента. Любой, кто пробовал натравить LLM-агента на сайт, знает это чувство: даешь простую задачу - собери данные с такой-то страницы. Агент бодро рапортует, что работа кипит. Проходит минута, две, он пошел по соседним ссылкам, начал сам что-то искать, что-то быстро перебирает, и в итоге половину сайтов он не смог открыть, половина второй половины - это мусор и только крупица нужной информации. В этой статье я предложу вам один способ, которым пользуюсь сам и который хорошо ( почти всегда ) решает эту проблему.

    habr.com/ru/articles/1020598/

    #claude_code #claude_code_skills #mcp #Firecrawl #вебскрапинг #aiагенты #llm #anthropic #вебпоиск

  7. 🔌 #Firecrawl integration for web scraping and search: standard API key setup for contributors, partner integration option for production deployments with per-user API key management, automatic key provisioning on signup, self-healing system with fallback for invalid keys

  8. 🔌 #Firecrawl integration for web scraping and search: standard API key setup for contributors, partner integration option for production deployments with per-user API key management, automatic key provisioning on signup, self-healing system with fallback for invalid keys

  9. 🔍 #OpenScouts: Create #AI scouts that continuously search the web and notify you when they find what you're looking for #opensource #NextJS #React #TypeScript #Supabase #OpenAI #Firecrawl

    ⚡ Built with cutting-edge tech stack: #NextJS 16 with App Router & Turbopack, #React 19, #TypeScript, #TailwindCSS v4, #Supabase for database, auth & edge functions, #pgvector for vector embeddings and semantic search, #OpenAI API for AI agent & embeddings, and Resend for email notifications

    🧵 👇

  10. 🔍 #OpenScouts: Create #AI scouts that continuously search the web and notify you when they find what you're looking for #opensource #NextJS #React #TypeScript #Supabase #OpenAI #Firecrawl

    ⚡ Built with cutting-edge tech stack: #NextJS 16 with App Router & Turbopack, #React 19, #TypeScript, #TailwindCSS v4, #Supabase for database, auth & edge functions, #pgvector for vector embeddings and semantic search, #OpenAI API for AI agent & embeddings, and Resend for email notifications

    🧵 👇

  11. 🎯 Supported models include #GPT-OSS-120B, #GPT-OSS-20B, #Llama4 Maverick, #Llama4 Scout, #Llama33-70B, #Llama31-8B, #KimiK2, #Qwen3-32B

    🔧 Key features: deterministic inference for faster tool-using agents, cost-effective scaling, approved tool use with clear allowlists, seamless migration capability

    📋 Ready-to-use cookbook tutorials with #BrowserBase #MCP, #BrowserUse #MCP, #Exa #MCP, #Firecrawl #MCP, #HuggingFace #MCP, #Parallel #MCP, #Stripe #MCP, #Tavily #MCP

  12. 🎯 Supported models include #GPT-OSS-120B, #GPT-OSS-20B, #Llama4 Maverick, #Llama4 Scout, #Llama33-70B, #Llama31-8B, #KimiK2, #Qwen3-32B

    🔧 Key features: deterministic inference for faster tool-using agents, cost-effective scaling, approved tool use with clear allowlists, seamless migration capability

    📋 Ready-to-use cookbook tutorials with #BrowserBase #MCP, #BrowserUse #MCP, #Exa #MCP, #Firecrawl #MCP, #HuggingFace #MCP, #Parallel #MCP, #Stripe #MCP, #Tavily #MCP

  13. Making the most out of a small LLM

    Yesterday i finally built my own #AI #server. I had a spare #Nvidia RTX 2070 with 8GB of #VRAM laying around and wanted to do this for a long time.

    The problem is that most #LLMs need a lot of VRAM and i don't want to buy another #GPU just to host my own AI. Then i came across #gemma3 and #qwen3. Both of these are amazing #quantized models with stunning reasoning given that they need so less resources.

    I chose huihui_ai/qwen3-abliterated:14b since it supports #deepthinking, #toolcalling and is pretty unrestricted. After some testing i noticed that the 8b model performs even better than the 14b variant with drastically better performance. I can't make out any quality loss there to be honest. The 14b model sneaked in chinese characters into the response very often. The 8b model on the other hand doesn't.

    Now i've got a very fast model with amazing reasoning (even in German) and tool calling support. The only thing left to improve is knowledge. #Firecrawl is a great tool for #webscraping and as soon as i implemented websearching, the setup was complete. At least i thought it was.

    I want to make the most out of this LLM and therefore my next step is to implement a basic #webserver that exposes the same #API #endpoints as #ollama so that everywhere ollama is supported, i can point it to my python script instead. This way it feels like the model is way more capable than it actually is. I can use these advanced features everywhere without being bound to it's actual knowledge.

    To improve this setup even more i will likely switch to a #mixture_of_experts architecture soon. This project is a lot of fun and i can't wait to integrate it into my homelab.

    #homelab #selfhosting #privacy #ai #llm #largelanguagemodels #coding #developement

  14. Making the most out of a small LLM

    Yesterday i finally built my own #AI #server. I had a spare #Nvidia RTX 2070 with 8GB of #VRAM laying around and wanted to do this for a long time.

    The problem is that most #LLMs need a lot of VRAM and i don't want to buy another #GPU just to host my own AI. Then i came across #gemma3 and #qwen3. Both of these are amazing #quantized models with stunning reasoning given that they need so less resources.

    I chose huihui_ai/qwen3-abliterated:14b since it supports #deepthinking, #toolcalling and is pretty unrestricted. After some testing i noticed that the 8b model performs even better than the 14b variant with drastically better performance. I can't make out any quality loss there to be honest. The 14b model sneaked in chinese characters into the response very often. The 8b model on the other hand doesn't.

    Now i've got a very fast model with amazing reasoning (even in German) and tool calling support. The only thing left to improve is knowledge. #Firecrawl is a great tool for #webscraping and as soon as i implemented websearching, the setup was complete. At least i thought it was.

    I want to make the most out of this LLM and therefore my next step is to implement a basic #webserver that exposes the same #API #endpoints as #ollama so that everywhere ollama is supported, i can point it to my python script instead. This way it feels like the model is way more capable than it actually is. I can use these advanced features everywhere without being bound to it's actual knowledge.

    To improve this setup even more i will likely switch to a #mixture_of_experts architecture soon. This project is a lot of fun and i can't wait to integrate it into my homelab.

    #homelab #selfhosting #privacy #ai #llm #largelanguagemodels #coding #developement

  15. #Firecrawl, an #opensource #webcrawler for #developers and #AIagents, raised $14.5 million in a Series A round led by Nexus Venture Partners. The company, which is already profitable, plans to use the funds to expand its team and develop tools to help website owners get paid when AI uses their content. techcrunch.com/2025/08/19/ai-c #Pirates #Tech #Startup #News

  16. How come all these AI companies have founders that look like this? White, young, (generally male), and totally have not ONE CARE for what this technology is doing to others! ps: I'm surprised one of them doesn't have a man-bun!

    #AI #Firecrawl

    techcrunch.com/2025/08/19/ai-c

  17. How come all these AI companies have founders that look like this? White, young, (generally male), and totally have not ONE CARE for what this technology is doing to others! ps: I'm surprised one of them doesn't have a man-bun!

    #AI #Firecrawl

    techcrunch.com/2025/08/19/ai-c

  18. #Firecrawl launches /search endpoint for web scraping and data extraction 🔥

    #Firecrawl new /search #API combines web search results with full page content in single call, designed for #AI agents and developers who need clean data quickly 🔍

    🧵👇 #webscraping

  19. #Firecrawl launches /search endpoint for web scraping and data extraction 🔥

    🔍 New /search #API endpoint combines web search results with full page content in one call
    🤖 Built specifically for #AI agents and developers who need clean, structured data quickly

  20. #Firecrawl launches /search endpoint for web scraping and data extraction 🔥

    #Firecrawl new /search #API combines web search results with full page content in single call, designed for #AI agents and developers who need clean data quickly 🔍

    🧵👇 #webscraping

  21. #Firecrawl launches /search endpoint for web scraping and data extraction 🔥

    🔍 New /search #API endpoint combines web search results with full page content in one call
    🤖 Built specifically for #AI agents and developers who need clean, structured data quickly

  22. Turn any website into LLM-ready text with the new Alpha endpoint from #Firecrawl. The service crawls sites and generates both concise and full-text versions that can be used with any language model.
    Key features:
    • 🔍 Crawls websites and extracts clean, meaningful text content

  23. Turn any website into LLM-ready text with the new Alpha endpoint from #Firecrawl. The service crawls sites and generates both concise and full-text versions that can be used with any language model.
    Key features:
    • 🔍 Crawls websites and extracts clean, meaningful text content

  24. Transform Web Data into Structured APIs with #LLM #API Engine 🎯

    🤖 Create custom #APIs using natural language with #OpenAI powered schema generation and #Firecrawl web scraping

    github.com/developersdigest/ll

    🧵 ↓

  25. Transform Web Data into Structured APIs with #LLM #API Engine 🎯

    🤖 Create custom #APIs using natural language with #OpenAI powered schema generation and #Firecrawl web scraping

    github.com/developersdigest/ll

    🧵 ↓

  26. Introducing llms.txt Generator ✨

    You can now concatenate any website into a single text file that can be fed into any #LLM.

    It crawls the whole website with #firecrawl and extract data with gpt-4o-mini.

    Create your own llms.txt at llmstxt.firecrawl.dev!

    #Scraping #ai #API

  27. Introducing llms.txt Generator ✨

    You can now concatenate any website into a single text file that can be fed into any #LLM.

    It crawls the whole website with #firecrawl and extract data with gpt-4o-mini.

    Create your own llms.txt at llmstxt.firecrawl.dev!

    #Scraping #ai #API

  28. For those who want to "farm" the open internet for LLM content, all kind of tools are available, Firecrawl is a good example, partly opensource. Most people are negative about this probably but i think if a website is openly accessible/available for a human we almost can't prevent it to be crawled/scraped and used for AI training.
    docs.firecrawl.dev/introductio
    #AI #crawling #scraping #firecrawl #llm

  29. For those who want to "farm" the open internet for LLM content, all kind of tools are available, Firecrawl is a good example, partly opensource. Most people are negative about this probably but i think if a website is openly accessible/available for a human we almost can't prevent it to be crawled/scraped and used for AI training.
    docs.firecrawl.dev/introductio
    #AI #crawling #scraping #firecrawl #llm