Skip to main content
buildradar
Sign in
Topic · web-scraping

web-scraping

Tracked open-source repos tagged web-scraping, sorted by stars.

119 repos
  • a stealthy browser automation framework

    861-2Star change over the last 7 days
  • spidr@postmodern

    A versatile Ruby web spidering library that can spider a site, multiple domains, certain links or infinitely. Spidr is designed to be fast and easy to use.

    836+0Star change over the last 7 days
  • Uscrapper@z0m31en7

    Uscrapper Vanta: Dive deeper into the web with this powerful open-source tool. Extract valuable insights with ease and efficiency, from both surface and deep web sources. Empower your data mining and analysis with Vanta's advanced capabilities. Fast, reliable, and user-friendly, Uscrapper Vanta is the ultimate choice for researchers and analysts.

    795+4Star change over the last 7 days
  • trawl@germondai

    Self-hosted scraping engine — bypasses any JS challenge & captcha: Cloudflare, Turnstile, reCAPTCHA, hCaptcha, GeeTest. FlareSolverr & Byparr alternative and drop-in replacement for your *arr stack.

    786+50Star change over the last 7 days
  • patchright-nodejs@Kaliiiiiiiiii-Vinyzu

    Undetected NodeJS version of the Playwright testing and automation library.

    777+5Star change over the last 7 days
  • Guide to Using Google Sheets for Basic Web Scraping

    764+8Star change over the last 7 days
  • Google Search Results via SERP API pip Python Package

    755+2Star change over the last 7 days
  • neo@4ier

    Turn any web app into an API. Chrome extension captures browser traffic, auto-generates schemas, lets AI replay APIs directly. No official API needed.

    751+1Star change over the last 7 days
  • gpt-promo-scanner@JUk1-GH

    ChatGPT Team(Business) 促销码自动扫描工具 — 批量发现/验证/价格收集,支持 17 国 34 个码,最高折扣 71% | ChatGPT Business promo code scanner — batch discovery, validation, price collection, 34 codes across 17 countries, up to 71% off

    733+11Star change over the last 7 days
  • scrapecraft@ScrapeGraphAI

    🤖 AI-powered web scraping editor with visual workflow builder. Build, test & deploy web scrapers using natural language. Powered by ScrapeGraphAI & LangGraph.

    695+11Star change over the last 7 days
  • figranium@figranium

    Stack blocks visually to build complex browser workflows and execute them via API

    685+61Star change over the last 7 days
  • deepcrawl@lumpinif

    100% free and full open-source edge Firecrawl alternative with better links extraction for agents - that you can deploy to cloudflare or vercel by yourself.

    668+6Star change over the last 7 days
  • Execute complex automation scripts via remote cloud sessions , featuring integrated residential proxies , automated CAPTCHA solving , and native JavaScript rendering for the toughest dynamic websites.

    666+4Star change over the last 7 days
  • Complete-Life-Cycle-of-a-Data-Science-Project

    653+0Star change over the last 7 days
  • 一款适用于央视网的网络视频流解析处理工具

    635+16Star change over the last 7 days
  • google-search@web-agent-master

    A Playwright-based Node.js tool that bypasses search engine anti-scraping mechanisms to execute Google searches. Local alternative to SERP APIs with MCP server integration.

    621+0Star change over the last 7 days
  • PHPScraper@spekulatius

    A universal web-util for PHP.

    588+0Star change over the last 7 days
  • primp@deedy5

    HTTP client that can impersonate web browsers

    584+6Star change over the last 7 days
  • social-media-profile-scrapers@shaikhsajid1111

    Fetch user's data across social media

    573+3Star change over the last 7 days
  • Stealth-Requests@jpjacobpadilla

    Undetected web-scraping & seamless HTML parsing in Python!

    563+2Star change over the last 7 days
  • reader@vakra-dev

    Open source web infrastructure for AI. Scrape, crawl, and automate the web, clean markdown, browser sessions, ready for your agents.

    559+1Star change over the last 7 days
  • NBA Stats API via Basketball Reference

    556+0Star change over the last 7 days
  • jekyll@programminghistorian

    Jekyll-based static site for The Programming Historian

    548+0Star change over the last 7 days
  • ketch@1broseidon

    Fast, stateless CLI for web search and scrape. Built for AI agents.

    544+20Star change over the last 7 days
  • Automate scraping hotel data including pricing, availability, ratings, and booking providers with a Google Hotels API integration.

    535Star change over the last 7 days
  • 🎯 哔哩哔哩(bilibili)评论区数据可视化分析软件-- up主可用于指导自己的题材选择,明确自己的粉丝群体

    528+2Star change over the last 7 days
  • browser-search@Johell1NS

    A skill for AI agents: search the web with SearXNG, browse with Camofox, bypass protections with CloakBrowser. Anti-hallucination by design. Self-hosted, free, unlimited.

    513+7Star change over the last 7 days
  • Python quick start guides to get the most out of Oxylabs' Web Scraper API free trial.

    512+0Star change over the last 7 days
  • scrapple@AlexMathew

    A framework for creating semi-automatic web content extractors

    504+1Star change over the last 7 days
← Back to topics