apify
apify's tracked open-source repos, sorted by stars.
- #1
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
★ 25,556+98Star change over the last 7 days - #2
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
★ 9,477+19Star change over the last 7 days - #3
The Apify MCP server enables your AI agents to extract data from social media, search engines, maps, e-commerce sites, or any other website using thousands of ready-made scrapers, crawlers, and automation tools available on the Apify Store.
★ 5,618+820Star change over the last 7 days - #4
Browser fingerprinting tools for anonymizing your scrapers. Developed by Apify.
★ 2,587+11Star change over the last 7 days - #5
Collection of Apify agent skills
★ 2,362+7Star change over the last 7 days - #6
Node.js implementation of a proxy server (think Squid) with support for SSL, HTTP/HTTPS, SOCKS5, authentication, and upstream proxy chaining.
★ 1,020+4Star change over the last 7 days - #7
HTTP client made for scraping based on got.
★ 770+1Star change over the last 7 days - #8
A universal CLI client for MCP. mcpc supports persistent sessions, stdio/HTTP, OAuth 2.1, tasks, JSON output for code mode, proxy for AI sandboxes, x402, and more.
★ 757+1Star change over the last 7 days - #9★ 583+9Star change over the last 7 days