Skip to content
@2scraper

2scraper

2scraper logo

2scraper

Open-source web scrapers for real-world websites.

Turn products, listings, comments, prices, and market data into clean JSON or CSV.

Browse scrapers · Request a website · Contribute

30+ scrapers Python powered JSON and CSV output Open source


Web data, without rebuilding the plumbing

2scraper is a collection of focused, ready-to-run scrapers for popular websites. Each repository targets one platform and documents the supported pages, captured fields, setup, and output—so you can spend less time reverse-engineering websites and more time using the data.

⚡ Ready to run
Clone a repository, follow its quick start, and collect data on your own infrastructure.
📦 Structured by default
Predictable records in JSON or CSV, with schemas and sample output where available.
🧰 Multiple execution paths
Playwright, Selenium, Puppeteer, the site's own JSON endpoints, or the 2Captcha Scraping Browser API—depending on the target.
🛡️ Built for real websites
Pagination, dynamic content, proxies, fingerprints, and CAPTCHA flows where the site requires them.

Start in three steps

  1. Pick a target from the scraper directory.
  2. Open its repository and follow the quick-start guide.
  3. Run locally, then add the optional browser, proxy, or CAPTCHA setup documented for that target.

Support differs by repository. The README in each scraper is the source of truth for engines, locales, fields, and infrastructure requirements.

Featured scrapers

Repository What it extracts
Amazon Search results, best sellers, product pages, and public reviews across 21 marketplaces — runs with no key or proxy
YouTube Comment threads and replies, video metadata, and video search — runs with no key or proxy
Catawiki Auction lots, bids, reserves, estimates, and seller data — runs with no key or proxy
StockX Products, asks, bids, last sales, and market statistics
Transfermarkt Player profiles, market values, club squads, and transfers
Medium Stories, authors, publications, tag feeds, and full article text — runs with no key or proxy

Scraper directory

All public platform scrapers, grouped by what they are used for.

🟢 runs with no API key and no proxy · 🔵 needs the 2Captcha Scraping Browser API — both measured and dated in that repository's README.

One platform can take more than one scraper. TikTok gates each route differently, so it is covered by four that work side by side: profiles 🟢, videos 🟢 and the EU Ad Library 🟢 need no key at all; TikTok Shop 🔵 is behind a captcha and needs the Scraping Browser.

🛒 E-commerce & retail

Scraper What it extracts
Amazon 🟢 Search results, best sellers, product pages and reviews
Andie Swim 🟢 Swimwear listings, per-size stock and prices
Bershka Inditex catalogue, one row per SKU
Catawiki 🟢 Auction lots, bids, reserves, estimates, sellers
Etsy 🔵 Search, category, shop and listing pages
Farfetch Fashion listings and product pages with prices
Givenchy 🟢 Beauty products and prices
Home Depot Category listings and product pages, prices, specs
Maison KOSÉ 🟢 Japanese cosmetics: products, prices, brands, stock
LG 🟢 Catalogue models, sizes and categories
Lidl US grocery products, prices, unit prices
MediaMarkt Electronics listings and product pages, prices
Montblanc 🟢 Per-market prices, stock, collections, variants
Rakuten 🟢 Ichiba products, prices, points, shops, reviews
Sleep Number Smart beds and mattresses: per-size prices, ratings
StockX Sneaker listings and products: asks, bids, last sale
TikTok Shop 🔵 Products, prices, units sold, sellers
Tokopedia 🔵 Indonesian marketplace: search, category, product pages
Woolworths 🟢 Supermarket products, prices, unit prices, specials

🏠 Classifieds, property & travel

Scraper What it extracts
Autotrader US car listings and detail pages: price, KBB fair price, VIN, dealers
Craigslist 🟢 Classified listings and postings
dubizzle UAE classifieds: cars, property, jobs
Flippa 🟢 Online businesses, websites, apps and domains for sale
MakeMyTrip 🔵 Indian hotel listings: nightly prices with taxes and fees, star and guest ratings
Spinny 🟢 Used-car listings and car pages, prices
Vrbo Vacation-rental search grids and property pages
Zimmo Belgian property listings: prices, area, bedrooms, EPC

🍔 Food & delivery

Scraper What it extracts
foodpanda Restaurant listings, ratings, cuisines, deals

📱 Apps, social & video

Scraper What it extracts
Google Play 🟢 App listings, search, installs, ratings and reviews
Snapchat 🟢 Public profiles, subscriber counts, Spotlight views and engagement, stories and highlights
TikTok Ad Library 🟢 EU ads: advertisers, creatives, run dates, audience bucket
TikTok profiles 🟢 Exact follower, like and video counts, bio
TikTok videos 🟢 Captions, engagement, hashtags, subtitles, media URLs
Weibo 🟢 Hot feed, account timelines, comments, engagement
YouTube 🟢 Comment threads and replies, video metadata, search

📝 Publishing & Q&A

Scraper What it extracts
Medium 🟢 Tag feeds, archives, author pages, full story text
Quora 🟢 Answers from questions, profiles and topics

💼 Jobs & business directories

Scraper What it extracts
BBB Business listings, BBB ratings, accreditation, complaints
Just Join IT 🟢 IT job offers with salaries, skills, seniority
Mercor 🟢 Contract roles, rates, eligibility, corporate openings
Wellfound Startup jobs with salary and equity ranges

📈 Finance, markets & fundraising

Scraper What it extracts
Binance 🟢 P2P adverts, copy-trading lead portfolios, announcements
Google Finance 🟢 Quotes, financials, analyst ratings, OHLCV, FX
Indiegogo 🟢 Campaigns, funding totals, backers, reward tiers
OpenSea 🟢 NFT floor prices, offers, sales history, rankings
Polymarket 🟢 Prediction-market prices, order books, token ids

⚽ Sports

Scraper What it extracts
Transfermarkt Market values, squads, transfers, player profiles

View all repositories →

Built to fit your workflow

Most repositories include:

  • a runnable Python implementation and command-line examples;
  • JSON and CSV output with documented fields;
  • sample records for a quick look at the data;
  • pagination and dynamic-content handling tailored to the target;
  • optional 2Captcha integrations — captcha solving, the Scraping Browser API, 2prx residential proxies and fingerprints — when a site needs them.

Every scraper can be used on your own infrastructure. Paid services are optional unless a repository explicitly says otherwise.

2Captcha products, one account

Every scraper runs on your own machine first. When a site pushes back, each repository says which of these helps — and which does not — with the measurement behind it.

Product What it gives a scraper Where it matters here
Scraping Browser API A managed Chrome over CDP with its own exit country, persistent profiles and captcha auto-solve — no browser or residential address of your own The 🔵 scrapers, and any server-side pipeline that cannot run a headful browser from a home address
Captcha solving Tokens for reCAPTCHA, Cloudflare Turnstile and other challenge widgets Sites that put a challenge widget in front of their pages
Residential proxies Residential exits by country Sites that refuse datacentre addresses — the repositories without 🟢 say so
Fingerprints A consistent, self-consistent browser identity Volume across many sessions

The four are billed separately; one 2Captcha account covers them.

Need another website?

If the target is not listed, open a scraper request with the website, pages you need, desired fields, and expected scale. For a private or custom extraction project, start an inquiry.

Contributing

  • Found a bug? Open an issue in the affected scraper repository and include the URL, command, and relevant log output.
  • Want to improve a scraper? Fork the repository and send a focused pull request.
  • Missing a platform? Request it here.

Please use scraped data responsibly and follow the target website's terms and applicable laws.


Built for developers and data teams who would rather use the data than fight the page.

Popular repositories Loading

  1. amazon-scraper amazon-scraper Public

    Amazon scraper: search results, best sellers, product pages and reviews across 21 marketplaces (Playwright, Selenium, Puppeteer, CDP) — captcha solving, proxies, fingerprints

    Python 7

  2. farfetch-scraper farfetch-scraper Public

    Farfetch listing-page scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) — JSON-LD parsing, reCAPTCHA solving, proxies, fingerprints

    Python 4

  3. mediamarkt-scraper mediamarkt-scraper Public

    MediaMarkt listing and product-page scraper (Playwright, Selenium, pyppeteer, or the 2Captcha Scraping Browser API via CDP) — JSON-LD parsing, proxies, captcha solving, fingerprints

    Python 3

  4. catawiki-scraper catawiki-scraper Public

    Catawiki auction scraper (Playwright, Selenium, pyppeteer, or the 2Captcha Scraping Browser API over CDP) — lots, bids, reserves, estimates, sellers, JSON/CSV

    Python 3

  5. tokopedia-scraper tokopedia-scraper Public

    Tokopedia product scraper (Playwright, Selenium, pyppeteer, or the 2Captcha Scraping Browser API over CDP) — search grids, category listings, product pages, lazy-load scrolling, JSON/CSV

    Python 2

  6. lg-scraper lg-scraper Public

    LG.com catalogue scraper (the site's own catalogue API, plus Playwright, Selenium and Puppeteer over the rendered grid) — no-browser primary path, proxies, fingerprints, run-to-run diffs

    Python 2

Repositories

Showing 10 of 51 repositories
  • kohls-scraper Public

    Kohl's (kohls.com) listing and product scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) — prices, price ranges, sale/clearance labels, ratings, per-SKU stock, JSON/CSV

    2scraper/kohls-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 25, 2026
  • lg-scraper Public

    LG.com catalogue scraper (the site's own catalogue API, plus Playwright, Selenium and Puppeteer over the rendered grid) — no-browser primary path, proxies, fingerprints, run-to-run diffs

    2scraper/lg-scraper's past year of commit activity
    Python 2 MIT 0 0 1 Updated Sep 25, 2026
  • makemytrip-scraper Public

    MakeMyTrip hotel listing scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) — nightly prices with taxes and fees, star and guest ratings, property types, locations, proxies, fingerprints

    2scraper/makemytrip-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 25, 2026
  • .github Public
    2scraper/.github's past year of commit activity
    0 3 0 4 Updated Sep 25, 2026
  • autotrader-scraper Public

    Autotrader.com car listings scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) — VIN, price, KBB fair price, deal rating, dealers, proxies, fingerprints

    2scraper/autotrader-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 25, 2026
  • screener-scraper Public

    screener.in stock screener scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) — screens, sector listings, P/E, market cap, ROCE to JSON/CSV, proxies, fingerprints

    2scraper/screener-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 25, 2026
  • youtube-scraper Public

    YouTube comment scraper (Playwright, Selenium, Puppeteer, or the 2Captcha Scraping Browser API via CDP) - comment threads and replies, video metadata, search, captcha solving, proxies

    2scraper/youtube-scraper's past year of commit activity
    Python 0 MIT 0 0 1 Updated Sep 25, 2026
  • bershka-scraper Public

    Bershka product scraper — Playwright, Selenium, pyppeteer or the Scraping Browser API via CDP; Inditex /itxrest API, one row per SKU, Akamai interstitial handling, proxies

    2scraper/bershka-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 24, 2026
  • homedepot-scraper Public

    Home Depot listing and product scraper (HTTP, Playwright, Selenium, Puppeteer or a remote browser over CDP) — Apollo-state and JSON-LD parsing, captcha solving, proxies, fingerprints

    2scraper/homedepot-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 24, 2026
  • givenchy-scraper Public

    Givenchy Beauty product and price scraper — Playwright, Selenium, Puppeteer or the Scraping Browser API via CDP; JSON-LD parsing, captcha solving, proxies, fingerprints

    2scraper/givenchy-scraper's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 24, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…