made by 0x1da49.comblog
← all reports

Top 46 Web Scraping, Crawlers & Browser Automation Repos (February 2026)

repos

Browse 46 web scrapers, data extractors, Playwright/Puppeteer automation scripts, and crawler frameworks published in February 2026.

Scrapers & Automation — February 28, 2026

GitHub
Playwright Best Practices AI SKillcurrents-dev/playwright-best-practices-skill
GitHub
CLI for Common Playwright Actionsmicrosoft/playwright-cli
GitHub
Pure Go PostgresSQL ParserValkDB/postgresparser
GitHub
Git worktree automation.svenmalvik/manifold

+ 6 more verified repositories in full dataset...

Executive Intelligence Summary

During February 2026, the Repos tracking engine indexed 46 open-source repositories matching the Scrapers & Automation cluster.

These projects represent active engineering teams and independent open-source developers shipping production code, architectural prototypes, and developer tooling.


Key Sector Telemetry & Maintainer Trends

  • Primary Topic Signals: automation (18), parser (12), browser (6), playwright (6), end (4), web (4)
  • Active Maintainers & Orgs: lekt9 (2), vakra-dev (2), microsoft (2), ValkDB (2), memvid (2)
  • Sample Records Displayed: 40 of 46
  • Remaining Available in Dataset: 6 records

3 Ways to Monetize & Leverage This Developer Intelligence

1. DevTool Sales & Technical Outbound

Reach out to repository creators who just launched tools in the Scrapers & Automation space. Maintainers actively shipping code are prime candidates for cloud infrastructure, CI/CD automation, API credits, and developer tools.

2. High-Signal Tech Recruitment

Identify skilled engineers working on bleeding-edge projects. Sourcing talent directly from verified open-source releases gives you verified code samples and active commits.

3. Competitive Intelligence & Directory Building

Import structured repository telemetry into internal dashboards, market landscape charts, or developer directory websites to track where open-source velocity is concentrating.


Developer Usage Recipe

import pandas as pd

# Load Repos Master Index
df = pd.read_csv("repos.csv")

# Filter for February 2026 records in this sector
sector_repos = df[df["month"] == "2026-02"]
print(f"Loaded {len(sector_repos):,} open-source repositories from February 2026")

If you need the full dataset of 19,400+ developer repository records formatted for instant use in spreadsheets, pipelines, or databases, check out the Repos Master Dataset.

related reports →