Skip to content

Repository files navigation

Hi, I'm Minner

Data Collection Engineer — I turn the messy web into clean, structured data, at scale.

Blog · Telegram · Email


What I do

  • Web scraping — robust, distributed crawlers built on Scrapy, Requests, httpx
  • Browser automation — Playwright, Selenium, Puppeteer for the hard-to-reach pages
  • Reverse engineering — JS & app reverse, captcha solving, anti-bot countermeasures
  • Data pipelines — ETL from raw HTML to queryable datasets with Pandas & Spark
  • Distributed systems — Kafka, Redis, and message queues keeping crawlers in sync

Tech stack

Languages     Python
Scraping      Scrapy · Requests · httpx · aiohttp
Automation    Playwright · Selenium · Puppeteer
Distributed   Kafka · Redis · RabbitMQ · Scrapy-Redis
Processing    Pandas · Spark · Airflow
Storage       MySQL · MongoDB · Elasticsearch · ClickHouse

Projects

Open-source work in progress — crawler frameworks, anti-bot toolkits, and pipeline templates coming soon.

Contact


Data is out there — I just fetch it, politely and at scale.

About

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages