Web Scraping Engineer | Data Extraction Specialist | API Reverse Engineering
Focused on building automated data extraction systems with Python. I build automated solutions for competitor price and inventory monitoring, foreclosure tracking, change detection, alerts, business directory data extraction, and recurring or one-off web data collection. I deliver clean, structured, and business-ready data at scale.
Languages & Frameworks:
- Python (primary)
- Playwright (modern browser automation)
- Scrapy (high-performance web scraping)
- SeleniumBase (modern Selenium wrapper for browser automation)
- BeautifulSoup & httpx (parsing & HTTP)
Advanced Techniques & Specializations:
- API reverse engineering & network analysis
- curl_cffi (TLS/JA3 fingerprinting for anti-bot bypass)
- Dynamic user-agent, proxy rotation & session management
- JavaScript rendering & dynamic content handling
- Data transformation (Pandas, regex)
- Automated monitoring & scheduled tasks (GCP Compute Engine, GitHub Actions, APScheduler, cron jobs)
- Structured data delivery (CSV, JSON, XLSX, databases)
High-performance e-commerce data extraction for Konga (one of Nigeria's leading marketplaces). Delivers structured product listings with pricing, inventory, and availability data at scale.
Foreclosure monitoring actor for real estate investment sourcing. Automated daily collection with structured output for market analysis and ROI tracking.
Business directory extraction actor for Nigeria. Clean structured data ideal for lead generation, market research, and business intelligence.
Multi-website price monitoring automation with intelligent change detection and alerts. Delivers structured data for data-driven purchasing decisions.
High-performance real estate data extraction engine using Playwright. Handles dynamic content, pagination, realistic user agent rotation, and outputs ETL-ready datasets.
Automated daily foreclosure collection for property sourcing, market monitoring, ROI analysis, and investment decision-making.
End-to-end property scraper with flexible search filters, error-resilient extraction, and automated data normalisation with Pandas across multiple pages and cities.
Daily price monitoring using Playwright + GitHub Actions. Features historical tracking, change detection, and automated scheduling.
- Extract valuable data from complex, JavaScript-heavy websites
- Reverse-engineer APIs to uncover hidden data sources
- Build scalable scrapers that run reliably in production
- Transform raw data into clean, structured, business-ready datasets
- Automate repetitive workflows with scheduled tasks and monitoring
- Solve data sourcing problems for e-commerce, real estate, market research, and more
I'm available for:
- Contract web scraping projects
- Data scraping consultations
- Data extraction pipeline development
- Freelance & full-time roles
Feel free to reach out if you have a data extraction challenge!
