Zakria
Available for freelance work

Automated Data Pipelines. Qualified Leads. Real Revenue.

Hi, I'm a Web Data Scraper who builds Automated Data Pipelines, Web Scrapers & Premium B2B Lead Workflows.

Zakria Shahbaz Portfolio profile photo
2+
Years
Experience
100% Client
Satisfaction

About Me

My Story

I help Founders, Sales Leaders, and Agencies automate their operations, scrape high-intent data, and build intelligent workflows. As a Python Developer and Data Specialist, I don't just write code — I build automated revenue machines that bridge the gap between complex data engineering and your bottom line.

  • Advanced Web Scraping & B2B Lead Generation
  • Data Pipelines & Workflow Automations
  • Custom AI Agents & RAG Workflows
2+Years Experience
7+Projects Completed
4+Happy Clients

Web Scraping

Scrapy
Selenium
Playwright
BeautifulSoup
Requests
aiohttp

Data Processing

Python
Pandas
NumPy
SQL
Jupyter

Databases

MySQL
PostgreSQL

Automation & Workflow

Dify AI
Git
Linux
AsyncIO
DNS Verification

Services I Offer

End-to-end web scraping, data pipeline automation, and B2B lead generation — built to eliminate manual work and deliver clean, actionable data at scale.

Web Scraping

Extracting valuable data from websites efficiently and reliably at any scale.

  • Custom Scrapers
  • Anti-Bot Bypassing
  • Dynamic Content (JS)
  • Daily Data Extraction

Data Automation Pipelines

Automate your workflows by connecting data sources, transforming data, and loading it where you need it.

  • ETL Pipelines
  • API Integration
  • Cloud Deployment
  • Scheduled Tasks

Lead Generation Scraping

We extract leads from Google Maps and then from your website in a two-phase process.

  • Google Maps Lead Extraction
  • Website Lead Extraction
  • B2B Contact Enrichment
  • Verified & Clean Data

Work Experience

Jan 2024 - Present

Data Automation Expert

Self-Employed (zaktecs.dev)

Delivering custom Python automation solutions for agencies and founders. Building web scrapers, ETL pipelines, and AI-powered workflows that eliminate manual processes and generate qualified B2B leads.

Feb 2024 - May 2025

Data Scraper

ITEH LIMITED

Developed and maintained web scrapers for extracting structured data from diverse online sources. Implemented anti-bot bypassing techniques and delivered clean datasets for business intelligence.

Jan 2025 - Present

Data Analyst

TEVTA Punjab

Performing exploratory and statistical data analysis to support workforce development programs. Building dashboards and reports to drive data-informed decision-making across the organization.

My Working Process

A proven methodology ensuring high quality delivery from concept to deployment.

01

Discovery & Requirements

We start by understanding your data needs, target sources, and the core problem you want to solve. This sets the foundation for the project.

02

Architecture & Strategy

I design the data pipeline or scraping strategy, ensuring it's scalable, robust, and capable of handling edge cases efficiently.

03

Development & Extraction

Once the strategy is approved, I build the scrapers or ETL pipelines using Python and cloud technologies, ensuring clean and accurate data.

04

Testing & Delivery

Rigorous testing to handle rate limits and format changes. Finally, we deliver the data in your preferred format or integrate it into your database.

Featured Projects

A selection of my recent work showcasing modern tech stacks and premium design.

Premium B2B Multi-Niche Lead Generation Pipeline cover
Featured
Lead Generation

Premium B2B Multi-Niche Lead Generation Pipeline

Engineered an async-first, production-grade lead enrichment engine that scraped Google Maps across 12 high-value B2B niches — including Corporate Law Firms, Industrial Cleaning, Property Management, Commercial Roofing, and Med Spas — covering 5,011 cities nationwide. The async architecture combined Playwright for discovery, aiohttp at 15-concurrency for high-throughput website crawling, and async DNS MX-resolution for email verification. Each record was enriched with 27 fields: business email, social URLs (FB/IG/LI), CMS platform detection, GA4, FB Pixel, SSL, DMARC, SPF, local schema, booking widgets, and live chat. A 100+ mega-corp blocklist filtered out national chains. Delivered 186,662 verified, deduplicated B2B leads with 48% email capture rate and 100% GMB-claimed coverage.

PythonPlaywrightaiohttpBeautifulSoupPandasdnspython
US Healthcare Premium Lead Generation Pipeline cover
Featured
Lead Generation

US Healthcare Premium Lead Generation Pipeline

Built a production-grade B2B healthcare lead scraper targeting four core verticals — Dentists, Medical Spas, Chiropractors, and Veterinarians — across all 50 US states and 2,740 cities. The Playwright-driven system executed 1,528 search queries on Google Maps, crawled each business's website (homepage + 3 internal pages), and enriched every lead with 27 fields including verified business email, social profiles, treatment offerings, insurance acceptance, online booking, and telehealth capability. Implemented DNS MX-record verification on every candidate email with multi-layer filtering (14 dummy domains, 10 webmail providers, 14 asset extension traps). Batch processing at 150 queries with JSON checkpoint crash recovery. Output: 67,230 records with 62.8% scoring 90+ on the composite lead quality scale and 98.9% phone coverage.

PythonPlaywrightRequestsdnspythonPandasurllib3
Funda.nl Real Estate Data Scraper cover
Featured
Web Scraping

Funda.nl Real Estate Data Scraper

Built a production-grade Playwright scraper that crawled 585 paginated search pages on Funda.nl, the Netherlands' largest real estate marketplace, extracting 15,345+ structured property listings — price, address, agent, energy label, and more. Engineered crash-safe checkpointing, automatic retries, and anti-bot evasion for stable, unattended long-running crawls.

PythonPlaywrightBeautifulSoupPandas
AliExpress Category & Pricing Scraper cover
Featured
Web Scraping

AliExpress Category & Pricing Scraper

Built an anti-detection AliExpress scraper using undetected-chromedriver that crawls 9 high-margin product categories, extracting exact JSON-embedded pricing, sold counts, ratings, and reviews. Forced PKR currency via cookie/geo manipulation from a DigitalOcean UK datacenter, and engineered atomic CSV writes with per-category checkpoint resume for crash-safe long-running crawls.

PythonSeleniumUndetected ChromeDriverPandas
Daraz.pk Product Intelligence Scraper cover
Featured
Web Scraping

Daraz.pk Product Intelligence Scraper

Built a headful Playwright scraper that connects to Chrome over the DevTools Protocol to extract structured product data — pricing, discounts, units sold, reviews, and seller location — across 10 categories on Daraz.pk, Pakistan's largest e-commerce marketplace. Engineered incremental CSV writes, auto-stop-on-empty-page logic, and human-like pacing to scale safely to ~15,000 product listings per run.

PythonPlaywrightPandasChrome DevTools Protocol

Frequent Questions

I start with a deep dive into your requirements and target data sources. Then, I design the scraping or pipeline architecture. Once approved, I begin development, followed by rigorous testing and deployment.
A standard single-site scraper usually takes 1-2 weeks. More complex pipelines involving multiple sources or anti-bot bypassing can take 3-6 weeks.
Yes! Websites change often, so I offer maintenance retainers to keep your scrapers updated and data flowing smoothly.
My primary stack revolves around Python, Scrapy, Selenium, and Playwright for scraping. For pipelines, I use PostgreSQL, cloud services, and workflow automation tools.
Yes, I provide end-to-end services. I can extract the data, clean it, analyze it, and deliver it in your preferred format — CSV, Excel, JSON, or directly into your database.

Contact Me

Email Me

zakria@zaktecs.dev

Call Me

+923206825657

Location

Lahore, Pakistan

Send a message

Book a 30-Minute Call

Schedule a free consultation

WhatsApp