Backend Engineer
Muhammad Hussnain Zia
I build
I build the systems nobody sees and everybody relies on — APIs, automation bots, and data-extraction pipelines that turn messy real-world documents into clean, structured data.
APIs & Services
Design → Deploy
Automation
Bots at scale
Extraction
PDF → Structured data

About
Backend by trade, automation by obsession.
I'm a backend developer who specializes in the unglamorous, high-stakes middle of the stack: reliable APIs, resilient scraping and bot infrastructure, and document-extraction pipelines that parse thousands of PDFs into data teams can actually use.
My day-to-day toolkit spans Python and Node.js / TypeScript, with automation built on Scrapy, Playwright, and Puppeteer, deployed and operated on Google Cloud Platform. I care about error handling, observability, and code the next engineer can read without a map.
BS Computer Science — University of Gujrat (2015 – 2019)
Quick facts
Focus
Backend · APIs · Automation
Core stack
Python · Node.js · TypeScript
Cloud
Google Cloud Platform
Based in
Pakistan
Skills
Tools I reach for.
Languages & Runtimes
- Python
- TypeScript
- JavaScript
- Node.js
Backend & APIs
- REST API design
- Express
- Microservices
- PostgreSQL
- MongoDB
Scraping & Automation
- Scrapy
- Playwright
- Puppeteer
- Bot automation
- PDF data extraction
DevOps & Cloud
- Google Cloud Platform
- Docker
- Terraform
- CI/CD
- Infrastructure management
- Monitoring
Experience
Where I've shipped.
2022 — Present
Senior Software Engineer · Arbisoft
Building document-extraction and bot-automation systems for healthcare and insurance data platforms: multi-carrier PDF extractors, resilient request runners, and the error handling that keeps them honest in production.
- Node.js
- TypeScript
- Python
- GCP
2020 — 2022
Backend Developer · Emblem Technology
Designed and built Django REST APIs powering client products — data modeling, authentication, and integrations — as part of a fast-moving software house team.
- Python
- Django
- REST APIs
Projects
Selected work.
Multi-carrier PDF Extraction Engine
Extraction pipeline that parses insurance EOB and claims PDFs from dozens of carriers into structured, validated JSON — with per-carrier extractors and defensive error handling for corrupted or malformed files.
- Node.js
- TypeScript
- PDF parsing
Distributed Scraping Infrastructure
Bot fleet built on Scrapy, Playwright, and Puppeteer for reliable large-scale data collection — session management, retries, and observability baked in.
- Python
- Scrapy
- Playwright
- Puppeteer
Cloud Automation Pipeline
GCP-based infrastructure for scheduling, running, and monitoring automation workloads — containerized services, CI/CD, and cost-aware scaling.
- GCP
- Docker
- CI/CD
Contact
Let's build something reliable.
Have a data-extraction problem, an automation idea, or an API that needs a steady pair of hands? Drop me a message and I'll get back to you.