About Hasan Butt
SYSTEMS ENGINEER // MLOPS // SRE PLATFORMS // FINOPS
I'm Hasan. I do three things, and I do them for real production systems — not demos.
First, I keep things running.
I build Kubernetes platforms on EKS, design SLOs that engineers actually respect, and set up incident response so that when things break at 3am — and they will — the fix takes minutes, not hours. Reliability isn't a dashboard. It's a system of decisions made before the outage.
Second, I cut cloud costs.
Most AWS and GCP bills are 30–50% waste: idle compute nobody remembers launching, storage nobody reads, and egress nobody planned for. I find it, kill it, and set up guardrails so it stays dead. This work pays for itself in the first month, usually several times over.
Third, I run AI infrastructure.
Companies are burning money on per-token API pricing when a well-run vLLM deployment on the right GPUs does the job for a fraction of the cost. I build and operate self-hosted LLM serving and production RAG pipelines — with monitoring, evals, and cost tracking, because AI infra is still infra.
These three things sound separate. They're not. Reliability, cost, and AI workloads are the same problem viewed from three angles: run compute well.
Most teams have someone for one of these. Very few have someone for all three.
That's the gap I fill.
- M.S. in Computer ScienceBahria University (2020–2022)
- B.S. in Computer ScienceBahria University (2016–2020)
Professional Timeline
AEITCH
Nov 2025 – Present (Active)Senior DevOps & SRE Engineer · Remote
- Architected resilient multi-AZ AWS environments and CI/CD pipelines for high-concurrency SaaS applications, maintaining 99.9% uptime SLAs.
- Reduced monthly cloud infrastructure costs by 15%+ via right-sizing, reserved capacity planning, and automated staging scale-down.
- Implemented full-stack telemetry (Prometheus, Grafana, ELK) with SLO/SLI error budgets, cutting mean time to recovery (MTTR) by 45%.
Nextbridge
May 2025 – Nov 2025 (6 mos)Lead DevOps Engineer · Hybrid
- Led migration of enterprise services to containerized AWS ECS with automated CodePipeline deployments, cutting release cycles from days to 45 minutes.
- Enforced ISO 27001 and least-privilege IAM security hardening across all production VPC networks.
- Conducted load stress-testing using Python Locust to profile API bottlenecks under 5,000+ concurrent users, reducing p95 latency by 35%.
VentureDive
Mar 2024 – Apr 2025 (1 yr 1 mo)Senior DevOps Engineer · Hybrid
- Executed large-scale workload migration from Singapore to US East region, cutting monthly AWS infrastructure bills by 30% without downtime.
- Managed automated PostgreSQL RDS snapshots and automated staging database synchronization with sub-15-minute turnaround.
- Built automated Python and Bash data transfer pipelines handling multi-gigabyte S3 object synchronization across production accounts.
QisstPay
Aug 2022 – Mar 2024 (1 yr 8 mos)DevOps & Infrastructure Engineer · Onsite
- Architected hybrid AWS and GCP Cloud Run infrastructure hosting 17+ mission-critical financial microservices with 99.99% payment gateway availability.
- Achieved strict PCI DSS compliance with automated audit logging, network segmentation, and encrypted financial data vaults.
- Established IPsec VPN tunnels with top financial institutions for real-time credit decisioning APIs.
Khired Networks
Oct 2020 – Aug 2022 (1 yr 10 mos)DevOps & SRE Engineer · Onsite
- Containerized monolithic systems into Docker containers deployed across AWS EC2 and on-premise clusters.
- Built centralized logging and monitoring with Graylog, Prometheus, and Grafana, establishing incident runbooks and SLO targets.
- Deployed automated CI/CD pipelines via Jenkins and AWS CodeDeploy, cutting deployment failure rates by 70%.
