Senior DevOps Engineer · Toronto

Ashans
Khadka

I work across data platforms, cloud infrastructure, and delivery tooling, with a focus on systems teams can operate and verify.

CurrentlySenior DevOps Engineer, Ada
DepthMongoDB · Kubernetes · Terraform
StatusOpen to senior DevOps & data infrastructure roles

SELECTED WORKS

01
Compliance
2025 – 2026

I saved $376K annually by replacing a compliance vendor with an in-house deletion API.

Regulatory deletion reaches into every store that holds customer data. A request is not done until we can verify its result across those systems. The existing vendor workflow was expensive and put a core legal obligation outside our control, so I built the in-house replacement.

The requirements kept changing as new stores and retention rules arrived. I separated the business rules from the storage layer, which meant we could adapt the underlying systems without rewriting the core logic. The rate-limited GDPR/CCPA services now run in every production environment, process approximately 200K PII-based requests per month, preserve data residency, and protect production performance.

RoleTechnical lead
Delivery1 month
DomainGDPR / CCPA
Vendor cost$376K annually
Deletion serviceall production environments
Deliveryone month
Vendor dependencyremoved from the request path

What I'd do differently: I would define, validate, and version the data contract on day one. The first version relied too much on one system's document shape, which made later changes harder to reason about.

domain-driven design · python · mongodb · redshift · airflow
02
Reliability
2023 – 2025

Cutting MongoDB upgrade cycles from months to weeks across seven environments.

I led the 4.4 through 8.0 upgrades. Database upgrades have enough ways to fail without discovering one during the change, so I used maintenance windows, standby-cluster safety nets, and rollback paths we had tested before relying on them.

I also wrote reusable runbooks and brought teammates into the work so upgrades could move faster without becoming less safe.

ScopeMongoDB infrastructure
Environments7
Upgrade outcomeNo customer-facing downtime caused by the upgrades
Leadership scope4.4 through 8.0
Rollbacktested before each change
Knowledge sharingreusable runbooks
mongodb atlas · terraform · datadog · kafka / debezium
03
Data Platform
2023 – 2025

Catching CDC failures before they become incidents.

A dashboard can look healthy while a data pipeline is still missing records. I built continuous validation that compared change events with landing data so we could see whether the path was actually complete.

It caught five pre-incident issues, including schema-driven lakehouse gaps, source-database ingestion delays, MSK throttling, and oversized events rejected by MSK. I presented the result to the wider engineering organization.

RoleData infrastructure contributor
FocusChange-data validation
OutcomeFive issues caught before incidents
MethodEvent and landing-data comparison
Validation outcomeFive pre-incident detections
AudienceWider engineering organization
data infrastructure · validation · reliability
02
Experience
Senior DevOps Engineer, Ada
Cloud infrastructure, data compliance, MongoDB, Kubernetes, and incident response.
2023 to now
Software Engineer, Ada
GDPR/CCPA deletion services, change-data validation, Airflow ELT reliability, and MongoDB upgrades.
2020 to 2023
Data Engineer / Developer, Ada
Python data-export tooling and client reporting workflows.
2019 to 2020
Software Developer, The Nielsen Company
Co-op, C# analytics dashboard and presentation automation.
2016, 2018
03
Dispatches
Loading…
All dispatches →
04
About

I'm a Senior DevOps Engineer. For the last seven years, I have worked across data platforms, cloud infrastructure, infrastructure-as-code workflows, developer tooling, and production reliability at Ada.

I care about getting to the actual cause instead of the nearest symptom. That matters most when a workflow looks healthy while data is missing, so I have spent time finding long-lived silent failures, building the Datadog dashboard that supports network incident diagnosis from the CDN through Kubernetes, and making retention and deletion systems safe to run.

I also leave behind runbooks and systems that other people can operate confidently.

As a senior IC, I coordinate projects across teams and present project demos to executive management. I help manage production Kubernetes and cloud infrastructure serving millions of requests per hour. I also manage the developer platform, including DevSpace, and built an agentic laptop bootstrap that cuts onboarding from days to hours while standardizing local tools and setup.

I also upgraded Airflow from 1.9 to 2.2 and added automated monitoring and alerting for the ELT sync that powers most analytics workloads.

I write some of it down under Dispatches. If any of it is useful to you, I'd like to hear about it.

Now: leading discovery and design for multi-region disaster recovery, including data-access audits and a cell-based routing design, for a customer requirement of under-one-minute RTO and RPO.
Ashans Khadka · Toronto