Senior DevOps Engineer · Toronto

Ashans
Khadka

I work across data platforms, cloud infrastructure, and delivery tooling. Most of the work is figuring out what has to be true before a production change is safe to make.

CurrentlySenior DevOps Engineer, Ada
DepthMongoDB · Kubernetes · Terraform
StatusOpen to senior DevOps & data infrastructure roles

SELECTED WORKS

01
Compliance
2025 – 2026

The $376K annual vendor cost went away once I built the deletion workflow in-house.

A deletion request touches every store that holds customer data. I could not call one complete until the relevant systems had confirmed it. The vendor workflow was expensive and put that legal obligation outside our control, so I built the replacement.

New stores and retention rules kept changing the shape of the work. I kept the business rules separate from storage so a change underneath did not mean rewriting the core decision.

The rate-limited GDPR/CCPA services now run in every production environment. They process approximately 200K PII-based requests per month while preserving data residency and protecting production performance.

RoleTechnical lead
Delivery1 month
DomainGDPR / CCPA
Vendor cost$376K annually
Deletion serviceall production environments
Deliveryone month
Vendor dependencyremoved from the request path

What I'd do differently: I would define, validate, and version the data contract on day one. The first version relied too much on one system's document shape, which made later changes harder to reason about.

domain-driven design · python · mongodb · redshift · airflow
02
Reliability
2023 – 2025

The MongoDB fleet stopped taking months to upgrade across seven environments.

I led the 4.4 through 8.0 upgrades. I did not want an upgrade to become routine until we had a tested way back, so each change used a maintenance window, a standby-cluster safety net, and a rollback path we had already run.

I wrote reusable runbooks and brought teammates into the work. That cut the upgrade cycle from months to weeks without making the process less safe.

ScopeMongoDB infrastructure
Environments7
Upgrade outcomeNo customer-facing downtime caused by the upgrades
Leadership scope4.4 through 8.0
Rollbacktested before each change
Knowledge sharingreusable runbooks
mongodb atlas · terraform · datadog · kafka / debezium
03
Data Platform
2023 – 2025

The CDC pipeline could look healthy while data was still missing.

A green dashboard was not enough proof that a change event had reached the lakehouse. I added continuous validation between the events and landed data, so we could see whether the path was actually complete.

It caught five issues before they became incidents: schema-driven lakehouse gaps, source-database ingestion delays, MSK throttling, and oversized events rejected by MSK. I presented the result to the wider engineering organization.

RoleData infrastructure contributor
FocusChange-data validation
OutcomeFive issues caught before incidents
MethodEvent and landing-data comparison
Validation outcomeFive pre-incident detections
AudienceWider engineering organization
data infrastructure · validation · reliability
02
Experience
Senior DevOps Engineer, Ada
Cloud infrastructure, data compliance, MongoDB, Kubernetes, and incident response.
2023 to now
Software Engineer, Ada
GDPR/CCPA deletion services, change-data validation, Airflow ELT reliability, and MongoDB upgrades.
2020 to 2023
Data Engineer / Developer, Ada
Python data-export tooling and client reporting workflows.
2019 to 2020
Software Developer, The Nielsen Company
Co-op, C# analytics dashboard and presentation automation.
2016, 2018
03
Dispatches
Loading…
All dispatches →
04
About

I'm a Senior DevOps Engineer. For the last seven years, I have worked on data platforms, cloud infrastructure, infrastructure-as-code workflows, developer tooling, and production reliability at Ada.

I tend to work on the systems that look healthy until they are not. That has meant finding silent data failures, building the Datadog dashboard used to trace network incidents from the CDN through Kubernetes, and making retention and deletion workflows safe to run.

I coordinate work across teams and present project demos to executive management. I also help manage Kubernetes and cloud infrastructure serving millions of requests per hour. When I change something, I try to leave behind a runbook and a system the next person can operate without needing me.

I manage the developer platform, including DevSpace. The agentic laptop bootstrap I built cuts onboarding from days to hours and gives new engineers a consistent local setup. Earlier, I upgraded Airflow from 1.9 to 2.2 and added automated alerting for the ELT sync that powers most analytics workloads.

I write some of it down under Dispatches. If any of it is useful to you, I'd like to hear about it.

Now: leading discovery and design for multi-region disaster recovery, including data-access audits and a cell-based routing design, for a customer requirement of under-one-minute RTO and RPO.
Ashans Khadka · Toronto