Senior DevOps Engineer · Toronto

Ashans
Khadka

I work across data platforms, cloud infrastructure, and delivery tooling. I build the checks and workflows teams use to ship production changes safely and find the actual cause when something goes wrong.

CurrentlySenior DevOps Engineer, Ada
DepthMongoDB · Kubernetes · Terraform
StatusOpen to senior DevOps & data infrastructure roles

SELECTED WORKS

01
Compliance
2025 – 2026

Saved more than $300K annually in vendor licensing by replacing the deletion workflow with an in-house API.

A deletion request touches every store that holds customer data. I marked a request complete only after each relevant system confirmed the deletion. The vendor workflow was expensive and put that legal obligation outside our control, so I built the replacement.

New stores and retention rules kept changing the work. I kept the business rules separate from storage, so changing an underlying system did not mean rewriting the core decision.

The rate-limited GDPR/CCPA services now run in every production environment. They process approximately 200K PII-based requests per month while preserving data residency and protecting production performance.

RoleTechnical lead
Delivery1 month
DomainGDPR / CCPA
Vendor costmore than $300K annually
Deletion serviceall production environments
Deliveryone month
Vendor dependencyremoved from the request path

What I'd do differently: I would define, validate, and version the data contract on day one. The first version relied too much on one system's document shape, which made subsequent changes harder to trace and validate.

domain-driven design · python · mongodb · redshift · airflow
02
Reliability
2023 – 2025

The MongoDB fleet stopped taking months to upgrade across seven environments.

I led the upgrades from MongoDB 4.4 through 8.0. I treated an upgrade as routine only after the recovery path had been tested, so each change used a maintenance window, a standby-cluster safety net, and a rollback procedure we had already run.

I wrote reusable runbooks and brought teammates into the work. That cut the upgrade cycle from months to weeks without making the process less safe.

ScopeMongoDB infrastructure
Environments7
Upgrade outcomeNo customer-facing downtime caused by the upgrades
Leadership scope4.4 through 8.0
Rollbacktested before each change
Knowledge sharingreusable runbooks
mongodb atlas · terraform · datadog · kafka / debezium
03
Data Platform
2023 – 2025

The CDC pipeline could look healthy while data was still missing.

A green dashboard was not enough proof that a change event had reached the lakehouse. I added continuous validation between the events and landed data, so we could see whether the path was actually complete.

It caught five issues before they became incidents: schema-driven lakehouse gaps, source-database ingestion delays, MSK throttling, and oversized events rejected by MSK. I presented the result to the wider engineering organization.

RoleData infrastructure contributor
FocusChange-data validation
OutcomeFive issues caught before incidents
MethodEvent and landing-data comparison
Validation outcomeFive pre-incident detections
AudienceWider engineering organization
data infrastructure · validation · reliability
Experience
Senior DevOps Engineer, Ada
Cloud infrastructure, data compliance, MongoDB, Kubernetes, and incident response.
2023 to present
Software Engineer, Ada
GDPR/CCPA deletion services, change-data validation, Airflow ELT reliability, and MongoDB upgrades.
2020 to 2023
Data Engineer / Developer, Ada
Python data-export tooling and client reporting workflows.
2019 to 2020
Software Developer, The Nielsen Company
Co-op, C# analytics dashboard and presentation automation.
2016, 2018
Dispatches
Loading…
All dispatches →
About

I'm a Senior DevOps Engineer. For the last seven years, I have worked on data platforms, cloud infrastructure, infrastructure-as-code workflows, developer tooling, and production reliability at Ada.

I work on problems where the dashboard says healthy but the data or production behavior says otherwise. That has meant finding silent data failures, building the Datadog dashboard used to trace network incidents from the CDN through Kubernetes, and making retention and deletion workflows safe to run.

I coordinate work across teams and present project demos to executive management. I also help manage Kubernetes and cloud infrastructure serving millions of requests per hour. For upgrades and new workflows, I leave behind a runbook and an operating path the next person can follow without needing me.

I manage the developer platform, including DevSpace. The agentic laptop bootstrap I built cuts onboarding from days to hours and gives new engineers a consistent local setup. I upgraded Airflow from 1.9 to 2.2 and added automated alerting for the ELT sync that powers most analytics workloads.

Find my technical notes under Dispatches.

Currently: leading discovery and design for multi-region disaster recovery, including data-access audits and a cell-based routing design, for a customer requirement of under-one-minute RTO and RPO.
Ashans Khadka · Toronto