Sr. Cloud Operations Engineer
Led reliability initiatives for Autodesk's cloud-based products in the media and entertainment industry, combining hands-on systems engineering with strategic architectural support.
- Co-led replacement of an end-of-life routing layer carrying all customer traffic for a SaaS platform serving ~16,000 tenant sites. Authored the architecture decision records and building the automated configuration service that replaced manual routing management
- Delivered multi region SaaS Platform expansion enabling data-residency compliance and latency optimizations for European and Asia Pacific customers, building and validating the live customer site migration workflow, beta program, and publishing throughput benchmarks that made migration timelines predicable at scale
- Migrated production observability for a business-critical platform off two sunsetting vendor tools, leading the gap analysis that mapped legacy metrics and alerts to replacement equivalents and retiring obsolete monitoring credentials and dashboards
- Automated provisioning, teardown, and recovey workflows accross a fleet of single tenantcy customer environments, optimizing for cost, performance, and reliability while reducing toil and human error
- Established the operational readiness standards governing how the team assumes on-call ownership of new services, defining alert quality gates, escalation rules, runbook requirements, and mandatory knowledge transfer sessions
- Authored 13 production incident postmortems accross 6 years and modernized 11 service runbooks, including operational documentation review for partner engineer teams
- Quantified 2,400 in monthly avoidable logging spend accross 15,900 tenants and automated remediation, with modeleted annual savings of approximately $29,000