Rajkumar Ponraj

Sr Manager Site Reliability Engineering Aws Dynatrace Azure @BMO

Toronto, CA
MOBILE NUMBERS
+91 *********19

Signup · Get unlimited contacts

WORK HISTORY

Dec 2023 — Present

Sr Manager Site Reliability Engineering Aws Dynatrace Azure @BMO

View department →

CA

I lead and coordinate a broad ecosystem of engineering teams and external partners, collaborating closely with Finance portfolio to translate business needs into a unified technology strategy that accelerates enterprise goals.• I provide strategic leadership to the Deployment function, managing 23 diverse resources and supervising the full Finance application portfolio—spanning legacy systems and multi‑cloud platforms on AWS and Azure—to ensure all deployments meet enterprise resiliency and reliability standards.• Built an end-to-end observability framework for AWS MWAA, Glue, Lambda, Redshift, and Aurora, enabling real-time visibility across ETL pipelines, serverless workloads, and scheduler performance.• Reduced incident resolution time by 40–50% through actionable SLO dashboards, health metrics, and OpenTelemetry‑based tracing.• Designed a single‑pane‑of‑glass Dynatrace view by ingesting CloudWatch metrics and logs via OpenTelemetry.• Implemented Monitoring‑as‑Code (CloudFormation) for alarms, dashboards, SLOs, and traces across multi-account AWS environments.• Developed SLIs/SLOs and error‑budget policies for MWAA DAGs, Glue ETL jobs, and Aurora clusters to enforce reliability governance.• Enhanced MWAA reliability by monitoring worker scaling, queue depth, DAG timing, retries, and event‑driven SLA alerts.• Built auto‑healing automation using Lambda and SSM to remediate Aurora failovers, deadlocks, long queries, and connection exhaustion.

EDUCATION

2002 — 2006

Anna University Chennai

Bachelor's degree, Electrical and Electronics Engineering

ABOUT RAJKUMAR PONRAJ

I am a seasoned Site Reliability Engineering Leader with deep expertise in AWS Observability, Cloud Reliability, and Platform Engineering, specializing in data and serverless services such as MWAA (Airflow), Redshift, Glue, Lambda, Aurora, and enterprise monitoring platforms including Dynatrace.I design and deliver scalable, automated, and highly observable platforms that accelerate engineering velocity, reduce MTTR, improve system reliability, and provide real‑time visibility across distributed systems and data pipelines.With 16+ years of experience across SRE, strategic leadership, Cloud, DevOps, Infrastructure Engineering, and Automation, I have led major modernization programs including:• Resource management on Deployment & Production support for mission‑critical financial systems• Full‑stack Observability-as-Code frameworks• Unified, single‑pane‑of‑glass dashboards• SLO/SLI and error–budget governance• Drift detection & self‑healing automation• Enterprise CI/CD pipelines across cloud and on‑prem• Large‑scale patching, server automation, and configuration standardizationI am passionate about building reliable platforms, enabling developer productivity, and mentoring teams to adopt cloud-native and automation-first practices. My leadership style focuses on collaboration, clarity, ownership, and continuous improvement.

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.

Rajkumar Ponraj — Sr Manager Site Reliability Engineering Aws Dynatrace Azure at BMO in Toronto, CA | Unifers