Lakshmi Narayana Vommi
Senior Resilience Engineer @Goodnotes
Signup · Get unlimited contacts
WORK HISTORY
Senior Resilience Engineer @Goodnotes
Belfast, GB
Led resilience assessments for critical distributed services by analyzing architecture, dependencies, traffic patterns, and failure modes.•Identified key resilience gaps including SQS/Redis coupling, missing timeouts, unlimited retries, and lack of fault isolation across service dependencies.•Drove engineering improvements such as controlled retry strategies, fallbacks, session caching, and observability-led debugging to improve stability and recovery.•Used incident patterns, production telemetry, and traffic behavior to uncover systemic weaknesses and drive targeted resilience improvements.•Partnered with engineering teams to translate resilience findings into concrete service hardening actions across retries, dependency isolation, graceful degradation, and recovery behavior.
EDUCATION
Birla Institute of Technology and Science, Pilani
Master's degree, Software Systems
Vellore Institute of Technology
B.Tech, Mechanical
SKILLS
ABOUT LAKSHMI NARAYANA VOMMI
Engineering leader with 20+ years of experience across resilience engineering, platform reliability, performance engineering, and cloud infrastructure. I have led cross-functional teams building, scaling, and hardening distributed systems across AWS, GCP, Azure, and IBM Cloud. My work spans production readiness, incident analysis, dependency failure testing, scaling strategy, CI/CD quality gates, observability, and service resilience for cloud-native SaaS platforms. I focus on turning incidents, traffic patterns, alerts, and bottlenecks into engineering improvements that reduce failure risk and improve recovery. Recent work has included resilience hardening across Redis, SQS, Envoy, CockroachDB, auth proxies, and microservice traffic paths through circuit breakers, retry controls, graceful degradation, startup dependency isolation, health checks, timeout tuning, scaling policies, and resiliency validation in CI. Core focus areas: Resilience Engineering, Platform Reliability, Performance Engineering, Chaos Engineering, Kubernetes, Observability, Incident Management, Distributed Systems, CI/CD, and Cloud Infrastructure.
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.