Velmuruga S S
Site Reliability Engineer (SRE) | Ansible Automation Platform (AAP) | Linux | RHEL | Infrastructure Automation | IBM
- Role
- Site Reliability Engineer at IBM
- Location
- Chennai, TN, IN
- LinkedIn followers
- 500 followers
About Velmuruga S S
Site Reliability Engineer (SRE) with 6.5+ years of experience in managing, automating, and optimizing large-scale enterprise Linux infrastructure.Currently working at IBM, specializing in Ansible Automation Platform (AAP), infrastructure automation, and system reliability across development, testing, and production environments.I have hands-on experience managing 600+ Linux servers, with a strong focus on improving system availability, reducing manual operational effort through automation, and ensuring high-performance infrastructure operations.Core Expertise:• Ansible Automation Platform (AAP) – Automation design & implementation • Infrastructure Automation & Configuration Management • Linux / RHEL Administration at scale • Incident Management, Monitoring & System Reliability • Patching, Hardening & Performance Optimization Key Contributions:• Designed and implemented multiple automation solutions using AAP across patching, backup, monitoring, and system validation workflows • Automated end-to-end infrastructure operations, significantly reducing manual effort and improving operational efficiency • Supported and maintained highly available 24x7 production environments at enterprise scale • Contributed to large-scale infrastructure transformation and cloud migration initiatives Recognitions:• Best Performance Award – 2026 (IBM) • Transformation Champion – 2026 (IBM) I am passionate about building reliable, scalable systems and driving automation-led transformation in enterprise environments.Specialized in building automation-driven infrastructure solutions using Ansible Automation Platform (AAP) across enterprise environments.
Experience
Site Reliability Engineer
Jul 2025 — Present · Chennai, IN
Manage reliability, availability, and performance of large-scale enterprise Linux infrastructure (600+ servers) • Administer and optimize Ansible Automation Platform (AAP) for enterprise-wide automation initiatives • Implement proactive monitoring, alerting, and incident response strategies to ensure system stability • Perform root cause analysis (RCA) and implement preventive measures for recurring issues • Improve system performance, capacity planning, and infrastructure scalability • Drive automation-led improvements to reduce operational overhead and enhance efficiency • Collaborate with cross-functional teams to improve system reliability and deployment workflows • Ensure continuous availability of critical production systems in a 24x7 environment • Recognized with Best Performance Award – 2026 and Transformation Champion – 2026
Education
Simplilearn Alumni
Certified ScrumMaster (CSM®)
2020
Chennai Institute of Technology
Bachelor of Engineering (BE), Electrical and Electronics Engineering
2015 — 2019
University of Madras
Master of Business Administration - MBA, Human Resources Management/Personnel Administration, General
2019 — 2021
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.