Velmuruga S S

Site Reliability Engineer (SRE) | Ansible Automation Platform (AAP) | Linux | RHEL | Infrastructure Automation | IBM

Role
Site Reliability Engineer at IBM
Location
Chennai, TN, IN
LinkedIn followers
500 followers
Information TechnologyView LinkedIn profile

About Velmuruga S S

Site Reliability Engineer (SRE) with 6.5+ years of experience in managing, automating, and optimizing large-scale enterprise Linux infrastructure.Currently working at IBM, specializing in Ansible Automation Platform (AAP), infrastructure automation, and system reliability across development, testing, and production environments.I have hands-on experience managing 600+ Linux servers, with a strong focus on improving system availability, reducing manual operational effort through automation, and ensuring high-performance infrastructure operations.Core Expertise:• Ansible Automation Platform (AAP) – Automation design & implementation • Infrastructure Automation & Configuration Management • Linux / RHEL Administration at scale • Incident Management, Monitoring & System Reliability • Patching, Hardening & Performance Optimization Key Contributions:• Designed and implemented multiple automation solutions using AAP across patching, backup, monitoring, and system validation workflows • Automated end-to-end infrastructure operations, significantly reducing manual effort and improving operational efficiency • Supported and maintained highly available 24x7 production environments at enterprise scale • Contributed to large-scale infrastructure transformation and cloud migration initiatives Recognitions:• Best Performance Award – 2026 (IBM) • Transformation Champion – 2026 (IBM) I am passionate about building reliable, scalable systems and driving automation-led transformation in enterprise environments.Specialized in building automation-driven infrastructure solutions using Ansible Automation Platform (AAP) across enterprise environments.

Experience

  1. Site Reliability Engineer

    IBM

    Jul 2025 — Present · Chennai, IN

    Manage reliability, availability, and performance of large-scale enterprise Linux infrastructure (600+ servers) • Administer and optimize Ansible Automation Platform (AAP) for enterprise-wide automation initiatives • Implement proactive monitoring, alerting, and incident response strategies to ensure system stability • Perform root cause analysis (RCA) and implement preventive measures for recurring issues • Improve system performance, capacity planning, and infrastructure scalability • Drive automation-led improvements to reduce operational overhead and enhance efficiency • Collaborate with cross-functional teams to improve system reliability and deployment workflows • Ensure continuous availability of critical production systems in a 24x7 environment • Recognized with Best Performance Award – 2026 and Transformation Champion – 2026

Education

  • Simplilearn Alumni

    Certified ScrumMaster (CSM®)

    2020

  • Chennai Institute of Technology

    Bachelor of Engineering (BE), Electrical and Electronics Engineering

    2015 — 2019

  • University of Madras

    Master of Business Administration - MBA, Human Resources Management/Personnel Administration, General

    2019 — 2021

Find verified contacts for anyone on LinkedIn

Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.

Free plan included · No credit card required

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.

Velmuruga S S — Site Reliability Engineer at IBM in Chennai, TN, IN | Unifers