D Gopala Krishna Gunji
Application & Production Support Engineer | SRE / Systems Engineer | Monitoring & Observability | Automation | Incident & Problem Management (ITIL) | Trading Platforms
- Role
- Production Support Site Reliability Engineer at PayPal
- Location
- New Hyde Park, NY, US
- LinkedIn followers
- 500 followers
About D Gopala Krishna Gunji
With over 8 years of experience in Application Support, Production Support, and Site Reliability Engineering, I specialize in ensuring mission-critical systems perform seamlessly under pressure. My passion lies in maintaining uptime, optimizing performance, and driving operational excellence across hybrid cloud infrastructures.(Linux/Unix, Windows)I thrive in high-velocity enterprise environments where resilience, collaboration, and technical precision define success. Leveraging tools like AppDynamics, Splunk, Datadog, and Dynatrace, I proactively identify issues before they impact users, ensuring 24/7 service reliability. My expertise spans incident management, automation, observability, and infrastructure optimization, backed by strong foundations in Linux/Unix systems, SQL databases, and AWS/Azure/GCP cloud environments.A firm believer in continuous improvement, I combine ITIL, DevOps, and SRE principles to streamline operations, reduce mean time to recovery (MTTR), and enhance user experience. Whether leading war-room resolutions or automating deployment pipelines with Jenkins, Ansible, and Terraform, I’m driven to make systems more stable, predictable, and efficient.Core Strengths: Incident & Problem Management | Automation | Monitoring & Observability | Infrastructure Reliability | Root Cause Analysis | Continuous ImprovementOperating Systems: Linux (RHEL, Ubuntu), UNIX, Windows ServerCloud & Infrastructure: AWS (EC2, RDS, S3, CloudWatch), Azure, GCP (GKE, Pub/Sub)Containerization & IaC: Docker, Kubernetes, Helm, Terraform, Ansible, CloudFormationMonitoring & Observability: Splunk, AppDynamics, Dynatrace, Datadog, Grafana, Prometheus, ELKIncident & Problem Management: ServiceNow, Jira, PagerDuty, Opsgenie, ITIL v4, RCAAutomation & CI/CD: Shell, Python, PowerShell, Jenkins, GitLab CI/CD, Control-M, Azure DevOps, BigPanda, Tidal. Databases: Oracle, MySQL, PostgreSQL, MongoDB, Cassandra, DB2Messaging & Middleware: Kafka, WebSphere, JBoss, TomcatSecurity & Compliance: SIEM (Splunk ES, Sentinel), GDPR, SOX, DORA, HIPAACollaboration & Reporting: Confluence, Slack, MS Teams, Zoom.
Experience
Production Support Site Reliability Engineer
Oct 2023 — Present · San Jose, CA, US
As a Production Support and Site Reliability Engineer at PayPal, I provide 24/7 operational support for the company’s global payments infrastructure, ensuring continuous availability, low latency, and secure transaction processing across millions of daily transactions. I work extensively across AWS and GCP environments, maintaining large-scale, cloud-native systems that power PayPal’s digital payments ecosystem.My responsibilities include managing enterprise observability and incident response using Datadog, Dynatrace, Grafana, Kibana, Splunk, BigPanda, and Tidal. I define and monitor key SLOs, SLIs, and SLAs aligned with PayPal’s operational goals and compliance standards, ensuring consistent service reliability. By engineering proactive dashboards, synthetic monitoring, and intelligent alerting through Datadog and BigPanda, I enable early issue detection and faster recovery during incidents.I manage end-to-end incident lifecycles through PagerDuty and ServiceNow, facilitating real-time bridges, escalation management, and RCA reviews under ITIL and SRE frameworks. Using Python and Shell scripting, I automate health checks, deployment validations, and batch monitoring, improving operational efficiency and reducing mean time to recovery. I also oversee Tidal automation schedules for payment reconciliation and time-sensitive data integrations.Collaborating with SRE, DevOps, and development teams, I continuously improve platform performance, container reliability, and CI/CD deployment stability. Through data-driven reporting in Datadog, Grafana, and Jira, I deliver actionable insights into incident trends and service performance. My work ensures PayPal’s mission-critical payment services remain resilient, scalable, and compliant with the highest operational standards.Technologies: Datadog, Dynatrace, Grafana, Splunk, BigPanda, Kibana, Tidal, Prometheus, PagerDuty, ServiceNow, Jenkins, Kubernetes, AWS, GCP, Python, Shell, Postgres, Cassandra, Kafka, Redis, ITIL, SRE
Education
Pace Institute of Technology & Sciences
Bachelor's degree, Automobile Engineering
Rivier University
Master of Science - MS, Computer and Information Sciences, General
State Board of Technical Education and Training
Diploma, Computer Science
Pace Institute of Technology & Sciences
Bachelor of Engineering - BE, Computer Science
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.