Keshav Ramaiah
Sde Ii (Member of Technical Staff) @Salesforce
Signup · Get unlimited contacts
WORK HISTORY
Sde Ii (Member of Technical Staff) @Salesforce
Bengaluru, IN
Built a cross-region Kafka-based voice event relay system forreal-time SIP session routing, achieving 99.99% uptime and <300msmessage latency.* Designed and implemented AQM (Automated Quality Management)system using LLM-based model scoring to evaluate agentconversations (voice/email/chat) on empathy, tone, and resolution innear real-time — enabling quality coaching for contact center agents.* Developed a common monitoring framework for my org. This is aninstrumentation wrapper for unified logging and metrics emission whichcoordinates Splunk, Argus, Tracer(Zipkin wrapper).* Resolved container crash issues in the EKS-based transcriptionservice by introducing blue-green deployments and configuringterminationGracePeriodSeconds, ensuring seamless voice callexperience during rollouts.* Fixed critical OOMKilled issues in the Recording service caused byconcurrent executions by implementing Reactive Kafka andenforcing concurrency limits with processing thresholds.
EDUCATION
Amrita Vishwa Vidyapeetham
Bachelor of Technology - BTech, Computer Science
Kendriya Vidyalaya
High School, Science
ABOUT KESHAV RAMAIAH
4.5 Years | Backend Engineer | High-Scale Payments & Billing | System Design & Observability | Latency OptimizationI’m a backend engineer with 4.5 years of experience building and maintaining high-scale distributed systems in the Payments, Billing, and Invoicing domain (~400 TPS). I specialize in backend architecture, system design, observability, and infrastructure optimization for performance-critical applications.🧩 Domain Expertise:Payments Processing, Gift Cards, Instant Bank Discounts (IBD), Settlements, Taxation, Billing & Disbursement systems, Marketing Tech Focus:Kafka, SQS, AWS Lambda, MAWS, PostgreSQL, MSSQL, multithreading, performance tuning, and fault-tolerant microservices Key Highlights:Operational Excellence: As part of a high-ownership team (100+ services, 16 engineers), I’ve led incident resolution during game-day failures, handling 160K+ failed transactions and driving independent mitigations with real-time stakeholder communication.Latency & Infra Optimization: Performed detailed latency investigations across business logic, DB queries, and infrastructure. Led infra decisions (e.g, Lambda vs MAWS) and drove optimizations through performance testing and tuning.Database Expertise: Spearheaded migrations from MSSQL to PostgreSQL, ensuring zero downtime. Designed complex DB schemas for multi-product systems, balancing reuse and customization, while keeping scalability and future extensibility in mind.System Design & Observability: Architected fault-tolerant services with a strong focus on metrics, alerting, and monitoring to ensure reliability and traceability in production. Looking forward to contributing to high-impact teams building scalable, resilient systems.
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.