Sri Harshithrao N.

πŸš€ Cloud & Big Data Engineer | Expertise in AWS, Azure, Snowflake, and Spark (PySpark, Scala) | Financial & Healthcare Data Solutions

Role
Data Engineer at Capital One
Location
Dallas, TX, US
LinkedIn followers
500 followers
Information TechnologyView LinkedIn profile

About Sri Harshithrao N.

Data Engineer with 4+ years of experience building scalable, secure, and cloud-native data pipelines across AWS, Azure, and Snowflake ecosystems. Proven ability to architect and optimize end-to-end ETL workflows using PySpark, SQL, and Python for both real-time and batch processing.Specialized in performance tuning, CI/CD automation, and data modeling in highly regulated industries like healthcare (HIPAA) and finance (FINRA/SEC).Experienced in:β€’ Cloud Platforms: AWS (S3, Glue, Redshift, Lambda), Azure (ADF, Databricks, Synapse)β€’ Big Data & ETL: Spark (PySpark, Scala), Snowflake, Kafka, Kinesisβ€’ DevOps & API: Terraform, Jenkins, Flask, FastAPI, GitHub ActionsPassionate about delivering business-driven insights and improving data availability, quality, and compliance at scale. Open to solving complex data problems across industries.

Experience

  1. Data Engineer

    Capital One

    Oct 2023 β€” Present Β· Irving, TX, US

    Developed and optimized Azure Data Factory (ADF) pipelines to securely ingest and transform on-prem healthcare data into Azure Data Lake Storage (ADLS). Re-engineered SQL Server ETL workflows using Databricks (PySpark, Scala) and Snowflake Tasks/Streams, reducing processing time by 40%. Built real-time streaming pipelines with Azure Stream Analytics, Kafka, and ADF triggers to ensure low-latency access to patient and provider data. Tuned Databricks clusters with autoscaling and monitoring, reducing costs by 30%.Improved SQL-driven ETL logic using CTEs, window functions, indexing, and partitioning, achieving 98% data consistency and cutting integration time by 40%. Integrated Azure Key Vault for secure authentication and enforced HIPAA-compliant encryption across pipelines. Built REST APIs with Flask and Play Framework for secure access to data from Cassandra.Implemented CI/CD pipelines with Azure DevOps, GitHub Actions, and Terraform to streamline deployments. Adopted Delta Lake architecture on Databricks for versioned, scalable, and ACID-compliant data storage. Set up automated monitoring and alerting using Azure Logic Apps and Snowflake Alerts, supporting real-time issue resolution and compliance tracking.Configured Linked Services across Azure SQL, ADLS, Blob Storage, and REST APIs to enable seamless data flow. Built resilient error handling in ADF using retry logic, logging, and exception management, ensuring 99.9% pipeline reliability. Collaborated with architects, analysts, and global teams to align Azure-based healthcare solutions with business goals. Designed and optimized Synapse pipelines to process EHR, claims, and provider data, improving speed by 35%. Developed Synapse warehouses integrating FHIR, HL7, and billing data, reducing report generation time by 40%.

Education

  • University of Central Missouri

    Master of Science - MS, BIG DATA ANALYTICS

  • Osmania University

    Bachelor of Engineering - BE, Computer Science

Find verified contacts for anyone on LinkedIn

Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.

Free plan included Β· No credit card required

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.

Sri Harshithrao N. β€” Data Engineer at Capital One in Dallas, TX, US | Unifers