Sandeep Yannam
Data Engineering Associate @Cognizant
Signup · Get unlimited contacts
WORK HISTORY
Data Engineering Associate @Cognizant
Hyderabad, IN
Designed incremental load strategy to optimize ETL performance based on data changes.• Incremental load strategy by first ensuring data uniqueness; loaded new records only when key attributes (e.g, date fields) changed, preserving historical versions with start_date, end_date, and is_current flags.• Automated data pipelines for daily incremental loads using Pyspark and Python.• Designed and implemented a Sales & Energy Consumption Analytics data model for Florida Power & Light (FPL) by converting raw operational sales and meter-usage data into a structured star-schema consisting of a centralized fact table and multiple supporting dimension tables (customer, meter, billing cycle, geography, etc.).• Developed a Large Language Model (LLM) in Amazon Bedrock for extracting information from a knowledge base in AWS S3, utilizing Retrieval-Augmented Generation (RAG) to enhance accuracy with real-time retrieval.
EDUCATION
B V Raju Institution of Technology
Bachelor of Engineering, Computer Science Engineering
ABOUT SANDEEP YANNAM
Data Engineer with 3+ years of professional experience designing, building, and optimizing scalable data pipelines and analytics platforms on AWS. Strong hands-on expertise in data engineering, ETL/ELT frameworks, and big data processing, with a focus on building reliable, high-performance data solutions for analytics and business intelligence.Experienced in working across the full data lifecycle—from data ingestion and transformation to storage, optimization, and consumption—using cloud-native and distributed technologies. Adept at handling large-scale datasets, performance tuning, schema evolution, and production support in fast-paced enterprise environments.Driven by a strong interest in data analytics and data science, and currently preparing to deepen technical and analytical expertise through advanced studies.Data Engineering & Big Data:Apache Spark (PySpark), Databricks, AWS Glue, ETL/ELT Pipelines, Batch Processing, Data Quality & ValidationCloud & Data Platforms:AWS S3 (Data Lake), Amazon Redshift, AWS Glue Catalog, IAM, Cloud-based Analytics ArchitecturesProgramming & Querying:Python, SQL, PySpark, Window Functions, Query OptimizationArchitecture & Design:Data Lake Architecture, Medallion Architecture (Bronze/Silver/Gold), Schema Evolution, SCD Type 2, Incremental Loads, Upserts & Merge StrategiesBuild & CI/CD: Git, Job Scheduling, Production Deployments, Monitoring & TroubleshootingTools:Databricks, AWS Console, Informatica PowerCenter, SQL Developer, VS CodeDomains:Enterprise Data Platforms, Analytics & Reporting SystemsProfessional Strengths:End-to-End Data Pipeline Ownership, Production Support, Cross-Team Collaboration, Problem Solving, Performance OptimizationGen AI Tools:Amazon Bedrock, RAG, Amazon-titan Embedding LLM, Anthropic Claude LLM
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.