Talhaa A.
Data Engineer @Sky
Signup · Get unlimited contacts
WORK HISTORY
Data Engineer @Sky
Isleworth, GB
Build and operate batch and near real-time telemetry pipelines in Databricks using PySpark and Delta Lake, delivering curated, query-optimized datasets for product quality and incident triage.Design robust data models (partitioning, clustering, incremental loads, schema evolution) to support high-volume event analytics with predictable latency and cost.Implement data quality controls (validation rules, anomaly detection, lineage, SLAs) to ensure trust in KPIs used across Sky Group and partner stakeholders.Develop reusable frameworks and utilities for marker parsing, normalization, and correlation analysis, accelerating new feature monitoring and reducing time-to-insight.Optimize cloud spend and workload performance through storage layout, compute sizing, and query tuning, including ~$500k annual savings in GCP through pipeline and query optimization.Partner with engineering, operations, and product teams to define telemetry requirements, close instrumentation gaps, and productionize new metrics end-to-end.Own and improve reporting and dashboards (definition, governance, and adoption), ensuring consistent metrics and clear operational thresholds.
EDUCATION
University of Portsmouth
Bsc (Hons) - Computer Science
ABOUT TALHAA A.
Data Engineer at Sky with 5+ years’ experience building production-grade data platforms and observability for large-scale telemetry. I design and run end-to-end pipelines and analytics products using PySpark, Databricks and Delta Lake, alongside SQL and Python, turning raw event data into trusted datasets and KPIs for product quality, triage and decision-making. I work closely with engineering and operations teams to define instrumentation, close telemetry gaps, validate end-to-end data flows, and ensure dashboards and metrics remain reliable under change.Previously at Capgemini, I delivered data engineering and business analysis for external clients, including Waitrose, building ETL processes (AWS, Alteryx), implementing data quality and testing frameworks, and supporting end-to-end delivery across branded and own-label product domains. I’m strongest in SQL, Python, Spark, GCP and AWS, and I’m particularly focused on performance tuning, cost optimisation (including initiatives delivering ~US$500k annual cloud savings), and pragmatic data governance in agile teams.
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.