Gaurhari Dass
Senior Data Engineer @Quantexa
Signup · Get unlimited contacts
WORK HISTORY
Senior Data Engineer @Quantexa
London, GB
End-to-End Pipeline Ownership: Architected and monitored a fraud detection pipeline ingesting 1.5M financial transactions/sec via Kafka, transforming with Spark Scala/PySpark, and storing in Snowflake, achieving 99.9% uptime using Kubernetes and Terraform. Reduced false positives by 20% with ML-based anomaly detection (Spark MLlib, Isolation Forest). • Robust Data Pipelines: Orchestrated workflows with Airflow DAGs, integrating DBT for data transformation and NQL for domain-specific queries, processing multilingual financial data, monitored via Prometheus and Grafana. • Data Modeling: Designed sharded MongoDB schemas and Elasticsearch indexes for sub-second fraud pattern queries, improving search performance by 30%. Modeled complex relationships using Quantexa’s.qmodel/.qentity, enhancing detection accuracy by 15%. • AI Innovation: Integrated LLM and RAG with Hugging Face and vector databases in Azure, linking entities for enriched fraud insights, boosting contextual analytics accuracy. • Collaboration: Partnered with product managers and analysts, delivering testable Python FastAPI REST services via Jenkins CI/CD, mentoring an 8-member team on scalable architectures.
EDUCATION
Swami sant dass public school
10th
NIT Jalandhar
Bachelor of Technology - BTech
NIT Jalandhar
Btech
Dayanand Model School
12th
ABOUT GAURHARI DASS
Seasoned technologist with over 14 years of expertise in big data, cloud architecture, web development and machine learning, delivering transformative solutions across industries. As a Senior Data Engineer at Quantexa in London, I worked fusion-based systems for fraud detection, processing millions of events. At EMBL-EBI as a Senior Cloud Architect, I led multi omics pipelines, handling petabyte-scale data and cutting processing times by 40% with Spark and Kafka, ElasticSearch, Mongodb, Kibana. My career spans Samsung, Rivigo, and IntelligenceNode, where I’ve built analytics pipelines, optimized logistics with ML, and enhanced retail categorization—boosting engagement, efficiency, and accuracy. Proficient in Apache Spark, Kafka Streams, Kafka Connect, Flink, Beam, Scala, Python, Java, C#, HBase, MongoDB, Elasticsearch, Airflow, dbt, GCP Pub/Sub, AWS, Hadoop (MapReduce, Pig, Sqoop), Spring Apis and REST APIs, I excel at designing and integrating cutting-edge solutions for complex challenges. Certified in MongoDB and dbt, I’m passionate about driving innovation and excellence. As a contributor to OmicsDI, I’m based in Cambridge, UK—open to connecting on big data, real-time systems, RESTful architectures, and ML opportunities!
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.