sahil jain
sahil jain

Data Engineer- Manager

Actively looking · Member since 29 Sep 2026
Location
Gurgaon, India
Desired salary
Unspecified
Work preference
Remote Only
Experience level
Lead

About

Professional summary

I am a Data Architect and Engineering Leader with over 11 years of experience designing and delivering enterprise-scale data platforms. My expertise centers on Google Cloud Platform, data lakes, data warehouses, lakehouse architectures, and governed analytical data products.

I build scalable ELT and streaming solutions using BigQuery, Pub/Sub, Cloud Storage, Cloud Functions, Snowflake, dbt, Airflow, Python, SQL, and Spark. I have also worked across AWS environments to support cross-cloud data sharing and modernization initiatives.

I specialize in enterprise data modeling using Kimball and Data Vault methodologies. I translate complex business requirements into logical and physical data models, reusable data standards, and reliable data pipelines for analytics, AI/ML, and personalization use cases.

I have led teams of data engineers through architecture definition, sprint planning, code reviews, mentoring, and delivery. I collaborate closely with analytics, product, GTM, and data science stakeholders to create self-service, high-quality data platforms.

I am experienced in data governance, security, observability, lineage, IAM, encryption, and data quality frameworks. I also drive Infrastructure as Code and CI/CD practices using Terraform and automated testing to improve platform reliability and deployment efficiency.

My background spans advertising technology, healthcare, financial trading, and financial services. I am passionate about cloud modernization, distributed data systems, cost optimization, and building secure platforms that generate measurable business value.

Skills

33 capabilities

Tech stack & tools

Working toolkit

Data Stores

Languages & Frameworks

Experience

Career history

Data Architect Omnicom (Annalect)

I own enterprise data architecture strategy for multi-terabyte customer interaction datasets, defining data lake and lakehouse patterns that support LLM, generative AI, real-time personalization, and global analytics initiatives.

I lead a team of 11 data engineers across architecture definition, sprint planning, code reviews, mentoring, and platform delivery. I architect cloud-native ELT platforms centered on BigQuery with Snowflake, Snowpipe, Streams, dbt, Airflow, Python, and SQL.

I design Kimball and Data Vault models, implement governance, lineage, monitoring, IAM, and encryption practices, and partner with Analytics, Product, and GTM teams on self-service data warehouses and dashboards. I also drive Terraform-aligned Infrastructure as Code and CI/CD adoption.

Lead Data Engineer RxAdvance

I designed enterprise-scale data platform architecture for high-volume healthcare transaction processing and near-real-time analytics. I built scalable ETL and ELT pipelines using Python, SQL, Spark, Hive, HDFS, Airflow, and cloud-native data lake patterns.

I led modernization of legacy Hadoop workloads by migrating Hive and HDFS datasets to cloud-native analytical architectures. I optimized distributed processing and query performance, developed API-based integration frameworks, and established Terraform and CI/CD practices.

I collaborated with data science teams to deliver production-ready feature engineering datasets for machine learning models.

Data Engineer ION Group

I developed distributed backend data services and streaming architectures for financial trading experimentation platforms on AWS. I built real-time ingestion pipelines using Kafka, Spark, Hive, HDFS, Lambda, and DynamoDB.

I supported migration of legacy Hadoop datasets into cloud-native analytical environments, improving scalability and operational efficiency. I partnered with analytics and product teams to define KPIs and improve governed data availability.

I also automated end-to-end data quality and regression testing frameworks to ensure reliable experimentation outcomes.

Software / Data Engineer Moody's

I built and maintained internal financial data platforms used by analysts and risk managers for credit ratings, compliance, and reporting. I developed ETL pipelines and automated data validation workflows using Python, SQL Server, and shell scripting.

I created SQL-based dashboards for executive credit insights and portfolio analytics, and worked with global data teams to streamline ingestion and transformation of structured and unstructured financial data.

I supported modernization of legacy analytics services and contributed to governance, security, documentation, and regulatory compliance initiatives.

Education

Learning history

Uttar Pradesh University

Bachelor of Science, Computer Engineering

Bachelor of Science in Computer Engineering.

This professional hasn’t added portfolio projects yet.

This professional hasn’t listed any services yet.

Jobs Talent AI Tools Salaries
Menu