…

Senior Data Engineer

LTIMindtree · IT Services & Consulting

  • Pune, India
  • On-site
  • Posted today
  • Data & AI

About the job

Role Summary

Design develop test and deploy data pipelines that ingest transform and serve data across the Sirius enterprise data platform

Working under the Data Tech Leads direction they implement the reusable ingestion framework canonical data models and analyticsready datasets that power marketing attribution claims analytics executive dashboards and AIML workloads

Key Responsibilities

Build and maintain ETLELT pipelines using Snowflake services covering data extraction from multiple sources Marketing Platforms PartnerDistribution Feeds PAS Claims Systems

Implement the reusable ingestion framework with API connectors batch connectors and standardized data quality checks

Develop and optimize Snowflake objects tables views stored procedures Snowpipe Streams Tasks Dynamic Tables

Implement schemaonwrite data models aligned to the canonical data domain sequence Marketing Distribution Claims PricingUW

Build data pipelines in Parquet format with Apache Iceberg support for lakehouse integration

Develop Databricks notebooks and Spark jobs for advanced analytics and data science workbench use cases where Databricks is in scope

Integrate data from diverse sources

Implement data quality cleansing validation and enrichment logic within pipelines

Collaborate with the Data Governance Liaison on lineage catalog registration and compliance tagging Informatica

Write unit tests and partner with the Unit Test Engineer to maintain test coverage meeting ART quality gates

Follow CICD practices for pipeline deployment coordinate with the shared DevOps Engineer

MustHave Skills Qualifications

5 -8 years of handson data engineering experience with at least 2 years on Snowflake

Strong SQL skills complex queries performance tuning Snowflakespecific SQL

Proficiency with Python or Scala for data transformation and pipeline scripting

Experience with cloud platforms Azure ADF ADLS Azure DevOps andor AWS S3 Glue Lambda

Understanding of data modeling star schema snowflake schema data vault canonical models

Familiarity with Parquet Avro JSON CSV data formats and Apache Iceberg table format

Experience working in AgileSAFe teams with 2week sprint cadence

Good understanding of data quality principles and data testing frameworks

GoodtoHave Preferred

Experience building ETLELT pipelines batch and streaming using tools like Vincilium Informatica dbt or equivalent

Handson Databricks Apache Spark experience PySpark Delta Lake

Experience with Snowpipe Streaming Kafka or eventdriven data ingestion patterns

Familiarity with Power BI Tableau and building datasets and views for BI consumption

Insurance or financial services data experience claims policy distribution marketing data

SnowPro Core or SnowPro Advanced Data Engineer certification

Experience with Gitbased version control and CICD for data pipelines