Senior Data Engineer
LTIMindtree · IT Services & Consulting
- Pune, India
- On-site
- Posted today
- Data & AI
About the job
Role Summary
Design develop test and deploy data pipelines that ingest transform and serve data across the Sirius enterprise data platform
Working under the Data Tech Leads direction they implement the reusable ingestion framework canonical data models and analyticsready datasets that power marketing attribution claims analytics executive dashboards and AIML workloads
Key Responsibilities
Build and maintain ETLELT pipelines using Snowflake services covering data extraction from multiple sources Marketing Platforms PartnerDistribution Feeds PAS Claims Systems
Implement the reusable ingestion framework with API connectors batch connectors and standardized data quality checks
Develop and optimize Snowflake objects tables views stored procedures Snowpipe Streams Tasks Dynamic Tables
Implement schemaonwrite data models aligned to the canonical data domain sequence Marketing Distribution Claims PricingUW
Build data pipelines in Parquet format with Apache Iceberg support for lakehouse integration
Develop Databricks notebooks and Spark jobs for advanced analytics and data science workbench use cases where Databricks is in scope
Integrate data from diverse sources
Implement data quality cleansing validation and enrichment logic within pipelines
Collaborate with the Data Governance Liaison on lineage catalog registration and compliance tagging Informatica
Write unit tests and partner with the Unit Test Engineer to maintain test coverage meeting ART quality gates
Follow CICD practices for pipeline deployment coordinate with the shared DevOps Engineer
MustHave Skills Qualifications
5 -8 years of handson data engineering experience with at least 2 years on Snowflake
Strong SQL skills complex queries performance tuning Snowflakespecific SQL
Proficiency with Python or Scala for data transformation and pipeline scripting
Experience with cloud platforms Azure ADF ADLS Azure DevOps andor AWS S3 Glue Lambda
Understanding of data modeling star schema snowflake schema data vault canonical models
Familiarity with Parquet Avro JSON CSV data formats and Apache Iceberg table format
Experience working in AgileSAFe teams with 2week sprint cadence
Good understanding of data quality principles and data testing frameworks
GoodtoHave Preferred
Experience building ETLELT pipelines batch and streaming using tools like Vincilium Informatica dbt or equivalent
Handson Databricks Apache Spark experience PySpark Delta Lake
Experience with Snowpipe Streaming Kafka or eventdriven data ingestion patterns
Familiarity with Power BI Tableau and building datasets and views for BI consumption
Insurance or financial services data experience claims policy distribution marketing data
SnowPro Core or SnowPro Advanced Data Engineer certification
Experience with Gitbased version control and CICD for data pipelines