Big Data Engineer II

MetLife · Pune Division

  • Experience5–8 yrs
  • SalaryNot disclosed
  • Work modeonsite
  • Posted30 Sept 2026

About MetLife

MetLife is hiring in Pune Division in insurance. This role looks for around 5+ years of experience.

Skills

  • SQL
  • Python
  • Scala
  • HBase
  • Cosmos DB
  • ETL pipeline development
  • Apache Spark
  • Hadoop
  • Hive
  • Azure Data Factory
  • Event Hubs
  • Azure Functions
  • Azure Synapse
  • Databricks
  • Data warehouses
  • Data marts
  • Data lakes
  • Medallion architecture
  • Real-time data processing
  • Batch data processing
  • Streaming pipelines

The role

A data engineer at an insurance global capability center builds production-grade data and analytics solutions, designing Azure data engineering and Apache Spark pipelines for large-scale structured and unstructured data. Big Data frameworks, Python, and SQL support reliable batch and real-time processing across data lakes and warehouses.

Full job description

MGCC GG 10.2 – Data Engineer II

Position Title

MGCC GG 10.2 – Big Data Engineer II

Function, Responsibility Level

Reports to (Responsibility Level): GG 11 or above

Supervises

Location:

Global Grade: 10

Cost Center (85 Series)

Complexity: OSTG

PID/s Load Mapping

Position Summary

MetLife established a Global capability center (MGCC) in India to scale and mature Data & Analytics, technology and operations capabilities in a cost-effective manner and make MetLife future ready. The center is integral to Global Technology and Operations with a focus to protect & build MetLife IP, promote reusability and drive experimentation and innovation. The Data & Analytics team in India mirrors the Global D&A team with an objective to drive business value through trusted data, scaled capabilities, and actionable insights. The operating models consists of business aligned data officers- US, Japan and LatAm & Corporate functions enabled by enterprise COEs- data engineering, data governance and data science

Role Value Proposition

Data Engineer plays a critical role in data and analytics life cycle and significantly contributes to production grade data and analytics solutions. The role requires one to demonstrate Big Data, Engineering and Cloud expertise. It is an individual contributor role, expected to independently function

Job Responsibilities

Design, build, and maintain robust ETL/ELT pipelines on cloud(Azure) or on-prem to collect, ingest and store large volumes of structured and unstructured data for batch/real time processing Monitor, optimize, and troubleshoot data pipelines to ensure reliability, scalability, and performance Ensure data processing, quality, security, and compliance guidelines, policies and standards are followed Collaborate with multiple partners from Business, Technology, Operations and D&A capabilities (Data Governance, Data Quality, Data Modeling, Data Architecture, Data science, DevOps, BI & insights)

Knowledge, Skills And Abilities

Education

Bachelor’s degree in computer science, information technology or equivalent educational qualification

Technical Skills And Experience

5-8+ years of relevant experience SQL, Python/Scala NoSql and distributed databases (Hbase, Cosmos DB) ETL pipleine development Big Data Frameworks : Apache Spark, Hadoop, Hive Cloud platforms: Azure data factory, Eventhub, Azure functions, Synapse, Databricks Datawarehouses, data marts, data lakes Medallion architecture Performance tuning, optimization, and data quality validation Real-time and batch data processing , streaming pieplines with SparkCommunication skills, analytical skills, structured problem-solving skills.,Partner, Stakeholder engagement experience

Good To Have (Preferred)

DevOps practices: Git, AzureDevops, CI/CD pipelines Unix shell scripting, MongoDB, Nifi Exposure to Gen AI technology and tools