Data Engineer - AI Labs

IDFC FIRST Bank · Bengaluru

  • Experience2+ yrs
  • SalaryNot disclosed
  • Work modeonsite
  • Levelmid
  • Posted1 Oct 2026

About IDFC FIRST Bank

IDFC FIRST Bank is hiring in Bengaluru in financial services. This role looks for around 2+ years of experience.

Skills

  • Machine Learning
  • Data Architecture
  • ETL Design
  • Cloud Platforms
  • Big Data Technologies
  • API Integration
  • Git
  • Data Governance
  • Predictive Analytics
  • Data Visualization
  • Programming

The role

A data engineer at a banking and financial services company builds and optimizes machine learning pipelines for generative AI solutions across structured and unstructured data, applying data architecture and ETL design. The role integrates cloud platforms and big data technologies to support data scientists and reliable model applications.

Full job description

Job Requirements

About the Role

The Data Engineer GenAI within the Data & Analytics function will work closely with Data Scientists to support the development of generative AI solutions across text, audio, image, and tabular data domains. This role is responsible for managing large volumes of structured and unstructured data, ensuring its efficient storage, retrieval, and augmentation to power GenAI models and applications. The ideal candidate will be skilled in building robust data pipelines, optimizing data architecture, and ensuring high standards of data reliability and governance.

Key Responsibilities

Primary Responsibilities

Build and maintain data engineering pipelines, with a focus on unstructured data.Conduct requirements gathering and scoping sessions with business users and stakeholders to define GenAI-related data needs.Design, build, and optimize data architecture and ETL pipelines for accessibility by Data Scientists and GenAI products.Manage the full data lifecycle: ingestion, transformation, and consumption.Ensure high standards of data reliability, integrity, and governance.Work with APIs to enable seamless data integration and usability.Create technical design documentation for data pipelines and projects.Debug technical issues and manage code versioning using Git.Demonstrate experience with big data infrastructure such as MapReduce, Hive, HDFS, YARN, HBase, MongoDB, DynamoDB, etc.

Secondary Responsibilities

Apply machine learning and predictive analytics techniques where relevant.Leverage domain knowledge in banking or financial services to enhance data solutions.Present data insights using effective storytelling and visualization techniques.

What We Are Looking For

Education

Bachelor’s or master’s degree in computer science or Data Engineering or Information Systems or a related field.

Experience

2+ years of Proven experience in building and managing data pipelines and architectures.Hands-on experience with big data technologies and cloud platforms (AWS, GCP, or Azure).Exposure to GenAI applications and working with unstructured data formats.

Skills and Attributes

Strong programming and debugging skills.Proficiency in data architecture, ETL design, and pipeline optimization.Familiarity with API integration and cloud services.Ability to work collaboratively with cross-functional teams.Excellent documentation and communication skills.Strong problem-solving mindset and attention to detail.Ability to deliver high-quality outputs under tight timelines.