One million success stories. Start yours today.

Direct from employer

Data Engineer - DaAI

Date Posted: Jul 07, 2026

Job Detail

  • location_on
    Location Bengaluru, Karnataka, India
  • desktop_windows
    Job Type: Full Time/Permanent
  • schedule
    Shift:
  • analytics
    Career Level:
  • group
    Positions:
  • calendar_view_day
    Experience:
  • male
    Gender: No Preference
  • school
    Degree:
  • calendar_month
    Apply Before: Nov 14, 2026

Job Description

As a Data Engineer, you will help build the data foundation for our agentic AI platform. You will work with senior data architects, AI/ML engineers, and platform engineers to implement data ingestion, transformation, profiling, enrichment, validation, and preparation pipelines across structured and unstructured enterprise data sources.

This is a hands-on engineering role for someone who enjoys working with real-world enterprise data, building reliable pipelines, writing robust Python and SQL, and helping convert raw enterprise information into AI-ready data assets.

Responsibilities

  • Build and maintain data ingestion pipelines for structured enterprise systems such as ERP, CRM, billing, finance, HR, OSS/BSS, ServiceNow, Salesforce, SAP, Oracle, databases, and APIs.
  • Build pipelines for unstructured and semi-structured data sources such as documents, emails, logs, transcripts, PDFs, spreadsheets, and media metadata.
  • Develop ETL/ELT workflows using Python, SQL, PySpark, Apache Spark, Airflow, dbt, Dagster, cloud-native services, or equivalent technologies.
  • Support data profiling routines to identify missing values, duplicates, inconsistent formats, incomplete master data, schema changes, and conflicting records.
  • Implement data quality checks using frameworks such as Great Expectations, dbt tests, AWS Glue DataBrew, custom validation scripts, or equivalent tools.
  • Support data labelling, contextualization, harmonization, enrichment, and classification workflows required for AI agent configuration.
  • Prepare data outputs for downstream AI consumption, including embeddings, metadata, semantic tags, graph-ready datasets, and retrieval-ready document chunks.

Technical requirements

  • Working knowledge of data pipeline development using PySpark, Apache Spark, Airflow, dbt, Dagster, or equivalent technologies.
  • Experience working with structured data from databases, APIs, enterprise applications, data lakes, warehouses, or lakehouse platforms.
  • Exposure to cloud data platforms such as Databricks, Snowflake, BigQuery, Azure Data Lake, AWS S3, Google Cloud Storage, or equivalent platforms.
  • Understanding of data modelling, schema design, joins, keys, relationships, data validation, and data quality concepts.
  • Practical experience with data profiling, cleansing, transformation, and reconciliation.
  • Familiarity with Git, CI/CD basics, unit testing, and production-grade engineering practices.

Preferred skills

  • Technology › Big Data - Data Processing › PySpark
  • Technology › Big Data - Data Processing › Spark › Apache Storm

About the role

  • Role: Senior Consultant
  • Experience: 7 - 12 years
  • Education: Bachelor of Engineering

Source: Infosys careers — Read the original posting and apply

This role is listed by the employer on its own careers site. Kaam Ki Khoj does not process applications for it.

Company Overview

Infosys
Infosys · Bengaluru, Karnataka, India
1,811 open roles

Infosys is an employer in the IT/Computers - Software, Software Services sector with operations in India. This is a directory listing maintained by Kaam Ki Khoj so that candidates can find the organisation; it is not an official company page and Kaam... Read More

Related Jobs

One upload, every employer

Let the right employer find you

Upload your CV once — we read it, build your profile and put you in front of every employer hiring on Kaam Ki Khoj. No forms, no fees.

Upload your CV Browse jobs