Skip to content
GetuJobs

AI Data Engineers

  • Infosys Limited
  • Noida
  • 3 - 5 Years
  • Full Time
  • Artificial Intelligence - BASIC
  • Generative AI for Data Analytics
  • Machine Learning

Posted October 8, 2026 applications close November 7, 2026


Job Description

Responsibilities

Data Engineering & AI Platform Development

Design and develop scalable batch and real-time data pipelines using Databricks, Spark, and cloud-native services.

Build and maintain enterprise-grade data lakes, lakehouses, and data warehouses.

Create ingestion frameworks for structured, semi-structured, and unstructured data.

Develop AI-ready data pipelines supporting LLM, RAG, Agentic AI, and predictive analytics use cases.

Implement metadata management, data governance, lineage, and data quality controls.

AI & Generative AI Enablement

Prepare and transform data for LLM training, fine-tuning, and inference workloads.

Build Retrieval-Augmented Generation (RAG) pipelines integrating vector databases and enterprise knowledge sources.

Develop data workflows supporting DBRX and other foundation models.

Optimize AI data pipelines for performance, scalability, and cost efficiency.

Cloud & Data Platform Engineering

Design solutions on AWS, Azure, or GCP environments.

Leverage cloud-native data services for ingestion, storage, orchestration, and monitoring.

Implement CI/CD and Infrastructure as Code (IaC) practices for data platforms.

Ensure security, compliance, and governance across cloud data ecosystems.

Additional Responsibilities

Bachelor’s or Master’s degree in Computer Science, Data Engineering, Information Technology, Artificial Intelligence, or related field.

Key Skillset – Databricks, DBRX, PySpark, Spark, Python, Snowflake, AWS, Azure, GCP, Delta Lake, Unity Catalog, MLflow, RAG, LangChain, Vector Database, Pinecone, Feature Store, Airflow, CI/CD, Docker, Kubernetes, Data Lakehouse, GenAI, LLM, Data Engineering.

Technical and Professional Requirements

Experience building enterprise GenAI or AI Agent solutions.

Exposure to DBRX foundation models and Databricks AI platform.

Hands-on experience with vector search and semantic retrieval systems.

Knowledge of DataOps, MLOps, and LLMOps best practices.

Experience with multi-cloud data architectures.

Understanding of Responsible AI and model governance frameworks.

Preferred Skills

  • Machine Learning
  • Artificial Intelligence – BASIC
  • Generative AI for Data Analytics

Educational Requirements

Bachelor of Engineering,Bachelor Of Technology (Integrated)

AI Data Engineers Sign in to apply

Browse Job Vacancies