AI Data Engineers
-
Infosys Limited
- Noida
- 3 - 5 Years
- Full Time
- Artificial Intelligence - BASIC
- Generative AI for Data Analytics
- Machine Learning
Posted October 8, 2026 applications close November 7, 2026
Please sign in or register for free to apply.
Job Description
Responsibilities
Data Engineering & AI Platform Development
Design and develop scalable batch and real-time data pipelines using Databricks, Spark, and cloud-native services.
Build and maintain enterprise-grade data lakes, lakehouses, and data warehouses.
Create ingestion frameworks for structured, semi-structured, and unstructured data.
Develop AI-ready data pipelines supporting LLM, RAG, Agentic AI, and predictive analytics use cases.
Implement metadata management, data governance, lineage, and data quality controls.
AI & Generative AI Enablement
Prepare and transform data for LLM training, fine-tuning, and inference workloads.
Build Retrieval-Augmented Generation (RAG) pipelines integrating vector databases and enterprise knowledge sources.
Develop data workflows supporting DBRX and other foundation models.
Optimize AI data pipelines for performance, scalability, and cost efficiency.
Cloud & Data Platform Engineering
Design solutions on AWS, Azure, or GCP environments.
Leverage cloud-native data services for ingestion, storage, orchestration, and monitoring.
Implement CI/CD and Infrastructure as Code (IaC) practices for data platforms.
Ensure security, compliance, and governance across cloud data ecosystems.
Additional Responsibilities
Bachelor’s or Master’s degree in Computer Science, Data Engineering, Information Technology, Artificial Intelligence, or related field.
Key Skillset – Databricks, DBRX, PySpark, Spark, Python, Snowflake, AWS, Azure, GCP, Delta Lake, Unity Catalog, MLflow, RAG, LangChain, Vector Database, Pinecone, Feature Store, Airflow, CI/CD, Docker, Kubernetes, Data Lakehouse, GenAI, LLM, Data Engineering.
Technical and Professional Requirements
Experience building enterprise GenAI or AI Agent solutions.
Exposure to DBRX foundation models and Databricks AI platform.
Hands-on experience with vector search and semantic retrieval systems.
Knowledge of DataOps, MLOps, and LLMOps best practices.
Experience with multi-cloud data architectures.
Understanding of Responsible AI and model governance frameworks.
Preferred Skills
- Machine Learning
- Artificial Intelligence – BASIC
- Generative AI for Data Analytics
Educational Requirements
Bachelor of Engineering,Bachelor Of Technology (Integrated)