U

AI Data Engineer II

UKG
NewPosted today

LOCATION

Bangalore · Hybrid

EXPERIENCE

2 - 5 Years

SALARY

Negotiable

SKILLS REQUIRED

Data validationPythonSQLData PipelinesPySparkPropagatable Data QualityDatabricksLakehouse Concepts

Job description

Overview

Build and support scalable data pipelines and cloud-based data solutions as an AI Data Engineer, using GCP, Python, PySpark, SQL, and cloud data platforms. Apply data engineering skills to emerging AI and Generative AI use cases.

What you'll do

  • Develop and maintain ETL/ELT data pipelines for batch data processing and analytics workloads.
  • Build data transformation solutions using Python, PySpark, and SQL.
  • Work with Databricks to ingest, transform, process, and curate enterprise datasets.
  • Integrate data from relational databases, APIs, files, and cloud storage.
  • Support migration of datasets and pipelines from legacy or existing platforms to modern cloud-based data platforms.
  • Perform data validation, reconciliation, and quality checks to ensure accuracy and completeness of migrated and processed data.
  • Develop and maintain datasets and data models used by reporting, analytics, and downstream applications.
  • Work with Azure Data Lake and related Azure data services for storing and processing enterprise data.
  • Troubleshoot data pipeline failures, performance issues, and data quality problems.
  • Participate in code reviews and follow standard development, testing, deployment, and CI/CD practices.
  • Collaborate with Data Engineers, Analysts, Data Scientists, and business teams to understand data requirements.
  • Support AI-related data requirements such as preparing and processing structured and unstructured datasets.
  • Gain hands-on exposure to Generative AI concepts, LLMs and Retrieval-Augmented Generation (RAG) as part of AI-enabled data initiatives.

What you'll need

  • 2–5 years of experience in Data Engineering, Data Analytics Engineering, or a related role.
  • Good hands-on experience with Python.
  • Strong working knowledge of SQL.
  • Experience with Spark / PySpark for data processing.
  • Hands-on experience with Databricks or a similar cloud data processing platform.
  • Experience developing and maintaining ETL/ELT pipelines.
  • Understanding of Data Lake / Lakehouse concepts.
  • Experience working with structured and semi-structured data such as CSV, JSON, and Parquet.
  • Understanding of data quality, data validation, and reconciliation techniques.
  • Experience with at least one cloud platform, preferably Microsoft Azure.
  • Familiarity with Git and CI/CD / DevOps practices.
  • Good analytical and troubleshooting skills.
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Strong willingness to learn new data and AI technologies.
  • Good communication, problem-solving, and collaboration skills.

Details

  • Location: Bangalore, India.

Read the full description and apply on the company’s own careers page.

Stay safe

Hiring on Abekus is free for applicants

We never charge a fee, and employers are prohibited from doing so. If a recruiter asks for payment, please report them right away.

AI Data Engineer II