Overview
Develop, enhance, and support scalable data platforms, pipelines, and data products for the investment management business.
What you'll do
- Develop and maintain batch and streaming pipelines using PySpark, Spark SQL, and Databricks.
- Build data processing solutions using Bronze/Silver/Gold architecture patterns.
- Implement ingestion pipelines from databases, APIs, event streams, and cloud storage.
- Develop incremental processing, CDC, and data transformation workflows.
- Participate in code reviews, testing, deployments, and production support.
- Monitor pipeline performance and troubleshoot operational issues.
- Support data quality, metadata, and governance requirements.
What you'll need
- 3–7 years of experience in software engineering, data engineering, or related technical roles.
- Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field, or equivalent experience.
- Hands-on experience with Databricks notebooks, workflows, clusters, and repositories.
- Proficiency in Python, PySpark, and SQL.
- Experience building production data pipelines in cloud environments.
- Knowledge of Apache Spark, data lake and lakehouse concepts, and cloud storage.
- Experience with orchestration tools and software engineering practices such as version control, testing, and CI/CD.
Nice to have
- Experience in financial services or asset management.
- Exposure to Apache Iceberg, Delta Lake, Unity Catalog, or similar technologies.
- Experience with Kafka, Event Hubs, Kinesis, or Structured Streaming.
- Experience supporting machine learning or analytics use cases.
Details
- Location: Bangalore, India.
Read the full description and apply on the company’s own careers page.