Overview
Software Engineer for the Data Platform team, focused on designing, building, and operating reliable, scalable, and secure data infrastructure.
What you'll do
- Design, build, and operate high-scale data ingestion and replication systems into a data lakehouse.
- Build and maintain modern data platform infrastructure handling petabytes of data.
- Improve reliability, observability, scalability, security, and developer experience for Spark/Databricks-based processing.
- Develop internal libraries, APIs, frameworks, and tooling in languages such as Go and Python.
- Work on foundational lake/lakehouse technologies including Delta Lake on S3, data catalogs, metadata services, and orchestration systems.
- Collaborate with multiple teams to understand platform needs and deliver durable solutions.
What you'll need
- 4+ years of professional software engineering experience in production environments.
- 4+ years building or maintaining large-scale production data infrastructure, data platforms, distributed systems, or data lake systems.
- Strong experience with Apache Spark or similar distributed data processing systems.
- Experience operating production infrastructure in AWS, including S3, RDS, and DynamoDB (and others listed).
- Experience designing, building, and operating reliable systems with scalability, observability, security, and operational excellence.
- Proficiency in at least one production programming language such as Go, Python, Scala, or Java.
Details
- Remote position open to candidates residing in the US.
Read the full description and apply on the company’s own careers page.