Overview
SDE III (Gen AI) to design and deploy production-ready generative AI applications and related AI systems.
What you'll do
- Design and implement generative AI applications from architecture through deployment and monitoring.
- Build RAG pipelines using vector databases, hybrid search, and caching for low-latency responses.
- Develop multimodal AI systems integrating text, vision, and audio capabilities.
- Architect scalable microservices for high-concurrency AI requests with cost, latency, and reliability optimization.
- Optimize large language models via fine-tuning and implement MLOps practices.
- Collaborate with product managers and stakeholders to translate requirements into AI solutions.
- Mentor junior engineers via pair programming, technical talks, and hands-on guidance.
What you'll need
- 5+ years hands-on experience building and deploying ML/AI systems, with 2+ years focused on generative AI and LLMs.
- Expert Python skills including async programming, multiprocessing, and performance optimization.
- Experience with AI frameworks including PyTorch, Transformers, LangChain, and vector databases.
- Experience deploying AI applications to production environments serving real users at scale.
- Experience with cloud platforms (GCP) and containerization technologies (Docker, Kubernetes).
- Bachelor’s degree in Computer Science, Mathematics, or a related field (Master’s preferred but not required).
Nice to have
- Published research papers or significant contributions to open-source AI projects.
- Experience with multimodal AI systems combining vision, language, and audio.
- Knowledge of AI safety, alignment, and constitutional AI principles.
Details
Read the full description and apply on the company’s own careers page.