Overview
Principal Data Scientist (Computer Vision) who serves as the primary technical architect for the AI/ML ecosystem and leads long-term technical vision for media understanding and generation.
What you'll do
- Define the long-term AI/ML technical vision for the content product ecosystem.
- Prototype and de-risk new initiatives and lead multimodal data foundation design.
- Architect end-to-end multimodal ML systems integrating visual, audio, and textual data.
- Build ML data foundations covering annotation, processing, and storage at petabyte scale.
- Lead scalable shared capability layers to unify systems across use cases.
- Oversee inference optimization using GPU acceleration and techniques like quantization and TensorRT.
- Establish evaluation and observability via offline/online evaluation and telemetry.
What you'll need
- Experience: 8+ years (as stated for the role).
- 10+ years of machine learning experience, with at least 2 years in LLMs, diffusion models, or other generative image/video models.
- Expert proficiency in Python, Java, or C++ plus multi-threading, memory management, and distributed computing (Spark, Flink).
- Deep learning framework proficiency in PyTorch or TensorFlow.
- Computer vision expertise including object detection (YOLO) and image segmentation (Mask R-CNN).
- Familiarity with NVIDIA DeepStream, Triton Inference Server, and TensorRT.
- Production deployment experience using Kafka, Airflow, Kubernetes, and cloud AI services (AWS Bedrock, SageMaker).
Details
- Location: Bangalore, India.
Read the full description and apply on the company’s own careers page.