Overview
Design, build, and maintain production-grade agent solutions and components using Python, OpenAI Agents SDK, LangGraph, and LangFuse. Deliver multi-tenant agent services with tenant isolation, testing, CI/CD, security, observability, and operational support.
What you'll do
- Design, implement, and maintain production agent components in Python using OpenAI Agents SDK, LangGraph, and LangFuse, following clean-code principles and team standards.
- Deliver features for multi-tenant services with built-in tenant isolation, baseline scalability, and secure defaults.
- Produce clear, testable user stories and decompose work into subtasks with measurable acceptance criteria.
- Write and maintain comprehensive unit tests and test automation; participate in mutation testing cycles and remediate identified gaps with guidance from senior engineers.
- Collaborate in backlog grooming, sprint planning, daily standups, and retrospectives; deliver committed sprint work and surface risks early.
- Work with DevOps to maintain and improve CI/CD pipelines, automate routine deployments, add relevant test stages, and troubleshoot build/test failures.
- Participate in incident response and post-mortem activities, provide technical analysis, implement remediation tasks, and document lessons learned.
- Incorporate information-security and data-privacy best practices in design, implementation, and code reviews.
- Monitor and improve observability by adding meaningful metrics, logs, and traces.
- Keep up to date with agent development patterns, testing techniques, and observability tooling, and propose practical improvements.
- Provide constructive code review feedback and incorporate reviewer suggestions.
- Produce and maintain onboarding documentation, runbooks, and troubleshooting guides.
- Identify and implement small refactors and focused performance improvements; escalate larger refactor or architectural changes for team prioritization.
- Fix broken builds, add automation for linting and tests, and improve pipeline reliability.
- Participate in technical interviews and candidate evaluations on request.
- Assist with benchmarking and profiling tasks and implement remediation under supervision.
- Ensure work aligns with the team vision and communicate impediments that could block delivery.
What you'll need
- BE/B.Tech, ME/M.Tech, or MCA.
- 3 to 6 years of experience.
- Proficiency with Python, OpenAI Agents SDK, RAG, Prompt Engineering, and Context Engineering.
- Experience in benchmarking evaluation sets and writing evaluators.
- Knowledge of multi-agent orchestration, including Worker (Supervisor) and Concurrent (Fan-Out/Fan-In).
- Skill writing.
- MCP implementation.
- Agent security and Responsible AI.
- Direct experience with LangGraph and LangFuse in production environments.
- Strong technical knowledge, excellent communication skills, and the ability to work effectively within a team.
Nice to have
- Knowledge of Retail Planning Products.
- DevOps experience deploying agents with performance, scale, and reliability in production.
Details
Read the full description and apply on the company’s own careers page.