





Tier-1 backing, mid-level (5+ years) band, and Bangalore metro increase competition.
Strong ML agent specialization and production evaluation focus make cross-industry transferability limited.
Mandates 5+ years, specific agentic systems experience, and rigorous evaluation skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the end-to-end performance of AI agents that plan, build, test, and ship production software at global scale.
Define, measure, and run high-leverage experiments to improve agent quality, reliability, and output, including prompt tuning, evaluation datasets, and experimentation.
Build and maintain evaluation frameworks and dashboards to detect regressions, manage staged rollouts, and make informed ship or rollback decisions.
5+ years of software building and shipping experience with real end-to-end ownership.
At least 1 year hands-on experience with agentic systems and their evaluation.
Proficiency in Python, SQL, or similar tools for experimentation, metrics, and debugging.
Work Experience Required: 5+ years in software development; notice period: Not explicitly mentioned in the JD.
Senior individual contributors such as Principal or Staff engineers, software architects, or lead builders with system end-to-end ownership.
Experience blending engineering, product judgment, and applied research in complex AI/agentic systems.
Ability to handle ambiguous, probabilistic signals and make data-driven decisions under uncertainty with velocity and rigor.