Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own the end-to-end AI document ingestion and extraction pipeline from heterogeneous formats (PDF, HTML, scanned documents) to structured actionable data with verbatim source text preservation.
Build and maintain evaluation systems (precision, recall), calibrate confidence thresholds, and design human-in-the-loop review workflows ensuring AI output accuracy.
Manage system performance including cost, latency, monitoring failures, production drift, and collaborate with domain experts on taxonomy and labeling standards.
Minimum Requirements
Bachelor's degree in Computer Science, Engineering, or related field; Master's degree preferred.
At least 5 years of experience building and owning production ML/AI systems, including NLP or document extraction.
Experience with OCR, layout-aware parsing, and processing heterogeneous document formats (PDF, HTML, scanned images).
Must currently reside within 150 km of Greater Delhi NCR, Greater Bangalore, or Greater Pune; hybrid work mode with core US overlap hours.
Ideal Candidate Profile
Experienced in building high-quality, production-grade AI systems focused on document understanding and extraction rather than research innovation.
Capable of end-to-end ownership including system design, evaluation, and operational monitoring with attention to unit economics and real-world performance metrics.
Comfortable collaborating closely with domain experts and cross-functional teams across time zones (US business hours overlap).
