Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and operate large-scale distributed software platforms for enterprise observability, automation, and AI-driven infrastructure reliability.
Develop reusable platform services, APIs, automation frameworks, and control planes that reduce operational toil and enable self-service across multiple teams.
Provide technical leadership influencing architecture and engineering direction spanning Storage, Compute, Network, and Platform domains with focus on scalability, security, and operational readiness.
Minimum Requirements
Bachelor's or Master's degree in Computer Science, Engineering, or equivalent practical experience.
10+ years of software engineering, SRE, infrastructure, or distributed-systems experience with demonstrated technical leadership.
Strong software engineering skills in Go, Python, or equivalent, with experience building production-grade distributed systems and automation.
Experience with distributed systems architectures, event-driven platforms, observability tools (e.g., OpenTelemetry, Prometheus), and infrastructure knowledge across Kubernetes/OpenShift, VMware, storage, and networking.
Ideal Candidate Profile
Proven track record owning complex software/platform initiatives across multiple infrastructure domains with measurable impact.
Experience building scalable observability and telemetry platforms for large-scale on-premise infrastructure including Storage, Compute, Networking, VMware, and Kubernetes/OpenShift.
Demonstrated ability to lead technical direction and mentoring in ambiguous problem spaces, delivering measurable operational and business outcomes including autonomous or AI-driven reliability capabilities.
