SRE (Site Reliability Engineer)
VXI Global SolutionsMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessSpecialized observability stack but popular SRE title yields moderate applicant competition.
SRE observability skills transfer across industries but require platform-specific tooling and cloud experience.
Mandatory observability tools and clear SRE/domain requirements create strict technical filtering.
Job Description
Structured overview of role & requirementsAbout This Role
Design, implement, and maintain observability pipelines using OpenTelemetry, Prometheus, and Grafana to monitor cloud infrastructure and applications.
Build dashboards and alerts to track system health, application performance, and business KPIs, integrating with Google Cloud Platform services and SolarWinds.
Correlate logs, metrics, and traces to proactively detect, diagnose, and resolve performance issues, collaborating with engineering teams to improve system observability.
Minimum Requirements
Strong experience with Prometheus and Grafana for monitoring and alerting.
Proficiency in OpenTelemetry for instrumenting distributed systems.
Working knowledge of Google Cloud observability tools (Cloud Monitoring, Logging, Trace) and exposure to SolarWinds.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experience in Site Reliability Engineering or Platform Engineering roles with a focus on observability solutions.
Demonstrated ability to integrate and correlate multi-source telemetry data (metrics, logs, traces) to reduce mean time to resolution (MTTR).
Familiarity with SLIs/SLOs and performance benchmarking, with some scripting skills (Python, Bash) and Infrastructure-as-Code knowledge considered valuable.
