FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
MaintainXSite Reliability Engineer improving reliability, observability, and developer autonomy for MaintainX’s industrial work execution platform. Building tooling and standards for resilient, self-service operations.
Posted 9/8/2026full-timeRemote • California • 🇺🇸 United StatesMid-LevelSenior💰 $120,000 - $249,260 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in observability practices and SRE concepts, with a focus on improving service reliability and operational readiness. Proficient in mentoring development teams and driving the adoption of best practices in cloud-native environments.
Highest-signal resume keywords
Observability PracticesSite Reliability Engineering (SRE)Cloud-Native PlatformsInfrastructure-as-CodeIncident Management
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Service Level Objectives (SLOs)Error BudgetsProduction Systems OperationProgramming Language ProficiencyTypeScriptNode.js
Soft Skills
Excellent CommunicationCollaboration AbilitiesMentoring
Tools & Technologies
Infrastructure-as-Code Tools
Industry Keywords
Service Health MetricsReliability InitiativesObservability StandardsIncident Response Practices
Tech Stack
Tools & technologiesCloudJavaScriptNode.jsTypeScript
About the role
Key responsibilities & impact- Assess service maturity and provide insights to development teams
- Partner with development teams to implement observability best practices
- Enable development teams to become autonomous with service deployment, support, and infrastructure
- Mentor developers on reliability practices and help them become self-sufficient
- Drive tooling and practice adoption across development teams as a bridge between Platform Division teams and development teams
- Improve the stability, resilience, and operational readiness of services
- Contribute to company-wide reliability initiatives, observability standards, incident response practices, and service health metrics
- Design for reliability, establish ownership and standards, and build shared tooling
Requirements
What you’ll need- Deep understanding of observability practices in distributed system environments
- Practical experience with SRE concepts, including SLOs, error budgets, and incident management
- 3–5+ years in software development, SRE, DevOps, or production development roles
- Experience operating production systems
- Proficiency in cloud-native platforms
- Knowledge of infrastructure-as-code concepts and tools
- Working knowledge of at least one programming language
- TypeScript/Node.js experience is a plus
- Excellent communication and collaboration abilities across technical and non-technical teams
- Ability to translate complex reliability concepts into actionable guidance
Benefits
Comp & perks- Equity may be included depending on the role
- Annual bonus may be included depending on the role
- Benefits differ by country
- Health coverage, retirement and leave benefits for roles in Canada and other countries
- Autodesk benefits for roles in the United States