FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Site Reliability Engineer
AbbottSenior Site Reliability Engineer for Abbott's cardiac monitoring platform, ensuring reliability and performance in a healthcare setting. Collaborating with teams to automate operations and improve system design.
Posted 7/24/2026full-timeSunnyvale • California • 🇺🇸 United StatesSenior💰 $90,000 - $180,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and maintaining highly available and fault-tolerant systems, with a strong focus on performance optimization and operational automation. Proficient in collaborating with cross-functional teams to enhance system reliability and implement effective monitoring solutions.
Highest-signal resume keywords
Site Reliability EngineeringPython ProgrammingKubernetes Container OrchestrationMicrosoft Azure ServicesMonitoring Tools Expertise
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Site Reliability EngineeringSoftware EngineeringPerformance OptimizationDistributed SystemsFault-Tolerant DesignCI/CD Pipeline DesignPython ProgrammingGo ProgrammingBash ScriptingPowerShell Scripting
Soft Skills
Excellent Communication SkillsAnalytical Problem-SolvingDebugging Skills
Tools & Technologies
Microsoft AzureKubernetesDockerPrometheusGrafanaAzure Monitor
Industry Keywords
SLIsSLOsError BudgetsOperational DocumentationBlameless Postmortem
Tech Stack
Tools & technologiesAzureDistributed SystemsDockerGoGrafanaKubernetesPrometheusPython
About the role
Key responsibilities & impact- Design, implement, and maintain highly available and fault-tolerant systems
- Identify and eliminate performance bottlenecks
- Define and monitor SLIs, SLOs, and error budgets
- Develop monitoring, logging, and alerting solutions
- Automate operational tasks
- Scale services and infrastructure for business demands
- Collaborate with engineering and security teams
- Create operational documentation and runbooks
- Lead blameless postmortem processes
- Work with a multi-disciplinary team on reliability roadmap planning
Requirements
What you’ll need- Bachelor's in Computer Science, Software Engineering, Systems Engineering, or equivalent professional experience
- Minimum 7 years of experience in site reliability, software engineering, or related fields
- Excellent communication skills for cross-functional collaboration
- Strong analytical, problem-solving, and debugging skills
- Proficiency in Python, Go, Bash, or PowerShell
- Expertise in Microsoft Azure services like AKS, Azure Monitor, etc.
- Container orchestration experience with Kubernetes and Docker
- Experience with monitoring tools such as Prometheus, Grafana
- Design and operate CI/CD pipelines for safe deployments
- Deep understanding of distributed systems and fault-tolerant design
Benefits
Comp & perks- Health insurance
- 401(k) matching
- Paid time off
- Flexible work arrangements
- Professional development
- Travel allowance