Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Abbott

Senior Site Reliability Engineer

Abbott

Senior Site Reliability Engineer for Abbott's cardiac monitoring platform, ensuring reliability and performance in a healthcare setting. Collaborating with teams to automate operations and improve system design.

Posted 7/24/2026full-timeSunnyvale • California • 🇺🇸 United StatesSenior💰 $90,000 - $180,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and maintaining highly available and fault-tolerant systems, with a strong focus on performance optimization and operational automation. Proficient in collaborating with cross-functional teams to enhance system reliability and implement effective monitoring solutions.

Highest-signal resume keywords
Site Reliability EngineeringPython ProgrammingKubernetes Container OrchestrationMicrosoft Azure ServicesMonitoring Tools Expertise

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Site Reliability EngineeringSoftware EngineeringPerformance OptimizationDistributed SystemsFault-Tolerant DesignCI/CD Pipeline DesignPython ProgrammingGo ProgrammingBash ScriptingPowerShell Scripting
Soft Skills
Excellent Communication SkillsAnalytical Problem-SolvingDebugging Skills
Tools & Technologies
Microsoft AzureKubernetesDockerPrometheusGrafanaAzure Monitor
Industry Keywords
SLIsSLOsError BudgetsOperational DocumentationBlameless Postmortem

Tech Stack

Tools & technologies
AzureDistributed SystemsDockerGoGrafanaKubernetesPrometheusPython

About the role

Key responsibilities & impact
  • Design, implement, and maintain highly available and fault-tolerant systems
  • Identify and eliminate performance bottlenecks
  • Define and monitor SLIs, SLOs, and error budgets
  • Develop monitoring, logging, and alerting solutions
  • Automate operational tasks
  • Scale services and infrastructure for business demands
  • Collaborate with engineering and security teams
  • Create operational documentation and runbooks
  • Lead blameless postmortem processes
  • Work with a multi-disciplinary team on reliability roadmap planning

Requirements

What you’ll need
  • Bachelor's in Computer Science, Software Engineering, Systems Engineering, or equivalent professional experience
  • Minimum 7 years of experience in site reliability, software engineering, or related fields
  • Excellent communication skills for cross-functional collaboration
  • Strong analytical, problem-solving, and debugging skills
  • Proficiency in Python, Go, Bash, or PowerShell
  • Expertise in Microsoft Azure services like AKS, Azure Monitor, etc.
  • Container orchestration experience with Kubernetes and Docker
  • Experience with monitoring tools such as Prometheus, Grafana
  • Design and operate CI/CD pipelines for safe deployments
  • Deep understanding of distributed systems and fault-tolerant design

Benefits

Comp & perks
  • Health insurance
  • 401(k) matching
  • Paid time off
  • Flexible work arrangements
  • Professional development
  • Travel allowance