Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Semrush

Senior Site Reliability Engineer, SRE Team

Semrush

Site Reliability Engineer improving Semrush’s brand-visibility platform reliability. Designing resilient infrastructure, automating operations, and leading incident response across hybrid locations.

Posted 8/12/2026full-time🇪🇸 SpainSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering with a focus on building scalable and reliable systems. Proficient in debugging applications, establishing SLOs, and mentoring engineering teams while collaborating effectively across departments.

Highest-signal resume keywords
Site Reliability EngineeringPython ProgrammingGo ProgrammingKubernetes ExperienceApplication Debugging

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Application DebuggingSystem Architecture DesignSLO EstablishmentOperational Tooling DevelopmentFailure Recovery
Soft Skills
Good Communication AbilitiesTeam-Oriented CollaborationMentoring Engineers
Tools & Technologies
KubernetesGCPCloud Providers
Industry Keywords
ObservabilityMetricsSecurity HardeningCost Dashboards

Tech Stack

Tools & technologies
CloudGoGoogle Cloud PlatformKubernetesPython

About the role

Key responsibilities & impact
  • Lead changes in common engineering practices across the company
  • Induce application failures and recover systems from failure states
  • Debug applications using metrics and add traces or metrics as needed
  • Establish and refine SLOs, cost dashboards, and security hardening initiatives with stakeholders
  • Collaborate with development teams on scalable, reliable, and efficient system architecture
  • Design full-stack platform solutions from concept through production
  • Build sophisticated operational tooling in Go and Python
  • Mentor engineers, interview candidates, and lead critical incidents
  • Participate in an on-call rotation, typically one week every 2–3 weeks, potentially including overnight incidents

Requirements

What you’ll need
  • 3+ years of experience as a Site Reliability Engineer
  • Experience with Kubernetes and cloud providers
  • Engineering experience in Python or Go
  • Strong understanding of application failures and how to handle them
  • Ability to debug applications using metrics
  • Familiarity with traces, observability, and implementation quirks in code
  • Willingness to be on call and work flexible hours
  • Good communication abilities and team-oriented collaboration
  • GCP knowledge is a plus, not required

Benefits

Comp & perks
  • Unlimited PTO
  • Hobby & team building budget allowance
  • Employee Support Program
  • Loss of family member financial aid
  • Employee Resource Groups