FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Site Reliability Engineering with a focus on building scalable and reliable systems. Proficient in debugging applications, establishing SLOs, and mentoring engineering teams while collaborating effectively across departments.
Highest-signal resume keywords
Site Reliability EngineeringPython ProgrammingGo ProgrammingKubernetes ExperienceApplication Debugging
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Application DebuggingSystem Architecture DesignSLO EstablishmentOperational Tooling DevelopmentFailure Recovery
Soft Skills
Good Communication AbilitiesTeam-Oriented CollaborationMentoring Engineers
Tools & Technologies
KubernetesGCPCloud Providers
Industry Keywords
ObservabilityMetricsSecurity HardeningCost Dashboards
Tech Stack
Tools & technologiesCloudGoGoogle Cloud PlatformKubernetesPython
About the role
Key responsibilities & impact- Lead changes in common engineering practices across the company
- Induce application failures and recover systems from failure states
- Debug applications using metrics and add traces or metrics as needed
- Establish and refine SLOs, cost dashboards, and security hardening initiatives with stakeholders
- Collaborate with development teams on scalable, reliable, and efficient system architecture
- Design full-stack platform solutions from concept through production
- Build sophisticated operational tooling in Go and Python
- Mentor engineers, interview candidates, and lead critical incidents
- Participate in an on-call rotation, typically one week every 2–3 weeks, potentially including overnight incidents
Requirements
What you’ll need- 3+ years of experience as a Site Reliability Engineer
- Experience with Kubernetes and cloud providers
- Engineering experience in Python or Go
- Strong understanding of application failures and how to handle them
- Ability to debug applications using metrics
- Familiarity with traces, observability, and implementation quirks in code
- Willingness to be on call and work flexible hours
- Good communication abilities and team-oriented collaboration
- GCP knowledge is a plus, not required
Benefits
Comp & perks- Unlimited PTO
- Hobby & team building budget allowance
- Employee Support Program
- Loss of family member financial aid
- Employee Resource Groups
