FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
P
Senior Site Reliability Engineer, SRE
PrizePicksSenior SRE ensuring reliable, scalable infrastructure for PrizePicks’ daily fantasy sports platform. Leading incident response, observability, Kubernetes operations, and reliability improvements.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and maintaining reliable production systems, with a strong focus on cloud computing, infrastructure as code, and application deployment. Proven ability to lead incident response, mentor engineers, and foster a culture of reliability and security.
Highest-signal resume keywords
Cloud Computing (AWS, Azure, GCP)Infrastructure as Code (Terraform, Crossplane)Application Development (Python, Ruby, Go)Kubernetes DeploymentMonitoring Implementation (Grafana, New Relic, Datadog)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Production System DesignIncident ResponsePerformance MonitoringSystem Failure AnalysisApplication SecuritySLO GovernanceDebugging Production IssuesBottleneck IdentificationVendor Integration SupportResilient Systems
Soft Skills
MentoringCross-Functional Collaboration
Industry Keywords
Reliability EngineeringScalabilityResilienceSecurityFast-Paced Environment
Tech Stack
Tools & technologiesAWSAzureCloudGoGoogle Cloud PlatformGrafanaKubernetesPythonRubyTerraform
About the role
Key responsibilities & impact- Design, implement, maintain, and monitor reliable production systems at scale
- Lead incident response, mitigate production issues, and conduct post mortem analysis
- Proactively monitor performance, analyze system failures, identify bottlenecks, and propose solutions
- Create and support observability/monitoring tools and vendor integrations
- Drive a reliability culture focused on system reliability, scalability, resilience, and security
- Train and mentor other engineers
Requirements
What you’ll need- 5+ years of experience as a reliability-focused engineer in a fast-paced, rapidly growing enterprise environment
- Deep understanding of cloud computing such as AWS, Azure, and/or GCP
- Experience with infrastructure as code tools such as Terraform or Crossplane
- Experience developing applications in Python, Ruby, or Go
- Experience deploying and supporting applications in Kubernetes at scale
- Experience implementing monitoring with Grafana, New Relic, or Datadog
- Experience debugging live, critical production issues
- Familiarity with resilient systems, application and supply chain security, and SLO governance
- Ability to work cross-functionally with diverse engineering teams
- Must be authorized to work for any employer in the U.S.; employment visa sponsorship is unavailable
Benefits
Comp & perks- Company-subsidized medical, dental, & vision plans
- 401(k) plan with company match
- Annual bonus
- Flexible PTO to encourage a healthy work/life balance (2 weeks STRONGLY encouraged!)
- Generous paid leave programs, including 16-week paid parental leave and disability benefits
- Workplace flexibility and modern work schedules focused on getting the job done, not hours clocked
- Company-wide in-person events and team outings
- Lifestyle enhancement program
- Company equipment provided (Windows & Mac options)
- Annual performance reviews with opportunities for growth and career development