FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
Cracker BarrelSite Reliability Engineer ensuring operational reliability for Cracker Barrel's digital platforms and services. Collaborating with teams to enhance system availability, performance, scalability, and reliability.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Site Reliability Engineering and DevOps practices, focusing on improving system availability, performance, and operational maturity. Proficient in monitoring, incident management, and automation to enhance service reliability across digital platforms.
Highest-signal resume keywords
Site Reliability EngineeringDevOps PracticesMonitoring and LoggingCI/CD Pipeline ManagementIncident Management
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Site Reliability EngineeringDevOpsCloud OperationsProduction SupportSystems EngineeringSoftware EngineeringMonitoringScriptingAPIsInfrastructure Automation
Soft Skills
CollaborationProblem ManagementRoot-Cause AnalysisCommunication
Tools & Technologies
Observability ToolsDashboardsMetricsSynthetic MonitoringGit-Based WorkflowsContent Management SystemsCloud Platforms
Industry Keywords
High-AvailabilityDigital ApplicationsOperational SupportSecurity ComplianceChange Management
Tech Stack
Tools & technologiesCloudMicroservices
About the role
Key responsibilities & impact- Design, implement, and maintain reliability practices that improve availability, performance, scalability, resiliency, and operational maturity across digital platforms and supporting services.
- Monitor production systems using observability tools, dashboards, logs, metrics, traces, alerts, and synthetic monitoring to identify issues before they impact customers or associates.
- Provide Tier-2 and Tier-3 production support for customer-facing and internal digital applications, including web, mobile, commerce, CMS, APIs, integrations, and cloud-hosted services.
- Lead and participate in incident response, root-cause analysis, problem management, post-incident reviews, and follow-up actions that reduce recurrence and improve service reliability.
- Develop automation, scripts, runbooks, self-healing processes, and operational tools that reduce manual effort, accelerate recovery, and improve consistency across environments.
- Partner with development teams to improve CI/CD pipelines, deployment readiness, release validation, rollback procedures, feature monitoring, and environment stability.
- Collaborate with infrastructure, cloud, security, architecture, QA, and vendor teams to ensure systems meet company standards for security, privacy, compliance, resiliency, and operational support.
- Define and track service health indicators such as availability, latency, error rates, capacity, incident trends, deployment quality, and other reliability metrics.
- Create and maintain technical documentation, operational support guides, escalation paths, production readiness checklists, and disaster recovery procedures.
- Understand and comply with all company privacy, security, accessibility, change management, and technology standards.
Requirements
What you’ll need- Bachelor’s degree in Computer Science, Computer Information Systems, Software Engineering, Information Technology, or a related discipline is preferred; equivalent experience or training may be considered.
- 3–5+ years of experience in site reliability engineering, DevOps, cloud operations, production support, systems engineering, software engineering, or a related technology operations role.
- Experience supporting high-availability web, mobile, commerce, API, integration, or cloud-hosted application environments.
- Hands-on experience with monitoring, logging, alerting, incident management, root-cause analysis, CI/CD pipelines, Git-based workflows, and release support.
- Experience with cloud platforms, containers, infrastructure automation, scripting, APIs, microservices, content management systems, or restaurant/retail technology environments preferred.
Benefits
Comp & perks- Medical, Rx, Dental and Vision Benefits on Day 1
- Life Insurance and Disability Coverage
- Paid Vacation/Employee Assistance Program
- Business Resource Groups
- Tuition Reimbursement
- Professional Development
- Onboarding, training, and development to help you thrive
- Recognition programs and employee events that bring us together
- 401k Plan with Company Matching Contributions at 90 days
- Employee Stock Purchase Program
- 35% Discount on Cracker Barrel Food and Retail items
- Exclusive Biscuit Perks like discounts on home, travel, cell phones, and more!