FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
Arbor EducationSite Reliability Engineer ensuring world-class resilience and performance across the platform. Collaborating with teams to address performance bottlenecks and improve observability.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in performance monitoring, capacity planning, and automation, with a strong focus on high availability and resilience. Proficient in Infrastructure as Code using Terraform and familiar with SRE practices to ensure effective service delivery.
Highest-signal resume keywords
Performance MonitoringCapacity PlanningInfrastructure As CodeTerraformSRE Practices
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Performance MonitoringCapacity PlanningScriptingAutomationInfrastructure As CodeTerraformRelational Database TechnologiesNginxContainerizationMessaging Workloads
Soft Skills
CollaborationProblem-SolvingIncident ResponseDocumentation
Tools & Technologies
DataDogPrometheusAWS AuroraDocker
Industry Keywords
SRE ProcessesDevOps PrinciplesAgile DevelopmentKanbanSoftware Best Practices
Tech Stack
Tools & technologiesAWSCloudDockerNGINXPrometheusTerraform
About the role
Key responsibilities & impact- Proactively monitor and analyse platform performance.
- Collaborate with engineering teams to address performance bottlenecks and ensure scalability.
- Assist engineering teams with implementing and reviewing SLOs
- Continually improve observability through monitoring and alerting, and dashboards, using tools such as DataDog or Prometheus for example.
- Work with other teams to ensure it is effective and provides full coverage.
- Ensure the service is highly available and resilient
- Champion best practices in design for high availability
- Devise runbooks and run game sessions to test our DR plan, H/A and backups
- Conduct assessments of capacity and plan for scaling to meet current and future business needs.
- Work closely with the Head of Platform Engineering and Head of SRE to strategize and implement scalable solutions.
- Work closely with the Platform team, feature teams and, 2nd line support and other stakeholders to ensure a good level of service is provided for our customers and embed SRE practices.
- Key player in the response and troubleshooting of incidents, ensuring rapid resolution and minimising downtime.
- Participate in blameless postmortems to identify root cause and corrective actions
- Develop and maintain playbooks and documentation
Requirements
What you’ll need- 7-12 years of experience
- Experience in performance monitoring and analysis
- Capacity planning experience
- Scripting and automation skills, with experience in relevant technologies.
- Experience with Infrastructure as Code, in particular, Terraform
- Understanding of relational database technologies and their cloud versions (e.g. AWS Aurora)
- Experience with messaging and distributed asynchronous workloads
- Experience with nginx or similar technologies
- Familiarity with SRE processes.
- Aware of DevOps principles like the 3 ways and 5 ideals.
- Desired Skills
- Experience with other database technologies and cloud platforms.
- Past experience with Enterprise solutions running at scale
- Familiarity with Kanban and Agile development processes
- Experience with containerisation, for example Docker
- Familiarity with software best practices such as Refactoring, Clean Code, Domain-Driven Design and Test-Driven Development.
Benefits
Comp & perks- Hybrid work environment
- Group Term Life Insurance paid out at 3x Annual CTC (Arbor India)
- 32 days holiday (plus Arbor Holidays). This is made up of 25 days annual leave plus 7 extra companywide days given over Easter, Summer & Christmas
- Work time: 9.30 am to 6 pm (8.5 hours only)
- Compensation - 100% fixed salary disbursement and no variable components