FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
MizuhoSite Reliability Engineer at Mizuho maintaining production system reliability and performance. Collaborating with teams to automate workflows and monitor system health.
Posted 7/21/2026full-timeNew York City • New York • 🇺🇸 United StatesMid-LevelSenior💰 $111,000 - $160,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Site Reliability Engineering (SRE) with a strong focus on automation, monitoring, and infrastructure management. Proficient in implementing best practices for system reliability and performance while mentoring team members on SRE methodologies.
Highest-signal resume keywords
Site Reliability Engineering (SRE)Infrastructure as Code (IaC)Monitoring with GrafanaAutomation Tools (Ansible, Terraform, Jenkins)Cloud Providers (AWS, Azure, Google Cloud)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Site Reliability Engineering (SRE)Infrastructure as Code (IaC)Automation Tools (Ansible, Terraform, Jenkins)Monitoring with GrafanaContainerization (Docker)Orchestration (Kubernetes)CI/CD PipelinesScripting (Python, Bash, Go)Problem-SolvingTroubleshooting
Soft Skills
CommunicationTeamworkMentoringAdaptabilityProblem-Solving
Tools & Technologies
GrafanaAnsibleTerraformJenkinsAWSAzureGoogle CloudDockerKubernetes
Industry Keywords
Site Reliability Engineering (SRE)Infrastructure as Code (IaC)MonitoringAutomationCloud Computing
Tech Stack
Tools & technologiesAnsibleAWSAzureCloudDockerGoGrafanaJenkinsKubernetesPythonTerraform
About the role
Key responsibilities & impact- Design, implement, and manage automated deployment, monitoring, and alerting solutions.
- Build and support scalable infrastructure through Infrastructure as Code (IaC) tools.
- Use Grafana and other monitoring platforms to track system reliability and performance.
- Partner with development and operations for ongoing improvements to system reliability and efficiency.
- Diagnose and resolve production issues quickly to minimize downtime.
- Create and maintain best practices and guidelines for SRE processes.
- Enhance observability by improving logging, monitoring, and alert systems.
- Participate in on-call rotations to ensure round-the-clock support for critical systems.
- Lead post-incident reviews and put preventative measures in place.
- Mentor and educate team members on SRE methodologies and technologies.
Requirements
What you’ll need- Bachelor’s degree (or equivalent experience) in Computer Science, Engineering, or a related area.
- Demonstrated experience as a Site Reliability Engineer (SRE) or in a similar capacity.
- Strong background in automation tools and methodologies such as Ansible, Terraform, or Jenkins.
- Advanced skills in monitoring and visualization with Grafana.
- Experience working with cloud providers like AWS, Azure, or Google Cloud.
- In-depth knowledge of containerization and orchestration tools (e.g., Docker, Kubernetes).
- Familiarity with CI/CD pipelines and associated tools.
- Proficient scripting or programming abilities in languages like Python, Bash, or Go.
- Exceptional problem-solving and troubleshooting capabilities.
- Excellent communication and teamwork skills.
- Comfortable working in a fast-paced, ever-changing environment.
Benefits
Comp & perks- Discretionary bonus
- Employee benefits package