FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
MizuhoSite Reliability Engineer at Mizuho ensuring production systems' reliability and performance. Collaborating with teams to automate workflows and enhance system efficiency.
Posted 7/22/2026full-timeMetroPark NYC • New York • 🇺🇸 United StatesMid-LevelSenior💰 $111,000 - $160,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Site Reliability Engineering (SRE) with a strong focus on automation, monitoring, and infrastructure management. Proficient in utilizing tools like Grafana, Terraform, and cloud services to enhance system reliability and performance.
Highest-signal resume keywords
Site Reliability Engineering (SRE)Automation Tools (Ansible, Terraform, Jenkins)Monitoring and Visualization (Grafana)Cloud Providers (AWS, Azure, Google Cloud)Containerization and Orchestration (Docker, Kubernetes)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Site Reliability Engineering (SRE)Automation ToolsMonitoring and VisualizationCloud ComputingContainerizationOrchestrationScripting (Python, Bash, Go)CI/CD PipelinesDatabase AdministrationNetworking and Security Best Practices
Soft Skills
Problem-SolvingTroubleshootingCommunicationTeamworkAdaptability
Tools & Technologies
GrafanaAnsibleTerraformJenkinsAWSAzureGoogle CloudDockerKubernetesPrometheus
Industry Keywords
Infrastructure as Code (IaC)Monitoring PlatformsOn-Call SupportPost-Incident ReviewsBest Practices for SRE
Tech Stack
Tools & technologiesAnsibleAWSAzureCloudDockerGoGrafanaJenkinsKubernetesPrometheusPythonTerraform
About the role
Key responsibilities & impact- Design, implement, and manage automated deployment, monitoring, and alerting solutions.
- Build and support scalable infrastructure through Infrastructure as Code (IaC) tools.
- Use Grafana and other monitoring platforms to track system reliability and performance.
- Partner with development and operations for ongoing improvements to system reliability and efficiency.
- Diagnose and resolve production issues quickly to minimize downtime.
- Create and maintain best practices and guidelines for SRE processes.
- Enhance observability by improving logging, monitoring, and alert systems.
- Participate in on-call rotations to ensure round-the-clock support for critical systems.
- Lead post-incident reviews and put preventative measures in place.
- Mentor and educate team members on SRE methodologies and technologies.
Requirements
What you’ll need- Bachelor’s degree (or equivalent experience) in Computer Science, Engineering, or a related area.
- Demonstrated experience as a Site Reliability Engineer (SRE) or in a similar capacity.
- Strong background in automation tools and methodologies such as Ansible, Terraform, or Jenkins.
- Advanced skills in monitoring and visualization with Grafana.
- Experience working with cloud providers like AWS, Azure, or Google Cloud.
- In-depth knowledge of containerization and orchestration tools (e.g., Docker, Kubernetes).
- Familiarity with CI/CD pipelines and associated tools.
- Proficient scripting or programming abilities in languages like Python, Bash, or Go.
- Exceptional problem-solving and troubleshooting capabilities.
- Excellent communication and teamwork skills.
- Comfortable working in a fast-paced, ever-changing environment.
- Hands-on experience with Prometheus or comparable time-series databases (preferred).
- Solid understanding of networking and security best practices (preferred).
- Knowledgeable in database administration and optimization strategies (preferred).
Benefits
Comp & perks- employee benefits package
- discretionary bonus