Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Visa

Senior Director, Cloud Platform – Reliability Engineering

Visa

Senior Director leading cloud platform and reliability engineering at Visa. Responsible for strategy and execution of cloud environment ensuring reliability, scalability, and security.

Posted 7/29/2026full-timeAuckland • 🇳🇿 New ZealandSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates extensive experience in leading cloud platform strategies, focusing on infrastructure, reliability, and cost optimization. Proven ability to manage large-scale cloud environments and drive operational excellence through automation and effective incident management.

Highest-signal resume keywords
Cloud Platform LeadershipInfrastructure as CodeCI/CD PracticesIncident ManagementMonitoring and Observability

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Cloud Platforms (AWS, Azure, GCP)Infrastructure as CodeCI/CDMonitoring SystemsLinux EnvironmentsScripting (Python)Capacity PlanningPerformance ManagementDisaster Recovery StrategiesRoot Cause Analysis
Soft Skills
Excellent CommunicationLeadership in Cross-Functional Teams
Industry Keywords
Operational ExcellenceService Level Objectives (SLOs)Service Level Agreements (SLAs)Automation-First Operating ModelVendor ManagementChange ManagementIncident ResponseSecure-By-Design Practices

Tech Stack

Tools & technologies
AWSAzureCloudGoogle Cloud PlatformLinuxPython

About the role

Key responsibilities & impact
  • Define and lead the cloud platform strategy, including infrastructure, reliability, scalability, and cost optimization.
  • Drive run cost optimization strategy across infrastructure, platforms, and shared services, ensuring cost efficiency without compromising reliability, security, or release velocity.
  • Demonstrated experience operating and evolving large‑scale cloud environments, with measurable scope such.
  • Proven ownership of 24x7 production operations with strict SLOs/SLAs, including availability, latency, error rates, and recovery objectives (RTO/RPO).
  • Experience managing complex operational concerns at scale, including: Capacity planning and performance management, Incident response for high‑severity, customer‑impacting events, Change management and release safety at scale, Vendor and third‑party service dependencies.
  • Build, lead, and mentor high‑performing engineering teams across platform engineering, cloud operations, and site reliability.
  • Drive an automation‑first operating model, reducing manual work through Infrastructure as Code, CI/CD enablement, and standardized self‑service platforms.
  • Own service reliability and operational excellence, including incident management, root cause analysis, and long‑term remediation.
  • Partner with engineering, product, and security leaders to align platform capabilities with business priorities.
  • Establish and enforce secure‑by‑design practices for infrastructure, applications, and data.
  • Implement strong observability standards (monitoring, logging, metrics, alerting) to improve service health and decision‑making.
  • Define and validate resilience and disaster recovery strategies, including failure testing and recovery readiness.
  • Provide clear executive‑level reporting on platform health, risks, and roadmap progress.
  • Manage third‑party vendors and cloud service partners as needed.
  • Ensure effective 24x7 operational coverage, while continuously reducing operational burden through engineering improvements.

Requirements

What you’ll need
  • 15+ years of experience in engineering, infrastructure, cloud platform, or site reliability roles, with significant leadership responsibility.
  • Proven experience leading cloud platforms or cloud‑based services at scale (AWS, Azure, or GCP).
  • Demonstrated experience running production systems with high availability and reliability expectations.
  • Strong ability to lead in a matrixed, cross‑functional environment.
  • Excellent communication skills, including the ability to explain complex technical topics to senior leadership.
  • Strong technical foundation in: Infrastructure as Code and configuration management, CI/CD and release engineering practices, Monitoring, logging, and alerting systems, Linux environments and scripting or programming (e.g., Python).

Benefits

Comp & perks
  • Health insurance
  • Professional development opportunities