Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Opedia Technologies

Data Center Operations and Maintenance Engineering Leader

Opedia Technologies

Leading Operations & Maintenance for AI infrastructure company, developing strategies and high-performing teams in Bellevue, WA area.

Posted 7/21/2026full-timeBellevue • Washington • 🇺🇸 United StatesSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in building and leading high-performing Operations and Maintenance teams, with a focus on developing operational strategies and processes for large-scale AI infrastructure. Proven ability to drive operational excellence through incident management, service reliability, and cross-functional collaboration.

Highest-signal resume keywords
Operations LeadershipIncident ManagementService ReliabilityCloud InfrastructureExecutive Communication

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Operational StrategyChange ManagementProblem ManagementService Level Objectives (SLOs)Operational KPIsRoot Cause AnalysisAutomationObservabilityScalable Support ModelsInfrastructure Operations
Soft Skills
MentoringInfluencingCollaborationCommunication
Industry Keywords
AI InfrastructureHyperscale EnvironmentsDistributed SystemsProduction EngineeringData Center Operations

Tech Stack

Tools & technologies
CloudDistributed Systems

About the role

Key responsibilities & impact
  • Build, lead, mentor, and grow high-performing Operations & Maintenance teams.
  • Develop the operational strategy, organizational structure, and execution model supporting large-scale AI infrastructure.
  • Establish world-class operational processes, including incident response, change management, problem management, and service reliability.
  • Lead major incident management efforts and executive communications during production events.
  • Drive operational excellence through proactive monitoring, observability, automation, and continuous improvement initiatives.
  • Partner with Engineering to ensure operational readiness for new infrastructure deployments and platform launches.
  • Define service level objectives (SLOs), operational KPIs, and reliability metrics across the infrastructure portfolio.
  • Build scalable on-call programs, escalation models, runbooks, and operational governance.
  • Champion root cause analysis and long-term corrective actions to improve platform resilience.
  • Influence infrastructure architecture and operational tooling to improve availability, efficiency, and customer experience.
  • Help shape the long-term operations organization as the company expands globally.

Requirements

What you’ll need
  • Experience leading Operations, Site Reliability, Infrastructure Operations, Data Center Operations, or Production Engineering organizations.
  • Proven success building or scaling operations teams within cloud infrastructure, hyperscale environments, AI infrastructure, or large distributed systems.
  • Deep expertise in production operations, incident management, service reliability, and operational excellence.
  • Experience leading cross-functional teams during high-severity production incidents.
  • Strong understanding of infrastructure operations across compute, networking, storage, and hardware environments.
  • Demonstrated success building operational processes, organizational structure, and scalable support models in high-growth environments.
  • Executive-level communication skills with the ability to influence engineering and business leadership.
  • Comfortable operating in an early-stage organization where many systems and processes are being built for the first time.

Benefits

Comp & perks
  • Competitive base pay for Bellevue market
  • Certain roles are eligible for additional rewards, including merit increases, annual bonus, and stock. These awards are allocated based on individual performance
  • U.S. based employees have access to medical, dental, and vision insurance
  • 401(k) plan and company match
  • employees also receive per calendar year, paid holidays.