Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
GE Vernova

Principal AI Application Operations Engineer

GE Vernova

Principal AI operations engineer maintaining GE Vernova’s enterprise AI infrastructure, reliability, and incident response. Driving SLA/SLO governance, DevOps improvements, and cross-functional delivery.

Posted 8/21/2026full-time🇺🇸 United StatesLead💰 $137,700 - $229,600 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in AI application infrastructure management, incident management frameworks, and operational strategy formulation. Proficient in leading cross-functional teams, driving data-informed decisions, and implementing process improvements across deployment pipelines.

Highest-signal resume keywords
AI/ML Application OperationsSaaS Application Incident ManagementDevOps Lifecycle ManagementObservability Tools (Datadog, Splunk, Dynatrace)Project Management Skills

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Incident Management Frameworks (ITIL)Software Development Lifecycle (SDLC)Agile MethodologiesCI/CD PracticesNew Product Introduction (NPI)SLA/SLO GovernanceProcess ImprovementOperational Standards EnforcementData AnalysisCloud Infrastructure (Azure, AWS, GCP)
Soft Skills
Strong Oral CommunicationStrong Written CommunicationInterpersonal SkillsLeadership SkillsProblem-Solving Ability
Tools & Technologies
AI Ops ToolingMonitoring ToolsLog Management SystemsPagerDutySupport Workstreams
Industry Keywords
AI OpsMLOpsDevSecOpsOperational JudgmentResource Planning

Tech Stack

Tools & technologies
AWSAzureCloudGoogle Cloud PlatformSDLCSplunk

About the role

Key responsibilities & impact
  • Serve as the primary operational owner for AI application infrastructure
  • Monitor system health and manage incident detection, escalation, and timely resolution across the DevOps lifecycle
  • Lead Root Cause Analysis (RCA) reviews and Push to Production meetings
  • Ensure readiness criteria are met, communicate risks, and track corrective actions to closure
  • Own and report SLA/SLO performance metrics
  • Provide regular dashboards to leadership and drive data-informed decisions addressing trends and service gaps
  • Influence operational strategy for AI applications, including resource planning, policy formulation, and tooling roadmap alignment
  • Monitor industry trends in AI Ops, MLOps, and DevSecOps and translate best practices into improvements
  • Navigate complex technical and business challenges using high-level operational judgment
  • Identify and implement process improvements across deployment pipelines, monitoring frameworks, and support workflows
  • Define and enforce operational standards, quality policies, and best practices across the AI application support team
  • Lead cross-functional teams and projects
  • Communicate complex technical concepts to technical and non-technical audiences, including senior leadership and external partners
  • Mentor and guide team members
  • Bridge development, operations, and business stakeholders to align priorities and outcomes

Requirements

What you’ll need
  • Minimum of a Bachelor's Degree in a relevant discipline
  • 10+ years of relevant professional experience
  • Expertise in SaaS application incident management and enterprise-grade operational management
  • Knowledge of Software Development Lifecycle (SDLC), including Agile, DevOps, and CI/CD practices
  • Knowledge of New Product Introduction (NPI) processes and product roadmap rollout (push to prod)
  • Knowledge of incident management frameworks (e.g., ITIL) and SLA/SLO governance
  • AI/ML application operations or AI Ops tooling and practices
  • Experience with observability tools such as Datadog, Splunk, or Dynatrace and log management systems
  • Strong oral and written communication skills
  • Strong interpersonal and leadership skills
  • Ability to analyze and resolve problems
  • Ability to lead programs/projects
  • Ability to document, plan, market, and execute programs
  • Established project management skills
  • Hands-on familiarity with AI/ML platforms, monitoring tools, and cloud infrastructure such as Azure, AWS, or GCP
  • Experience establishing support through purchased services agreements and coordinating L1/L2/L3 support
  • Experience setting up PagerDuty rotations and support workstreams
  • Legally authorized to work in the United States
  • Successful completion of a drug screen, as applicable

Benefits

Comp & perks
  • Discretionary annual bonus
  • Medical coverage
  • Dental coverage
  • Vision coverage
  • Prescription drug coverage
  • Health Coach access
  • Employee Assistance Program
  • GE Vernova Retirement Savings Plan
  • Tax-advantaged 401(k) savings opportunity with company matching contributions
  • Company retirement contributions
  • Fidelity resources and financial planning consultants
  • Tuition assistance
  • Adoption assistance
  • Paid parental leave
  • Disability benefits
  • Life insurance
  • 12 paid holidays
  • Permissive time off
  • Relocation assistance