FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Principal AI Application Operations Engineer
GE VernovaPrincipal AI operations engineer maintaining GE Vernova’s enterprise AI infrastructure, reliability, and incident response. Driving SLA/SLO governance, DevOps improvements, and cross-functional delivery.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in AI application infrastructure management, incident management frameworks, and operational strategy formulation. Proficient in leading cross-functional teams, driving data-informed decisions, and implementing process improvements across deployment pipelines.
Highest-signal resume keywords
AI/ML Application OperationsSaaS Application Incident ManagementDevOps Lifecycle ManagementObservability Tools (Datadog, Splunk, Dynatrace)Project Management Skills
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Incident Management Frameworks (ITIL)Software Development Lifecycle (SDLC)Agile MethodologiesCI/CD PracticesNew Product Introduction (NPI)SLA/SLO GovernanceProcess ImprovementOperational Standards EnforcementData AnalysisCloud Infrastructure (Azure, AWS, GCP)
Soft Skills
Strong Oral CommunicationStrong Written CommunicationInterpersonal SkillsLeadership SkillsProblem-Solving Ability
Tools & Technologies
AI Ops ToolingMonitoring ToolsLog Management SystemsPagerDutySupport Workstreams
Industry Keywords
AI OpsMLOpsDevSecOpsOperational JudgmentResource Planning
Tech Stack
Tools & technologiesAWSAzureCloudGoogle Cloud PlatformSDLCSplunk
About the role
Key responsibilities & impact- Serve as the primary operational owner for AI application infrastructure
- Monitor system health and manage incident detection, escalation, and timely resolution across the DevOps lifecycle
- Lead Root Cause Analysis (RCA) reviews and Push to Production meetings
- Ensure readiness criteria are met, communicate risks, and track corrective actions to closure
- Own and report SLA/SLO performance metrics
- Provide regular dashboards to leadership and drive data-informed decisions addressing trends and service gaps
- Influence operational strategy for AI applications, including resource planning, policy formulation, and tooling roadmap alignment
- Monitor industry trends in AI Ops, MLOps, and DevSecOps and translate best practices into improvements
- Navigate complex technical and business challenges using high-level operational judgment
- Identify and implement process improvements across deployment pipelines, monitoring frameworks, and support workflows
- Define and enforce operational standards, quality policies, and best practices across the AI application support team
- Lead cross-functional teams and projects
- Communicate complex technical concepts to technical and non-technical audiences, including senior leadership and external partners
- Mentor and guide team members
- Bridge development, operations, and business stakeholders to align priorities and outcomes
Requirements
What you’ll need- Minimum of a Bachelor's Degree in a relevant discipline
- 10+ years of relevant professional experience
- Expertise in SaaS application incident management and enterprise-grade operational management
- Knowledge of Software Development Lifecycle (SDLC), including Agile, DevOps, and CI/CD practices
- Knowledge of New Product Introduction (NPI) processes and product roadmap rollout (push to prod)
- Knowledge of incident management frameworks (e.g., ITIL) and SLA/SLO governance
- AI/ML application operations or AI Ops tooling and practices
- Experience with observability tools such as Datadog, Splunk, or Dynatrace and log management systems
- Strong oral and written communication skills
- Strong interpersonal and leadership skills
- Ability to analyze and resolve problems
- Ability to lead programs/projects
- Ability to document, plan, market, and execute programs
- Established project management skills
- Hands-on familiarity with AI/ML platforms, monitoring tools, and cloud infrastructure such as Azure, AWS, or GCP
- Experience establishing support through purchased services agreements and coordinating L1/L2/L3 support
- Experience setting up PagerDuty rotations and support workstreams
- Legally authorized to work in the United States
- Successful completion of a drug screen, as applicable
Benefits
Comp & perks- Discretionary annual bonus
- Medical coverage
- Dental coverage
- Vision coverage
- Prescription drug coverage
- Health Coach access
- Employee Assistance Program
- GE Vernova Retirement Savings Plan
- Tax-advantaged 401(k) savings opportunity with company matching contributions
- Company retirement contributions
- Fidelity resources and financial planning consultants
- Tuition assistance
- Adoption assistance
- Paid parental leave
- Disability benefits
- Life insurance
- 12 paid holidays
- Permissive time off
- Relocation assistance