Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Gen

Staff Platform Machine Learning Engineer – Engine

Gen

Staff Platform Machine Learning Engineer developing and maintaining infrastructure for ML systems at Gen. Focused on reliability, performance, and tooling for machine learning workloads.

Posted 7/28/2026full-timeNew York City • New York • 🇺🇸 United StatesLead💰 $197,600 - $218,400 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and maintaining infrastructure for machine learning workloads, with a strong focus on reliability, performance, and cost efficiency. Proficient in cloud infrastructure, CI/CD, and operational tooling to enhance developer workflows and support model deployment.

Highest-signal resume keywords
Cloud InfrastructureKubernetesCI/CDMachine Learning InfrastructureSystems Engineering

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Software EngineeringPlatform EngineeringDevOpsBackend ServicesProduction OperationsDistributed SystemsModel ServingData PlatformsInfrastructure-as-CodeAutomation
Soft Skills
Problem-SolvingOperations-First MindsetAdaptability
Tools & Technologies
AWSDockerTerraformMonitoring Tools
Certifications & Qualifications
Bachelor’s DegreeMaster’s Degree
Industry Keywords
Machine LearningObservabilityOperational ToolingData Workflows

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsDockerKubernetesTerraform

About the role

Key responsibilities & impact
  • Design, build, and maintain infrastructure supporting ML training, deployment, and inference workloads.
  • Own and improve CI/CD, infrastructure-as-code, observability, and operational tooling for ML systems.
  • Operate and evolve production systems with a strong focus on reliability, performance, and cost efficiency.
  • Build and maintain Kubernetes- and cloud-based services that support model execution and data workflows.
  • Partner with data scientists and engineers to support model deployment and operational needs.
  • Improve developer workflows through automation, tooling, and platform abstractions.
  • Participate in on-call and operational ownership for platform components.

Requirements

What you’ll need
  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field
  • 7+ years of experience in software engineering, platform engineering, or DevOps/SRE roles
  • Strong systems engineering background with production operations experience
  • Strong experience with backend or platform services
  • Hands-on experience with cloud infrastructure and tooling (e.g., AWS, Docker, Kubernetes, Terraform, CI/CD, monitoring)
  • Experience operating distributed systems in production
  • Experience with machine learning infrastructure, model serving, or data platforms is a strong plus
  • Strong problem-solving skills and an operations-first mindset
  • Ability to thrive in a fast-paced, high-tech environment and manage complex problems

Benefits

Comp & perks
  • Flexible working options
  • Time off
  • Competitive pay
  • Benefits
  • Well-being programs