Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Lambda

Senior Platform Engineer – Core Infrastructure

Lambda

Senior Platform Engineer developing AI cloud infrastructure solutions for Lambda, enhancing Kubernetes clusters and cloud-native services. Leading incident response and mentoring engineering teams.

Posted 7/21/2026full-timeSan Francisco • California, Washington • 🇺🇸 United StatesSenior💰 $230,000 - $340,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in architecting, deploying, and managing Kubernetes clusters in production environments, with a strong focus on automation, security, and observability. Proven ability to mentor engineers and lead incident response while collaborating with product teams to design scalable cloud-native services.

Highest-signal resume keywords
Kubernetes InternalsInfrastructure-As-Code (Terraform, Pulumi)Observability Stacks (Prometheus, Grafana, OpenTelemetry)Coding Skills in Go or PythonHelm or Kustomize

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
KubernetesAutomationCI/CD PipelinesNetworkingService MeshesContainer RuntimesGitOpsNetwork PoliciesSecrets ManagementImage Scanning
Soft Skills
MentoringIncident ResponseRoot-Cause Analysis
Tools & Technologies
AWSLambdaTerraformPulumiPrometheusGrafanaOpenTelemetry
Industry Keywords
Platform EngineeringInfrastructureSite Reliability Engineering (SRE)Cloud-Native Services

Tech Stack

Tools & technologies
AWSCloudGoGrafanaKubernetesPrometheusPythonTerraform

About the role

Key responsibilities & impact
  • Architect, deploy, and operate Kubernetes clusters across AWS and Lambda's bare-metal datacenters.
  • Build and maintain automation for cluster lifecycle management — provisioning, upgrades, and scaling.
  • Own the reliability, performance, and security of Kubernetes workloads in production.
  • Implement observability, logging, and alerting for clusters and critical workloads.
  • Partner with product teams to design scalable, cloud-native services and CI/CD pipelines.
  • Set the standards for resource management, networking, and RBAC across the platform.
  • Lead incident response, root-cause analysis, and post-mortems for platform issues.
  • Mentor engineers and raise the bar for platform engineering across the org.

Requirements

What you’ll need
  • 5+ years in Platform, Infrastructure, or SRE roles, including running Kubernetes in production at scale.
  • Deep knowledge of Kubernetes internals and day-2 operations (upgrades, scaling, troubleshooting).
  • Strong with Helm, Kustomize, or similar, and GitOps-based delivery.
  • Proficient with infrastructure-as-code (Terraform, Pulumi, or equivalent).
  • Solid grounding in networking, service meshes, and container runtimes.
  • Hands-on with observability stacks (Prometheus, Grafana, OpenTelemetry).
  • Strong coding skills in Go or Python for automation and tooling.
  • Practical security experience: network policies, secrets management, and image scanning.

Benefits

Comp & perks
  • Health, dental, and vision coverage for you and your dependents
  • Wellness and commuter stipends for select roles
  • 401k Plan with 2% company match (USA employees)
  • Flexible paid time off plan that we all actually use