FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Platform Engineer – Core Infrastructure
LambdaSenior Platform Engineer developing AI cloud infrastructure solutions for Lambda, enhancing Kubernetes clusters and cloud-native services. Leading incident response and mentoring engineering teams.
Posted 7/21/2026full-timeSan Francisco • California, Washington • 🇺🇸 United StatesSenior💰 $230,000 - $340,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in architecting, deploying, and managing Kubernetes clusters in production environments, with a strong focus on automation, security, and observability. Proven ability to mentor engineers and lead incident response while collaborating with product teams to design scalable cloud-native services.
Highest-signal resume keywords
Kubernetes InternalsInfrastructure-As-Code (Terraform, Pulumi)Observability Stacks (Prometheus, Grafana, OpenTelemetry)Coding Skills in Go or PythonHelm or Kustomize
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
KubernetesAutomationCI/CD PipelinesNetworkingService MeshesContainer RuntimesGitOpsNetwork PoliciesSecrets ManagementImage Scanning
Soft Skills
MentoringIncident ResponseRoot-Cause Analysis
Tools & Technologies
AWSLambdaTerraformPulumiPrometheusGrafanaOpenTelemetry
Industry Keywords
Platform EngineeringInfrastructureSite Reliability Engineering (SRE)Cloud-Native Services
Tech Stack
Tools & technologiesAWSCloudGoGrafanaKubernetesPrometheusPythonTerraform
About the role
Key responsibilities & impact- Architect, deploy, and operate Kubernetes clusters across AWS and Lambda's bare-metal datacenters.
- Build and maintain automation for cluster lifecycle management — provisioning, upgrades, and scaling.
- Own the reliability, performance, and security of Kubernetes workloads in production.
- Implement observability, logging, and alerting for clusters and critical workloads.
- Partner with product teams to design scalable, cloud-native services and CI/CD pipelines.
- Set the standards for resource management, networking, and RBAC across the platform.
- Lead incident response, root-cause analysis, and post-mortems for platform issues.
- Mentor engineers and raise the bar for platform engineering across the org.
Requirements
What you’ll need- 5+ years in Platform, Infrastructure, or SRE roles, including running Kubernetes in production at scale.
- Deep knowledge of Kubernetes internals and day-2 operations (upgrades, scaling, troubleshooting).
- Strong with Helm, Kustomize, or similar, and GitOps-based delivery.
- Proficient with infrastructure-as-code (Terraform, Pulumi, or equivalent).
- Solid grounding in networking, service meshes, and container runtimes.
- Hands-on with observability stacks (Prometheus, Grafana, OpenTelemetry).
- Strong coding skills in Go or Python for automation and tooling.
- Practical security experience: network policies, secrets management, and image scanning.
Benefits
Comp & perks- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use