FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Software Engineer – Core Cloud Platform
LambdaSenior Software Engineer building control-plane systems for Lambda’s GPU cloud infrastructure. Collaborating with multiple teams for operational excellence in AI cloud capacity.
Posted 7/28/2026full-timeSan Francisco • California • 🇺🇸 United StatesSenior💰 $296,000 - $346,000 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in building and operating cloud platform services, with a strong focus on backend systems, API design, and operational readiness. Proficient in Python or Go, with a solid understanding of cloud infrastructure and distributed systems.
Highest-signal resume keywords
Backend DevelopmentAPI DesignCloud InfrastructureProduction DebuggingLinux and Kubernetes
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
PythonGoAPI DesignDistributed SystemsFault ToleranceState MachinesCI/CDObservabilityInfrastructure AutomationCapacity Management
Soft Skills
Clear CommunicationMentoringCollaboration
Tools & Technologies
KubernetesContainersCloud ServicesNetworkingStorage
Industry Keywords
Compute LifecycleOrchestration ServicesOperational ReadinessProduction ServicesIncident Management
Tech Stack
Tools & technologiesCloudDistributed SystemsGoKubernetesLinuxPython
About the role
Key responsibilities & impact- Build and operate core cloud platform services for compute lifecycle, bare metal hosts, capacity, placement, and maintenance workflows.
- Design reliable APIs, backend services, state machines, and orchestration systems that power Lambda’s GPU cloud.
- Work on bare metal lifecycle systems including launch, terminate, restart/reboot, host reclaim, validation, quarantine, and return-to-pool workflows.
- Improve deployment, observability, testing, alerting, runbooks, and operational readiness for business-critical control-plane services.
- Debug complex production issues across distributed services, infrastructure dependencies, networking, and cloud workflows.
- Partner with infrastructure, networking, fleet, security, support, and product teams to define cross-system contracts and deliver end-to-end cloud capabilities.
- Contribute to architecture, design docs, code reviews, incident follow-through, and mentoring across the team.
Requirements
What you’ll need- Bachelor's degree or equivalent working experience.
- Have 6+ years of professional software engineering experience building production backend or distributed systems.
- Are strong in Python, Go, or a similar backend/system language.
- Have experience designing and operating APIs, workflow engines, schedulers, orchestration services, or other distributed systems.
- Understand reliability fundamentals: fault tolerance, idempotency, retries, state machines, failure handling, and production debugging.
- Have experience with cloud or cloud-like infrastructure primitives such as compute, networking, storage, capacity management, identity, or fleet operations.
- Are comfortable with Linux, containers, Kubernetes, infrastructure automation, and service deployment patterns.
- Have owned production services, participated in on-call, and improved systems based on operational learnings.
- Care about testability, CI/CD, observability, metrics, logging, alerting, and supportable operations.
- Can take ambiguous infrastructure problems and drive them to clear designs, implementation plans, and production outcomes.
- Communicate clearly across engineering, product, support, infrastructure teams, and leadership.
Benefits
Comp & perks- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use