Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Pinterest

Senior Site Reliability Engineer

Pinterest

Senior Site Reliability Engineer at Pinterest ensuring reliability of cloud-native platforms on AWS and Kubernetes. Collaborating with teams to improve operational practices and incident responses.

Posted 7/30/2026full-timeRemote • California • 🇺🇸 United StatesSenior💰 $139,764 - $287,749 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering and DevOps practices, with a strong focus on Kubernetes management, AWS operations, and CI/CD automation. Proven ability to enhance infrastructure reliability and performance through effective incident management, observability, and collaboration across teams.

Highest-signal resume keywords
Kubernetes ManagementAWS OperationsGitOps with ArgoCDTerraform/TerragruntCI/CD Automation with GitHub Actions

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Site Reliability EngineeringDevOpsKubernetesAWSTerraformArgoCDHelmBash ScriptingPython ScriptingCI/CD Pipelines
Soft Skills
CollaborationCommunicationOwnership MindsetTroubleshootingCritical Evaluation
Tools & Technologies
GitHub ActionsTerraform/TerragruntArgoCDKubernetesAWS
Industry Keywords
Infrastructure ReliabilityObservabilityIncident ResponseMulti-TenancyIAM

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsKubernetesLinuxPythonTerraform

About the role

Key responsibilities & impact
  • Ensuring the reliability, availability, and performance of production infrastructure and platform services
  • Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads
  • Managing GitOps-based deployment workflows using ArgoCD and Helm
  • Driving infrastructure provisioning and change management through Terraform/Terragrunt
  • Building and supporting CI/CD automation and deployment workflows using GitHub Actions
  • Leading incident response efforts, root cause analysis, and post-incident improvement initiatives
  • Reducing operational toil through scripting, tooling, and process automation
  • Advancing observability practices across logs, metrics, traces, dashboards, and alerting
  • Supporting secure secrets integration, IAM-aware operations, and platform guardrails
  • Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes

Requirements

What you’ll need
  • 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
  • Strong hands-on experience operating AWS in production environments
  • Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
  • Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
  • Experience implementing and operating ArgoCD within a GitOps delivery model
  • Strong hands-on experience with Helm
  • Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management
  • Solid scripting and automation skills using Bash and/or Python
  • Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
  • Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
  • Experience with monitoring, alerting, and observability in production environments
  • Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages
  • Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
  • Bachelor’s degree in computer science, engineering, a related field or equivalent experience
  • Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs
  • Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review)
  • High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables.

Benefits

Comp & perks
  • Equity
  • Flexibility to do your best work
  • Professional development opportunities