Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Plaud

SRE Engineer

Plaud

SRE Engineer ensuring reliability and performance of Plaud.ai's AI products at scale. Designing cloud-native systems and managing reliability practices for AI workloads.

Posted 7/22/2026full-timeSeattle • Washington • 🇺🇸 United StatesMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering (SRE) with a focus on building and maintaining scalable cloud-native systems. Proficient in incident management, observability, and reliability automation to enhance operational maturity.

Highest-signal resume keywords
Site Reliability Engineering (SRE)Cloud Platforms (AWS/GCP/Azure/OCI)KubernetesIncident ManagementProgramming (Go, Python, Java)

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Site Reliability EngineeringCloud PlatformsKubernetesDistributed SystemsIncident ManagementReliability AutomationSLOsSLIsError BudgetsObservability

Tech Stack

Tools & technologies
AWSAzureCloudDistributed SystemsGoGoogle Cloud PlatformJavaKubernetesPython

About the role

Key responsibilities & impact
  • Ensure reliability and performance of Plaud.ai’s AI products at scale
  • Design and operate highly available, scalable cloud-native systems for AI workloads
  • Own production reliability, incident response, and on-call practices
  • Build observability (metrics, logs, tracing) and reliability automation
  • Define and manage SLOs, SLIs, and error budgets with engineering teams
  • Drive postmortems and reliability improvements across the platform
  • Lead incident response and continuous reliability improvement
  • Partner with product and engineering teams on reliability design
  • Improve observability and operational maturity

Requirements

What you’ll need
  • 5+ years in SRE, Infra, or Platform Engineering roles
  • Strong experience with cloud platforms (AWS/GCP/Azure/OCI)
  • Hands-on with Kubernetes and distributed systems
  • Experience in on-call rotation and incident management
  • Proficient in at least one programming language (Go, Python, Java)

Benefits

Comp & perks
  • Meaningful Ownership An Employee Stock Ownership Plan (ESOP) that gives a real stake in Plaud’s long-term success.
  • High-Impact Environment Work in a fast-moving, product-driven environment where your ideas directly shape the future of AI productivity.
  • Comprehensive Health & Retirement Benefits Top-tier medical, dental, and vision insurance for employees and dependents, supported by a generous employer subsidy, plus a 401(k) retirement plan with company matching for full-time employees.
  • Time Off & Workplace Benefits Unlimited PTO, plus 13 paid holidays, 12 weeks of fully paid parental leave for all parents, a hybrid work model with a minimum of three in-office days per week, and access to high-quality office snacks, drinks, and equipment.
  • Cutting-Edge AI Tools for Productivity Access to best-in-class AI tools, including Cursor, GPT models, Gemini, Claude, and other frontier AI systems to maximize engineering and execution efficiency.
  • Best-in-Class Equipment Choice of top-spec laptops, high-performance workstation setups, and cutting-edge Plaud devices for all new hires.
  • Team & Culture Annual company offsites, team events, and a culture that values craftsmanship, ownership, and velocity.