Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Alpaca

Senior Site Reliability Engineer

Alpaca

Senior SRE strengthening PostgreSQL, Kubernetes, and observability for Alpaca’s global brokerage infrastructure. Operating production systems, improving reliability, and mentoring engineers across cloud and database operations.

Posted 8/5/2026full-timeRemote • 🏈 Anywhere in North AmericaSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering (SRE) and DevOps practices, with a strong focus on PostgreSQL reliability, Kubernetes operations, and incident response. Proficient in shipping infrastructure as code using GitOps workflows and enhancing observability across production services.

Highest-signal resume keywords
Site Reliability Engineering (SRE)Kubernetes OperationsPostgreSQL ReliabilityGitOps WorkflowIncident Response

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
PostgreSQLKubernetesGitOpsGoPythonLinuxCloud NetworkingObservabilityPerformance TuningSchema Review
Soft Skills
Strong CommunicationMentoring
Tools & Technologies
Cloud ResourcesVPCsLoad BalancingDNSTLS
Industry Keywords
Production OperationsIncident ResponseStructured DebuggingPostmortemsRegulated Environments

Tech Stack

Tools & technologies
CloudDNSGoKubernetesLinuxPostgresPythonSQL

About the role

Key responsibilities & impact
  • Operate production day-to-day, including on-call, incident response, postmortems, and follow-up actions
  • Define and refine SLIs/SLOs and error budgets, helping product teams operate within them
  • Strengthen observability across metrics, logs, traces, and alerting
  • Ship cloud resources and Kubernetes workloads through code in a GitOps workflow
  • Own PostgreSQL reliability through performance tuning, schema and migration review, online migrations on large tables, HA/DR, and CDC pipelines
  • Mentor engineers on reliability and database fundamentals through code review, design review, and pairing

Requirements

What you’ll need
  • 4+ years in SRE, DevOps, Platform/Infrastructure, or backend engineering with significant production operations ownership
  • Hands-on experience operating production services on Kubernetes and shipping infrastructure as code in a GitOps workflow
  • Solid working knowledge of PostgreSQL in production, including query plans, pg_stat_*, indexing, schema trade-offs, and safe online migrations
  • Cloud networking fundamentals: VPCs, routing, L4/L7 load balancing, DNS, and TLS
  • Comfortable with a modern observability stack
  • Proficient with Linux at the operator level
  • Practiced in incident response, structured debugging, and postmortems
  • At least working proficiency in Go or Python
  • Strong written and verbal communication
  • Genuine interest in databases and growing PostgreSQL/DBA expertise
  • Nice-to-haves include deeper PostgreSQL experience, typed SQL access layers in Go, large-scale messaging systems, security and compliance in regulated environments, and trading, brokerage, or regulated fintech familiarity

Benefits

Comp & perks
  • Competitive Salary & Stock Options
  • Health Benefits
  • New Hire Home-Office Setup: One-time USD $500
  • Monthly Stipend: USD $150 per month via a Brex Card