Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Ad Hoc LLC

Site Reliability Engineer

Ad Hoc LLC

Site Reliability Engineer ensuring the availability, performance, and reliability of a federal enterprise cloud platform. Collaborating with teams to enhance government technology and services.

Posted 7/1/2026full-timeRemote • 🇺🇸 United StatesMid-LevelSenior💰 $125,000 - $135,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in cloud platform reliability, performance monitoring, and observability tooling, with a strong focus on automation and infrastructure as code using Terraform. Capable of collaborating with government partners to meet security and performance standards while participating in incident response and capacity planning.

Highest-signal resume keywords
Cloud Infrastructure (AWS)Infrastructure as Code (Terraform)Monitoring and Observability ToolingIncident ResponseDevOps Concepts

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Performance TuningCapacity PlanningAutomation of Operational TasksService Level Objectives (SLOs)Error BudgetsKubernetes (Amazon EKS)ContainerizationNetworking
Soft Skills
CollaborationProblem-SolvingCommunication
Tools & Technologies
MetricsLoggingAlertingDashboards
Certifications & Qualifications
U.S. Public Trust / Suitability Determination
Industry Keywords
Federal Enterprise Cloud PlatformBlameless PostmortemsOn-Call OperationsService Level Agreements (SLAs)

Tech Stack

Tools & technologies
AWSCloudKubernetesTerraform

About the role

Key responsibilities & impact
  • Help ensure the availability, performance, and reliability of a large federal enterprise cloud platform
  • Monitor platform health and support service level objectives (SLOs), service level indicators, and error budgets
  • Build and maintain observability tooling, including metrics, logging, alerting, and dashboards
  • Participate in on-call rotations and incident response, helping restore service and reduce time to recovery
  • Contribute to blameless postmortems and drive follow-up actions
  • Automate repetitive operational tasks to reduce toil
  • Support capacity planning and performance tuning across cloud infrastructure (AWS) and Kubernetes (Amazon EKS)
  • Implement reliability improvements as infrastructure as code (Terraform)
  • Work with government partners and application teams to meet security, SLA, and performance requirements
  • Support recruiting efforts by evaluating exercises and assisting with interviews

Requirements

What you’ll need
  • Bachelor's and 5+ years of experience; relevant experience may be substituted for education
  • Experience with monitoring and observability tooling and on-call operations
  • Proficient with at least one infrastructure-as-code tool (Terraform preferred)
  • Background in key DevOps concepts: containerization, networking, and cloud infrastructure
  • Must be able to obtain and maintain a U.S. Public Trust / suitability determination

Benefits

Comp & perks
  • Company-subsidized health, dental, and vision insurance
  • Flexible PTO
  • 401K with employer match
  • Paid parental leave after one year of service
  • Employee Assistance Program