Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Vynca

Site Reliability Engineer

Vynca

Site Reliability Engineer designing and managing scalable AWS infrastructure for Vynca's healthcare tech platform. Collaborating with Software Engineers and improving system reliability through automation and observability practices.

Posted 7/30/2026full-timeRemote • Arizona, California, Colorado, Florida, Illinois, Nevada, North Carolina, Oregon, Texas, Utah, Washington • 🇺🇸 United StatesMid-LevelSenior💰 $140,000 - $150,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in AWS infrastructure management using Terraform, Kubernetes operations, and observability practices. Proficient in deploying and managing applications with Helm while ensuring compliance with security standards such as HIPAA and SOC 2.

Highest-signal resume keywords
AWS Infrastructure ManagementTerraform Infrastructure as CodeKubernetes Production SupportObservability Practices ImplementationCI/CD Pipeline Development

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
AWSTerraformKubernetesHelmDistributed SystemsEvent-Driven ArchitecturesLinux Systems AdministrationMonitoringLoggingCI/CD
Soft Skills
Problem-SolvingCommunicationCollaborationOwnershipAccountability
Industry Keywords
Site Reliability EngineeringDevOps EngineeringPlatform EngineeringCloud Infrastructure EngineeringHIPAA ComplianceSOC 2 Compliance

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsKubernetesLinuxTerraform

About the role

Key responsibilities & impact
  • Design, provision, and manage AWS infrastructure using Terraform as the source of truth
  • Operate, maintain, and scale production workloads running on Kubernetes
  • Package, deploy, and manage applications using Helm and infrastructure automation tools
  • Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms
  • Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity
  • Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience
  • Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions
  • Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability
  • Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design
  • Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements
  • Participate in an on-call rotation and provide operational support for production systems

Requirements

What you’ll need
  • Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, Cloud Infrastructure Engineering, or similar infrastructure-focused roles
  • Bachelor’s degree in Computer Science, Information Systems, Software Engineering, or a related technical field; equivalent professional experience will also be considered
  • Strong hands-on experience operating production workloads within AWS environments
  • Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation
  • Experience operating and supporting production Kubernetes environments
  • Hands-on experience deploying and managing applications using Helm
  • Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance
  • Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response
  • Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals
  • Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation
  • Strong problem-solving skills with the ability to troubleshoot complex infrastructure and application issues
  • Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams
  • High level of ownership, accountability, and initiative with a proactive approach to reliability and operational excellence
  • Ability and willingness to participate in an on-call rotation supporting production systems

Benefits

Comp & perks
  • medical, dental, and vision insurance
  • income protection benefits
  • flexible PTO
  • company holidays
  • 401k
  • access to other wellness benefits