Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Datavant

Site Reliability Engineer

Datavant

Site Reliability Engineer at Datavant focused on improving cloud infrastructure reliability and security. Collaborating across teams to enhance operational efficiency and effectiveness in a cloud environment.

Posted 7/29/2026full-timeRemote • 🇺🇸 United StatesMid-LevelSenior💰 $100,000 - $130,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering and DevOps practices, with a strong focus on cloud infrastructure, automation, and security compliance. Proficient in Infrastructure as Code and capable of leading teams to improve operational efficiency and reliability.

Highest-signal resume keywords
Site Reliability EngineeringInfrastructure as Code (Terraform, Ansible)Cloud Infrastructure ExpertiseSecurity Knowledge (IAM, Encryption, Compliance)CI/CD Pipeline Development

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Site Reliability EngineeringInfrastructure as CodeCloud InfrastructureSystems Language (Python, Go)CI/CD Pipeline DevelopmentCode ReviewOperational Problem SolvingAutomationNetwork SecurityWorkload Migration
Soft Skills
Strong Communication SkillsTeam LeadershipMentoring
Tools & Technologies
TerraformAnsibleCI/CD ToolsAI Agents
Certifications & Qualifications
SOC2HITRUSTNIST Compliance
Industry Keywords
Cloud EnvironmentsOperational ConsistencyGovernance FrameworksHybrid-Cloud StrategiesMulti-Cloud Footprint

Tech Stack

Tools & technologies
AnsibleAWSAzureCloudDNSGoGoogle Cloud PlatformLinuxPythonTerraform

About the role

Key responsibilities & impact
  • Lead the design and implementation of reliability improvements across assigned services, with minimal guidance
  • Identify systemic inefficiencies in architecture, implementation, and operational process and drive solutions
  • Implement customized solutions to complex operational problems derived from technical requirements
  • Review code, systems, and configuration with a focus on efficiency gains, optimization, and best practices and hold peers to those standards
  • Own SLO/SLI definitions for assigned services and drive teams toward meeting and improving those targets
  • Lead incident response, facilitate postmortems, and ensure action items result in durable reliability improvements
  • Support the integration of newly acquired cloud environments into Datavant's existing infrastructure, ensuring reliability, security, and operational consistency from day one
  • Contribute to the implementation of hybrid-cloud and cross-cloud connectivity strategies that ensure interoperability across a growing multi-cloud footprint
  • Help maintain and expand the modular network security edge, keeping it flexible enough to absorb additional environments as M&A activity requires
  • Drive standardization of infrastructure and operational practices across consolidated environments, reducing fragmentation and toil
  • Partner with security and platform engineering teams to ensure newly integrated environments adhere to governance frameworks and automation-first principles
  • Develop tools and processes to improve team service delivery including scaling, resiliency, efficiency, visibility, quality, and operations management
  • Analyze service delivery data and team feedback to drive meaningful improvements to development processes
  • Participate in and help evolve on-call practices, runbooks, and alerting strategy for assigned teams
  • Teach and lead more junior engineers on team processes, technical implementations, and SRE best practices
  • Build and maintain Infrastructure as Code (Terraform, Ansible) to support scalable, repeatable, and secure deployments
  • Develop and improve CI/CD pipelines, automated testing, and deployment tooling.

Requirements

What you’ll need
  • 5+ years of experience in site reliability engineering, DevOps, or platform/infrastructure engineering
  • Strong expertise in cloud infrastructure, including networking, security, compute, storage, and IAM
  • Experience supporting workload migrations and integrating new cloud environments into existing architectures
  • Hands-on experience with Infrastructure as Code and automation (Terraform and Ansible)
  • Strong security knowledge, including IAM, encryption, network security, and compliance frameworks (SOC2, HITRUST, NIST)
  • Demonstrated ability to solve complex operational problems independently and drive solutions end-to-end
  • Strong proficiency in at least one systems language (Python, Go, or similar) and comfort across multiple languages and configuration formats
  • Proven ability to conduct meaningful code and system reviews not just for correctness, but for efficiency and architectural soundness
  • Strong communication skills, including the ability to document technical decisions, address gaps in understanding, and build team alignment
  • Experience leveraging AI agents to accelerate daily workload.

Benefits

Comp & perks
  • N/A 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score