Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Elicit

Infrastructure Engineer

Elicit

Own and evolve Elicit’s cloud infrastructure, enhancing reliability and cost-efficiency for enterprise clients in AI. Drive infrastructure improvements focusing on observability and compliance, supporting significant enterprise deployments.

Posted 7/23/2026full-timeRemote • California • 🇺🇸 United StatesMid-LevelSenior💰 $185,000 - $260,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in managing cloud infrastructure across AWS and GCP, with a strong focus on Kubernetes, Terraform, and CI/CD practices. Capable of driving compliance and security operations while enhancing developer experience through effective observability and incident response strategies.

Highest-signal resume keywords
AWS ExperienceKubernetes ExpertiseTerraform ProficiencyCI/CD FluencySRE Mindset

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Infrastructure As Code (IaC)Kubernetes Clusters ManagementTerraformCI/CD Pipeline OptimizationObservability Stack DevelopmentIncident Response ProcessesDatabase Restoration DrillsSecurity Event MonitoringCost ManagementCapacity Planning
Soft Skills
Problem-SolvingCollaborationAdaptability
Tools & Technologies
AWSGCPCloudflareDataDogArgo CDGitHub ActionsSIEM ToolingKarpenterAurora PostgreSQLMongoDB Atlas
Certifications & Qualifications
SOC 2NIST AI Framework
Industry Keywords
Cloud InfrastructureComplianceSecurity OperationsDisaster RecoveryMonitoring

Tech Stack

Tools & technologies
AWSCloudGoogle Cloud PlatformKubernetesMongoDBPostgresRedisTerraform

About the role

Key responsibilities & impact
  • Own our cloud infrastructure across AWS and GCP — Kubernetes clusters, networking, databases (Aurora PostgreSQL, Redis, MongoDB Atlas), Cloudflare, and our CI/CD pipeline.
  • Scale single-tenant deployments from a handful to many — each with distinct data retention, geographic, monitoring, and compliance requirements. Make a private cloud deployment a repeatable, low-overhead operation.
  • Build our observability and incident response practice — proactive monitoring, alerting, SLA tracking, and structured post-mortems that make the whole team better at diagnosing and resolving issues.
  • Drive compliance and security operations — ensure we follow through on the policies we've written (SOC 2, NIST AI framework, EU Cyber Resilience). Own disaster recovery exercises, database restoration drills, and security event monitoring (SIEM).
  • Manage infrastructure cost and capacity — make smart decisions about where we run workloads (AWS, CoreWeave, Parasail), optimize spend, and plan capacity as usage grows.
  • Improve developer experience — CI/CD pipeline performance, preview environments, local development tooling, and deployment confidence.
  • Contribute to backend systems where infrastructure and application intersect — circuit breakers, inference routing, data connector infrastructure for enterprise customers bringing their own data.

Requirements

What you’ll need
  • 5+ years of hands-on infrastructure/SRE/platform engineering experience.
  • An AI-native way of working. Agentic coding tools (Claude Code, Cursor, Devin, etc.) are how we build at Elicit, and infrastructure is no exception: agents help us write IaC and investigate incidents. You should be an enthusiastic practitioner who uses AI to multiply your impact, and excited to find new places agents can safely take on infrastructure work.
  • Solid Terraform experience. this is our primary infrastructure-as-code layer and the most important technical requirement.
  • Strong Kubernetes expertise. You've operated production clusters, not just deployed to them. Comfortable with EKS, networking, autoscaling (Karpenter), and debugging cluster-level issues.
  • AWS experience (primary), with GCP familiarity a plus.
  • GitOps and CI/CD fluency. Argo CD, GitHub Actions, or equivalent. You understand deployment automation, rollback strategies, and change management.
  • SRE mindset. You've built or significantly improved observability stacks (DataDog or equivalent), incident response processes, and on-call practices.
  • Security and compliance awareness. Experience with SOC 2 or similar frameworks, SIEM tooling, and translating compliance requirements into engineering practice.
  • Ability to write software. You can contribute to our backend codebases where infrastructure meets application logic.

Benefits

Comp & perks
  • Flexible work environment: work from our office in Oakland or remotely with time zone overlap (between GMT and GMT-8), as long as you’re comfortable traveling for quarterly in-person offsites.
  • Fully covered health, dental, vision, and life insurance for you, generous coverage for the rest of your family (FSA/HSA, too).
  • Flexible vacation policy, with a minimum recommendation of 20 days / year and plenty of company holidays.
  • Every Elician receives a $200 monthly wellbeing stipend to spend on whatever supports your health and wellbeing.
  • 401K with a 6% employer match.
  • A new Mac + $1,000 budget to set up your workstation or home office in your first year, then $500 every year thereafter.
  • $1,000 quarterly AI Experimentation & Learning budget, so you can freely experiment with new AI tools to incorporate into your workflow, take courses, purchase educational resources, or attend AI-focused conferences and events.
  • A team administrative assistant who can help you with personal and work tasks.