Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Demandbase

DevOps Engineer, AI Runtime Services

Demandbase

DevOps Engineer managing AI Runtime Services for Demandbase, a leading pipeline AI platform that's empowering GTM teams. Building and supporting internal AI solutions while ensuring system reliability.

Posted 7/24/2026full-timeHyderabad • 🇮🇳 IndiaMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in production infrastructure management, particularly with Kubernetes, AWS, and Terraform, while ensuring reliability and observability in LLM systems. Proficient in Python for service ownership and tooling, with a strong background in incident response and cost control.

Highest-signal resume keywords
KubernetesAWSTerraformPythonIncident Response

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Production InfrastructureOn-Call ExperienceObservability FluencyCost ControlLLM Systems Familiarity
Tools & Technologies
GitOps
Industry Keywords
SLOsIncident ResponseCachingGuardrailsEvalsExperimentation InfraTracesSpansPrompt/Response Capture

Tech Stack

Tools & technologies
AWSKubernetesPythonTerraform

About the role

Key responsibilities & impact
  • The LLM gateway. The single front door to every model provider we use
  • Reliability, for real. SLOs, on-call, incident response, and the postmortems for the runtime
  • Cost control. Per-team and per-model attribution
  • LLM observability. Traces, spans, prompt/response capture
  • Evals and experimentation infra
  • Caching and guardrails

Requirements

What you’ll need
  • Strong production infrastructure background: Kubernetes, AWS, Terraform, GitOps
  • Python, at a level where you're comfortable owning services and tooling in it
  • Real on-call and incident-response experience
  • Observability fluency beyond "we have dashboards"
  • Hands-on familiarity with how LLM systems actually work

Benefits

Comp & perks
  • Group Medical
  • Personal Accident
  • Term Life Insurance
  • Preventive healthcare covers dental, vision, and OPD needs
  • Strong mental health support
  • Fitness benefit
  • Car lease policy
  • Gratuity for long-term financial well-being