Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
NVIDIA

Senior Software Engineer, DGX Cloud Orchestration

NVIDIA

Senior Software Engineer designing scalable automation solutions for NVIDIA’s GPU infrastructure. Collaborating across teams and optimizing cloud operations for global workflows.

Posted 7/23/2026full-timeRemote • California • 🇺🇸 United StatesSenior💰 $184,000 - $356,500 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and developing APIs, cloud infrastructure, and container orchestration tools to enhance operational workflows and system efficiency. Proven ability to lead technical projects while ensuring quality and scalability in high-reliability environments.

Highest-signal resume keywords
API Design And DevelopmentCloud Infrastructure (AWS, GCP, Azure)Container Orchestration (Kubernetes, Docker)Programming Languages (Go, Java, Python)High-Scale Distributed Systems

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
API DesignWorkflow AutomationSchema-Driven PlatformsCloud Operations OptimizationDistributed Systems Architecture
Soft Skills
Outstanding CommunicationCollaboration SkillsProblem Solving
Tools & Technologies
KubernetesPrometheusOpenTelemetryGrafanaTelemetry Systems
Industry Keywords
Operational WorkflowsInfrastructure Lifecycle ProcessesManual Process AutomationReliability Engineering

Tech Stack

Tools & technologies
AWSAzureCloudDistributed SystemsDockerGoGoogle Cloud PlatformGrafanaJavaKubernetesPrometheusPython

About the role

Key responsibilities & impact
  • Design and develop APIs to orchestrate and integrate operational workflows.
  • Build state management and workflow automation systems that streamline infrastructure lifecycle processes.
  • Collaborate across teams to codify business processes into scalable, self-measuring systems.
  • Develop extensible, schema-driven platforms for reducing manual toil and ensuring operational consistency.
  • Drive integrations with container orchestration tools like Kubernetes and observability systems such as Prometheus, OpenTelemetry, Grafana.
  • Optimize the reliability and efficiency of cloud operations through automated workflows and telemetry systems.
  • Lead and ship impactful technical projects, ensuring quality and scalability at every stage.

Requirements

What you’ll need
  • 8+ years of industry experience with a Bachelor’s or Master’s degree (or equivalent experience), or 2+ years with a PhD.
  • Expertise in designing, building, and operating services in a high reliability environment.
  • Proficiency in programming languages such as Go, Java, or Python.
  • Strong understanding of cloud infrastructure (AWS, GCP, Azure) and container technologies like Docker and Kubernetes.
  • Experience with high-scale distributed systems, including architectural patterns for APIs and data pipelines.
  • Outstanding communication and collaboration skills, with a focus on solving complex operational challenges.
  • A passion for automating manual processes and driving system efficiency.

Benefits

Comp & perks
  • equity
  • benefits 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score