FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing and optimizing cloud operations within a microservices architecture, with a strong focus on Infrastructure as Code (IaC) and Site Reliability Engineering (SRE) practices. Proficient in deploying and maintaining secure production infrastructure using modern tools and methodologies.
Highest-signal resume keywords
Kubernetes ExpertiseDocker ProficiencyInfrastructure As Code (IaC) SkillsAWS ExpertiseLinux/Unix Administration
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Microservices ArchitectureCloud OperationsNetworking (VLANs, Routing, VPNs)Relational Databases (MySQL, SQL)PrometheusGrafanaAPM Tools (Dynatrace, Datadog, AppDynamics, New Relic)CI/CD Tools (Git, Jenkins)Automation (Ansible, Terraform)Metrics Collection and Alerting
Soft Skills
Problem-SolvingAdaptabilityCollaboration
Tools & Technologies
AWSKubernetesDockerLinuxTerraformAnsibleGitJenkinsPrometheusGrafana
Industry Keywords
SaaSIaaSSite Reliability Engineering (SRE)Infrastructure ManagementCloud Infrastructure
Tech Stack
Tools & technologiesAnsibleAWSCloudDockerGrafanaJenkinsKubernetesLinuxMicroservicesMySQLPrometheusSQLTerraformUnix
About the role
Key responsibilities & impact- Join the NVIDIA AIR team building a SaaS/IaaS platform for digital twins of AI data centers
- Handle DevOps, infrastructure, and Site Reliability Engineering (SRE) requirements for NVIDIA AIR
- Automate repetitive workflows to improve efficiency
- Work on a microservices-based architecture
- Deploy and troubleshoot non-disruptive cloud operations with emphasis on secure production infrastructure
- Continuously evaluate existing systems and drive improvements
- Manage deployment and upgrades for operating systems, Kubernetes clusters, and other orchestration tools
- Provide day-to-day support for engineering activities using CI/CD tools such as Git and Jenkins
- Manage multiple work tracks and evolving priorities
- Build or maintain metrics collection and alerting infrastructure
Requirements
What you’ll need- B.Sc. in Computer Science or equivalent experience
- 5+ years of experience in complex microservices based architectures
- Highly skilled in Kubernetes and Docker
- Experience in IaaS environments, including deploying, configuring, and administering Linux-based bare metal servers
- Strong networking background in VLANs, routing, and VPNs
- Experience with relational databases, MySQL, and SQL
- Experience with modern deployment architecture for non-disruptive cloud operations, including blue-green and canary rollouts
- Infrastructure as code (IaC) skills with frameworks such as Ansible and Terraform
- Expertise in AWS
- Knowledge of best practices for managing and monitoring highly available and secure production infrastructure
- Strong expertise in Infrastructure as a Service (IaaS)
- Linux/Unix administration skills
- Experience with Prometheus/Grafana
- Experience with APM tools such as Dynatrace, Datadog, AppDynamics, or New Relic
- Experience implementing robust metrics collection and alerting infrastructure
Benefits
Comp & perks- NVIDIA is described as a desirable employer with creative, passionate, and self-motivated teams
