Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
TensorWave

Senior Solutions Engineer

TensorWave

Senior Solutions Engineer resolving complex customer technical issues in high-performance AI environments. Collaborating closely with operations and engineering teams for seamless solutions.

Posted 7/28/2026full-timeRemote • Nevada • 🇺🇸 United StatesSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Infrastructure Engineering and Platform Engineering with a focus on high-performance computing and large-scale AI stacks. Proficient in Kubernetes, AI/GPU infrastructure, and Linux, with strong capabilities in problem-solving, customer engagement, and technical communication.

Highest-signal resume keywords
Kubernetes ExpertAI/GPU Infrastructure SpecialistNetwork PathologistLinux Power UserBuilder Mindset

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Kubernetes AdministrationGPU Workload OrchestrationPython ProgrammingAnsible AutomationKernel NetworkingDebugging at OS LayerRDMA/RoCEv2SRIOVBGPCustom Diagnostic Tool Development
Soft Skills
Executive CommunicationActive CollaborationProblem SolvingCustomer Engagement
Industry Keywords
Infrastructure EngineeringPlatform EngineeringSREHigh-Performance ComputingLarge-Scale AI StacksProduction EnvironmentsSystem Reliability24/7/365 Availability

Tech Stack

Tools & technologies
AnsibleKubernetesLinuxPython

About the role

Key responsibilities & impact
  • Resolve Complex Escalations: Act as the final authority on issues exceeding GOC scope, utilizing code-level debugging and architectural investigation.
  • Direct Customer Engagement: Partner with customer technical leads to diagnose production issues, ensuring transparency and rapid resolution through active collaboration.
  • Iterative Problem Solving: Develop diagnostic scripts and workarounds to maintain customer operations while long-term patches are in development.
  • Drive Root Cause Analysis: Own end-to-end P1 resolution, partnering with TAMs to deliver clear, actionable post-incident analysis.
  • Bridge to Engineering: Convert recurring customer pain points into evidence-based feature requests, influencing product roadmap to resolve systemic failures.
  • Build Scalable Knowledge: Document non-obvious platform behaviors and refine GOC runbooks, ensuring institutional knowledge grows with every incident.

Requirements

What you’ll need
  • 5–9 years in Infrastructure Engineering, Platform Engineering, or SRE, with a specific focus on high-performance computing or large-scale AI stacks. Proven track record of managing complex production environments where system reliability is mission-critical.
  • Kubernetes Expert: Deep experience in cluster administration and scheduler internals; comfortable reading/modifying controller code.
  • AI/GPU Infrastructure Specialist: Proficient in orchestrating GPU workloads and diagnosing training job failures using ROCm or CUDA.
  • Network Pathologist: Skilled in RDMA/RoCEv2, SRIOV, and BGP; capable of interpreting switch telemetry to identify silent packet drops.
  • Linux Power User: Expert in kernel networking, hugepages, and cgroups; able to debug at the OS layer when applications are silent.
  • Builder Mindset: Proficient in Python and Ansible; capable of writing custom diagnostic tools to automate remediation.
  • Executive Communicator: Strong technical rigor when presenting findings to VPs of Engineering, maintaining trust while delivering difficult updates.
  • Prior experience in a customer-facing engineering role (e.g., Solutions Engineering, Technical Support Engineering).
  • Experience in high-uptime environments where 24/7/365 availability is required.

Benefits

Comp & perks
  • Stock Options
  • 100% paid Medical, Dental, and Vision insurance for Employees
  • Company Health Savings Account Contributions
  • 100% paid Short Term and Long Term Disability Insurance for Employees
  • Life and Voluntary Supplemental Insurance Options
  • Other Insurance Options, such as Pet & Legal Insurance
  • Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support
  • Flexible Spending Account
  • 401(k)
  • Employee Assistance Program
  • Flexible PTO
  • Paid Holidays
  • Parental Leave
  • Other In-Office Perks