Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
10a Labs

Software Engineer – Infrastructure & Platform

10a Labs

Software Engineer building secure, scalable infrastructure for 10a Labs’ AI evaluations. Developing sandboxes, backend services, and agentic evaluation systems for frontier-model safety research.

Posted 8/18/2026full-timeRemote • 🇺🇸 United StatesMid-LevelSenior💰 $110,000 - $160,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in building and managing backend services and infrastructure for AI evaluations, with a strong focus on security, scalability, and reproducibility. Proficient in utilizing container orchestration technologies and cloud platforms to create efficient evaluation environments.

Highest-signal resume keywords
Python ProgrammingDocker ContainerizationKubernetes OrchestrationAWS Cloud InfrastructureInfrastructure-as-Code (Terraform)

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Backend DevelopmentAPI DesignDistributed Systems EngineeringDebugging SkillsLinux SystemsNetworking SecurityAutomation ToolsObservability TechniquesState ManagementMulti-Agent Workflows
Soft Skills
Problem-SolvingCollaborationAdaptability
Tools & Technologies
TerraformAWSGCPDockerKubernetesVirtual Machines
Industry Keywords
AI SystemsAgentic WorkflowsModel EvaluationsInfrastructure SecurityEvaluation Reliability

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsDockerGoogle Cloud PlatformKubernetesLinuxPythonTerraform

About the role

Key responsibilities & impact
  • Design and build sandboxed evaluation environments where AI models can safely execute code, use tools, interact with services, and complete complex tasks
  • Build backend services and infrastructure supporting large-scale, repeatable AI and agentic evaluations
  • Develop agent scaffolding and evaluation harnesses, including tool-use loops, context management, retries, state management, token budgets, and multi-agent or subagent workflows
  • Build systems for provisioning and orchestrating isolated environments using Docker, Kubernetes, VMs, and cloud infrastructure
  • Design secure approaches to networking, permissions, secrets, credentials, and resource isolation for model-driven environments
  • Develop APIs, internal tools, and automation that allow researchers, engineers, and subject-matter experts to create and run evaluations efficiently
  • Improve evaluation reliability and reproducibility through logging, observability, snapshotting, debugging tools, and automated testing
  • Build systems capable of running thousands of evaluation tasks reliably and capturing artifacts and telemetry needed to understand model behavior
  • Partner with analysts, red teamers, and domain experts to translate complex evaluation ideas into robust technical systems
  • Investigate failures across the evaluation stack and distinguish model limitations from infrastructure, harness, or environment failures

Requirements

What you’ll need
  • 3–5+ years of professional software engineering experience, particularly in backend, infrastructure, platform, SRE, or distributed systems engineering
  • Strong programming skills in Python and experience building production-quality software
  • Experience designing and operating backend services, APIs, or distributed systems
  • Hands-on experience with Docker, Kubernetes, virtual machines, or other container/orchestration technologies
  • Experience working with AWS, GCP, or similar cloud infrastructure
  • Strong understanding of Linux systems, networking, authentication, permissions, and infrastructure security
  • Experience with infrastructure-as-code or automation tools such as Terraform
  • Strong debugging skills across application, infrastructure, and networking layers, especially in agentic loops
  • Ability to build systems that are reproducible, observable, scalable, and secure
  • Comfort working on ambiguous technical problems where architecture and requirements may evolve quickly
  • Interest in AI systems, agentic workflows, AI security, or model evaluations; prior professional AI experience is helpful but not required

Benefits

Comp & perks
  • Performance-based annual bonus
  • Support for conferences, continuing education, or leadership training
  • Fully remote work environment
  • Comprehensive health, dental, and vision coverage
  • Generous PTO and paid holiday schedule
  • 401(k) plan