10a Labs Logo

10a Labs

Software Engineer, Infrastructure & Platform

Posted 15 Days Ago
Remote
Hiring Remotely in USA
110K-160K Annually
Mid level
Remote
Hiring Remotely in USA
110K-160K Annually
Mid level
Build and maintain secure, reproducible infrastructure and backend services for large-scale AI and agentic model evaluations. Design sandboxed execution environments, orchestration, observability, and APIs to run thousands of evaluation tasks while ensuring isolation, security, and reproducibility. Collaborate with researchers and red teamers to translate evaluation ideas into reliable systems and debug failures across the stack.
The summary above was generated by AI

About 10a Labs: 10a Labs is the safety and threat-intelligence layer trusted by frontier AI labs, AI unicorns, Fortune 10 companies, and leading global technology platforms. Our adversarial red teaming, model evaluations, and intelligence collection enable engineering, safety, and security teams to stay ahead of evolving threats and deploy AI systems safely.

Software Engineer, Infrastructure & PlatformAbout the Role

We are seeking a Software Engineer, Infrastructure & Platform to build the systems and infrastructure that power advanced AI evaluations, including evaluations focused on autonomous model behavior, agentic systems, and loss-of-control risks.

This is a hands-on engineering role at the intersection of backend systems, infrastructure, and AI. You will build secure and reproducible environments where frontier models can interact with tools, execute code, complete complex tasks, and operate across realistic multi-step workflows, including machine learning research and engineering.

The ideal candidate has strong backend and infrastructure fundamentals with attention to security, enjoys debugging complex distributed systems, and is excited to apply those skills to difficult problems in AI safety and evaluation.

What You'll Do
  • Design and build sandboxed evaluation environments where AI models can safely execute code, use tools, interact with services, and complete complex tasks.
  • Build backend services and infrastructure supporting large-scale, repeatable AI and agentic evaluations.
  • Develop agent scaffolding and evaluation harnesses, including tool-use loops, context management, retries, state management, token budgets, and multi-agent or subagent workflows.
  • Build systems for provisioning and orchestrating isolated environments using technologies such as Docker, Kubernetes, VMs, and cloud infrastructure.
  • Design secure approaches to networking, permissions, secrets, credentials, and resource isolation for model-driven environments.
  • Develop APIs, internal tools, and automation that allow analysts, engineers, and subject-matter experts to create and run evaluations efficiently.
  • Improve the reliability and reproducibility of evaluations through logging, observability, snapshotting, debugging tools, and automated testing.
  • Build systems capable of running thousands of evaluation tasks reliably and capturing the artifacts and telemetry needed to understand model behavior.
  • Partner with analysts, red teamers, and domain experts to translate complex evaluation ideas into robust technical systems.
  • Investigate failures across the evaluation stack and distinguish between model limitations and infrastructure, harness, or environment failures.
What We're Looking For
  • 3–5+ years of professional software engineering experience, particularly in backend, infrastructure, platform, SRE, or distributed systems engineering.
  • Strong programming skills in Python and experience building production-quality software.
  • Experience designing and operating backend services or distributed systems.
  • Hands-on experience with Docker, Kubernetes, virtual machines, or other container/orchestration technologies.
  • Experience working with GCP, AWS, or similar cloud infrastructure.
  • Strong understanding of Linux systems, networking, authentication, permissions, and infrastructure security.
  • Experience with infrastructure-as-code or automation tools such as Terraform.
  • Strong debugging skills and comfort diagnosing failures across application, infrastructure, and networking layers, especially in agentic loops.
  • Ability to build systems that are reproducible, observable, scalable, and secure.
  • Comfort working on ambiguous technical problems where the architecture and requirements may evolve quickly.
  • Interest in AI systems, agentic workflows, AI security, or model evaluations. Prior professional AI experience is helpful but not required.
Nice to Have
  • Experience building developer platforms, CI/CD systems, test infrastructure, sandboxes, or ephemeral compute environments.
  • Experience with agent frameworks, LLM APIs, tool-calling systems, or AI evaluation infrastructure.
  • Experience designing secure execution environments for untrusted or semi-trusted code.
  • Background in SRE, platform engineering, cloud infrastructure, cybersecurity, or developer tooling.
  • Experience with distributed task execution, queues, workflow orchestration, or large-scale automated testing.
  • Familiarity with AI safety, adversarial testing, model evaluations, or autonomous-agent systems.
  • Familiarity with agentic AI fundamentals, including common harnesses, Model Context Protocol, agent benchmarks, and security risks to AI agents.
Compensation & Benefits
  • Salary Range: $110K–$160K, depending on experience and location
  • Bonus: Performance-based annual bonus
  • Professional Development: Support for conferences, continuing education, or leadership training
  • Work Environment: Fully remote, U.S.-based
  • Health Benefits: Comprehensive health, dental, and vision coverage
  • Time Off: Generous PTO and paid holiday schedule
  • Retirement: 401(k) plan

Similar Jobs

Yesterday
Remote
United States
135K-165K Annually
Senior level
135K-165K Annually
Senior level
Aerospace • Manufacturing
Build and operate multicloud Kubernetes infrastructure and platform capabilities, including cloud foundations, networking, identity, cluster lifecycle, disconnected deployments, packaging, GitOps delivery, infrastructure as code, observability, upgrades, and incident response. Partner with engineering teams to provide reusable self-service infrastructure primitives across commercial, government, secure, classified, and disconnected environments.
Top Skills: Argo CdAWSCloud NetworkingDnsGateway ApiGCPGitopsHelmIamInfrastructure As CodeIpamKubernetesOpaOpentofuRbacService MeshTerraformWorkload Identity
Yesterday
Remote
United States
165K-195K Annually
Senior level
165K-195K Annually
Senior level
Aerospace • Manufacturing
Build and operate multicloud Kubernetes infrastructure and platform capabilities across commercial, government, classified, and disconnected environments. Responsibilities include cloud foundations, networking, identity, cluster lifecycle, packaging, GitOps delivery, infrastructure-as-code, policy enforcement, staged rollouts, observability, incident response, and backup/recovery. The engineer will also create self-service infrastructure primitives for internal engineering teams and support reliable fleet-scale deployments.
Top Skills: Argo CdAWSDnsFluxGateway ApiGCPGitopsHelmIamIpamKubernetesOpaOpentofuRbacService MeshTerraform
18 Days Ago
Remote
United States
190K-220K Annually
Senior level
190K-220K Annually
Senior level
Artificial Intelligence • Healthtech • Software
Own and evolve the AWS and Kubernetes platform, build Terraform modules, Helm charts, and GitHub Actions workflows, enhance platform capabilities with Go/Python, implement monitoring and runbooks, ensure secure CI/CD and data segregation, and support Node.js/React/Next.js engineering teams.
Top Skills: ArgocdAWSDatadogDockerDynamoDBElasticsearchGithub ActionsGoHelmJavaScriptKubernetesNext.JsNode.jsPandasPostgresPythonReactSnowflakeTerraformTypescript

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account