Jobs Companies Field Ai DevOps Engineer

Sobre esta vaga de DevOps Engineer na Field Ai

Field Ai · Presencial · Boston, MA
FieldAI is transforming how robots interact with the real world. Our growing R&D team is based in Boston, where we develop risk-aware, reliable, field-ready AI systems that tackle the hardest problems in robotics and unlock the potential of embodied intelligence. We take a pragmatic approach that goes beyond off-the-shelf, purely data-driven methods or transformer-only architectures, combining cutting-edge research with real-world deployment. Our solutions are already deployed globally, and we continuously improve model performance through rapid iteration driven by real field use.

About the Role

Field AI is transforming how robots interact with the real world. Our R&D team, the FieldAI Research Institute (FAIRI), is based in Cambridge, MA, where we build risk-aware, field-ready AI systems that unlock general purpose intelligence for robotics.

FAIRI is looking for a DevOps Engineer to own the infrastructure our humanoid research runs on. Today that infrastructure is borrowed from Field AI's commercial platform teams and held together by researchers doing it in the margins: the humanoid monorepo has no CI, collected robot data lands in S3 and stops there, and there is no registry telling us what any given dataset actually contains.

You will be the first dedicated infrastructure hire inside FAIRI. You will build the CI/CD, data pipelines, and IaC that let a small research team ship reliably — partnering with Field AI's platform, cloud, and data-processing teams rather than rebuilding what they already run well. This is a hands-on ownership role, not a coordination role.

What You'll Do

CI/CD and Build Infrastructure — 30%

  • Stand up CI/CD for the humanoid monorepo: containerize, push to ECR, run unit tests, build, and gate on simulation system tests before promotion.

  • Work with the platform team's self-hosted GitHub Actions runners (ARM, AMD, CUDA, Jetson-class targets) rather than standing up parallel infrastructure.

  • Cut build times through change detection and remote caching — full builds are currently ~45 minutes uncached.

  • Build test infrastructure that lets the same test run against simple sim, Isaac Sim, or real hardware, driven over ROS 2 messages or the robot REST API.

  • Establish per-automation integration tests so shared-library and output-format changes cannot silently break pipelines.

  • Data Pipelines and Orchestration — 30%

    • Stand up and own FAIRI's Airflow stack for humanoid data processing.

    • Build the ingest path from robot to usable dataset: rosbag/MCAP capture, episode segmentation, format conversion, and delivery to training.

    • Implement data lifecycle guardrails — filtering, review-for-deletion, and retention — so idle-robot and failed-run data does not accumulate indefinitely.

    • Build and operate the dataset and mission registry so every dataset is attributable to a subject, session, robot, and purpose.

    • Support MoCapDB in production: ECS Fargate services, AWS Batch retargeting workers, RDS Postgres, and S3, integrated with FieldAI Auth.

    • Cloud, IaC, and Security — 25%

      • Own FAIRI's AWS footprint as code: ECR, S3, IAM roles and cross-account trust policies, VPC and networking, Kubernetes/EKS workloads.

      • Close the gaps where infrastructure is not yet in code, and bring permissions changes under review.

      • Own compliance posture for research tooling — SOC 2 constraints on SaaS, experiment tracking, and data-sharing controls — in partnership with IT and Security.

      • Eliminate person-owned infrastructure: documented owners, runbooks, and access paths for every FAIRI-owned service.

      • Manage secrets, VPN/Tailscale access paths, and hardware-in-the-loop connectivity to robots on the floor.

      • Enablement and Documentation — 15%

        • Write and maintain runbooks, onboarding guides, and architecture documentation so a new engineer can test and deploy on day one rather than learning it from a teammate.

        • Be the interface between FAIRI and Field AI's platform, cloud, and data-processing teams — negotiating what FAIRI reuses versus owns.

        • Support researchers and systems engineers directly when pipelines, builds, or environments break, including live troubleshooting during demos.

        • Bring reproducibility discipline to research workflows: versioned configs, pinned environments, traceable runs.

        • What You Bring

          You don't need every item below, but you should bring real, hands-on depth in several of them:

          • 4+ years in DevOps, platform, infrastructure, or SRE roles

          • AWS in production: ECR, S3, IAM, VPC, EKS, Batch, ECS/Fargate

          • Infrastructure as code (Terraform, CDK, Pulumi, or equivalent) with a review-and-version discipline

          • CI/CD design and operation at scale — GitHub Actions strongly preferred, including self-hosted runners

          • Kubernetes in production, including workload scheduling and resource governance

          • Container tooling and build optimization: Docker, BuildKit or daemonless alternatives (Buildah, Kaniko), multi-arch builds, remote caching

          • Workflow orchestration — Apache Airflow or equivalent

          • Python, plus comfort in Bash and reading C++

          • Linux systems administration and networking fundamentals

          • Observability: logging, metrics, tracing, and alerting you actually built

          • Beyond the toolkit:

            • Comfortable being the only infra person in the room. You can take an open-ended ask and run with it without much hand-holding.

            • Bias toward reuse. You would rather integrate a platform team's runners than build a parallel stack, and you can negotiate that boundary well.

            • Strong documentation habits — you leave runbooks and processes better than you found them.

            • Pragmatic about research velocity. You know when to enforce rigor and when it would just slow the team down.

            • What Sets You Apart

              • Infrastructure experience in robotics, autonomous vehicles, or ML research environments

              • ROS 2, rosbag/MCAP, or Foxglove familiarity

              • Large-scale data pipeline work — TB-scale sensor or video data, lifecycle and retention policy design

              • ML infrastructure: experiment tracking, GPU scheduling, training pipelines, simulation infrastructure (Isaac Sim / Isaac Lab)

              • Hardware-in-the-loop CI — running tests against physical devices from a pipeline

              • Compliance and audit experience: SOC 2, access reviews, data governance

              • Having been the first infrastructure hire on a team before

              • Interest in humanoid robotics and how machines learn from human movement

              • Why FAIRI

                This role sits inside FAIRI's Research Operations function, supporting the Humanoid Program. It is a foundational hire: you will define what FAIRI's infrastructure looks like rather than inherit it, with the leverage of Field AI's existing platform teams behind you and a research team that will feel the difference immediately.

                FAIRI's humanoid work runs on real deadlines with real customers among them. The infrastructure you build ships to the floor.


Why Join Field AI?
FieldAI is tackling one of robotics’ hardest problems: deploying robots in unstructured, previously unknown environments. Our Field Foundational Models™ advance perception, planning, localization, and manipulation with an emphasis on explainability and safety, so our systems can be trusted where it matters most.

You will work alongside a world-class team that values creativity, resilience, and bold thinking. We bring a decade-long track record of real-world deployments, strong performance in DARPA challenges, and experience from organizations such as DeepMind, NASA JPL, Boston Dynamics, NVIDIA, Amazon, Tesla Autopilot, Cruise, Zoox, Toyota Research Institute, and SpaceX.

Our R&D organization is growing and anchored in Boston, with close collaboration across our teams in Southern California and with colleagues around the US and globally.

Be Part of the Next Robotics Revolution
Solving problems at this scale takes a team as unique as the mission. We are looking for people who push beyond conventional approaches, enjoy tackling tough and ambiguous questions, and bring interdisciplinary perspective. Our success depends on exceptional AI researchers and engineers, as well as strong software developers, product designers, field deployment experts, and communicators who can turn breakthroughs into real capability.
We are headquartered in Mission Viejo (Irvine adjacent), Southern California, with teammates across the US and around the world. Join us to shape the future of embodied intelligence as part of a fun, close-knit team building systems that work in the real world.

Equal Opportunity
FieldAI celebrates diversity and is committed to creating an inclusive environment for all employees. Candidates and employees are evaluated based on merit, qualifications, and performance. We do not discriminate on the basis of race, color, religion, sex, gender, national origin, ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, or any other legally protected status.
Pronto para se candidatar à Field Ai?
Candidatar-se à Field Ai

Vagas semelhantes

Recorded Future
Senior Platform Engineer
Recorded Future
⚡ Candidate-se cedo Boston, MA Presencial $129,000–$193,500
● Nova 👁 Vista ✓ Candidatada há 3h
Klaviyo
Senior Lead Software Engineer - Data Platform
Klaviyo
⚡ Candidate-se cedo Boston, MA Presencial $216,000–$324,000
● Nova 👁 Vista ✓ Candidatada há 14h
LangChain
Senior Frontend Engineer, AI Observability & Evals Platform
LangChain
⚡ Candidate-se cedo San Francisco, CA Presencial $175,000–$240,000
● Nova 👁 Vista ✓ Candidatada há 6d
LangChain
Frontend Engineer, AI Observability & Evals Platform
LangChain
⚡ Candidate-se cedo San Francisco, CA Presencial $160,000–$195,000
● Nova 👁 Vista ✓ Candidatada há 1sem
Speechify
Software Engineer, Platform - Boston, MA, USA
Speechify
⚡ Candidate-se cedo Boston, MA, USA Presencial
● Nova 👁 Vista ✓ Candidatada há 1sem
Anduril Industries
Staff Software Engineer, Developer Platform
Anduril Industries
⚡ Candidate-se cedo Seattle, Washington, United St... Presencial $191,000–$253,000
● Nova 👁 Vista ✓ Candidatada há 1sem
Anduril Industries
Staff Software Engineer, Developer Platform
Anduril Industries
⚡ Candidate-se cedo Costa Mesa, California, United... Presencial $191,000–$253,000
● Nova 👁 Vista ✓ Candidatada há 1sem
Anduril Industries
Staff Software Engineer, Developer Platform
Anduril Industries
⚡ Candidate-se cedo Boston, Massachusetts, United... Presencial $191,000–$253,000
● Nova 👁 Vista ✓ Candidatada há 1sem
EI
Senior Software Engineer, AI-Native Platform (Remote)
ezCater, Inc
⚡ Candidate-se cedo Boston, MA Presencial $158,000–$190,000
● Nova 👁 Vista ✓ Candidatada há 1sem

Cadastre-se para receber sugestões sob medida com base nas vagas que você abre e nas buscas que você salva.

Mais vagas na Field Ai

Ver todas as vagas na Field Ai →

Candidatar-se agora
🤖

Opa — calma aí

A JobsRadar foi feita para pessoas de verdade passando por um momento difícil na busca por emprego — não para requisições automatizadas. Você está clicando rápido demais e agora está temporariamente bloqueado.

Volte mais tarde. Se você está mesmo procurando emprego, estamos com você — apenas aja como um ser humano.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Ganhe vantagem na sua busca por emprego.

Entre no nosso canal do Telegram para o que ajuda você a conseguir a vaga — referências salariais, o pulso semanal do mercado e avisos de novos recursos. Sem spam, só sinal.

Entre no canal — é grátis