Jobs Companies CrewAI Software Engineer, Infrastructure & Reliability

Sobre este puesto de Software Engineer, Infrastructure & Reliability en CrewAI

CrewAI · Presencial · San Francisco, California, United States

About CrewAI

CrewAI is the leading framework and enterprise platform for building and orchestrating multi-agent AI systems, powering 300M+ agent executions per month across thousands of companies. The Agent Management Platform is our control plane for deploying, monitoring, governing, and scaling agents in production. This role owns the infrastructure foundation that keeps it reliable, secure, and fast.

The Role

You'll build and operate the platform infrastructure behind CrewAI's cloud and enterprise deployments. You'll work across multiple hyperscalers - AWS, Azure, and GCP. You’ll work on containers, CI/CD, deployment automation, observability, secrets, networking, and runtime reliability. Your job is to make the product and runtime teams faster while making customer’s production environments safer.

This is not a pure DevOps support role. You'll write code, improve systems, design deployment paths, harden production, and build the internal platform that lets CrewAI scale and scale our customer deployments.

What You'll Do

  • Own and improve the infrastructure that runs CrewAI's platform: AWS, ECS/ECR, Docker, Kubernetes/Helm, networking, secrets, databases, Redis, and related services.
  • Build and maintain CI/CD pipelines for build, test, image publishing, migrations, environment promotion, rollbacks, and deploy safety.
  • Improve reliability across cloud and enterprise deployments: health checks, alerting, incident response, capacity planning, recovery paths, and operational runbooks - and own the front-line on-call rotation and its SLAs.
  • Partner with runtime engineers on Celery/FastAPI/Redis workloads and with product engineers on Rails/Solid Queue/Postgres production behavior.
  • Manage production observability and telemetry infrastructure: logs, metrics, traces, dashboards, Sentry/OpenTelemetry plumbing, actionable alerts, and telemetry export to customers' own monitoring systems.
  • Harden security and compliance posture across IAM, workload identity, secrets management, vulnerability scanning, dependency/image hygiene, and least-privilege access.
  • Build the tooling and automation that lets field engineers and customers run self-hosted installs themselves - Helm charts, environment config, release artifacts, pre-flight checks, and install runbooks - so engineering does fewer hands-on installs over time.
  • Reduce operational toil by automating recurring workflows and making deployments boring.

Requirements

What We're Looking For

  • Strong infrastructure/platform engineering experience in production SaaS environments.
  • Deep practical experience with AWS, Docker, CI/CD, GitHub Actions, and containerized services.
  • Experience with ECS and/or Kubernetes; Helm experience is a strong plus.
  • Comfort operating PostgreSQL, Redis, background job systems, queues, and web services in production.
  • Strong debugging instincts across app, infra, network, deploy, and dependency layers.
  • Security-minded approach to IAM, secrets, workload identity, vulnerability management, and production access.
  • Ability to write reliable automation in Python, Ruby, Go, Bash, or similar.
  • Calm, rigorous approach to incidents, rollbacks, migrations, and production change management.

Bonus

  • Experience with AI/agent platforms, workflow runtimes, or high-volume async execution systems.
  • Experience supporting enterprise/self-hosted deployments.
  • Terraform or other IaC experience.
  • SRE background: SLOs, incident review, capacity planning, load testing.
  • Familiarity with Rails, FastAPI, Celery, OpenTelemetry, or multi-service observability.
¿Listo para postularte en CrewAI?
Postúlate en CrewAI

Sobre CrewAI

The leading AI automation suite that enables enterprises to orchestrate Agents at Scale

Ver todos los empleos en CrewAI →

Empleos similares

Redwood Materials
Embedded Software Engineer – Power Electronics, Energy Storage
Redwood Materials
⚡ Postúlate pronto San Francisco, California, Uni... Presencial $180,000–$237,500
● Nuevo 👁 Visto ✓ Postulado hace 48m
Redwood Materials
Infrastructure Software Engineer, Energy Storage
Redwood Materials
⚡ Postúlate pronto San Francisco, California, Uni... Presencial $180,000–$237,500
● Nuevo 👁 Visto ✓ Postulado hace 1h
Brex
Senior Software Engineer, Full Stack
Brex
⚡ Postúlate pronto San Francisco, California, Uni... Híbrido $192,000–$240,000
● Nuevo 👁 Visto ✓ Postulado hace 1h
Redwood Materials
Software Engineer - Site Controller, Energy Storage
Redwood Materials
⚡ Postúlate pronto San Francisco, California, Uni... Presencial $180,000–$237,500
● Nuevo 👁 Visto ✓ Postulado hace 2h
Workato
Staff Software Engineer - AI Platform / Distributed Systems Engineer
Workato
⚡ Postúlate pronto San Francisco, California Presencial
● Nuevo 👁 Visto ✓ Postulado hace 16h
Workato
Software Engineer - AI Agent Platforms
Workato
⚡ Postúlate pronto San Francisco, California Presencial
● Nuevo 👁 Visto ✓ Postulado hace 16h
Workato
Senior Software Engineer
Workato
⚡ Postúlate pronto San Francisco, California Presencial
● Nuevo 👁 Visto ✓ Postulado hace 16h
TU
Senior AI Solutions Engineer - Software Engineering
Turing
⚡ Postúlate pronto New York, New York, United Sta... Presencial $260,000–$320,000
● Nuevo 👁 Visto ✓ Postulado hace 18h
StubHub
Senior Software Engineer - Supply Platform
StubHub
⚡ Postúlate pronto San Francisco, California, Uni... Híbrido $200,000–$250,000
● Nuevo 👁 Visto ✓ Postulado hace 20h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en CrewAI

Ver todos los empleos en CrewAI →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis