Software Engineer, Infrastructure & Reliability
CrewAI · San Francisco, Estados Unidos De América · Hybrid
10000 Trabajos a distancia y desde casa en línea
CrewAI · San Francisco, Estados Unidos De América · Hybrid
StradIT · New York, Estados Unidos De América · Hybrid
StradIT · Jersey City, Estados Unidos De América · Hybrid
Resource Innovations · San Francisco, Estados Unidos De América · Remote
Resource Innovations · Salt Lake City, Estados Unidos De América · Remote
CrewAI · San Francisco, Estados Unidos De América · Hybrid
CrewAI is the leading framework and enterprise platform for building and orchestrating multi-agent AI systems, powering 300M+ agent executions per month across thousands of companies. The Agent Management Platform is our control plane for deploying, monitoring, governing, and scaling agents in production. This role owns the infrastructure foundation that keeps it reliable, secure, and fast.
You'll build and operate the platform infrastructure behind CrewAI's cloud and enterprise deployments. You'll work across multiple hyperscalers - AWS, Azure, and GCP. You’ll work on containers, CI/CD, deployment automation, observability, secrets, networking, and runtime reliability. Your job is to make the product and runtime teams faster while making customer’s production environments safer.
This is not a pure DevOps support role. You'll write code, improve systems, design deployment paths, harden production, and build the internal platform that lets CrewAI scale and scale our customer deployments.