Member of Technical Staff — DevOps
Causal Labs San Francisco, California, United States
Research Services · 11-50 employees
About the role
You will design and operate the internal infrastructure for agentic development to improve engineering velocity. This includes managing CI/CD pipelines, self-hosting LLMs, and optimizing compute costs across the development stack.
What they look for
Requirements
The ideal candidate has a strong background in systems engineering, container orchestration, and infrastructure-as-code. You must have experience building developer tooling and a proven track record of improving engineering velocity.
Full description
Member of Technical Staff — DevOps
We look for infrastructure engineers who are obsessed with engineering velocity. Everything we do — research, training, serving, customer deployments — moves only as fast as our tooling. Your mission is to design, build, and operate the platform underneath it all, including the internal infrastructure for agentic development, so every engineer and researcher can iterate rapidly with agents doing more of the work.
Responsibilities
- Measure developer velocity, and partner with researchers, product engineers, and forward-deployed engineers to find what slows them down and remove it
- Build internal infrastructure for agentic development: sandboxes, tool access, context, and eval harnesses that let autonomous research and coding agents run end to end
- Self-host and serve open-source LLMs for internal agent use
- Own CI/CD, build and test infrastructure, reproducible dev environments, and infrastructure-as-code so research and product code ship quickly and safely
- Build the monitoring, alerting, and on-call for our product and customer-facing services to catch failures before users do
- Own cost across the dev stack: Own cost across the dev stack: track and cut spend on compute, tools, and inference
What we're looking for
We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains.
- AI-pilled: uses coding agents daily and has built internal infrastructure for agentic development, not just used it
- Experience self-hosting and serving open-source LLMs (e.g. vLLM, SGLang)
- Strong systems background: Linux, networking, containers and orchestration (e.g. Docker, Kubernetes), infrastructure-as-code (e.g. Terraform)
- Experience building CI/CD and developer tooling (e.g. GitHub Actions, Buildkite, Bazel, Nix)
- Knowledge of cloud platforms (GCP, AWS, or Azure) and best practices for observability and security
- Track record of measurably improving developer velocity
- Owns deliverables end-to-end, from requirements through autonomous execution
Similar roles
-
Senior DevOps Engineer
Heavy Construction Systems Specialists, LLC United States
-
Senior DevOps Engineer (Platform Operations, Security & Observability) [626]
D-Wave Connecticut, United States · $125K–$187K/yr
-
Director, DevOps/ Full-Stack Engineering
BNY Mellon Pittsburgh, Pennsylvania, United States · $147K–$225K/yr
-
DevOps Engineer [Zeal]
Livefront Peru
-
DevOps Engineer
Exostar Fairfax County, Virginia, United States · $100K–$130K/yr
-
Azure DevOps Cloud Engineer
Derex Technologies Inc Mechanicsville, Virginia, United States