C

Member of Technical Staff — DevOps

Causal Labs San Francisco, California, United States

Research Services · 11-50 employees

Yesterday
devops Senior (5-10 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will design and operate the internal infrastructure for agentic development to improve engineering velocity. This includes managing CI/CD pipelines, self-hosting LLMs, and optimizing compute costs across the development stack.

What they look for

DevOps Infrastructure Engineering Linux Networking Docker Kubernetes Terraform CI/CD GitHub Actions Buildkite Bazel Nix GCP AWS Azure LLM Serving

Requirements

The ideal candidate has a strong background in systems engineering, container orchestration, and infrastructure-as-code. You must have experience building developer tooling and a proven track record of improving engineering velocity.

Full description

Member of Technical Staff — DevOps

We look for infrastructure engineers who are obsessed with engineering velocity. Everything we do — research, training, serving, customer deployments — moves only as fast as our tooling. Your mission is to design, build, and operate the platform underneath it all, including the internal infrastructure for agentic development, so every engineer and researcher can iterate rapidly with agents doing more of the work.

Responsibilities

  • Measure developer velocity, and partner with researchers, product engineers, and forward-deployed engineers to find what slows them down and remove it
  • Build internal infrastructure for agentic development: sandboxes, tool access, context, and eval harnesses that let autonomous research and coding agents run end to end
  • Self-host and serve open-source LLMs for internal agent use
  • Own CI/CD, build and test infrastructure, reproducible dev environments, and infrastructure-as-code so research and product code ship quickly and safely
  • Build the monitoring, alerting, and on-call for our product and customer-facing services to catch failures before users do
  • Own cost across the dev stack: Own cost across the dev stack: track and cut spend on compute, tools, and inference

What we're looking for

We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains.

  • AI-pilled: uses coding agents daily and has built internal infrastructure for agentic development, not just used it
  • Experience self-hosting and serving open-source LLMs (e.g. vLLM, SGLang)
  • Strong systems background: Linux, networking, containers and orchestration (e.g. Docker, Kubernetes), infrastructure-as-code (e.g. Terraform)
  • Experience building CI/CD and developer tooling (e.g. GitHub Actions, Buildkite, Bazel, Nix)
  • Knowledge of cloud platforms (GCP, AWS, or Azure) and best practices for observability and security
  • Track record of measurably improving developer velocity
  • Owns deliverables end-to-end, from requirements through autonomous execution

Similar roles