Site Reliability Engineer, Actions
Keep a CI/CD fleet running billions of workflow minutes reliable, fast and cost-efficient.
914 roles across 160 companies — filter by region or category.
Showing 50 of 61 matching positions (1–50)
Clear filtersKeep a CI/CD fleet running billions of workflow minutes reliable, fast and cost-efficient.
Keep a high-throughput build platform available and predictable across multiple cloud regions.
Operate the managed cloud platform running Terraform, Vault and Consul for enterprise customers.
Own infrastructure-as-code, networking and capacity planning for Airbyte's open-source projects.
Design and run the multi-region cloud footprint behind Amplitude's analytics platform, with cost and resilience both in view.
Design and run the multi-region cloud footprint behind Automattic's open-source projects, with cost and resilience both in view.
Own model deployment tooling so research work reaches Buffer's product suite without bespoke glue each time.
Own Kubernetes, deployment tooling and paved-road abstractions underneath CircleCI's developer platform.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Close's product suite.
Own infrastructure-as-code, networking and capacity planning for Coinbase's financial platform.
Build the pipelines, feature stores and monitoring that keep models behind Confluent's data platform reproducible and observable.
Design and run the multi-region cloud footprint behind Confluent's data platform, with cost and resilience both in view.
Own infrastructure-as-code, networking and capacity planning for CrowdStrike's security platform.
Own Kubernetes, deployment tooling and paved-road abstractions underneath DuckDuckGo's open-source projects.
Own infrastructure-as-code, networking and capacity planning for DuckDuckGo's open-source projects.
Build the internal platform that lets product teams ship to Fly.io's developer platform safely and often.
Own Kubernetes, deployment tooling and paved-road abstractions underneath GitHub's developer platform.
Own infrastructure-as-code, networking and capacity planning for GitLab's developer platform.
Build the internal platform that lets product teams ship to Gymdesk's product suite safely and often.
Own Kubernetes, deployment tooling and paved-road abstractions underneath LangChain's developer platform.
Design and run the multi-region cloud footprint behind Linear's developer platform, with cost and resilience both in view.
Own infrastructure-as-code, networking and capacity planning for Lokker's security platform.
Own infrastructure-as-code, networking and capacity planning for MongoDB's developer platform.
Build the pipelines, feature stores and monitoring that keep models behind MongoDB's developer platform reproducible and observable.
Design and run the multi-region cloud footprint behind Oyster's employment platform, with cost and resilience both in view.
Build the internal platform that lets product teams ship to Pinecone's developer platform safely and often.
Run the training and serving infrastructure behind Pinecone's developer platform, including GPU scheduling, artefacts and rollouts.
Build the internal platform that lets product teams ship to PlanetScale's developer platform safely and often.
Run the training and serving infrastructure behind PlanetScale's developer platform, including GPU scheduling, artefacts and rollouts.
Build the pipelines, feature stores and monitoring that keep models behind Railway's developer platform reproducible and observable.
Own infrastructure-as-code, networking and capacity planning for Remote's employment platform.
Own infrastructure-as-code, networking and capacity planning for Rippling's employment platform.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Sentry's developer platform.
Build the pipelines, feature stores and monitoring that keep models behind Snowflake's data platform reproducible and observable.
Own infrastructure-as-code, networking and capacity planning for Temporal's developer platform.
Build the pipelines, feature stores and monitoring that keep models behind Vercel's developer platform reproducible and observable.
Own infrastructure-as-code, networking and capacity planning for Weights & Biases's developer platform.
Build the internal platform that lets product teams ship to Wikimedia Foundation's programmes and platforms safely and often.
Run the training and serving infrastructure behind Zapier's product suite, including GPU scheduling, artefacts and rollouts.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Zapier's product suite.
Automate the delivery pipeline behind Cal.com's open-source projects and keep deployments boring.
Set service level objectives and build the observability that makes ClickHouse's data platform debuggable under pressure.
Own availability, latency and incident response for Close's product suite, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes CrowdStrike's security platform debuggable under pressure.
Own availability, latency and incident response for GitLab's developer platform, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes Grammarly's product suite debuggable under pressure.
Own CI/CD, environments and release automation for LangChain's developer platform.
Automate the delivery pipeline behind Linear's developer platform and keep deployments boring.
Own CI/CD, environments and release automation for PlanetScale's developer platform.
Own availability, latency and incident response for Remote's employment platform, and remove the causes rather than the symptoms.
No jobs match your search and filters. Try a different keyword or clear the filters.