Cloud Infrastructure Engineer
Own infrastructure-as-code, networking and capacity planning for Remote's employment platform.
1638 roles across 160 companies — filter by region or category.
Showing 48 of 98 matching positions (51–98)
Clear filtersOwn infrastructure-as-code, networking and capacity planning for Remote's employment platform.
Run the training and serving infrastructure behind Replit's developer platform, including GPU scheduling, artefacts and rollouts.
Own infrastructure-as-code, networking and capacity planning for Replit's developer platform.
Own infrastructure-as-code, networking and capacity planning for Rippling's employment platform.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Sentry's developer platform.
Build the pipelines, feature stores and monitoring that keep models behind Snowflake's data platform reproducible and observable.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Supabase's developer platform.
Build the pipelines, feature stores and monitoring that keep models behind Swiggy's delivery network reproducible and observable.
Own infrastructure-as-code, networking and capacity planning for Temporal's developer platform.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Toggl's product suite.
Build the pipelines, feature stores and monitoring that keep models behind Vercel's developer platform reproducible and observable.
Own infrastructure-as-code, networking and capacity planning for Weights & Biases's developer platform.
Build the internal platform that lets product teams ship to Wikimedia Foundation's programmes and platforms safely and often.
Run the training and serving infrastructure behind Zapier's product suite, including GPU scheduling, artefacts and rollouts.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Zapier's product suite.
Build the pipelines, feature stores and monitoring that keep models behind Zoho's product suite reproducible and observable.
Automate the delivery pipeline behind Cal.com's open-source projects and keep deployments boring.
Set service level objectives and build the observability that makes ClickHouse's data platform debuggable under pressure.
Own availability, latency and incident response for Close's product suite, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes CrowdStrike's security platform debuggable under pressure.
Own availability, latency and incident response for Freshworks's product suite, and remove the causes rather than the symptoms.
Own availability, latency and incident response for GitLab's developer platform, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes Grammarly's product suite debuggable under pressure.
Own CI/CD, environments and release automation for LangChain's developer platform.
Automate the delivery pipeline behind Linear's developer platform and keep deployments boring.
Own availability, latency and incident response for Palo Alto Networks's security platform, and remove the causes rather than the symptoms.
Own CI/CD, environments and release automation for PlanetScale's developer platform.
Own availability, latency and incident response for Postman's developer platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Remote's employment platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Snowflake's data platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Tailscale's developer platform, and remove the causes rather than the symptoms.
Keep a large multi-tenant DevSecOps platform reliable, working across Kubernetes, observability tooling and incident response.
Deploy and troubleshoot Ubuntu, OpenStack and Kubernetes environments alongside enterprise customers.
Keep large-scale, multi-tenant observability infrastructure healthy across Kubernetes, Go and the company's own open-source stack.
Maintain the availability and performance of a large-scale observability platform running across multiple cloud regions.
Build and scale the cloud platform infrastructure supporting all Atlassian products across global regions.
Scale the serverless deployment infrastructure that builds and serves frontend applications globally.
Keep thousands of managed Postgres instances healthy across global regions with automated tooling.
Build and operate the Atlas managed database platform serving hundreds of thousands of clusters worldwide.
Design cloud-native security architectures for customers deploying workloads across AWS, Azure and GCP.
Scale the infrastructure supporting flash sales and peak traffic events for the global commerce platform.
Keep Wikipedia and sister projects available worldwide by managing the infrastructure behind free knowledge.
Operate the cloud infrastructure platform serving droplets, managed databases and Kubernetes clusters at scale.
Operate the anycast network and edge infrastructure that routes traffic for a large share of the internet.
Lead the platform team scaling infrastructure for the incident management product across cloud regions.
Build the deployment tooling and orchestration layer that lets developers ship apps globally in seconds.
Keep PostHog Cloud running reliably at scale across ClickHouse, Kafka and Kubernetes infrastructure.
Operate and scale the managed database platform across multiple cloud regions with high availability targets.
No jobs match your search and filters. Try a different keyword or clear the filters.