Cloud Cost Engineer
Design and run the multi-region cloud footprint behind Vivasoft's product suite, with cost and resilience both in view.
1638 roles across 160 companies — filter by region or category.
Showing 50 of 158 matching positions (101–150)
Clear filtersDesign and run the multi-region cloud footprint behind Vivasoft's product suite, with cost and resilience both in view.
Own infrastructure-as-code, networking and capacity planning for Weights & Biases's developer platform.
Build the internal platform that lets product teams ship to Wikimedia Foundation's programmes and platforms safely and often.
Run the training and serving infrastructure behind Zapier's product suite, including GPU scheduling, artefacts and rollouts.
Own Kubernetes, deployment tooling and paved-road abstractions underneath Zapier's product suite.
Build the pipelines, feature stores and monitoring that keep models behind Zoho's product suite reproducible and observable.
Own availability, latency and incident response for Atoms's robotics products, and remove the causes rather than the symptoms.
Own CI/CD, environments and release automation for Augmedix Bangladesh's healthcare platform.
Automate the delivery pipeline behind BJIT's product suite and keep deployments boring.
Set service level objectives and build the observability that makes BRAC University's learning platform debuggable under pressure.
Automate the delivery pipeline behind Cal.com's open-source projects and keep deployments boring.
Set service level objectives and build the observability that makes ClickHouse's data platform debuggable under pressure.
Own availability, latency and incident response for Close's product suite, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes CrowdStrike's security platform debuggable under pressure.
Own CI/CD, environments and release automation for Figma's design platform.
Own availability, latency and incident response for Freshworks's product suite, and remove the causes rather than the symptoms.
Own availability, latency and incident response for GitLab's developer platform, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes Grammarly's product suite debuggable under pressure.
Own availability, latency and incident response for Gusto's employment platform, and remove the causes rather than the symptoms.
Automate the delivery pipeline behind Instacart's commerce platform and keep deployments boring.
Own availability, latency and incident response for Klarna's financial platform, and remove the causes rather than the symptoms.
Own CI/CD, environments and release automation for LangChain's developer platform.
Automate the delivery pipeline behind Linear's developer platform and keep deployments boring.
Own availability, latency and incident response for Monzo's financial platform, and remove the causes rather than the symptoms.
Own CI/CD, environments and release automation for Notion's product suite.
Own availability, latency and incident response for Palo Alto Networks's security platform, and remove the causes rather than the symptoms.
Own CI/CD, environments and release automation for Pathao's delivery network.
Own CI/CD, environments and release automation for PlanetScale's developer platform.
Own availability, latency and incident response for Postman's developer platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Remote's employment platform, and remove the causes rather than the symptoms.
Set service level objectives and build the observability that makes Scale AI's AI products debuggable under pressure.
Set service level objectives and build the observability that makes ShopUp's marketplace debuggable under pressure.
Own availability, latency and incident response for Snowflake's data platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Tailscale's developer platform, and remove the causes rather than the symptoms.
Own availability, latency and incident response for Typeform's product suite, and remove the causes rather than the symptoms.
Keep a large multi-tenant DevSecOps platform reliable, working across Kubernetes, observability tooling and incident response.
Deploy and troubleshoot Ubuntu, OpenStack and Kubernetes environments alongside enterprise customers.
Keep large-scale, multi-tenant observability infrastructure healthy across Kubernetes, Go and the company's own open-source stack.
Plan, operate and optimise mobile network infrastructure serving the country's largest subscriber base.
Own CI/CD pipelines, container orchestration and cloud infrastructure across multiple client projects.
Maintain the availability and performance of a large-scale observability platform running across multiple cloud regions.
Build and scale the cloud platform infrastructure supporting all Atlassian products across global regions.
Scale the serverless deployment infrastructure that builds and serves frontend applications globally.
Keep thousands of managed Postgres instances healthy across global regions with automated tooling.
Build and operate the Atlas managed database platform serving hundreds of thousands of clusters worldwide.
Design cloud-native security architectures for customers deploying workloads across AWS, Azure and GCP.
Scale the infrastructure supporting flash sales and peak traffic events for the global commerce platform.
Keep Wikipedia and sister projects available worldwide by managing the infrastructure behind free knowledge.
Set up CI/CD pipelines and cloud infrastructure for client projects on AWS and Azure.
Maintain campus network infrastructure, learning management systems and student information systems.
No jobs match your search and filters. Try a different keyword or clear the filters.