Data Infrastructure Engineer
Build the ingestion, storage and query layers that every team uses to reason about Rippling's employment platform.
914 roles across 160 companies — filter by region or category.
Showing 50 of 101 matching positions (51–100)
Clear filtersBuild the ingestion, storage and query layers that every team uses to reason about Rippling's employment platform.
Own the lakehouse and streaming infrastructure underneath Sanity's product suite, with an eye on cost and freshness.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Sanity's product suite.
Design red-teaming and evaluation methodology for the models behind Snowflake's data platform, and turn results into mitigations.
Run experiments that push the model capabilities behind Supabase's developer platform, and publish or ship what works.
Design red-teaming and evaluation methodology for the models behind Temporal's developer platform, and turn results into mitigations.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Together AI's AI products.
Own the modelling loop behind Toptal's marketplace — features, training pipelines, offline evaluation and online experiments.
Own the transformation layer and semantic definitions behind reporting on Twilio's product suite.
Build the ingestion, storage and query layers that every team uses to reason about Vanta's security platform.
Own the transformation layer and semantic definitions behind reporting on Vercel's developer platform.
Own the transformation layer and semantic definitions behind reporting on Weights & Biases's developer platform.
Own the lakehouse and streaming infrastructure underneath Wikimedia Foundation's programmes and platforms, with an eye on cost and freshness.
Own the transformation layer and semantic definitions behind reporting on Zoom's product suite.
Optimise inference throughput and cost for the models running behind Zoom's product suite.
Own funnel, retention and feature adoption analysis for Amplitude's analytics platform.
Own funnel, retention and feature adoption analysis for Atlassian's product suite.
Own funnel, retention and feature adoption analysis for BrowserStack's developer platform.
Design schemas and pipelines that keep analytics and machine learning on Buffer's product suite accurate and timely.
Build and operate the batch and streaming pipelines that move data across Coinbase's financial platform.
Build and operate the batch and streaming pipelines that move data across Drata's security platform.
Instrument, measure and interpret how people actually use Drata's security platform.
Design schemas and pipelines that keep analytics and machine learning on Fivetran's data platform accurate and timely.
Design experiments and causal analyses that decide what ships next across Gymdesk's product suite.
Design schemas and pipelines that keep analytics and machine learning on Hotjar's analytics platform accurate and timely.
Design schemas and pipelines that keep analytics and machine learning on Resend's developer platform accurate and timely.
Model user and system behaviour on Rootly's product suite and turn the findings into decisions the team acts on.
Instrument, measure and interpret how people actually use Sanity's product suite.
Own funnel, retention and feature adoption analysis for Supabase's developer platform.
Instrument, measure and interpret how people actually use Turso's developer platform.
Build and operate the batch and streaming pipelines that move data across Twilio's product suite.
Build dashboards and deep dives that show how Cohere's AI products is actually performing.
Build dashboards and deep dives that show how GitLab's developer platform is actually performing.
Build dashboards and deep dives that show how Miro's product suite is actually performing.
Answer the recurring commercial and operational questions about Oyster's employment platform with clear, reproducible analysis.
Answer the recurring commercial and operational questions about Remote's employment platform with clear, reproducible analysis.
Build dashboards and deep dives that show how Toptal's marketplace is actually performing.
Build the privacy-respecting telemetry and analytics pipelines that inform Firefox and Mozilla product decisions.
Model merchant and buyer behaviour at very large scale to guide commerce product decisions.
Build data pipelines processing billions of communication events daily for analytics and billing systems.
Analyse product usage patterns across Jira and Confluence to surface insights for growth and retention teams.
Build data pipelines processing events from WordPress sites representing a large share of the web.
Develop anomaly detection and natural language processing models integrated into the Elastic Stack.
Build pipelines that process billions of security events daily for threat intelligence and detection models.
Model marketplace dynamics including pricing, demand forecasting and host acquisition to grow supply.
Analyse engagement and retention metrics across a self-serve subscription product to guide roadmap decisions.
Build fraud detection and risk models protecting a large cryptocurrency exchange from bad actors.
Build pipelines processing session recordings and clickstream data at scale for behaviour analytics.
Build the billing and usage metering pipelines tracking resource consumption across the cloud platform.
Analyse usage patterns across time tracking and hiring products to guide feature prioritisation.
No jobs match your search and filters. Try a different keyword or clear the filters.