Research Scientist, Multimodal
Investigate training, alignment and evaluation methods that make the models behind Glean's AI products more useful.
1638 roles across 160 companies — filter by region or category.
Showing 50 of 150 matching positions (51–100)
Clear filtersInvestigate training, alignment and evaluation methods that make the models behind Glean's AI products more useful.
Build the ingestion, storage and query layers that every team uses to reason about Grafana Labs's developer platform.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Groq's AI products.
Own the transformation layer and semantic definitions behind reporting on Gusto's employment platform.
Own the lakehouse and streaming infrastructure underneath Harvey's AI products, with an eye on cost and freshness.
Model warehouse data into tested, documented datasets the whole company trusts when analysing HashiCorp's developer platform.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Hightouch's product suite.
Build the ingestion, storage and query layers that every team uses to reason about HubSpot's product suite.
Build the ingestion, storage and query layers that every team uses to reason about Instacart's commerce platform.
Optimise inference throughput and cost for the models running behind Instacart's commerce platform.
Build the ingestion, storage and query layers that every team uses to reason about Khan Academy's learning platform.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Khan Academy's learning platform.
Train, evaluate and serve models that power Linear's developer platform at production scale.
Own the transformation layer and semantic definitions behind reporting on Mercury's financial platform.
Build the ingestion, storage and query layers that every team uses to reason about Mercury's financial platform.
Train, evaluate and serve models that power Modal's developer platform at production scale.
Own the transformation layer and semantic definitions behind reporting on Modal's developer platform.
Own the modelling loop behind MongoDB's developer platform — features, training pipelines, offline evaluation and online experiments.
Own the transformation layer and semantic definitions behind reporting on Okta's security platform.
Own the lakehouse and streaming infrastructure underneath Palo Alto Networks's security platform, with an eye on cost and freshness.
Own the modelling loop behind Perplexity's AI products — features, training pipelines, offline evaluation and online experiments.
Probe model behaviour for failure modes and build the evaluations that gate releases across Perplexity's AI products.
Run experiments that push the model capabilities behind Pinecone's developer platform, and publish or ship what works.
Design red-teaming and evaluation methodology for the models behind PostHog's analytics platform, and turn results into mitigations.
Own the lakehouse and streaming infrastructure underneath PostHog's analytics platform, with an eye on cost and freshness.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Postman's developer platform.
Own the modelling loop behind Remote's employment platform — features, training pipelines, offline evaluation and online experiments.
Own the transformation layer and semantic definitions behind reporting on Retool's developer platform.
Build the ingestion, storage and query layers that every team uses to reason about Rippling's employment platform.
Build the ingestion, storage and query layers that every team uses to reason about Robinhood's financial platform.
Optimise inference throughput and cost for the models running behind Robinhood's financial platform.
Own the transformation layer and semantic definitions behind reporting on Runway's AI products.
Own the lakehouse and streaming infrastructure underneath Runway's AI products, with an eye on cost and freshness.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Scale AI's AI products.
Investigate training, alignment and evaluation methods that make the models behind Sierra's AI products more useful.
Design red-teaming and evaluation methodology for the models behind Snowflake's data platform, and turn results into mitigations.
Train, evaluate and serve models that power Spade's financial platform at production scale.
Own the transformation layer and semantic definitions behind reporting on Spade's financial platform.
Train, evaluate and serve models that power Stripe's financial platform at production scale.
Design red-teaming and evaluation methodology for the models behind Temporal's developer platform, and turn results into mitigations.
Model warehouse data into tested, documented datasets the whole company trusts when analysing Together AI's AI products.
Own the modelling loop behind Toptal's marketplace — features, training pipelines, offline evaluation and online experiments.
Own the transformation layer and semantic definitions behind reporting on Twilio's product suite.
Build the ingestion, storage and query layers that every team uses to reason about Vanta's security platform.
Own the transformation layer and semantic definitions behind reporting on Vercel's developer platform.
Own the transformation layer and semantic definitions behind reporting on Weights & Biases's developer platform.
Own the lakehouse and streaming infrastructure underneath Wikimedia Foundation's programmes and platforms, with an eye on cost and freshness.
Own the transformation layer and semantic definitions behind reporting on Zoom's product suite.
Optimise inference throughput and cost for the models running behind Zoom's product suite.
Design experiments and causal analyses that decide what ships next across Abridge's healthcare platform.
No jobs match your search and filters. Try a different keyword or clear the filters.