Member of Technical Staff, AI Inference
Squeeze latency and cost out of the serving stack behind a consumer answer engine handling very high query volume.
538 roles across 160 companies — filter by region or category.
Showing 50 of 186 matching positions (1–50)
Clear filtersSqueeze latency and cost out of the serving stack behind a consumer answer engine handling very high query volume.
Build multi-step agents that browse, reason over sources and complete tasks on behalf of users.
Own the indexing and retrieval pipelines that keep the answer engine fresh across the open web.
Ship experiments across onboarding, sharing and subscription flows to move activation and retention.
Sell Perplexity Enterprise to mid-market teams, running discovery through to close.
Own paid, lifecycle and partnership channels driving qualified pipeline for the enterprise product.
Build the query and storage layer of the lakehouse platform, working across Spark, Delta Lake and cloud object storage.
Join a platform team building distributed data systems, with structured mentoring through your first year.
Build model training, serving and governance features that let enterprises run their own AI workloads on their own data.
Design lakehouse architectures with enterprise customers and prove them out in technical evaluations.
Own governance, lineage and access control across data and AI assets for large regulated customers.
Document pipelines, SQL warehousing and workflow orchestration for practitioners across cloud providers.
Build the AI agent that scaffolds, edits and deploys full applications from a natural-language brief.
Run the container and sandbox fleet that gives every user an instant, isolated development environment.
Design an IDE that stays approachable for first-time programmers without frustrating professionals.
Build the access, gateway and browser isolation controls enterprises use to retire their VPNs.
Build ingestion and query systems handling trillions of log events with predictable cost.
Train, evaluate and serve models that power Canva's design platform at production scale.
Turn model capabilities into dependable product surfaces across Canva's design platform, balancing latency, cost and quality.
Build the pipelines, feature stores and monitoring that keep models behind Canva's design platform reproducible and observable.
Own the AI surfaces inside Canva's design platform, deciding what to automate, what to assist and what to leave alone.
Own the transformation layer and semantic definitions behind reporting on Cloudflare's developer platform.
Deliver working implementations of Cloudflare's developer platform inside customer environments.
Build detections and response automation covering the infrastructure behind Cloudflare's developer platform.
Partner with engineering teams on threat models and secure defaults throughout Cloudflare's developer platform.
Own model deployment tooling so research work reaches Databricks's data platform without bespoke glue each time.
Run experiments that push the model capabilities behind Databricks's data platform, and publish or ship what works.
Run the large multi-team efforts behind Databricks's data platform from kickoff through launch.
Build agentic workflows and tool-calling systems on top of Databricks's data platform, and own their accuracy in production.
Review designs, harden services and fix classes of vulnerability across Datadog's developer platform.
Probe model behaviour for failure modes and build the evaluations that gate releases across Datadog's developer platform.
Build the internal platform that lets product teams ship to Datadog's developer platform safely and often.
Train, evaluate and serve models that power Datadog's developer platform at production scale.
Design and run the multi-region cloud footprint behind Freshworks's product suite, with cost and resilience both in view.
Run the large multi-team efforts behind Freshworks's product suite from kickoff through launch.
Review designs, harden services and fix classes of vulnerability across Freshworks's product suite.
Own the transformation layer and semantic definitions behind reporting on Freshworks's product suite.
Prototype and ship interface work for GoTo Group's delivery network, living between design and front-end code.
Partner with engineering teams on threat models and secure defaults throughout GoTo Group's delivery network.
Investigate intrusions and turn each one into durable detection coverage for GoTo Group's delivery network.
Turn model capabilities into dependable product surfaces across GoTo Group's delivery network, balancing latency, cost and quality.
Build detections and response automation covering the infrastructure behind Grab's delivery network.
Make the build, test and deploy loop fast for everyone working on Grab's delivery network.
Own infrastructure-as-code, networking and capacity planning for Grab's delivery network.
Build the internal platform that lets product teams ship to Grab's delivery network safely and often.
Partner with engineering teams on threat models and secure defaults throughout Harvey's AI products.
Own the lakehouse and streaming infrastructure underneath Harvey's AI products, with an eye on cost and freshness.
Prototype and ship interface work for Harvey's AI products, living between design and front-end code.
Build the internal platform that lets product teams ship to HubSpot's product suite safely and often.
Review designs, harden services and fix classes of vulnerability across HubSpot's product suite.
No jobs match your search and filters. Try a different keyword or clear the filters.