Inference & model-serving jobs
Serving roles — GPU, Triton, TensorRT, quantization, latency and throughput.
510 open roles · showing 501–510
- Senior Machine Learning Platform EngineerCharlie Health · New York, NY$170k–220k/yr
- Head of Engineering, ComputeTemporal Technologies · United States$214k–352k/yr
Head of Engineering, ComputeTemporal · United States$320k–335k/yr- Member of Technical Staff, InferenceInferact · San Francisco$200k–400k/yr
Director, ML Services EngineeringAdobe · San Jose$266k–385k/yr
AI Innovation and Development, Lead ScientistFair Isaac (FICO) · San Diego, CA$123k–193k/yr- Member of Technical Staff, Developer RelationsInferact · San Francisco$200k–400k/yr
- Member of Technical Staff, TPU Performance EngineeringInferact · San Francisco$200k–400k/yr
Senior Software Engineer - Big Data/ML OpsTruveta · Seattle, WA$155k–190k/yr
ML Solution Architect (Early Talent)Nebius · United States
No roles match those filters.