Inference & model-serving jobs
Serving roles — GPU, Triton, TensorRT, quantization, latency and throughput.
472 open roles · showing 451–472
Lead Full Stack Machine Learning EngineerCerebras · India Office
AI Compiler EngineerEnCharge AI · Remote-US, Canada, Germany and Norway$190k–255k/yr- AI Field Engineer - AI NativesFireworks AI · San Mateo$200k–260k/yr
Staff Software Engineer, AI RuntimeDatabricks · Mountain View, California; San Francisco, California$190k–265k/yr
Senior Software Engineer, AI RuntimeDatabricks · Mountain View, California; San Francisco, California$160k–225k/yr
Lead ML/Perception EngineerMay Mobility · USA - Remote$235k–275k/yr
Agentic Search Infrastructure Engineer - MoveworksServiceNow · Mountain View, CALIFORNIA, United States
Principal Engineer, AIAnaplan · Pennsylvania-Remote, United States
Machine Learning Engineer II - Autonomous Driving Training InfrastructureMay Mobility · USA - Remote$160k–210k/yr
Forward Deployed Engineer, Enterprise - TavilyNebius · New York, United States; United States$180k–224k/yr- [AI Research Div.] ML Infrastructure Engineer (5년 이상)Krafton · Seoul
Machine Learning Engineer 4Adobe · Bangalore
Principal Software Engineer, GPU ComputeRoblox · San Mateo, CA, United States$345k–399k/yr
Senior Software Engineer - ML InfrastructureSambaNova Systems · Remote - US$200k–275k/yr- Performance Engineer, On-Device InferenceSarvam AI · Bengaluru
Staff Engineer, Distributed Storage and HPC & AI InfrastructureTogether AI · San Francisco$250k–300k/yr
ASIC ArchitectCerebras · Headquarters/Sunnyvale Office- ML EngineerMach9 · San Francisco$180k–300k/yr
Senior Perception Learning Engineer - SLAMApptronik · Sunnyvale, C A$190k–235k/yr
Research Engineer, Pre-TrainingJump Trading · New York, London, Chicago$300k–350k/yr- Quantitative ResearcherJane Street · New York, New York, United States
Software Engineer - Baseten Inference StackBaseten · San Francisco$180k–360k/yr
No roles match those filters.