Amazon Web Services (AWS)
Seattle / Global
Senior ML Inference Engineer — High-Performance LLM Serving
- $168.000 - $227.000
You have blocked notifications
Oops! You have blocked notifications. Click here for more info
You have blocked notifications, please check your browser settings.
You're currently subscribed to job notifications
Subscribe to notifications
You will no longer receive notifications
Seattle / Global
Annapurna Labs (U.S.) Inc. in Seattle seeks a senior software engineer to advance AWS Neuron, the complete software stack for AWS Inferentia and Trainium.
You will deliver high-performance model inference for customer workloads on Inferentia- and Trainium-powered instances and lead core serving tech within open-source frameworks such as vLLM and SGLang. You will collaborate with model development, compiler, runtime, and performance teams to ensure production-ready accuracy, scalability, and
#J-18808-Ljbffr
Seattle / Global
Seattle / Global
Seattle / Global
Seattle / Global
Bellevue / Global
Seattle / Global