Staff ML Performance Engineer - GPU & Inference
Modal
Modal is building an infrastructure layer for AI and is seeking strong engineers to optimize ML systems for performance at scale. You will contribute to Modal’s container runtime and open-source projects, pushing language and diffusion models toward higher throughput and lower latency. The role emphasizes working with Torch, TensorRT, CUDA, and NVIDIA GPU architectures to maximize efficiency, while exploring low-level OS foundations to improve performance and reliability. #J-18808-Ljbffr Modal
Vacancy posted more than 2 months ago
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff ML Performance Engineer - GPU & Inference. Be the first to apply!
Related searches
- assistant mechanical engineer San Francisco, CA
- assistant engineer San Francisco, CA
- staff engineer San Francisco, CA
- staff data engineer San Francisco, CA
- software engineer staff San Francisco, CA
- assistant electrical engineer San Francisco, CA
- staff design engineer San Francisco, CA
- senior staff engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
