Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Model Inference Engineer for Production-Scale AI

$325k

Jobleads-US

A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments. The ideal candidate has over 5 years of software engineering experience, strong familiarity with ML architectures, and experience with distributed systems. This role involves collaboration with researchers and focus on performance optimization. Compensation ranges from $325K to $490K. #J-18808-Ljbffr Jobleads-US

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Model Inference Engineer for Production-Scale AI in San Francisco, CA vacancy
  • Databricks Mosaic AI is seeking an experienced backend/infrastructure engineer to build platforms powering AI workloads, including model training, serving, and vector search. You will join...  ...visibility team connected to research, product, and enterprise use cases, helping... 
    Senior

    Databricks

    San Francisco, CA
    10 hours ago
  • MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference systems, optimizing for performance and reliability. The ideal candidate will have 3+ years of experience in production... 
    Senior

    MakerMaker.AI

    San Francisco, CA
    44 minutes ago
  • $220k - $320k

     ...Help us make inference blazingly fast. If you...  ...optimization techniques into production systems, we'd love...  ...language models for companies that...  ...frontier-quality AI at a fraction of the...  ..., and planet-scale hosting. We are...  ...ten-person team of engineers who work in-person... 
    Senior
    Full time
    Work at office

    Inference

    San Francisco, CA
    3 days ago
  • $190k - $265k

     ...enabling data and AI teams to solve the...  ...business. Founded by engineers — and customer-...  ...interfacing with data to scaling our services and...  ...the platforms and products that power...  ...apps, AI agents, model training, model serving...  ...Foundation Model Inference team is the backbone... 
    Suggested
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  • $300 per month

     ...vertically integrated AI infrastructure...  ...believe in the scale of our ambition and...  ...:At Crusoe, our Production Engineering team ensures the...  ...re looking for a Senior Production Engineer...  ...large language models to help us build...  ...AI pipelines and inference servicesDefine, measure... 
    Senior
    Temporary work

    Crusoe

    San Francisco, CA
    2 days ago
  • Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize... 
    Senior

    Causal Labs

    San Francisco, CA
    23 hours ago
  • MakerMaker in San Francisco is seeking a Senior ML systems engineer to build and operate production inference systems for large models. You will own performance, profiling, and optimizations to ensure high throughput and low latency in production. You will collaborate with... 
    Senior

    MakerMaker

    San Francisco, CA
    1 day ago
  • $300 per month

     ...vertically integrated AI infrastructure company...  ...urgency, who believe in the scale of our ambition and...  ...RoleAt Crusoe, our Production Engineering team ensures the reliability...  .... We’re looking for a Senior Production Engineer...  ...-scale training and inference clustersAutomate... 
    Senior
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...seeking a talented software engineer to join their dynamic Inference team. This role involves...  ...implementing infrastructure for large-scale multimodal models, focusing on high-...  ...closely with researchers and product teams to push the boundaries of AI technology, ensuring... 

    Jobleads-US

    San Francisco, CA
    3 days ago
  • Anyscale is seeking a Distributed LLM Inference Engineer in San Francisco, California. This pivotal role involves pushing the boundaries of performance for ML inference at scale. You'll work closely with product teams to deliver end-to-end solutions while leveraging open... 

    Anyscale

    San Francisco, CA
    3 days ago
  • $150k - $237.5k

     ...low-cost and large-scale energy storage...  ...we already have. Senior Software Engineer, Energy Storage This...  ...telemetry, asset modeling, alerting,...  ...boundaries to ensure the product system works well...  ...time. Apply AI tools to accelerate...  ...information, and inferences drawn from your... 
    Senior
    Full time

    Redwood Materials

    San Francisco, CA
    1 day ago
  • $161.3k - $241.9k

     ...combining frontier agentic AI, an enterprise-grade...  ...point. We have strong product-market fit and world-...  ...support. We’re scaling fast and defining a new...  ...customer interaction, every model inference, and every production...  ...for a Production Engineer to help build and operate... 
    Senior

    Neura Market

    San Francisco, CA
    3 days ago
  • $216k - $270k

    About Scale AIScale AI is the data foundation for AI, helping organizations...  ...and deploy reliable production AI applications. We...  ...than ever. New foundation models, reasoning techniques, agent...  ...one of the hardest engineering challenges.As a Senior Frontier Agent Engineer (... 
    Senior
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  • $350k

    Mirendil is looking for engineers to build infrastructure for frontier reasoning models at their San Francisco location. This role focuses on large-scale reinforcement learning (RL) model training and requires a solid understanding of engineering principles. The ideal candidate... 
    Senior

    Mirendil

    San Francisco, CA
    3 days ago
  • $225k

    Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed...  ...engineering fundamentals and experience with large-scale systems. Compensation includes a competitive salary range from... 

    Dormont Manufacturing Co

    San Francisco, CA
    2 days ago
  • Engineering Manager, Foundation Model Inference (FMAPI)RDQ427R519At Databricks, we are driven by a...  ...operate the premier data and AI infrastructure platform,...  ...to lead a cutting-edge product to success, we invite you...  ...our customers to serve, scale, and optimize frontier models... 
    Worldwide

    DataBricks

    San Francisco, CA
    3 days ago
  •  ...a Member of Technical Staff focused on AI Safety to lead red-teaming efforts and ensure...  ..., partner with researchers to define production safety standards, and research advanced...  ...expertise in LLM safety, strong software engineering skills, and relevant academic... 
    Senior

    Xcede

    San Francisco, CA
    2 days ago
  • $192k - $240k

     ...visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service...  ...resources, and support you need to grow your career.Engineering at BrexEngineering at Brex is about building systems that scale with speed and intention. Our teams span... 
    Senior
    Work at office
    Remote work
    Work from home

    Brex

    San Francisco, CA
    1 day ago
  •  ...fighting fires. You'll partner closely with Engineering, IT, Legal, and Business leaders to...  ...querying and alert tuning). • Ability to AI tools to streamline security operations and...  ...playbooks, and operational metrics that scale with company growth. • Strong cross-functional... 
    Senior

    Vailexa

    San Francisco, CA
    1 day ago
  • A tech startup focused on AI workloads is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and...  ...Ideal candidates should have strong software engineering skills and experience with ML inference... 
    Senior

    Gimlet Labs

    San Francisco, CA
    23 hours ago
  • $166k - $225k

     ...world's best data and AI infrastructure...  ...business. Databricks’ Model Serving product provides enterprises...  ...real-time, low-latency inference, governance, monitoring...  ...operationalize models at scale with strong SLAs and cost efficiency.As a Senior Engineer, you’ll play a... 
    Senior
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  • $172.43k - $230.95k

     ...vertically integrated AI infrastructure...  ...who believe in the scale of our ambition and...  ...About This Role:The Senior Software Engineer for the AI Model Lifecycle team will...  ...building cutting-edge AI products and solving challenging...  ...on GPU systems and inference frameworks.Benefits:... 
    Senior
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • Magnitude, located in San Francisco, is seeking a Sr. GTM Engineer to reimagine and scale the go-to-market motion. In this vital role, you will architect AI-powered experiments and integrate tools to drive growth. The ideal candidate has over 6 years in GTM or growth engineering... 
    Senior
    Flexible hours

    Magnitude

    San Francisco, CA
    1 day ago
  • Commure is building the AI Operating System for healthcare, enabling automated revenue...  ...and autonomous RCM processing at scale. The Claims Submissions team focuses on getting...  ...workflows and collaborate with clinicians and engineers to deliver reliable healthcare financial... 
    Senior

    Athelas

    San Francisco, CA
    1 day ago
  •  ...About the Team Our Inference team brings OpenAI’s most...  ...the world through our products. We empower consumers,...  ...our start-of-the-art AI models, allowing them to do things...  ...We are looking for an engineer who wants to take the...  ...to rapidly increasing scale. Are self-directed... 
    Full time

    OpenAI

    San Francisco, CA
    10 hours ago
  • $190k - $282k

     ...Production Engineer, Security Engineering Join to apply for the Production Engineer, Security Engineering...  ...role at CoreWeave . CoreWeave is the AI Hyperscaler, delivering a cloud platform...  ...Bash, or Go. Experience managing large-scale distributed systems and ensuring their... 
    Senior
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    San Francisco, CA
    1 day ago
  • Brighterway is hiring a Senior Software Engineer to design, build, and scale production systems across frontend, backend, and ML-adjacent workflows. You will own data models, APIs, and deployment pipelines in a hands-on role with high ownership. The role focuses on building... 
    Senior

    Brighterway

    San Francisco, CA
    4 days ago
  • Attentive is the AI marketing platform enabling 1:1 personalization at scale. The Senior Software Engineer in Intelligent Messaging will architect robust backend systems powered...  ...LLMs, collaborating with data scientists, product managers, and designers to empower marketers... 
    Senior

    Socket.dev

    San Francisco, CA
    40 minutes ago
  • $200k - $320k

    A leading AI company in insurance automation is looking for a Senior Software Engineer who will ship end-to-end features, diagnose and debug across technology stacks, and tune performance for scale. This role requires 4+ years of software engineering experience with a strong... 
    Senior

    Raydar

    San Francisco, CA
    4 days ago
  • $179k - $218k

     ...only vertically integrated AI infrastructure company...  ...urgency, who believe in the scale of our ambition and...  ...bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture...  ...hardware failures in the production environment. Lead Root Cause... 
    Senior
    Temporary work

    Crusoe

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Model Inference Engineer for Production-Scale AI. Be the first to apply!