Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote AI Inference Engineer — Edge Model Deployment & Optimization (Burlingame)

Full-time

Quadric Inc.

A leading technology company in California is seeking an AI Inference Engineer to bridge AI models with unique platforms. Key responsibilities include model optimization, deployment, and performance profiling. Candidates should have a Bachelor’s or Master’s degree, 5+ years' experience in AI frameworks, and proficiency in C/C++ and Python. Competitive benefits included, such as health care, retirement plans, and work from home options.
#J-18808-Ljbffr

Vacancy posted 2 hours ago
Similar jobs that could be interesting for youBased on the Remote AI Inference Engineer — Edge Model Deployment & Optimization (Burlingame) in Burlingame, CA vacancy
  •  ...-modality foundation model to drive the next generation...  .... As a Model Optimization & Deployment Engineer, you will focus on...  ...highly concurrent inference code to ensure real-time...  ...execution on edge devices. In this role...  ...memory bandwidth on AI accelerators. Write... 
    Suggested
    Temporary work
    Relocation package

    Zoox

    San Diego, CA
    15 days ago
  •  ...seeking a highly motivated and technically skilled Edge AI/Model Optimization Engineer to support the deployment, optimization, and sustainment of AI and agentic...  ...Language Models (LLMs), embedding models, and AI inference services for constrained hardware platforms,... 
    Suggested
    Local area

    NextGen Federal Systems

    Aberdeen, MD
    4 days ago
  •  ...architecture. Quadric's co-optimized software and...  ...network (NN) inference workloads in a wide variety of edge and endpoint...  ...Role: The AI Inference Engineer in Quadric is the...  ...world of AI/LLM models and Quadric unique...  ...optimize the model deployment for efficient inference... 
    Suggested
    Full time
    Temporary work
    Work from home

    Quadric Inc.

    Burlingame, CA
    2 hours ago
  • $110k - $270k

     ...Quadric's co-optimized software and...  ...network (NN) inference workloads in...  ...wide variety of edge and endpoint...  ...Role The AI Inference Engineer in Quadric is...  ...world of AI/LLM models and Quadric unique...  ...the model deployment for efficient...  ...week at our Burlingame office, the... 
    Suggested
    Work at office
    Local area
    Immediate start
    Flexible hours
    2 days per week

    quadric, Inc

    Burlingame, CA
    2 days ago
  • $100k - $150k

     ...Edge Computing AI Engineer – Remote Bright Vision Technologies is a technology consulting...  ...Computing AI Engineer to design, optimize, and deploy machine learning models that run efficiently on...  ...with at least one major edge inference framework. Solid understanding... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    Sunnyvale, CA
    2 days ago
  • $30 - $90 per hour

     ...Role Overview As an AI Engineer, you will play a pivotal...  ...train next-generation models. Your contributions...  ...Design, build, and optimize robust machine learning...  ...infrastructure and model deployment. Orchestrate containerized...  ...- $90 Eligibility Remote work opportunity... 
    Remote work
    Hourly pay
    Contract work

    SaidGig

    United States
    5 days ago
  • $100k - $150k

    Role Description We are looking for an Edge AI Engineer to design, optimize, and deploy machine learning models that run efficiently on resource-constrained edge devices...  ...on target hardware. ~Build cross-platform inference runtimes leveraging frameworks such as... 
    Full time
    Local area
    Immediate start

    Bright Vision Technologies

    Remote
    12 hours ago
  • $197.5k - $272k

     ...transformation to AI-enabled...  ...proven by global deployment, we’re solving...  ...grade AI on the Edge. We are looking...  ...great Staff AI Engineer to join our seasoned...  ...and deploy AI models that analyze...  ...devices and model optimization. You will work...  ...(C++14/17 for inference). ~ Deep... 
    Full time
    Work at office
    Worldwide
    Flexible hours
    Shift work
    3 days per week

    Sonatus, Inc

    Remote
    8 hours ago
  •  ...Travel Required: No Remote Type: Hybrid...  ...LLC is looking for a AI Engineer who will support the...  ...platforms. Apply cutting-edge techniques in statistical...  ...research, build, and deploy complex, user-...  ...Python development GPU inference optimization Preferred... 
    Remote work
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Sunayu

    Remote
    8 hours ago
  •  ...a Senior Forward Deployed AI Engineer to support our Public...  ...on building and optimizing production ready...  ...prototype models into scalable, efficient...  ...to edge devices in restricted...  ...will consider fully remote candidates. Qualified...  ...stack—from model inference on consumer hardware... 
    Remote work
    Full time
    Casual work
    Live out
    Work at office
    Local area
    Flexible hours

    Webai, Inc.

    Austin, TX
    8 hours ago
  • $147.2k - $210.3k

     ...highly skilled AI Engineer to design, develop, and deploy advanced AI solutions...  ...building and optimizing AI technologies...  .... ~AI Model Development – Design...  ...and inference efficiency....  ...Requirements ~Fully Remote Opportunity – Work...  ...Work on Cutting-Edge AI Solutions in... 
    Remote work
    Full time
    Flexible hours

    Gainwell Technologies LLC

    Remote
    5 days ago
  • $175k - $250k

     ...work EST timezone. Remote | Full-time...  ...developing a cutting-edge autonomous agent...  ...outcomes. The Staff AI Engineer will be...  ...Responsibilities: Learning & Optimization Feedback Loop...  ...candidates for deployment autonomously....  ...strategies. Model & Inference Infrastructure... 
    Remote job
    Full time
    Immediate start
    Shift work

    Mlabs

    Massachusetts
    8 hours ago
  • $45 - $60 per hour

     ...architecture. Quadric's co-optimized software and...  ...network (NN) inference workloads in a wide variety of edge and endpoint devices...  ...be based out of our Burlingame, California office....  ...Responsibilities: Model pruning: Prune the...  ...industry experts in AI and semiconductor technology... 
    Hourly pay
    Temporary work
    Internship
    Work at office
    Relocation

    quadric, Inc

    Burlingame, CA
    3 days ago
  • $150k - $250k

     ...in-class Voice AI models powering the...  ...models serve 600M+ inference calls monthly,...  ...an Applied AI Engineer to join our...  ...of cutting-edge AI research and...  ...to production deployment. Continue to provide...  ...is performing optimally and meeting...  ...is a remote-first company,... 
    Remote work
    Full time
    Work at office
    Local area
    Relocation

    Assemblyai, Inc.

    New York, NY
    8 hours ago
  •  ...infrastructure for AI-driven...  ...and inference workloads reliable...  ...style compute, deployment workflows,...  ...researchers and engineers, and make...  ...supporting model inference, model...  ..., image optimization, CI/CD systems...  .... ~Remote-first work with...  ...to cutting-edge ideas across... 
    Remote work
    Full time

    FirstPrinciples

    Remote
    3 days ago
  • $117.8k - $168.3k

     ...highly skilled AI Engineer to design, develop, and deploy advanced AI solutions...  ...building and optimizing AI technologies...  .... ~AI Model Development: Design...  ..., and inference efficiency....  ...Requirements ~Fully Remote Opportunity – Work...  ...Work on Cutting-Edge AI Solutions in... 
    Remote work
    Full time
    Flexible hours

    Gainwell Technologies LLC

    Remote
    7 days ago
  • $70k - $200k

     ...are We? Field AI is transforming...  ...-globally-deployed solutions delivering...  ...improving models through real-field...  ...customers and engineers to deliver value...  ...evaluate models, optimize inference, and deploy...  ...deploying AI in edge or resource-constrained...  ...a hybrid or remote option. Why... 
    Remote work
    Full time

    Field AI

    Irvine, CA
    8 hours ago
  •  ...are currently seeking a AI Foundational Model Engineer to join our team in Jersey...  ...purpose Design, build, deploy, and optimize enterprise-grade AI...  ...improvement. Model adaptation, inference optimization, APIs,...  ...While many positions offer remote or hybrid work options, these... 
    Remote work
    Full time
    Temporary work
    Work at office
    Flexible hours

    NTT Ltd

    Jersey City, NJ
    22 days ago
  • $110k

     ...innovative digital engineering company...  ...advanced sensing, AI/ML, and...  ...operational deployment—supporting the...  ...developing and optimizing machine learning models, conducting...  ...turning cutting-edge research into...  ...testing, and inference pipelines for...  ...to imagery, remote sensing, or... 
    Remote work
    Full time
    Flexible hours

    Prime Solutions Group, Inc.

    Remote
    8 hours ago
  •  ...Applied AI / Machine Learning Engineer, are you there?...  ...reading! 100% Remote Salary in USD...  ...specialized multimodal models and are...  ...production inference across server and edge environments....  ...will Optimize multimodal AI...  ...Package and deploy transformer and... 
    Remote work
    Contract work
    Flexible hours

    Athyna

    Miami, FL
    4 days ago
  • $219.3k - $274.1k

     ...Deepgram's speech AI models are among the...  ...different engineering problem, and it...  ...embedded and edge platforms. You...  ...the stack: ~Optimizing and compiling...  ...for on-device inference. ~Writing performance...  ...for on-device deployment, including...  ...hours and remote work options.... 
    Remote work
    Full time
    Flexible hours

    Deepgram

    Remote
    1 day ago
  •  ...is seeking an AI Infra Engineer, to support the...  ...of cutting‑edge and energy efficient...  ...cooling optimization, and more. Our...  ...support business deployment. You will also...  ...with two days remote....  ...and forecasting models for time‑series...  ...networks, causal inference, Markov models... 
    Remote work
    Work at office
    Local area
    3 days per week

    Lenovo

    Morrisville, NC
    5 days ago
  •  ...hiring a Full-Stack AI Engineer to design, build, and deploy production-ready...  ...Large Language Models (LLMs), machine...  ...~ChromaDB ~Optimize prompt engineering and inference workflows. ~...  ...to cutting-edge AI engineering,...  ...ownership. ~Fully remote role with long-term... 
    Remote work
    Full time

    Pavago

    Remote
    13 days ago
  • $150k - $200k

     ...on the cutting edge of the...  ...Summary of the Lead AI Software Engineer**This senior-level...  ...predictive models. You will co-design...  ...& guardrails, deployment, and...  ...This position is remote and based in the...  ...and GPU-aware inference where applicable...  ...budgets, and optimize prompts,... 
    Remote work
    Local area
    Flexible hours

    Streamline Healthcare Solutions

    Springfield, IL
    5 days ago
  •  ...Computer Vision/AI Engineer BrightAI is...  ...building a cutting-edge AI platform...  ...intelligence and deploy next-generation...  ...Vision and ML model development —...  ...and real-time inference. Drive CV projects...  ...performance optimization. Prioritize...  ...perception problems, remote sensing, or... 
    Remote work
    Full time
    Immediate start

    Brightai

    Remote
    8 hours ago
  • $117.7k - $221.4k

     ...for embodied AI systems. We...  ...on stronger models, but also on...  ...questions, new edge cases, and...  ...how Cola engineers think: build...  ...instead of optimizing any one of them...  ...featurization, and inference foundations...  ...of deploying them in production...  ...is based remotely, but if the... 
    Remote work
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  •  ...Focused on optimizing Deepgram's speech models for low-power consumer hardware...  ...-time Embedded AI Engineer will enhance on-device inference across various...  ...accelerators while working remotely. Key...  ...techniques for on-device deployment Familiarity with edge inference runtimes... 
    Remote work
    Full time

    Virtual Vocations Inc

    United States
    1 day ago
  •  ...Our team analyzes inference stack performance across...  ...the application, model, and fleet layers to...  ...understanding into performance optimizations and models that...  ...collaborating with engineering and research teams...  ...OpenAI is an AI research and deployment company dedicated to... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...training and deploying frontier models for developers...  ...are building AI systems to power...  ...researchers, engineers, designers, and...  ...teams to deploy optimized NLP models to...  ...throughput of inference. ~ Strong...  ...on the cutting edge of AI research...  ...improvement Remote-flexible, offices... 
    Remote work
    Full time
    Work experience placement
    Work at office
    Flexible hours

    Cohere

    San Francisco, CA
    8 hours ago
  • $100k - $150k

    AI Research Engineer - Remote Bright Vision Technologies is...  ...bridge cutting-edge applied research...  ...experimentation to deployment and continuous...  ...large language models, and adjacent areas...  ...training and inference pipelines using...  ...accelerator resources. * Optimize models for... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    Hoboken, NJ
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote AI Inference Engineer — Edge Model Deployment & Optimization (Burlingame). Be the first to apply!