Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML engineer

Xforia Inc

What you will do:

  • Optimize trained models for inference, ensuring they are ready for deployment.
  • Work closely with the internal AI research team to integrate inference optimization into the training process.
  • Provide expertise to help customers deploy customized inference services on our cloud.
  • Design and develop tools and systems that enable ML cycles, such as CI/CD/CT pipelines, with an emphasis on inference.
  • Deploy code/solutions to different environments, including batch and streaming/microservices architectures.
  • Serve as domain experts in ML inference, working closely with both internal teams and external customers.
  • Ensure models are efficiently deployed across various platforms, including cloud and standalone environments.
  • Collaborate closely with MLOps team members to design and implement scalable, efficient ML systems.
  • Continuously improve deployment processes and infrastructure to enhance model performance and reliability.
  • Improve the reliability, security, scalability, and observability of our distributed inference infrastructure.
  • Build tools to give visibility into bottlenecks and sources of instability and design and implement solutions to address high-priority issues.
  • Support benchmarking across different GPU accelerators and capture benchmarking data by writing tools and capturing data.
You will succeed in this role if you have:
  • Strong Python skills and experience in model deployment.
  • Experience with ML systems, particularly in deploying and optimizing modern LLMs.
  • Familiarity with transformer architectures and fine-tuning models for inference.
  • Ability to build and deploy ML models in cloud environments (e.g., AWS, Azure, GCP) and standalone systems.
  • Experience with batch processing and streaming/microservices architectures.
  • Proficiency in testing, debugging, and maintaining systems.
  • Familiarity with MLPerf Benchmarks and Language Model Evaluation Harness is a big plus.
  • Ability to collaborate effectively with cross-functional teams and provide technical guidance.
  • Excellent problem-solving skills and a proactive attitude towards challenges.
  • Own problems end-to-end and willingness to pick up new knowledge as needed to get the job done.
  • Work experience at a cloud provider or AI compute/sub-system company.
  • Experience with deep learning frameworks (such as PyTorch, TensorFlow).
  • Experience with deep learning runtimes (such as ONNX Runtime, TensorRT).
  • Experience with inference servers/model serving frameworks (such as Triton, TFServ, CubeFlow).
  • Experience with distributed systems collectives such as NCCL, OpenMPI.
  • Experience deploying ML workloads on distributed systems.
  • Understanding of the nuances in training and deployment of distributed ML models, and familiarity with techniques like quantization, sparsity, etc.
  • Good intuition for when off-the-shelf solutions will work and the ability to build tools quickly if they won't.
  • A humble attitude, eagerness to help colleagues, and a commitment to team success.
What do we offer:
  • Opportunity to work with cutting-edge technology and make a significant impact on ML inference capabilities.
  • Collaborative and inclusive work environment.
  • Competitive compensation and benefits package.
  • Access to state-of-the-art research facilities and resources.
  • Encouragement and support for continuous learning and professional development.
Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the ML engineer in United States vacancy
  •  ...and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier...  ...AI models and real-time applications. About the role As an ML Engineer at Sciforium, you will operate at the intersection of... 
    Suggested
    Full time
    Flexible hours

    Sciforium

    San Francisco, CA
    1 day ago
  • $195k - $300k

     ...Menlo Ventures, and Lightspeed. Built by a world-class team: Engineers, designers, and operators from places like Scale, Meta, Airbnb,...  ...Role: AI at Eve isn't a feature — it's the foundation. As an ML Engineer, you'll be working at the intersection of cutting-edge... 
    Suggested
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours

    Eve

    Remote
    1 day ago
  •  ...modes, and improve through agent architecture changes, prompt engineering, fine-tuning, or rule-based post-processing Go deep into the...  ...onboarding takes days, not weeks. Requirements ~2+ years building ML/AI systems in production ~ Built and deployed AI agents or... 
    Suggested
    Full time

    Triomics

    New York, NY
    1 day ago
  •  ...We are seeking a skilled ML Engineer with deep expertise in developing, optimizing, and deploying machine learning models, with a focus on healthcare applications. This role will center on creating AI models to analyze clinical data, including structured and unstructured... 
    Suggested
    Full time

    FSS Gov Solutions

    United States
    1 day ago
  •  ...Machine Learning Engineer (Llama AI Platform) Location: Remote (Preferred U.S. Time Zones) Employment Type: Full-Time Company...  ...of the following: Fine-tuning open-source LLMs. ML Engineering and MLOps practices. LangChain, LlamaIndex,... 
    Suggested
    Full time
    Remote work

    Performacentric

    Remote
    1 day ago
  •  ...WWC Global, an operating firm of Command Holdings, is seeking a Machine Learning (ML) Engineer to serve on a potential contract supporting USSOCOM's mission to transform the SOF Enterprise into a decisive data-centric organization by applying scientific methods, algorithms... 
    Full time
    Contract work
    Work at office
    Local area
    Visa sponsorship
    Work visa
    Flexible hours

    Command Holdings

    Remote
    1 day ago
  • $153.2k - $183.3k

     ...Meet the Team   As a Machine Learning Engineer II – Road & Lane, you will help develop next‑generation models that estimate road surfaces...  ...Master’s with 2+ years.   ~ Hands‑on experience developing ML models for perception tasks such as lane detection, road surface... 
    Remote job
    Full time

    Torc Robotics

    Remote
    1 day ago
  •  ...About the Role We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model... 
    Full time
    Work at office
    3 days per week

    Pika

    Remote
    1 day ago
  •  ...of speed, flexibility, and ingenuity to strengthen and protect our nation’s vital interests. Requisition #: 1418 Job Title:  ML Engineer  Location: Omaha, NE Clearance Level:   Active DoD Top Secret Overview: At Agile Defense we know that action defines... 
    Full time
    Contract work

    Agile Defense

    Omaha, NE
    1 day ago
  •  ...leave a clean trail of tests and dashboards behind. Required Qualifications: Expert-level PyTorch. Proven software engineer who loves ML; comfortable writing production code across the stack. Hands-on experience training or fine-tuning large language or other... 
    Full time
    Contract work
    Flexible hours
    Shift work

    Sesame, L.l.c.

    San Francisco, CA
    1 day ago
  •  ...We’re a team of AI researchers, designers, growth experts, and engineers rethinking human-computer interaction from the ground up. We value...  ...A, this is just the beginning. About the Role As a ML engineer at Wispr, you’ll play a crucial role in building the first... 
    Full time

    Wispr Flow

    San Francisco, CA
    1 day ago
  •  ...with technology, giving them more time to focus on what matters most: their patients. Position Overview We are hiring two ML Engineers / Researchers to help build the next generation of Knowtex's AI stack. We are looking for researchers with deep expertise in... 
    Full time

    Knowtex

    San Francisco, CA
    1 day ago
  •  ...About the Role We’re looking for an Applied ML Engineer to design, evaluate, and scale recommendation and ranking systems that power how content, ads, and interactive experiences are selected and surfaced in real time. This role focuses on decision-making systems, with... 
    Full time

    Darwin

    Palo Alto, CA
    1 day ago
  • $100k - $250k

     ...generative materials science. This innovative field blends AI, engineering, and materials science, revolutionizing how materials are created...  ...in technology and sustainability. The opportunity As an ML Engineer at Radical AI, you will be responsible for developing... 
    Full time

    Radical Ai

    New York, NY
    1 day ago
  •  ...CAEmployment Type: 1099, C2C, W-2Industry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: ML Engineer with LLMLocation: Sunnyvale, CA(onsite)Job Description:6-8 years of experience in machine learning and LLM, with a proven track... 

    SRI Tech

    Sunnyvale, CA
    3 days ago
  • TypeContract to Hire Job Description Job Title: Machine Learning Engineer (MLE I / MLE II / Senior) Location: United States (Team...  ...insights and opportunities for optimization. Design and build scalable ML systems and data-driven applications. Participate in technical discussions... 
    Full time
    Contract work
    Internship

    TalentsBridge

    Atlanta, GA
    2 days ago
  • Job DescriptionOverview: Our company is seeking a detail-oriented and highly analytical ML Engineer who will assist in driving our AI-driven product development initiatives. The successful candidate will possess a strong understanding of data analysis, along with the ability... 
    Full time
    Work at office

    Keystone Cooperative

    Indianapolis, IN
    3 days ago
  • $207k - $301k

     ...prediction, ranking, embedding) in production and experience building architecture in different modeling domains.5 years of experience with ML design and ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).5 years of... 

    Google

    Mountain View, CA
    2 days ago
  •  ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: ML EngineerLocation: Charlotte , NC (hybrid 2 days a week)JD:...  ...predictive models using machine learning algorithmsPerform feature engineering, model selection, and hyperparameter tuning to optimize model... 
    Hourly pay
    2 days per week

    SRI Tech

    Charlotte, NC
    3 days ago
  • $106.9k - $160.4k

     ...timberlands, wood products, and corporate functions. As we continue to scale AI across the enterprise, we are seeking a skilled ML Engineer to design, build, and operationalize machine learning solutions that are reliable, scalable, secure, and delivering measurable business... 
    Full time
    Temporary work

    Weyerhaeuser

    Seattle, WA
    2 days ago
  • $100.4k - $180.7k

    Posting TitleML and Optimization Engineer.LocationCO - Golden.Position TypeLimited Term (Fixed Term).Hours Per Week40.Working at NLRNLR is...  ..., with core strengths in high‑performance computing, AI/ML, modeling and simulation, and visualization. We steward state‑of... 
    Full time
    Fixed term contract
    Live in
    Local area
    Remote work
    Relocation
    Shift work

    National Renewable Energy Laboratory

    Golden, CO
    5 days ago
  •  ...Position Overview We are seeking an experienced machine learning engineer to develop advanced post-training quantization methods for large language models (LLMs), diffusion models, and other ML applications for our revolutionary optical inference engines. This role... 
    Full time

    Neurophos

    Austin, TX
    1 day ago
  •  ...first commercially available AI Co-Scientist. It is a discovery engine that transforms messy biological data into insights in minutes. Scientists...  ...to patient outcomes. ABOUT THE ROLE We are hiring an ML Engineer, Discovery Applications to build the high level, end-to-... 
    Full time
    Work at office

    Mithrl

    San Francisco, CA
    1 day ago
  • $213k - $263k

    Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access...
    Full time
    Temporary work
    Remote work

    Waymo

    New York, NY
    1 day ago
  • $160k - $220k

     ...at scale. This is a hands-on role where you'll see direct impact on business metrics daily. What You'll Do Build and deploy ML models serving 100M+ predictions daily Develop ranking algorithms that balance relevance, diversity, and revenue Implement real... 
    Remote job
    Full time

    Launch Potato

    Remote
    1 day ago
  • $156k - $234k

     ...execution.  Stay up to date with the latest developments in AI and ML for autonomous driving. Independently develop offline...  ...team technical solutions and drives consensus.  Mentors and guides engineers within the group.   Bachelor’s Degree in Computer Science, Robotics... 
    Remote job
    Full time
    Immediate start
    Relocation

    Torc Robotics

    Remote
    1 day ago
  •  ...first commercially available AI Co-Scientist. It is a discovery engine that transforms messy biological data into insights in minutes. Scientists...  ...to patient outcomes. ABOUT THE ROLE We are hiring an ML Engineer, Analysis and Simulation to build the core analytical... 
    Full time
    Work at office

    Mithrl

    San Francisco, CA
    1 day ago
  •  ...pretraining science. We build foundational understanding of models to advance the frontier of intelligence. About the role: As a ML Engineer, you’ll build and operate the infrastructure that makes cutting-edge machine learning research possible. At Tilde, we believe... 
    Full time
    Internship

    Tilde Research

    Remote
    1 day ago
  • Entefy’s vision is simplifying how people interact digitally. To make it happen, we’re seeking computer vision-aries. Our hyper-talented Product team is seeking a Machine Learning and AI expert with the skills to redefine how computers make sense of the visual world. ...
    Remote job
    Full time

    Entefy

    Remote
    1 day ago
  • $143k - $197k

     ...combining the highest-quality care and technology solutions. About the Role We are looking for an experienced Senior AI/ML Engineer to spearhead the design, development, and deployment of next-generation AI capabilities for our platform. In this role, you won't... 
    Full time

    Lyra Health

    United States
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML engineer. Be the first to apply!