ML engineer
Xforia Inc
What you will do:
- Optimize trained models for inference, ensuring they are ready for deployment.
- Work closely with the internal AI research team to integrate inference optimization into the training process.
- Provide expertise to help customers deploy customized inference services on our cloud.
- Design and develop tools and systems that enable ML cycles, such as CI/CD/CT pipelines, with an emphasis on inference.
- Deploy code/solutions to different environments, including batch and streaming/microservices architectures.
- Serve as domain experts in ML inference, working closely with both internal teams and external customers.
- Ensure models are efficiently deployed across various platforms, including cloud and standalone environments.
- Collaborate closely with MLOps team members to design and implement scalable, efficient ML systems.
- Continuously improve deployment processes and infrastructure to enhance model performance and reliability.
- Improve the reliability, security, scalability, and observability of our distributed inference infrastructure.
- Build tools to give visibility into bottlenecks and sources of instability and design and implement solutions to address high-priority issues.
- Support benchmarking across different GPU accelerators and capture benchmarking data by writing tools and capturing data.
- Strong Python skills and experience in model deployment.
- Experience with ML systems, particularly in deploying and optimizing modern LLMs.
- Familiarity with transformer architectures and fine-tuning models for inference.
- Ability to build and deploy ML models in cloud environments (e.g., AWS, Azure, GCP) and standalone systems.
- Experience with batch processing and streaming/microservices architectures.
- Proficiency in testing, debugging, and maintaining systems.
- Familiarity with MLPerf Benchmarks and Language Model Evaluation Harness is a big plus.
- Ability to collaborate effectively with cross-functional teams and provide technical guidance.
- Excellent problem-solving skills and a proactive attitude towards challenges.
- Own problems end-to-end and willingness to pick up new knowledge as needed to get the job done.
- Work experience at a cloud provider or AI compute/sub-system company.
- Experience with deep learning frameworks (such as PyTorch, TensorFlow).
- Experience with deep learning runtimes (such as ONNX Runtime, TensorRT).
- Experience with inference servers/model serving frameworks (such as Triton, TFServ, CubeFlow).
- Experience with distributed systems collectives such as NCCL, OpenMPI.
- Experience deploying ML workloads on distributed systems.
- Understanding of the nuances in training and deployment of distributed ML models, and familiarity with techniques like quantization, sparsity, etc.
- Good intuition for when off-the-shelf solutions will work and the ability to build tools quickly if they won't.
- A humble attitude, eagerness to help colleagues, and a commitment to team success.
- Opportunity to work with cutting-edge technology and make a significant impact on ML inference capabilities.
- Collaborative and inclusive work environment.
- Competitive compensation and benefits package.
- Access to state-of-the-art research facilities and resources.
- Encouragement and support for continuous learning and professional development.
Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the ML engineer in United States vacancy
- ...and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier... ...AI models and real-time applications. About the role As an ML Engineer at Sciforium, you will operate at the intersection of...SuggestedFull timeFlexible hours
$195k - $300k
...Menlo Ventures, and Lightspeed. Built by a world-class team: Engineers, designers, and operators from places like Scale, Meta, Airbnb,... ...Role: AI at Eve isn't a feature — it's the foundation. As an ML Engineer, you'll be working at the intersection of cutting-edge...SuggestedFull timeTemporary workWork at officeLocal areaFlexible hours- ...modes, and improve through agent architecture changes, prompt engineering, fine-tuning, or rule-based post-processing Go deep into the... ...onboarding takes days, not weeks. Requirements ~2+ years building ML/AI systems in production ~ Built and deployed AI agents or...SuggestedFull time
- ...We are seeking a skilled ML Engineer with deep expertise in developing, optimizing, and deploying machine learning models, with a focus on healthcare applications. This role will center on creating AI models to analyze clinical data, including structured and unstructured...SuggestedFull time
- ...Machine Learning Engineer (Llama AI Platform) Location: Remote (Preferred U.S. Time Zones) Employment Type: Full-Time Company... ...of the following: Fine-tuning open-source LLMs. ML Engineering and MLOps practices. LangChain, LlamaIndex,...SuggestedFull timeRemote work
- ...WWC Global, an operating firm of Command Holdings, is seeking a Machine Learning (ML) Engineer to serve on a potential contract supporting USSOCOM's mission to transform the SOF Enterprise into a decisive data-centric organization by applying scientific methods, algorithms...Full timeContract workWork at officeLocal areaVisa sponsorshipWork visaFlexible hours
$153.2k - $183.3k
...Meet the Team As a Machine Learning Engineer II – Road & Lane, you will help develop next‑generation models that estimate road surfaces... ...Master’s with 2+ years. ~ Hands‑on experience developing ML models for perception tasks such as lane detection, road surface...Remote jobFull time- ...About the Role We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model...Full timeWork at office3 days per week
- ...of speed, flexibility, and ingenuity to strengthen and protect our nation’s vital interests. Requisition #: 1418 Job Title: ML Engineer Location: Omaha, NE Clearance Level: Active DoD Top Secret Overview: At Agile Defense we know that action defines...Full timeContract work
- ...leave a clean trail of tests and dashboards behind. Required Qualifications: Expert-level PyTorch. Proven software engineer who loves ML; comfortable writing production code across the stack. Hands-on experience training or fine-tuning large language or other...Full timeContract workFlexible hoursShift work
- ...We’re a team of AI researchers, designers, growth experts, and engineers rethinking human-computer interaction from the ground up. We value... ...A, this is just the beginning. About the Role As a ML engineer at Wispr, you’ll play a crucial role in building the first...Full time
- ...with technology, giving them more time to focus on what matters most: their patients. Position Overview We are hiring two ML Engineers / Researchers to help build the next generation of Knowtex's AI stack. We are looking for researchers with deep expertise in...Full time
- ...About the Role We’re looking for an Applied ML Engineer to design, evaluate, and scale recommendation and ranking systems that power how content, ads, and interactive experiences are selected and surfaced in real time. This role focuses on decision-making systems, with...Full time
$100k - $250k
...generative materials science. This innovative field blends AI, engineering, and materials science, revolutionizing how materials are created... ...in technology and sustainability. The opportunity As an ML Engineer at Radical AI, you will be responsible for developing...Full time- ...CAEmployment Type: 1099, C2C, W-2Industry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: ML Engineer with LLMLocation: Sunnyvale, CA(onsite)Job Description:6-8 years of experience in machine learning and LLM, with a proven track...
- TypeContract to Hire Job Description Job Title: Machine Learning Engineer (MLE I / MLE II / Senior) Location: United States (Team... ...insights and opportunities for optimization. Design and build scalable ML systems and data-driven applications. Participate in technical discussions...Full timeContract workInternship
- Job DescriptionOverview: Our company is seeking a detail-oriented and highly analytical ML Engineer who will assist in driving our AI-driven product development initiatives. The successful candidate will possess a strong understanding of data analysis, along with the ability...Full timeWork at office
$207k - $301k
...prediction, ranking, embedding) in production and experience building architecture in different modeling domains.5 years of experience with ML design and ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).5 years of...- ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: ML EngineerLocation: Charlotte , NC (hybrid 2 days a week)JD:... ...predictive models using machine learning algorithmsPerform feature engineering, model selection, and hyperparameter tuning to optimize model...Hourly pay2 days per week
$106.9k - $160.4k
...timberlands, wood products, and corporate functions. As we continue to scale AI across the enterprise, we are seeking a skilled ML Engineer to design, build, and operationalize machine learning solutions that are reliable, scalable, secure, and delivering measurable business...Full timeTemporary work$100.4k - $180.7k
Posting TitleML and Optimization Engineer.LocationCO - Golden.Position TypeLimited Term (Fixed Term).Hours Per Week40.Working at NLRNLR is... ..., with core strengths in high‑performance computing, AI/ML, modeling and simulation, and visualization. We steward state‑of...Full timeFixed term contractLive inLocal areaRemote workRelocationShift work- ...Position Overview We are seeking an experienced machine learning engineer to develop advanced post-training quantization methods for large language models (LLMs), diffusion models, and other ML applications for our revolutionary optical inference engines. This role...Full time
- ...first commercially available AI Co-Scientist. It is a discovery engine that transforms messy biological data into insights in minutes. Scientists... ...to patient outcomes. ABOUT THE ROLE We are hiring an ML Engineer, Discovery Applications to build the high level, end-to-...Full timeWork at office
$213k - $263k
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access...Full timeTemporary workRemote work$160k - $220k
...at scale. This is a hands-on role where you'll see direct impact on business metrics daily. What You'll Do Build and deploy ML models serving 100M+ predictions daily Develop ranking algorithms that balance relevance, diversity, and revenue Implement real...Remote jobFull time$156k - $234k
...execution. Stay up to date with the latest developments in AI and ML for autonomous driving. Independently develop offline... ...team technical solutions and drives consensus. Mentors and guides engineers within the group. Bachelor’s Degree in Computer Science, Robotics...Remote jobFull timeImmediate startRelocation- ...first commercially available AI Co-Scientist. It is a discovery engine that transforms messy biological data into insights in minutes. Scientists... ...to patient outcomes. ABOUT THE ROLE We are hiring an ML Engineer, Analysis and Simulation to build the core analytical...Full timeWork at office
- ...pretraining science. We build foundational understanding of models to advance the frontier of intelligence. About the role: As a ML Engineer, you’ll build and operate the infrastructure that makes cutting-edge machine learning research possible. At Tilde, we believe...Full timeInternship
- Entefy’s vision is simplifying how people interact digitally. To make it happen, we’re seeking computer vision-aries. Our hyper-talented Product team is seeking a Machine Learning and AI expert with the skills to redefine how computers make sense of the visual world. ...Remote jobFull time
$143k - $197k
...combining the highest-quality care and technology solutions. About the Role We are looking for an experienced Senior AI/ML Engineer to spearhead the design, development, and deployment of next-generation AI capabilities for our platform. In this role, you won't...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML engineer. Be the first to apply!
Related searches
- data scientist machine learning engineer United States
- machine learning ai engineer United States
- computer vision machine learning engineer United States
- machine learning engineer United States
- ai ml engineer United States
- graduate machine learning engineer United States
- machine learning software engineer United States
- entry level machine learning engineer United States
- junior machine learning engineer United States
- staff machine learning engineer United States

