Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Infrastructure Engineer

$100k - $150k
Full-time

Bright Vision Technologies

Role Description

We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated an impact on production of AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production.

Key Responsibilities

  • Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost.
  • Identify and eliminate bottlenecks across data loading, model compute, communication, and memory.
  • Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference.
  • Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding.
  • Tune attention implementations using Flash Attention, paged attention, and related techniques.
  • Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving.
  • Drive compiler-level optimizations using Triton, XLA, Torch Inductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains.
  • Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training.
  • Build and maintain rigorous benchmark suites and regression frameworks across workloads.
  • Collaborate with ML and platform engineering teams to embed best practices in standard pipelines.
  • Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies.
  • Evaluate new hardware and software offerings and advise on adoption.
  • Document performance tuning playbooks and share findings broadly across engineering teams.
  • Stay current with AI systems to research and translate advances into production improvements.

Qualifications

  • Bachelor's or master's degree in computer science, Computer Engineering, or related field.
  • Six or more years of experience in performance engineering, ML systems, or HPC.
  • Strong proficiency in Python and C++.
  • Hands-on experience optimizing deep learning workloads on modern GPUs.
  • Deep understanding of distributed training and inference techniques.
  • Experience with profiling tools across CPU, GPU, and distributed systems.
  • Familiarity with model compression techniques and their accuracy implications.
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies.
  • Excellent measurement, debugging, and analytical reasoning skills.
  • Strong communication and collaboration skills.

Preferred Qualifications

  • Experience optimizing LLM inference at production scale.
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects.
  • Familiarity with custom kernel authoring in Triton or CUTLASS.
  • Experience with FinOps for AI workloads.
  • Publications or talks on AI systems performance.

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at View phone number on remotive.com.

Learn more about Bright Vision Technologies at .

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the AI Infrastructure Engineer in Remote vacancy
  •  ...are poised to disrupt the billion-dollar engineering simulation industry with our fast-...  ...You will help develop and improve core infrastructure and user-facing capabilities across our...  ...generation of engineering software and AI-enabled workflows. What You’ll Work... 
    Suggested
    Full time

    Flexcompute

    Remote
    12 hours ago
  • ​GTSC seeks a   Cloud Architect Artificial Intelligence (AI) Engineer , Level IV, to support our customer in the   Annapolis Junction, MD   area. Location:   Annapolis Junction, Maryland All work is on-site. This is not a hybrid or remote position.   Mission... 
    Suggested
    Full time
    Temporary work

    Gtsc-Talent Solutions

    Remote
    12 hours ago
  • $191k - $315k

     ...Overview:  The Network Growth and Relationship AI team is at the forefront of creating...  ...close collaboration with the product, engineering and data science team and has a very...  ...Prior experience with large scale ML data infrastructure ~ Experience with developing and designing... 
    Suggested
    Full time
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Remote
    12 hours ago
  •  ...This role sits at the intersection of platform engineering, site reliability, and applied ML systems. The...  ...reliability, scalability, and operability of Meshy’s AI model serving stack, along with core engineering infrastructure. The team operates a conventional production... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    MeshyAI

    San Francisco, CA
    3 days ago
  •  ...TetraScience is the Scientific Data and AI company. We are catalyzing the Scientific...  ...players in compute, cloud, data, and AI infrastructure have converged on TetraScience as the de...  ...We’re looking for a Senior AI Platform Engineer to help design, build, and scale our AI... 
    Suggested
    Immediate start
    Remote work
    Flexible hours

    TetraScience

    New York, NY
    3 days ago
  •  ...Senior Azure AI Software Engineer Location: Hartford, Connecticut (Onsite) Employment Type: Full-time Experience Level: Mid–Senior...  ...expertise in Azure AI services, C#/.NET development, and cloud infrastructure who enjoys building scalable, production-grade systems.... 
    Full time

    Core Talent Finder

    Remote
    12 hours ago
  • $125k - $220k

     ...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. AI ENGINEER, PLATFORM INFRASTRUCTURE, SPECIAL PROGRAMS As an AI Engineer, Platform Infrastructure you will build the tooling, and work with our... 
    Permanent employment
    Full time
    Temporary work
    Immediate start
    Weekend work

    Spacex

    Remote
    12 hours ago
  •  ...impact by investing in advanced analytics, AI, and new capabilities that help...  ...The Role We are looking for an AI Engineer to help build and evolve the core application...  ...understand how application logic, workflows, infrastructure, and internal tooling work together, and... 
    Full time
    Local area

    Medisolv, Inc.

    Remote
    12 hours ago
  • $160k - $235k

     ...opportunities in their team's network.  This role is part of the  AI Platform team, which owns the AI services that power Affinity's...  ...to deliver actionable insights to customers. As a  Senior AI Engineer , you will collaborate with machine learning engineers, data... 
    Remote job
    Full time
    Work at office
    Worldwide
    Flexible hours
    2 days per week
    3 days per week

    Affinity.co

    San Francisco, CA
    12 hours ago
  • $286.2k - $326.7k

    Senior. Distinguished AI Engineer - Agentic AI Platform (Remote Eligible) At Capital One, we are creating responsible and reliable...  ...customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine... 
    Full time
    Part time
    Work at office
    Local area
    Remote work

    Capital One Financial Corporation

    San Francisco, CA
    12 hours ago
  • $150k - $300k

     ...About Traversal Traversal is the AI Site Reliability Engineer (SRE) for the enterprise—already trusted by some of the largest companies in...  ...constantly. This is a place to grow your career, make a real impact, and help define a new category of infrastructure software.... 
    Full time
    Work at office
    Flexible hours

    Traversal

    Remote
    12 hours ago
  • $160k - $220k

     ...of:     The role   We are looking for an experienced AI Engineer to lead the implementation of Azure AI Foundry within an established...  ...our existing data platform, governance model, and analytics infrastructure.   You will work closely with data engineering,... 
    Full time
    Remote work

    Valtech Se

    New York, NY
    12 hours ago
  •  ...Founders Fund-backed NVIDIA cloud partner building the infrastructure platform that powers AI at scale. We connect AI Factories—high-performance GPU...  .... Your job is to change that. As an AI Infrastructure Engineer, you'll work directly with AI platform customers to get... 
    Remote work

    Hydra Host

    New York, NY
    4 days ago
  •  ...HIPAA) environment, and we are building toward being an AI-native company. We are past the experimentation phase but...  ...stage is exactly why this role exists. The Staff AI Engineer owns ShiftKey’s AI infrastructure layer: the organizational knowledge platform, the infrastructure... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    Shift work

    Shift Key

    Remote
    12 hours ago
  • A leading U.S. technology firm is hiring an AI Infrastructure Engineer for a full-time remote position with H-1B visa sponsorship available. This role involves designing and optimizing AI platforms, deploying Kubernetes and Docker container environments, and enhancing... 
    Remote job
    Full time
    H1b
    Visa sponsorship

    NewsNowGh

    New York, NY
    4 days ago
  • AI Infrastructure Engineer Job in USA 2026 with H-1B Visa Sponsorship AI Infrastructure Engineer Job in USA 2026 with H-1B Visa Sponsorship A leading U.S.-based technology firm is hiring an AI Infrastructure Engineer for a full-time remote position with H-1B visa sponsorship... 
    Full time
    H1b
    Remote work
    Visa sponsorship

    NewsNowGh

    New York, NY
    4 days ago
  • $60 per hour

     ...A leading AI development company is looking for proficient programmers to contribute to cutting-edge AI systems while enjoying the flexibility of remote work. Responsibilities include designing coding problems for AI systems, writing clear code, and evaluating AI-generated... 
    Remote work

    DataAnnotation

    Oklahoma City, OK
    4 days ago
  •  ...A technology company is seeking proficient programmers to contribute to AI development remotely. You will design coding tasks, evaluate AI code, and help shape future technologies while enjoying a flexible schedule. Ideal candidates possess fluency in English and are... 
    Hourly pay
    Remote work
    Flexible hours

    DataAnnotation

    Wisconsin
    1 day ago
  • $86.8k - $198k

     ...Job Title AWS AI Cloud Engineer The Opportunity As an AWS AI Cloud Engineer, you can resolve a problem with a complete end‑to‑end GenAI solution in a fast, agile environment. If you’re looking for the chance to not just develop software, but to create a system that will... 
    Full time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    McLean, VA
    3 days ago
  • $60 per hour

     ...An AI development firm is seeking proficient programmers to contribute to developing cutting-edge AI systems. Enjoy the flexibility of remote work and the ability to set your own schedule while engaging in diverse coding challenges that help refine intelligent systems... 
    Remote work

    DataAnnotation

    Springfield, IL
    2 days ago
  • $60 per hour

     ...A leading AI development company is looking for proficient programmers to join its remote coding team. You'll solve engaging coding problems, write high‑quality code, and provide feedback on AI models. Fluency in English and proficiency in programming languages like Kotlin... 
    Remote work
    Flexible hours

    DataAnnotation

    Helena, MT
    2 days ago
  • $60 per hour

     ...A technology company working on AI is seeking proficient programmers. Work from anywhere with a flexible schedule and earn up to $60/hour. Responsibilities include designing coding problems, writing quality code, evaluating AI-generated code, and contributing feedback... 
    Remote work
    Flexible hours

    DataAnnotation

    Brooklyn, NY
    2 days ago
  • $60 per hour

     ...A dynamic AI development company is seeking proficient programmers to contribute to innovative AI systems. This fully remote position allows flexibility in project selection and work schedule, offering competitive pay that can reach up to $60 USD/hour. Responsibilities... 
    Remote work

    DataAnnotation

    Providence, RI
    3 days ago
  • $150k - $210k

     ...empowers members to perform at a higher level and live longer by using AI to transform continuous physiological data into clear insights...  ...members can act on every day. WHOOP is hiring a Senior AI/ML Engineer to help scale the intelligence layer behind WHOOP’s AI-powered... 
    Full time
    Work at office
    Relocation

    Whoop

    Boston, MA
    12 hours ago
  • $94.2k - $223.5k

     ...a compelling opportunity for a Software Engineer to join their dynamic team. OHDAP combines...  ...strengths within Oracle, including OCI AI/ML, Oracle Cerner healthcare technology,...  ..., OHDA, Oracle Health Applications & Infrastructure, and the mission of the Oracle Healthcare... 
    Full time

    Career-Mover

    Remote
    12 hours ago
  • $110k - $140k

     ...Description Vultr is seeking a highly skilled and experienced AI Platform Engineer to own the strategy and execution for embedding AI into...  ...with hands-on experience deploying LLM inference infrastructure and a genuine passion for accelerating how engineers work.... 
    Work at office
    Immediate start
    Remote work

    Remotive

    New York, NY
    2 days ago
  • Role Description The AI Platform team is seeking an AI Platform Engineer to design and build the foundational systems, APIs, and tooling that empower our...  .... In this role, you will architect the AI-powered infrastructure layer that development and product teams rely on—... 
    Remote work
    Flexible hours

    Remotive

    New York, NY
    3 days ago
  •  ...and more so that they can focus on making great games. About the Role As a Senior AI Platform Engineer in the AI Lab team, you will take ownership of Playamp's AI platform infrastructure and play a key role in enabling teams across the company to build and scale AI-powered... 
    Full time
    Local area

    Plarium

    Poland, NY
    3 days ago
  • $190k - $230k

     ...Exposure Management (CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world's largest community of...  ..., inclusion, respect, and accountability. Senior Software Engineer, AI Platform Location: Seattle, WA; Austin, TX; Boston, MA; Washington... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    Shift work
    1 day per week

    HackerOne

    Washington DC
    3 days ago
  •  ...Mercor is building a groundbreaking AI-native platform and seeking a software engineer to spearhead the end-to-end product delivery. You will work on integrations, real-time analytics dashboards, and support pilot launches through technical execution. The ideal candidate... 
    Remote work

    Mercor Inc

    Arizona City, AZ
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!