AI Infrastructure Engineer
$100k - $150kBright Vision Technologies
AI Infrastructure Engineer- Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: AI Infrastructure Engineer Location: 100% Remote (United States) Position Type: Full-time, Direct W2 Salary Range: $100,000–$150,000 Annually Experience: 6+ years Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. Job Summary We are seeking an AI Performance Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated an impact on production of AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. Key Responsibilities * Profile and optimize end-to-end AI training and inference pipelines for throughput, latency, and cost. * Identify and eliminate bottlenecks across data loading, model compute, communication, and memory. * Implement and tune quantization, sparsity, and pruning strategies to reduce model footprint and accelerate inference. * Optimize distributed training using tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding. * Tune attention implementations using Flash Attention, paged attention, and related techniques. * Implement KV cache optimization, continuous batching, and speculative decoding for LLM serving. * Drive compiler-level optimizations using Triton, XLA, Torch Inductor, or TVM, working with the broader ML framework community to land improvements that translate into measurable end-to-end performance gains. * Optimize data pipelines, sharding strategies, and storage access patterns for high-throughput training. * Build and maintain rigorous benchmark suites and regression frameworks across workloads. * Collaborate with ML and platform engineering teams to embed best practices in standard pipelines. * Drive cost-efficiency improvements through model architecture, hardware selection, and scheduling strategies.
- Evaluate new hardware and software offerings and advise on adoption.
- Document performance tuning playbooks and share findings broadly across
HPC.
- Strong proficiency in Python and C++.
- Hands-on experience optimizing deep learning workloads on modern GPUs.
- Deep understanding of distributed training and inference techniques.
- Experience with profiling tools across CPU, GPU, and distributed systems.
- Familiarity with model compression techniques and their accuracy
- Excellent measurement, debugging, and analytical reasoning skills.
- Strong communication and collaboration skills.
- Experience optimizing LLM inference at production scale.
- Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects.
- Familiarity with custom kernel authoring in Triton or CUTLASS.
- Experience with FinOps for AI workloads.
- Publications or talks on AI systems performance.
$94.5k - $212.5k
...continuously invest in innovative ideas, such as AI‑enabled insights and technology‑powered... ...GitOps structures with FluxCD. Engineer advanced Azure DevOps pipeline patterns... ...security, and consistency. Mentor junior AI Infrastructure Engineers, fostering growth and...SuggestedLocal area- ...AI Infrastructure Engineer At BNY, our culture allows us to run our company better and enables employees' growth and success. As a leading global financial services company at the heart of the global financial system, we influence nearly 20% of the world's investible...SuggestedWork experience placementWorldwideFlexible hours
$170k - $210k
...AI Infrastructure Engineer Utilidata is a fast-growing AI company enabling AI data centers to dynamically orchestrate power and unlock more compute capacity from existing energy infrastructure. For over a decade, we have applied AI to the electric grid — bringing real...SuggestedLocal areaRemote workFlexible hours- ...A leading U.S. technology firm is hiring an AI Infrastructure Engineer for a full-time remote position with H-1B visa sponsorship available. This role involves designing and optimizing AI platforms, deploying Kubernetes and Docker container environments, and enhancing...SuggestedFull timeH1bRemote workVisa sponsorship
- ...Seekr is building the infrastructure that powers the next generation of enterprise AI. As a Senior AI Infrastructure Engineer, you will design, build, and operate the platforms that enable large‑scale training, serving, evaluation, and deployment of foundation models and...SuggestedWork experience placementFlexible hours
$177.5k - $248k
...deeper understanding in healthcare. Our AI-powered platform was purpose-built for medical... ..., PhDs, creatives, technologists, and engineers working together to empower people and... ...in Pittsburgh. The Role As an AI Infrastructure Engineer at Abridge, you’ll play a pivotal...Hourly payFull timeFlexible hours- A leading AI research firm in San Francisco seeks a Staff Infrastructure Engineer to identify and resolve infrastructure bottlenecks and design large-scale systems for AI training. The ideal candidate has over 3 years of experience in infrastructure engineering and strong...
- ...limits of what's possible. As a Lead Software Engineer at JP Morgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an integral part of an... ...software applications and systems. Collaborate with AI teams to translate computational requirements...
- ...technology products. As a Senior Lead Software Engineer at JPMorgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an integral part of an... ...scalable cloud infrastructure platforms optimized for AI and machine learning workloads. Collaborate...For contractors
$147.06k - $191.52k
...architecture, development, and operations of our ML engineering and GenAI systems, enabling scalable and responsible AI solutions across the business. This role requires a deep understanding of ML infrastructure and MLOps, combined with hands-on or architectural experience...Hourly pay- ...A tech company focused on AI is seeking proficient programmers to join their remote team. Contributing to cutting-edge AI development, you will design and solve coding problems, evaluate AI-generated code, and provide valuable feedback. Applicants should have strong programming...Hourly payRemote workFlexible hours
$60 per hour
...A leading AI development company is looking for proficient programmers to join its remote coding team. You'll solve engaging coding problems, write high‑quality code, and provide feedback on AI models. Fluency in English and proficiency in programming languages like Kotlin...Remote workFlexible hours$60 per hour
...A leading AI development company is seeking proficient programmers to join their remote team. This role involves designing and solving coding problems for training AI systems, particularly in Android development, and evaluating AI-generated code. Ideal candidates should...Remote workFlexible hours$60 per hour
...A cutting-edge AI development company is looking for proficient programmers to tackle programming tasks such as developing coding solutions, building apps, and improving intelligent systems. This fully remote position, available to candidates in the US and select countries...Remote work$60 per hour
...A leading AI development company is looking for proficient programmers to contribute to cutting-edge AI systems while enjoying the flexibility of remote work. Responsibilities include designing coding problems for AI systems, writing clear code, and evaluating AI-generated...Remote work$350k
Mirendil, located in San Francisco, is seeking an Infrastructure Engineer to develop the foundational infrastructure for their frontier AI research initiatives. The ideal candidate will work on crucial areas such as Kubernetes, secure execution environments, and scalable...$190k - $270k
AI Chopping Block, Inc. in San Francisco is seeking an AI Infrastructure Engineer to maintain user-facing services and production systems. The role involves building and managing infrastructure with tools like Ansible and Kubernetes, ensuring reliability and scalability...$163k - $237k
Google is seeking a Developer Relations Engineer for our AI Infrastructure team in Seattle. This role involves writing and running AI workloads on Google’s accelerators and creating engaging developer content. Ideal candidates have a strong technical background and a public...- Gravity Engineering Services Pvt Ltd. is seeking an experienced AI Infrastructure Engineer in Seattle, Washington. The role involves architecting and maintaining scalable GPU cluster environments for large-scale LLM training and optimizing distributed training pipelines...
- Job Title Founding AI Infrastructure Engineer Salary Not Disclosed + Equity Company Description Piris Labs (YC W26) is a San Francisco startup raising a ~$20M seed round to build the full‑stack networking layer for AI inference. Led by engineering veterans from Meta...
$60 per hour
A leading AI development firm seeks proficient programmers to join their remote team. You will tackle coding challenges that contribute to AI systems, with a focus on Android development. Ideal candidates are fluent in English and have experience in Kotlin and other programming...Remote jobHourly pay- Litmus, a growth-stage software company powering industrial AI, seeks an IT Infrastructure Engineer in Santa Clara. You’ll manage AI services, on‑prem and cloud infrastructure, and secure access for a fast‑moving production environment. You’ll deploy and maintain VMware...
- Gravity Engineering Services Pvt Ltd. is looking for a Software Engineer specializing in AI infrastructure in Raleigh, North Carolina. The role requires proven experience in software engineering and proficiency in Python and AWS. You will design and deploy AI systems,...
$163k - $237k
Google is seeking a Developer Relations Engineer for our AI Infrastructure team in New York, NY. This role focuses on building and engaging the AI developer community through events and educational resources. The ideal candidate has a technical background, with experience...- NVIDIA is seeking a Senior Systems Software Engineer in Santa Clara, California, to innovate and improve AI infrastructure. This role involves leading performance analysis, enhancing scalability on the Kubernetes stack, and collaborating with teams on workload automation...
- Accenture is seeking a seasoned AI Infrastructure Architect in San Francisco to design and implement scalable AI infrastructure, including accelerated computing clusters, model serving endpoints, and secure governance. You will collaborate across cloud, on‑premises, and...
£50.25k - £58.23k per year
AI Supercomputing Infrastructure Engineer The Bristol Centre for Supercomputing (BriCS) operates the Isambard‑AI National Artificial Intelligence Research Resource, the AI Data facility and the Isambard3 Tier‑2 Supercomputer. Isambard‑AI is the most powerful supercomputer...Full timeContract workMonday to Friday- Jack & Jill in San Francisco is seeking a Founding Engineer to own the technical core of an AI infrastructure platform. You will shape the architecture and build essential components from ground up, surrounding pricing intelligence, demand forecasting, and inventory optimization...
$182k - $242k
CoreWeave is seeking an experienced professional to contribute to building distributed systems and ML infrastructure. The successful candidate will play a pivotal role in designing an optimal research cluster experience, including a Python SDK, while collaborating closely...$60 per hour
...A cutting-edge technology firm seeks proficient programmers to advance AI development. In this fully remote role, you will tackle diverse coding problems, evaluate AI code, and contribute to intelligent systems. The position offers competitive pay of up to $60/hour and...Remote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!
- ai engineer United States
- ai ml engineer United States
- ai research engineer United States
- senior ai engineer United States
- machine learning ai engineer United States
- ai developer United States
- ai prompt engineer United States
- ai engineer remote United States
- principal infrastructure engineer United States
- infrastructure engineering manager United States


