AI Infrastructure Engineer
$170k - $210kKARMAN INC
Karman is a fast-growing NVIDIA-backed AI company enabling AI data centers to dynamically orchestrate power and unlock more compute capacity from existing energy infrastructure. For over a decade, we have applied AI to the electric grid — bringing real-time visibility and power-flow control to complex energy infrastructure. Our Karman platform, built on a custom NVIDIA module, brings that same capability to AI data centers, giving operators a way to better use the power already available to them.
The AI Infrastructure Engineer is responsible for designing, building, and owning the end-to-end infrastructure that serves Utilidata's AI and ML models across edge deployments, cloud environments, and data center integrations. They are also responsible for designing, building, and owning the integration of power data with AI inference software. This is Utilidata's first dedicated role of this kind, and will serve as the foundational function for how the company deploys and operates AI capabilities in production. The role requires deep technical expertise in ML model serving, distributed systems, and GPU infrastructure, with a strong emphasis on reliability, performance, and scalability. This position works cross-functionally with product, engineering, and data science teams and is open to fully remote candidates, with periodic travel expected for company retreats and key on-site engagements.
Responsibilities
- Lead the design and build of Utilidata's AI inference platform — establishing architecture patterns, deployment standards, and operational practices that will scale with the company
- Own end-to-end model serving infrastructure for Utilidata's AI infrastructure (on-prem and datacenter)
- Build and maintain fault-tolerant, high-performance systems for serving AI models at scale, with a focus on low latency, reliability, and cost efficiency
- Collaborate closely with algorithms engineers to integrate AI inference data and configuration with power optimization algorithms
- Optimize GPU utilization and inference performance across our hardware fleet, including NVIDIA accelerators central to Utilidata's edge AI platform
- Establish MLOps best practices including CI/CD pipelines for model deployment, monitoring, and rollback across environments
- Contribute to infrastructure roadmap decisions, including build vs. buy tradeoffs, tooling selection, and platform evolution as the team grows
Minimum Qualifications
- 5+ years of software engineering experience with a strong focus on AI infrastructure, backend systems, or distributed systems
- Hands-on experience with AI model serving frameworks (e.g., vLLM, SGLang, Triton, TensorRT, TorchServe, or similar)
- Understanding of container orchestration and cluster management (Kubernetes, Docker)
- Experience deploying and operating infrastructure across both datacenter and on-prem environments
- Strong knowledge of GPU workloads and the tradeoffs that come with them — you understand how inference differs from training, and why it matters
- Proficiency in Python; C++, CUDA, Go, Rust a plus
- Excellent communication skills and comfort working cross-functionally in a lean, fast-moving environment
- Willingness to travel up to 10% of time
Enhanced Qualifications (Nice to Have)
- Dynamo experience a plus
- Experience with edge AI deployments or constrained compute environments
- Familiarity with infrastructure as code (Terraform, Helm)
- Experience with observability platforms (Datadog, Prometheus, Grafana)
- Background in energy, utilities, or industrial IoT
- Contributions to open-source ML infrastructure projects
Salary Range
$170,000 to $210,000 base compensation depending on experience plus stock options. Salary will be commensurate with an individual's skills, training, years of experience, and in line with internal compensation bands.
Location
This position can be performed remotely from anywhere in the United States.
Our Commitments
Karman values the diversity of our team. We provide equal employment opportunities without regard to race, color, religion, creed, sex, gender, sexual orientation, gender identity or expression, national origin, age, physical disability, mental disability, medical condition, pregnancy or childbirth, sexual orientation, genetics, genetic information, marital status, or status as a covered veteran or any other basis protected by applicable federal, state and local laws.
We are committed to:
- Creating a diverse and inclusive workplace that is welcoming, supportive, affirming and respectful
- Empowering employees to solve problems and work together to make a difference
- Providing mentorship and growth opportunities as part of a collaborative team
- A flexible work environment with flexible paid time off
- Competitive compensation and benefits, including health, dental, vision, and employer-match 401k
$100k - $150k
...Engineering San Francisco or remote · Full-time Automate the cloud infrastructure and operational workflows behind a fast-moving AI platform, and grow into the part of the stack that suits you About Coral Bricks Our mission is to make frontier intelligence affordable...SuggestedFull timeInternshipRemote workFlexible hours$140k - $165k
...As vCluster’s AI Infrastructure Specialist, you will work directly with customers at the earliest and most critical stage of their journey... ...role exists to make that happen. As an AI Infrastructure Engineer, your role will include: Lead Technical Deployments:...SuggestedRemote workFlexible hours$141k - $177k
...and quota management behind an AWS/Okta-SSO'd gateway. Make AI spend and usage visible - build the Bedrock-to-Snowflake pipeline... .... What you will bring with you ~6+ years in software engineering, platform engineering, or DevOps, including ownership of...SuggestedFull timeRemote work- ...Chime is seeking a Senior Software Engineer to join the AI Enablement team, building internal platforms, primitives, and guardrails to make AI a trusted co-pilot for employees. You will work on Archimedes, Bosun, and easyRAG, powering internal assistants, developer workflows...Suggested
- ...Autodesk is seeking a Principal Software Developer to design and implement AI features for the platform. You will collaborate with product managers, developers and operations to shape the platform roadmap and deliver robust AI-enabled software. You will drive API design...Suggested
$184k - $215k
...This role is responsible for designing, implementing, and maintaining the cloud infrastructure and CI/CD systems that power X-energy's AI-native application platform (APEX). The DevOps Engineer will work with the AI and Application Development team to accelerate...Full timeWork at officeMonday to Friday- ...sustainable, resilient, and diversified housing solutions. Summary of Position: Reporting to the Senior Director, the Senior AI Platform Engineer builds, deploys, secures, monitors, and maintains Quarterra's enterprise AI and automation environments. This senior...Work at officeLocal areaRemote work
- ...About this role We are hiring a Senior AI Platform Engineer. The full role brief is on its way. The confirmed facts are below, and our people team shares the complete spec at screening. /01 Seniority Senior /02 Workplace Remote /03 Contract length 12 months...Contract workRemote work
$139.7k - $232.9k
...Principal AI Platform Engineer page is loaded## Principal AI Platform Engineerremote type: Remote Eligible, USAlocations: Remote, USAtime type: Full timeposted on: Posted Yesterdayjob requisition id: R90667# Role SummaryThe Principal AI Platform Engineer serves as a senior...Work experience placementRemote work$133k - $180k
### Job Details#### Job Title:Principal AI Platform Engineer#### Location:,, ,#### Company:#### Industry Sector:#### Industry Type:#### Career Type:#### Job Type:Full Time#### Minimum Years Experience Required:N/A#### Salary:$133,000 - $180,000 USD$133,000.00-$180,000....Full timeWork experience placementWork at officeWork visa- Thinking Machines Lab Inc. in San Francisco, California is seeking an infrastructure research engineer to design, optimize, and scale the systems that power large AI models. Your work will make inference faster, more cost-effective, more reliable, and more reproducible...
$184k - $287.5k
AI Infrastructure Engineers at NVIDIA build the systems, tooling, and data infrastructure that enable operation of our GPU cloud services. We are enabling engineering teams to innovate while proactively identifying, tracking, and mitigating risks across the entire technical...- # AI Infrastructure Engineer, pAGIRemoteSan FranciscoAll jobsRemote jobs## SkillsDistributed Systems EngineeringPerformance OptimizationML InfrastructureInference SystemsGPU PerformanceInfrastructure ToolingCapacity ManagementHealth MonitoringCompute SchedulingResource...
- ## JobbeskrivelseWe Are:The Global AI Infrastructure team is at the center of enabling infrastructure reinvention for the next era of digital... ...(BCM), NGC, NCCL, NVLink, and CUDA along with LLM inference engines (TensorRT-LLM), production serving frameworks (vLLM, SGLang),...Work experience placementLive inWork at officeLocal area
- Fluidstack in New York, NY is seeking an experienced engineer to help build civilization-scale AI infrastructure. You will ship production-grade code using Go, Python or TypeScript and contribute to building scalable AI workloads that drive rapid decision-making. You will...
- SpaceX is seeking a Sr. Network Engineer for AI Infra (Starshield) in Washington, DC to lead network design and deployments for high-security... ...and software teams, and help define secure, scalable AI infrastructure for national security missions. #J-18808-Ljbffr SPACE...
- ...Inc. in the San Francisco Bay Area is seeking a Senior Software Engineer to join the Mk8s team. You will design and build scalable control... ..., Kubernetes operators, and GPU-aware orchestration that powers AI workloads across a distributed platform. As part of the...
- Nasdaq, Inc. is seeking a Software Engineer to contribute to the Nasdaq Questionnaires platform, a SaaS solution for governance, compliance, and board-related workflows for corporate clients worldwide. You will design and build APIs and backend services on a modern cloud...Worldwide
- xFigura in Boston is seeking a Founding Engineer to work directly with the CEO and Head of R&D to build the core systems behind an AI-native design platform used by architecture firms, designers, and universities. The role focuses on foundational work: orchestrating generative...
- Zizy Inc. is seeking a Backend Engineer to own and maintain the infrastructure powering our browsing capabilities and AI workflows. You will design, deploy, and operate backend services, manage Kubernetes configurations, and collaborate with a research team to support model...
- Lightning AI, the company behind PyTorch Lightning, is seeking a customer-facing Machine Learning Solutions Engineer. You will translate complex customer requirements into scalable AI/ML solutions on the Lightning AI platform, leading discovery, demos, POCs, and post-sales...
- Zof AI seeks an experienced MLOps Engineer to own the platform for AI systems, spanning model serving, scaling, and reliability. You will push the... ...San Francisco, you will design and operate the backend infrastructure, partner with software engineers, and drive incident practices...
- Intone Inc. is seeking a Senior AI Platform Engineer for a 12-month contract based in New York City. You will design, build, and secure AI platform infrastructure in a regulated environment, bridging technical implementation with security and compliance requirements. You...Contract work
- MeridianLink is seeking an AI Engineer for the Trust & Explainability layer of the AI Platform. You will implement tracing across multi-agent workflows, surface explanations, and integrate open-source observability tools to help lenders and borrowers understand model behavior...
- ByteDance's Volcano Ark training platform is seeking engineers to advance end-to-end large model post-training (SFT, RL) on a serverless... ...stack. Join the Data AML team driving scalable, secure ML infrastructure. You will design elastic, multi-tenant training across data...
- DRW, a diversified trading firm, is seeking an AI Inference Platform Engineer to build, operate, and optimize systems that serve large language and multimodal models across the firm. You will own the serving platform end-to-end, onboarding models, tuning runtimes and configurations...
- InterImage, Inc. seeks engineers to take Generative AI concepts from proof of concept to production. You will design, develop, deploy, and continuously improve AI-powered applications for mission-critical customers in Azure cloud environments. The role emphasizes building...
- Quovy builds an AI operating layer for specialty insurance—submission intake, underwriting, quote & bind, claims, and servicing—deployed into live carrier and MGA operations. Forward Deployed Engineers embed with carriers and MGAs, learn underwriters' workflows, and design...
- Komatsu America Corp. in Warrendale, PA is seeking a software engineer to contribute to analytics and AI-driven software platforms, developing for Snowflake, Palantir, and Databricks within the Azure cloud. The role focuses on software development and coaching newer engineers...
- Capital One is seeking an AI Engineer 5 to advance responsible AI systems and scalable AI infrastructure in the IFX team. You will design and deploy AI components, including foundation models, LLM inference, agents, and vector search, while ensuring governance and cost...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!
- senior ai engineer Eastern, KY
- ai developer Eastern, KY
- ai engineer Eastern, KY
- ai ml engineer Eastern, KY
- ai engineer remote Eastern, KY
- machine learning ai engineer Eastern, KY
- ai prompt engineer Eastern, KY
- remote infrastructure engineer Eastern, KY
- infrastructure engineer Eastern, KY
- principal infrastructure engineer Eastern, KY

