AI Infrastructure Engineer
Pragmatike
About the Role
Pragmatike is recruiting on behalf of a fast-scaling, well-funded distributed cloud infrastructure startup building next-generation AI-native cloud services. The company is redefining how compute is delivered by providing GPU-powered infrastructure for AI/ML workloads, secure storage, and high-speed data transfer through a decentralized architecture that significantly reduces environmental impact compared to traditional cloud providers.
We are seeking a AI Infrastructure Engineer with strong experience in production-grade model serving and infrastructure for AI systems. This is a highly technical, hands-on role focused on building scalable, reliable, and efficient ML inference platforms powering real-time AI applications.
You will be responsible for designing and operating the core infrastructure that serves machine learning models at scale. You will work closely with infrastructure, platform, and applied AI teams to ensure high availability, low latency, and cost-efficient inference systems. Strong ownership, production mindset, and experience with distributed GPU systems are essential.
Your Responsibilities
- Build and operate production-grade model serving infrastructure using frameworks such as vLLM , TGI , Triton , or equivalent
- Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models
- Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers
- Optimize GPU utilization , memory efficiency, network throughput, and model artifact storage performance
- Design observability systems for tracking inference latency, throughput, GPU usage , cost metrics, and system health
- Manage model registries and CI/CD pipelines enabling automated and reproducible model deployments
- Own the full lifecycle of ML systems from development through production, including operational support and on-call responsibilities
- Define engineering best practices and contribute to platform scalability in a fast-moving startup environment
Required Qualifications
- 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar infrastructure roles focused on ML systems
- Hands-on experience with model serving frameworks such as vLLM , TGI , Triton , or equivalent
- Strong background in container orchestration and operating GPU-based workloads in production
- Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines
- Proficiency in Python and infrastructure-as-code tools (e.g., Terraform , Helm , or similar)
- Strong understanding of distributed systems, performance tuning, and production reliability engineering
- Ability to effectively use AI coding assistants to accelerate development and debugging workflows
- Ownership mindset with the ability to operate independently in a remote-first environment
Preferred Qualifications
- Experience with ML platforms such as Kubeflow , MLflow , or KubeAI
- Knowledge of GPU scheduling , CUDA/ROCm optimization , or multi-tenant inference systems
- Experience with cost optimization across different GPU types and inference workloads
- Background in early-stage startups or greenfield infrastructure projects
- Proven experience building production systems from scratch rather than maintaining legacy platforms
Why Join Us
- Take ownership of critical infrastructure powering a rapidly scaling AI-native cloud platform
- Build foundational ML inference systems from the ground up in a high-growth, well-funded startup
- Work at the intersection of distributed systems, GPU computing , and sustainable cloud architecture
- Gain deep expertise in next-generation AI infrastructure and large-scale model serving systems
- Influence core engineering decisions and define best practices that will scale with the company.
$30 per hour
...Freelance AI Trainer - Civil Engineering & Python 1 day ago Be among the first 25 applicants This opportunity is only for candidates currently... ...Hydraulic ~ Construction Engineering & Management ~ Infrastructure, Coastal, Earthquake, Sustainable Engineering...SuggestedPart timeFreelanceRemote workFlexible hours- ...Junior Cloud Field Engineer Joinrs is looking for a Junior Cloud Field Engineer to help... ...adopt the latest private cloud infrastructure, Linux, cloud native operations, and open... ...publishes Ubuntu, it provides the platform for AI, IoT, and cloud. Canonical recruits globally...SuggestedWork at officeRemote workWork from home
- ...Role Overview As a Senior Infrastructure Engineer , you'll join Wonderland, Jimdo's internal Platform Engineering team responsible for the cloud... ...platform tooling and automation. Explore and integrate AI tools that improve engineering productivity and operational...SuggestedFull timeWork at officeRemote work
- Una startup tecnologica è alla ricerca di un Data Engineer freelance per collaborare con una realtà digitale. Il candidato ideale avrà esperienza in Python e ambienti cloud e sarà responsabile della progettazione di pipeline dati e dell'integrazione di flussi. Il contratto...SuggestedFreelance
$15 per hour
Summary The Wikimedia Foundation is seeking a Senior Software Engineer to join the team supporting the Wikidata Platform — the structured data backbone of Wikimedia projects and a key part of the global open knowledge ecosystem. You’ll help scale and sustain the Wikidata...SuggestedFull time- ...As a Software Engineer, you would will be part of a 140+ people Tech organisation. Each engineer is part of a squad focussing on a part of the product, reporting to a Team Lead or an Engineering Manager. Each squad is made with all the skills needed to build features...
- Enterprise Architect (Cloud & Infrastructure, TOGAF, Azure, AWS, SaaS, IaaS, PaaS, Active Directory, SSO, IT Strategies, Cloud Computing) in Eagle, Idaho Active Directory, AWS, Azure, Cloud Technology Architecture, SSO, TOGAF Location: Idaho Job Function: Cloud Technology...
- Un'azienda specializzata in data science cerca un Data Scientist per lavorare completamente in remoto dall'Italia. Il candidato ideale ha 3-4 anni di esperienza, forti competenze in machine learning e programmazione avanzata in Python. Offriamo un salario competitivo tra...Remote work
- ...Orbyta Group cerca un Data Engineer per contribuire allo sviluppo di una data platform in ambito cloud. Il candidato lavorerà alla progettazione e all’implementazione di pipeline dati end-to-end, supportando l’operatività e l’ottimizzazione dei processi. La posizione...
- ..., and financial goals. A recommendation engine needs well‑structured feature data. An automation... ...becomes a shared dependency that every AI feature builds on top of. This role owns... ...tells us if our AI is working, and the infrastructure that lets us iterate on data quality...
- ...functional collaboration with quality, production, and local teams. Your Background ~ Bachelor's or Master's degree in Electrical Engineering, Electromechanics, or a related technical discipline. ~3-4 years of experience in a technical engineering, field service, or...Full timeLocal areaRemote work
- ...About the project Hands-on Tech Lead for an AI Companion in an online mahjong game. The client is a social gaming company (web... ...Key design constraints: a valid-action contract with the game engine (the bridge supplies legal moves), win detection , and a 2-second...Contract workTemporary work
- ...implementing, and operating large scale, multi-cloud Kubernetes infrastructure driven by GitOps methodology. You understand that great... ...and implement with respect and kindness across multiple engineering teams. We embrace an empathetic, supportive, and communicative...Full timeRemote work
- ...entrepreneurial culture to unlock full potential by bringing energy to the world. Partner with the best As Junior Control Software Engineer, you will be involved in Control algorithms and software development for Gas Turbine and Steam Turbine driven Compressor and...Remote jobPermanent employmentTemporary workWorldwideFlexible hours
$120k - $160k
...Sr. Embedded Firmware Engineer JOB-10047352 Anticipated Start Date Sept. 1, 2026 Location Houston, TX Type of Employment Direct Hire Employer Info Our client designs and manufactures advanced protection relays and disturbance fault...Full timeLocal area- ...attiva della crescita di xTech , Practice che con oltre 500 professionisti e professioniste altamente specializzati negli ambiti Data & AI, Cloud, Digital Platforms and Solutions e nella creazione di soluzioni end-to-end per supportare le aziende clienti nella...Full timeRemote work
- ...innovator, digital commerce pioneer, and AI-powered software platform operating on... ...capabilities into a data-driven business engine—delivering personalized guidance that eliminates... ...and quality metrics. External Infrastructure Integration: Coordinate technical and...Full timeWork at officeRemote workWork from homeShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!





