Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Infrastructure Engineer

Full-time

Pragmatike

About the Role

Pragmatike is recruiting on behalf of a fast-scaling, well-funded distributed cloud infrastructure startup building next-generation AI-native cloud services. The company is redefining how compute is delivered by providing GPU-powered infrastructure for AI/ML workloads, secure storage, and high-speed data transfer through a decentralized architecture that significantly reduces environmental impact compared to traditional cloud providers.

We are seeking a AI Infrastructure Engineer with strong experience in production-grade model serving and infrastructure for AI systems. This is a highly technical, hands-on role focused on building scalable, reliable, and efficient ML inference platforms powering real-time AI applications.

You will be responsible for designing and operating the core infrastructure that serves machine learning models at scale. You will work closely with infrastructure, platform, and applied AI teams to ensure high availability, low latency, and cost-efficient inference systems. Strong ownership, production mindset, and experience with distributed GPU systems are essential.

Your Responsibilities

  • Build and operate production-grade model serving infrastructure using frameworks such as vLLM , TGI , Triton , or equivalent
  • Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models
  • Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers
  • Optimize GPU utilization , memory efficiency, network throughput, and model artifact storage performance
  • Design observability systems for tracking inference latency, throughput, GPU usage , cost metrics, and system health
  • Manage model registries and CI/CD pipelines enabling automated and reproducible model deployments
  • Own the full lifecycle of ML systems from development through production, including operational support and on-call responsibilities
  • Define engineering best practices and contribute to platform scalability in a fast-moving startup environment

Required Qualifications

  • 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar infrastructure roles focused on ML systems
  • Hands-on experience with model serving frameworks such as vLLM , TGI , Triton , or equivalent
  • Strong background in container orchestration and operating GPU-based workloads in production
  • Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines
  • Proficiency in Python and infrastructure-as-code tools (e.g., Terraform , Helm , or similar)
  • Strong understanding of distributed systems, performance tuning, and production reliability engineering
  • Ability to effectively use AI coding assistants to accelerate development and debugging workflows
  • Ownership mindset with the ability to operate independently in a remote-first environment

Preferred Qualifications

  • Experience with ML platforms such as Kubeflow , MLflow , or KubeAI
  • Knowledge of GPU scheduling , CUDA/ROCm optimization , or multi-tenant inference systems
  • Experience with cost optimization across different GPU types and inference workloads
  • Background in early-stage startups or greenfield infrastructure projects
  • Proven experience building production systems from scratch rather than maintaining legacy platforms

Why Join Us

  • Take ownership of critical infrastructure powering a rapidly scaling AI-native cloud platform
  • Build foundational ML inference systems from the ground up in a high-growth, well-funded startup
  • Work at the intersection of distributed systems, GPU computing , and sustainable cloud architecture
  • Gain deep expertise in next-generation AI infrastructure and large-scale model serving systems
  • Influence core engineering decisions and define best practices that will scale with the company.
Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the AI Infrastructure Engineer in United Arab Emirates vacancy
  •  ...We are seeking a highly skilled Software Engineer who is eager to contribute their expertise. The ideal candidate should have a strong background in software development, a passion for tackling complex challenges, possess a long-term vision, and the ability to work seamlessly... 
    Suggested
    Full time
    Relocation

    Infinite Field

    United Arab Emirates
    5 days ago
  •  ...continuing to expand its global presence while building the infrastructure for the next generation of financial services. Responsibilities...  ...development, testing, and deployment; deliver high-quality engineering work 4. Write technical and system documentation to ensure... 
    Suggested
    Full time

    Bybit

    United Arab Emirates
    7 days ago
  •  ...continuing to expand its global presence while building the infrastructure for the next generation of financial services. Job responsibilities...  ...Proficient in NoSQL cache, message queue, search engines, such as Redis, Kafka, Elasticsearch, etc Skilled in system... 
    Suggested
    Full time

    Bybit

    United Arab Emirates
    7 days ago
  •  ...About ElevenLabs ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in...  ...builders doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to work... 
    Suggested
    Full time
    Immediate start

    Valor Capital Group

    United Arab Emirates
    6 days ago
  • $80k - $128k

     ...implementing CI/CD pipelines, automating infrastructure, integrating enterprise applications,...  ...Power Automate flows. Strong software engineering experience with Python, including OOP, testing...  ..., Gotham, or Maven. Experience with AI/ML workflows. Strong understanding of... 
    Suggested
    Contract work
    Currently hiring
    Shift work

    Peraton

    United Arab Emirates
    6 hours ago
  •  ...industry, continuing to expand its global presence while building the infrastructure for the next generation of financial services. Job Summary:...  ...of common vulnerabilities. 8. Able to proficiently use AI tools to enhance security work efficiency. 9. Highly proactive... 
    Full time

    Bybit

    United Arab Emirates
    6 days ago
  •  ...such as ride-hailing and last-mile delivery. Building on this infrastructure, we are now introducing financial services to help our users...  ...infusing social values. About the role As a Backend Staff Engineer, you will be responsible for driving the technical standards... 
    Remote job

    Yassir

    United Arab Emirates
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!