Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer, Platform Infrastructure

$165k - $225k

Moonlite AI

Senior Software Engineer, Platform Infrastructure

Moonlite delivers high-performance AI infrastructure for organizations running intensive computational research, large-scale model training, and demanding data processing workloads. We provide infrastructure deployed in our facilities or co-located in yours, delivering flexible on-demand or reserved compute that feels like an extension of your existing data center. Our team of AI infrastructure specialists combines bare-metal performance with cloud-native operational simplicity, enabling research teams and enterprises to deploy demanding AI workloads with enterprise-grade reliability and compliance.

Your Role:

You will be foundational to building the comprehensive infrastructure platform that bridges our physical infrastructure – bare-metal servers, GPU clusters, high-performance storage, and networking fabric – with the systems our customers depend on for large-scale computation, inference, simulations, and training. Working closely with product, your platform team members, and infrastructure specialists, you'll design and implement the orchestration layer, APIs, and automation framework that make thousands of servers, petabytes of storage, and high-speed networks feel like a unified, programmable platform.

Job Responsibilities
  • Infrastructure Abstraction Layer: Design and build systems that bridge physical infrastructure (bare-metal servers, storage clusters, network fabric) with customer-facing services, enabling programmatic management of compute, networking, and storage at scale.
  • Research Cluster Provisioning: Design and implement systems for provisioning and managing research computing environments including Kubernetes and SLURM clusters, enabling automated deployment, resource scheduling, and workload orchestration for distributed AI training and HPC workloads.
  • Platform Orchestration: Implement comprehensive orchestration systems that coordinate across compute, storage and networking to deliver unified experience for complex research workloads.
  • Network Automation & Placement: Design and build network provisioning automation including intelligent VM placement decisions for optimal network topology, automated VLAN and subnet configuration, and software-designed networking orchestration for high-performance interconnects.
  • Enterprise APIs & SDKs: Develop robust APIs and SDKs that enable researchers and engineering teams to programmatically provision and manage infrastructure resources across all platform domains.
  • Observability & Telemetry: Implement comprehensive observability, telemetry, and logging systems that provide visibility into infrastructure health, performance, and utilization across the infrastructure footprint.
  • Performance Engineering: Build and optimize platform services that deliver consistent high-throughput low-latency networking for demand research applications and data-intensive workloads.
  • Cross-Team Collaboration: Work closely with engineering, infrastructure, and product to define requirements, drive infrastructure-product-rollouts, and improve resource lifecycle management.
  • Compliance & Security: Implement platform-wide compliance and security features supporting SOC 2, ISO 27001, and enterprise regulatory requirements including comprehensive audit logging, access controls, and data residency management.
Requirements
  • Experience: 5+ years in software engineering with a proven track record of infrastructure platforms, distributed systems, or cloud platforms for production environments.
  • Kubernetes & Container Orchestration: Strong familiarity with Kubernetes architecture, container orchestration concepts, and experience deploying workloads in Kubernetes environments. Understanding of pods, deployments, services, and basic Kubernetes operations.
  • Infrastructure Systems: Strong understanding of infrastructure fundamentals including compute orchestration, storage systems, networking technologies, and how they integrate together to deliver complete platform experiences.
  • Programming Skills: Experience with systems programming languages (Go, C/C++, Rust, Python) for performance-critical components is a strong plus.
  • Linux Production Experience: Strong experience with linux in production environments, including systems administration, performance tuning, and troubleshooting.
  • Bare-Metal & Virtualization: Deep knowledge of bare-metal infrastructure, provisioning systems, out-of-band management, and virtualization technologies (KVM, Kubernetes, etc).
  • API & Platform Design: Proven experience designing and building APIs, SDKs, and automation frameworks that enable programmatic infrastructure management.
  • Cloud Platform Knowledge: Strong familiarity with cloud environments (AWS, GCP, Azure) and understanding of how to translate cloud-native patterns to bare-metal infrastructure.
  • Infrastructure Automation: Experience with Infrastructure-as-code tools (Terraform, Ansible) and building automated deployment pipelines.
  • Problem Solving & Autonomy: Self-starter who can navigate ambiguity, balance pragmatic shipping with good long-term architecture, and independently drive complex technical initiatives.
  • Communication Skills: Strong written and verbal communication skills, including ability to write clear technical communication and collaborate across teams.
  • Commitment to Growth: Growth mindset with continuous focus on learning and professional development.
Preferred Qualifications
  • Background provisioning or managing research computing environments (Kubernetes, SLURM, or HPC clusters)
  • Experience building internal platforms, infrastructure-as-a-service, or developer tooling
  • Background with GPU computing platforms and AI/ML infrastructure requirements
  • Knowledge of high-performance networking technologies (InfiniBand, RDMA, SR-IOV)
  • Experience with observability and monitoring platforms (Prometheus, Grafana, ELK stack)
  • Familiarity with both cloud-native and bare-metal infrastructure deployment models
  • Understanding of enterprise compliance requirements and security best practices
  • Extra points for experience with financial services technology infrastructure and understanding of trading system requirements
Key Technologies
  • Go, Python, Kubernetes, Docker, Terraform, Ansible, Linux, Networking (BGP, VXLAN), Storage Systems, FastAPI, PostgreSQL, Redis, NVIDIA GPU Technologies, InfiniBand
Why Moonlite
  • Build Next-Generation Infrastructure: Your work will create the platform foundation that enables financial institutions to harness AI capabilities previously impossible with traditional infrastructure.
  • Hands-On Ownership: As an early engineer, you'll have end-to-end ownership of projects and the autonomy to influence our product and technology direction.
  • Shape Industry Standards: Contribute to defining how enterprise AI infrastructure should work for the most demanding regulated environments.
  • Collaborate with Experts: Work alongside seasoned engineers and industry professionals passionate about high-performance computing, innovation, and problem-solving.
  • Start-Up Agility with Industry Impact: Enjoy the dynamic, fast-paced environment of a startup while making an immediate impact in an evolving and critical technology space.

We offer a competitive total compensation package combining a competitive base salary, startup equity, and industry-leading benefits. The total compensation range for this role is $165,000 – $225,000, which includes both base salary and equity. Actual compensation will be determined based on experience, skills, and market alignment. We provide generous benefits, including a 6% 401(k) match, fully covered health insurance premiums, and other comprehensive offerings to support your well-being and success as we grow together.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer, Platform Infrastructure in United States vacancy
  •  ...Client is specifically looking for engineers who want to build systems that power...  ...just the products themselves. As a Senior Infrastructure Engineer, you will design and build the internal platform that enables the teams to ship software with speed, reliability, and... 
    Senior
    Software

    Siza Buso Consulting

    New York, NY
    1 day ago
  • $125k - $135k

     ...Infrastructure Engineer The Information Technology (IT) Alliance plays an important role throughout the HelloFresh business – from delivering the best hardware and software service for each employee to managing the advanced technology in the fulfillment centres and... 
    Senior
    Software
    Local area

    HelloFresh

    Newark, NJ
    3 days ago
  •  ...open‑source vector search engine powering the next generation...  ...’re building the retrieval infrastructure layer for modern AI. Recently...  ...the future of AI. As a Senior Software Engineer on the Cloud Operations...  ...at the intersection of platform engineering and reliability... 
    Senior
    Software
    Remote work
    Flexible hours

    Qdrant

    New York, NY
    3 days ago
  •  ...Hiring: Senior Software Engineer – Enterprise AI (Platform & Infrastructure) Palo Alto, CA (Hybrid – No Remote Option) Contract: 12+ Months (Possible Extension) We are looking for an experienced Backend/Infrastructure Engineer to join an Enterprise AI Platform team... 
    Senior
    Software
    Contract work
    Remote work

    Quebec Solution Inc

    Palo Alto, CA
    20 hours ago
  •  ...AfterQuery is seeking a Senior Software Engineer to design and build core infrastructure for its data generation platforms. Candidates should have 6-10 years of experience and a strong background in scalable system architecture. This role involves mentoring junior engineers... 
    Senior
    Software

    AfterQuery

    New York, NY
    20 hours ago
  • $174k - $252k

    Google Inc. is seeking a Senior Software Engineer for Infrastructure within Google Cloud to develop innovative technologies that enhance user interaction. The ideal candidate will possess a Bachelor's degree along with substantial hands-on experience in software development... 
    Senior
    Software

    Google Inc.

    San Francisco, CA
    4 days ago
  • $100k - $200k

     ...Senior Infrastructure Engineer (DevOps / Platform) Location Remote – United States (Occasional travel for team offsites 2–3 times per year) Compensation...  ...across infrastructure systems Requirements 5–8 years of software engineering experience 3+ years in a dedicated... 
    Senior
    Software
    Remote work

    MAP SSG Inc

    New York, NY
    3 days ago
  •  ...A leading technology firm in the United States seeks a Senior/Staff Software Engineer for the Data Infrastructure Group. This role involves contributing to and advancing the company's data infrastructure, building tools, and working collaboratively with various teams... 
    Senior
    Software

    The Resume Database

    New York, NY
    3 days ago
  • A cutting-edge tech company in California seeks a Senior Software Engineer specializing in Infrastructure. You'll design and maintain AWS-based production systems, manage Kubernetes clusters, and enhance observability stacks. This hands-on role requires 5+ years of experience... 
    Senior
    Software

    Ironflow AI

    San Diego, CA
    20 hours ago
  • $190.5k - $230k

     ...nTop is pioneering the future of engineering design with our advanced software that pushes the boundaries of...  ...how aircraft get designed. Our platform collapses months of configuration...  ...Role We are looking for a Senior Infrastructure Engineer (5-10yrs experience) to... 
    Senior
    Software
    Local area

    nTop

    New York, NY
    5 days ago
  • $136k - $200k

    A leading tech company based in San Jose is seeking a Software Engineer III for their Google TV division. The role requires expertise in software development and large-scale infrastructure. The ideal candidate will manage project priorities and have experience in distributed... 
    Senior
    Software

    Google Inc.

    San Jose, CA
    2 days ago
  • Bandwidth Recruitment in Raleigh, NC, is looking for a Sr. Software Developer (Infrastructure) to enhance AI integration and developer tooling. This role demands 5+ years in web services, strong skills in Python or Go, and experience with cloud infrastructure like AWS.... 
    Senior
    Software

    Bandwidth Recruitment

    Raleigh, NC
    20 hours ago
  • $148.5k - $313.7k

     ...efforts. Job Category Software Engineering Job Details About...  ...Note: By applying to the Senior / Lead / Principal Software...  ...and building foundational infrastructure that enables great customer...  ...on backend systems, cloud platforms, and infrastructure. Languages... 
    Senior
    Software
    Work experience placement
    Remote work

    Salesforce

    United States
    2 days ago
  • $156k - $211k

     ...Senior Software Engineer, Infrastructure Remote within U.S. or Remote within Ontario Afresh is the leading AI company in fresh food—partnering with grocers...  ...record-breaking 70% growth in 2025, we’ve expanded our platform to cover all fresh departments, launched our full store... 
    Senior
    Software
    Remote work

    Afresh

    New York, NY
    3 days ago
  •  ...Engineered to outperform, Teraswitch is on a mission...  ...provide high-performance infrastructure services for critical...  ..., storage, and platform infrastructure that powers...  ...operations. This senior/staff-level role will...  ...with and support the Software team and other... 
    Senior
    Software
    Full time

    TeraSwitch Inc

    Pittsburgh, PA
    20 hours ago
  • A leading technology company is seeking a Software Engineer to develop next-generation technologies that change how billions of users connect...  ...Science or related fields. Join a dynamic team at the forefront of AI and infrastructure innovation. #J-18808-Ljbffr Google Inc.
    Senior
    Software

    Google Inc.

    Sunnyvale, CA
    3 days ago
  •  ..., we’re transforming the developer platform to offer a unified developer experience...  ...us achieve that, we’re seeking a Senior Staff Software Engineer for a new key role that will drive...  ...Platform sits on top of our Infrastructure Platform, and its APIs must provide... 
    Senior
    Software
    Contract work

    GrabJobs

    San Jose, CA
    1 day ago
  • $196k - $220.5k

     ...thing that nearly everyone does on our platform: play video games. Over 90% of our...  ...playing games. Our Platform Infrastructure teams are responsible for building and...  ...reliable, efficient, and scalable. As a Senior Software Engineer on these teams, you will... 
    Senior
    Software
    Full time
    Relocation
    Relocation package

    Discord

    San Francisco, CA
    3 days ago
  • $119.8k - $234.7k

     ...building a large-scale, productized data platform that powers critical insights and...  ...across Azure-based services for AI Infrastructure. This platform will process...  ...scalability, and long-term evolution. As a Senior Software Engineer - Data Platform, AI Framework you... 
    Senior
    Software
    Ongoing contract
    Local area

    Microsoft Corporation

    Redmond, WA
    20 hours ago
  •  ...Software Engineer, AI/ML (AI Infrastructure & Platform) Role: Software Engineer, AI/ML (AI Infrastructure & Platform) Location: Hybrid, NYC About Us Wealth.com is the industry's leading estate planning platform, empowering more than 1,000 wealth management... 
    Senior
    Software
    Temporary work
    Flexible hours

    Wealth LLC

    New York, NY
    20 hours ago
  • $216k - $270k

     ...As a Software Engineer on the Machine Learning Infrastructure team, you will build the "Operating System" for our large-scale GPU clusters. You will architect a high-performance training platform that handles the immense complexity of multi-thousand GPU workloads, ensuring... 
    Senior
    Software
    Full time

    Scale AI

    Seattle, WA
    4 days ago
  • $216k - $270k

     ...As a Software Engineer on the ML Infrastructure team, you will design and build platforms for scalable, reliable, and efficient serving of LLMs. Our platform powers cutting-edge research and production systems, supporting both internal and external use cases across various... 
    Senior
    Software
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $130k - $280k

     ...integrated, privacy-sensitive AI-powered platform that includes solutions for video...  .... About the Role As a Senior Software Engineer on this team, you will help...  ...projects targeted at making Verkada's infrastructure the most reliable and cost efficient... 
    Senior
    Software
    Hourly pay
    Full time
    Work at office
    Work visa
    Flexible hours
    Shift work

    Verkada

    San Mateo, CA
    6 days ago
  • $155.42k - $395.9k

     ...About the Team: The ML Inference Platform is part of the AV ML Infrastructure organization. Our team owns...  ...About the Role: We are seeking a Senior ML Infrastructure engineer to help build and scale robust...  ...core platform backend software components. Collaborate with ML... 
    Senior
    Software
    Remote work
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Austin, TX
    20 hours ago
  • $174k - $252k

    Senior Software Engineer, Performance, Platforms Infrastructure Engineering Google, Sunnyvale, CA, USA Bachelor’s degree or equivalent practical experience. 5 years of experience programming in C++ or Python. 3 years of experience testing, maintaining, or launching software... 
    Senior
    Software
    Full time
    Worldwide

    Google Inc.

    Sunnyvale, CA
    2 days ago
  • $272k - $408k

     ...digital future. The Sr. Director of Engineering for Identity Security Platform Infrastructure will lead the organization...  ...develop leadership depth by coaching senior leaders and proactively shaping...  ...for Experience: 10+ years of software engineering experience, including... 
    Senior
    Software
    Currently hiring
    Local area
    Immediate start
    Remote work
    Work from home

    1Password

    New York, NY
    3 days ago
  • The Opportunity ModernFi is hiring a Senior/Staff Platform & Infrastructure Engineer to join a talented and collaborative team building the foundation of...  ...opportunity to define how we run our infrastructure, release software, and approach security as we scale from serving... 
    Senior
    Software
    Full time

    Modernfi

    New York, NY
    20 hours ago
  • Job Recommendation When you upload your resume, we provide job recommendations to you. Please confirm you have read and understand how your data may be processed pursuant to the Microsoft Data Privacy Notice and Transparency FAQ.
    Senior
    Software

    Microsoft Corporation

    Mountain View, CA
    20 hours ago
  • $280k - $380k

     ...Roku is the #1 TV streaming platform in the U.S., Canada, and...  ...reliability and automation, engineering systems that perform under...  ...can, and turning complex infrastructure into reliable, well‑documented...  ...Reliability Engineering) Senior Software Engineer to join our dynamic... 
    Senior
    Software
    Full time
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    20 hours ago
  • An AI automation platform in San Francisco is seeking a Software Engineer to build and scale distributed systems for complex IT workflows. This role involves writing Terraform modules, supporting self-hosted deployments, and ensuring the reliability of production systems... 
    Senior
    Software

    Serval

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer, Platform Infrastructure. Be the first to apply!