Cloud Platforms and Infrastructure Engineer, TPU/GPU
$152k - $221kPropose solution architectures and manage the deployment of cloud networking solutions according to customer requirements and implementation best practices.Design and implement cloud-based technical architectures, migration approaches, and application optimizations that enable business objectives in collaboration with customers.Work with internal Specialists, Product, and Engineering teams to package approaches, best practices, and lessons learned into thought leadership, methodologies, and published assets.Interact with sales, partners, and customer technical stakeholders to manage project scope, priorities, deliverables, risks and issues, and timelines.Travel up to 30% for in-region for meetings, technical reviews, and onsite delivery activities.Minimum qualifications:Bachelor's degree in Computer Science or equivalent practical experience.6 years of experience automating infrastructure provisioning, Developer Operations (DevOps), continuous integration, or delivery utilizing Kubernetes and Linux-based systems. 3 years of experience in project management and technical solution delivery.Experience coding in one or more general purpose languages (e.g., Python, Java, Go, C or C++) including data structures, algorithms, software design, Linux environments and Kubernetes orchestration.Experience working with Cloud Providers such as Google Cloud Platform (GCP).Ability to travel 30% of the time, as needed, for client engagements.Preferred qualifications:Experience with third-party networking (e.g., PANW, Fortinet, VMWare) and design, including redundancy and load balancing.Experience troubleshooting networking protocols including TCP/IP, Hypertext Transfer Protocol, and Border Gateway Protocol (BGP).Experience in customer-facing migration, including service discovery, assessment, planning, execution, and operations.Experience with standard IT security practices, including IAM, data protection, encryption, and certificate/key management.Experience running AI/ML training and inference workloads on GPU/TPU using frameworks such as PyTorch, JAX, TensorFlow, or Slurm.Knowledge of containerization and orchestration technologies, including Google Kubernetes Engine (GKE) and related cloud-native services.The Google Cloud Consulting Professional Services team guides customers through the moments that matter most in their cloud journey to help businesses thrive. We help customers transform and evolve their business through the use of Google’s global network, web-scale data centers, and software infrastructure. As part of an innovative team in this rapidly growing business, you will help shape the future of businesses of all sizes and use technology to connect with customers, employees, and partners.As a Cloud Platform and Infrastructure Engineer, you will provide technical guidance to customers adopting Google Cloud Platform (GCP) services, including providing best practices on secure foundational cloud implementations, automated provisioning of infrastructure and applications, cloud-ready application architectures, and more. You will also provide guidance in ensuring that customers receive the best of what GCP can offer and have the best experience in migrating, building, modernizing, and maintaining applications in GCP. Additionally, you will work with Product Management and Product Engineering to drive excellence in Google Cloud products and features.Google Cloud accelerates every organization’s ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google’s cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $152000 - $221000 (USD) + 15% bonus target + equity + benefitsLearn more about benefits at Google.Applicants in San Francisco: Qualified applications with arrest or conviction records will be considered for employment in accordance with the San Francisco Fair Chance Ordinance for Employers and the California Fair Chance Act.The application window will be open until at least September 24, 2026. This opportunity will remain online based on business needs which may be before or after the specified date.Note: By applying to this position you will have an opportunity to share your preferred working location from the following: San Francisco, CA, USA; Atlanta, GA, USA; Austin, TX, USA; Boulder, CO, USA; Chicago, IL, USA; Addison, TX, USA; Miami, FL, USA; Sunnyvale, CA, USA.Bachelor's degree in Computer Science or equivalent practical experience.6 years of experience automating infrastructure provisioning, Developer Operations (DevOps), continuous integration, or delivery utilizing Kubernetes and Linux-based systems. 3 years of experience in project management and technical solution delivery.Experience coding in one or more general purpose languages (e.g., Python, Java, Go, C or C++) including data structures, algorithms, software design, Linux environments and Kubernetes orchestration.Experience working with Cloud Providers such as Google Cloud Platform (GCP).Ability to travel 30% of the time, as needed, for client engagements.
- ...Lambda's Superintelligence Cloud Lambda, The... ...a leader in AI cloud infrastructure serving tens of thousands... ...superintelligence. One person, one GPU. If you'd like to... ...hardware and network engineers building one of the... ...network, hardware, platform and product teams...PlatformCloudLocal areaFlexible hours
- Infrastructure Engineer We're looking for an Infrastructure Engineer to architect... ...the stack to make sure the platform can support whatever we're building... ...Design, build, and own the cloud infrastructure (AWS/GCP)... ...serving and performance of GPU-intensive workloads. Drive...PlatformCloudFull timeWork at office
- ...Principal Network Engineer Houston; New York; San Francisco; Seattle Nscale is the GPU cloud engineered for AI. We provide cost-... ...effective, high-performance infrastructure for AI start-ups and large... ...network challenges in the platform. What You'll Be Doing...PlatformCloudFlexible hours
$225k
...exciting new opportunity? Join a rapidly scaling AI cloud infrastructure provider building next-generation GPU platforms designed for large-scale AI training,... ...The company is looking for a Senior Network Engineer to design, deploy, and support ultra-low-latency...PlatformCloudFull timeRemote work$120k - $200k
...About the Role As a Senior Infrastructure Engineer at Bland, you'll help us to... ...directly determines whether our platform can handle business-defining... ..., with deep knowledge of cloud infrastructure (AWS/GCP preferred... ..., model serving, or GPU computing. Experience with real...PlatformCloudWork at officeNight shift$150k - $250k
...software that runs factories. Our platform spans cloud backends, on-prem services,... ...Role You will work on the infrastructure that our services, our... ..., the deployment pipeline, GPU workload scheduling, and... ...manufacturing operators, mechanical engineers, robotics researchers, and...PlatformCloudFull time- ...next generation of AI infrastructure: large-scale AI datacenters... ...and the orchestration platform that coordinates them.... ...is seeking a Network Engineer to design, build, and... ...with AI/HPC, GPU, or large‑scale distributed... ...performance. Experience with cloud networking on GCP, AWS...PlatformCloud
- Senior Network Engineer Houston; New York; San Francisco; Seattle About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise... ...and software platforms. We thrive on a culture...PlatformCloud
- Network Engineer Gimlet is building the first multi-silicon neocloud... ...large-scale compute infrastructure with an execution platform that partitions AI... ...: Experience with AI/HPC, GPU, or large-scale distributed... ...performance. Experience with cloud networking on GCP, AWS, Azure...PlatformCloud
- ...ultrafast AI inference platform. We built a serverless... ...runtime that launches GPU‑backed containers in... ...someone to stand up that infrastructure and drive deployments... ...Hardware, and Network Engineering to identify blockers... ...in events across the cloud native community Fitness...PlatformCloud
$165k - $250k
...Overview We're looking for a hands-on Infrastructure Security Engineer to own security across our cloud and ML infrastructure — the platform that trains our models and serves... ...tenant isolation for user-deployed apps, GPU cluster security, and model protection from...PlatformCloudWork at officeImmediate startRemote workWorldwide3 days per week$180k - $250k
...products. We build the infrastructure, tools, and model... ...practical: a unified platform where high-performance... ...You are a hands-on engineer who builds the software... ...keep a large fleet of GPU servers healthy and productive... ...bare-metal and cloud based server fleets at...PlatformCloudLocal areaRelocation package- ...Whoever deploys frontier compute infrastructure fastest will decide whether... ...world. The Production Engineering Team Examples of key... ...time: build the observability platform that turns raw telemetry into... ...platforms, so every new site and GPU generation lands cleanly...PlatformCloudLocal area
- ...journey , is looking for an AI Infrastructure Engineer to join us in building out... ...to run our models on all platforms. What you’ll do: Design... ...on both on-prem and multi-cloud clusters. But most... ...excellent understanding of GPU’s handling large workloads...PlatformCloudWork experience placementWork at officeVisa sponsorship
- ...Senior Software Engineer, Developer Platform Houston, TX or San Francisco Bay Area... ...scale on Kubernetes across cloud (AWS) and on-prem data... ...handling scheduling, autoscaling, GPU and heterogeneous resources... ...ML, simulation, data, and infrastructure teams to understand...PlatformCloudTemporary work
$350k
Thinking Machines Lab is seeking a Network Engineer in San Francisco to manage and improve our GPU network fabric. The role requires in-depth knowledge of large... ..., where initiative and effective communication with cloud providers are key. The position offers a competitive...CloudVisa sponsorship$150k - $190k
...we build an end-to-end platform for developing,... ...seeking a Senior Network Engineer with hands-on Cumulus... ...backbone behind our AI infrastructure platform. You’ll play... ...Ethernet fabrics supporting GPU clusters and AI... ...Junos. Experience with cloud networking technologies...PlatformCloudWork at officeRemote workWork from homeFlexible hours2 days per week$140k - $200k
...from robots to smart infrastructure - building a safer and... ...Infrastructure / Quality Engineering role will play a... ...building, and validating the cloud infrastructure and... ...infrastructure, cloud platforms, and data quality validation... ...high-performance GPU cloud inference services...PlatformCloudWork experience placementLocal area- ...Job Description The AI Infrastructure team at Zensors builds the engine that powers our visual sensing platform. We provide the tools to automate... .... Deep understanding of GPU hardware performance ,... ...distributed training systems or cloud-scale inference serving (e.g...PlatformCloud
- ...California. The Role: As a Platform Engineer , you’ll be responsible for designing... ...the systems that keep Zyphra’s infrastructure robust, observable, secure, and... ...environments, such as ML clusters or GPU farms as well as hyperscaler cloud environments (i.e. AWS, GCP, etc....PlatformCloudWork at officeRelocation package
- Platform Engineer Department: Engineering Employment Type: Full Time Location: San... ...Francisco Description You'll own the infrastructure platform that our AI models run on... ...DevOps role. You'll work across GPU orchestration, multi-cloud Kubernetes, real-time networking,...PlatformCloudFull timeRelocationVisa sponsorship
$240k - $280k
...looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns... ...care of the rest. The platform is manifest-driven such... ..., inference bring-up to GPU driver/CUDA stack, health... ...work at a hyperscaler, GPU cloud, or datacenter-scale...PlatformCloudFull time- About the TeamOpenAI’s Infrastructure Operations team is... ...deliver highly available GPU infrastructure for AI... ...Infrastructure Operations Engineer to operate and improve... ...data center, cloud, AI, or HPC networks and... ...more of the following platforms: Cisco NX-OS, Arista EOS...PlatformCloudPermanent employment
$220k - $290k
...Calico seeks a Senior / Staff Cloud Engineer to lead the execution and... ...machine learning (ML) platform. As the technical authority... ...environment, you will build infrastructure-as-Code (IaC), design secure... ...workloadsTroubleshoot node-level GPU/TPU issues, manage scheduling...PlatformCloud- ...We run an agentic AI platform that acts as a personal... ...insurance navigation. The infrastructure challenge is unlike... ...a Staff Platform Engineer to build and own the infrastructure... ...Citizen Health - the cloud platform, deployment... ...model serving, manage GPU and compute capacity,...PlatformCloudRemote workFlexible hours
$170k - $250k
...Join an early-stage infrastructure company operating in the... ...company is looking for an engineer to work directly with the CTO on complex GPU virtualisation... ...involvement with a production platform, providing the opportunity... ...more efficient for cloud customers You will be...PlatformCloudFull timeVisa sponsorshipFlexible hours- ...problems. About The Role As a Platform Engineer at Phonic, you'll build the... ...every day. You'll own the infrastructure, tooling, and deployment systems... ...-time voice AI platform - cloud infrastructure, deployment... ...infrastructure — model serving, GPU workloads, or inference...PlatformCloudWork at office
$170k - $205k
...vertically integrated AI infrastructure company built from the... ...construction, and cloud services. If you... ...experienced Senior Software Engineer to join their team.... ...to define the AI platform roadmap. Influence... ...performance optimizations on GPU systems and inference...PlatformCloudTemporary workWork at office$176k - $220k
...Work together with engineers, scientists, operators... ...Human data is the core infrastructure to AI advancement. Frontier... ...hiring a Senior LLM Platform Engineer to join our... ...model serving, shared cloud infrastructure, and developer... ..., Triton, PyTorch, or GPU-backed serving....PlatformCloudFull timeWork at officeRemote workFlexible hours- ...Head of GPU Cloud About the Company Developing software foundation for... ...-generation accelerated computing platform for large-scale AI infrastructure. Industry Information Technology... ...of GPU Cloud to spearhead the engineering team dedicated to the development...PlatformCloud
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cloud Platforms and Infrastructure Engineer, TPU/GPU. Be the first to apply!
- senior cloud solutions architect San Francisco, CA
- cloud engineer San Francisco, CA
- senior principal cloud computing engineer San Francisco, CA
- aws cloud architect San Francisco, CA
- software engineer - cloud services San Francisco, CA
- google cloud engineer San Francisco, CA
- cloud engineering manager San Francisco, CA
- senior cloud network engineer San Francisco, CA
- principal cloud computing engineer San Francisco, CA
- cloud security engineer San Francisco, CA



