Cloud Platforms and Infrastructure Engineer, TPU/GPU
$152k - $221kPropose solution architectures and manage the deployment of cloud networking solutions according to customer requirements and implementation best practices.Design and implement cloud-based technical architectures, migration approaches, and application optimizations that enable business objectives in collaboration with customers.Work with internal Specialists, Product, and Engineering teams to package approaches, best practices, and lessons learned into thought leadership, methodologies, and published assets.Interact with sales, partners, and customer technical stakeholders to manage project scope, priorities, deliverables, risks and issues, and timelines.Travel up to 30% for in-region for meetings, technical reviews, and onsite delivery activities.Minimum qualifications:Bachelor's degree in Computer Science or equivalent practical experience.6 years of experience automating infrastructure provisioning, Developer Operations (DevOps), continuous integration, or delivery utilizing Kubernetes and Linux-based systems. 3 years of experience in project management and technical solution delivery.Experience coding in one or more general purpose languages (e.g., Python, Java, Go, C or C++) including data structures, algorithms, software design, Linux environments and Kubernetes orchestration.Experience working with Cloud Providers such as Google Cloud Platform (GCP).Ability to travel 30% of the time, as needed, for client engagements.Preferred qualifications:Experience with third-party networking (e.g., PANW, Fortinet, VMWare) and design, including redundancy and load balancing.Experience troubleshooting networking protocols including TCP/IP, Hypertext Transfer Protocol, and Border Gateway Protocol (BGP).Experience in customer-facing migration, including service discovery, assessment, planning, execution, and operations.Experience with standard IT security practices, including IAM, data protection, encryption, and certificate/key management.Experience running AI/ML training and inference workloads on GPU/TPU using frameworks such as PyTorch, JAX, TensorFlow, or Slurm.Knowledge of containerization and orchestration technologies, including Google Kubernetes Engine (GKE) and related cloud-native services.The Google Cloud Consulting Professional Services team guides customers through the moments that matter most in their cloud journey to help businesses thrive. We help customers transform and evolve their business through the use of Google’s global network, web-scale data centers, and software infrastructure. As part of an innovative team in this rapidly growing business, you will help shape the future of businesses of all sizes and use technology to connect with customers, employees, and partners.As a Cloud Platform and Infrastructure Engineer, you will provide technical guidance to customers adopting Google Cloud Platform (GCP) services, including providing best practices on secure foundational cloud implementations, automated provisioning of infrastructure and applications, cloud-ready application architectures, and more. You will also provide guidance in ensuring that customers receive the best of what GCP can offer and have the best experience in migrating, building, modernizing, and maintaining applications in GCP. Additionally, you will work with Product Management and Product Engineering to drive excellence in Google Cloud products and features.Google Cloud accelerates every organization’s ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google’s cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $152000 - $221000 (USD) + 15% bonus target + equity + benefitsLearn more about benefits at Google.Applicants in San Francisco: Qualified applications with arrest or conviction records will be considered for employment in accordance with the San Francisco Fair Chance Ordinance for Employers and the California Fair Chance Act.The application window will be open until at least August 25, 2026. This opportunity will remain online based on business needs which may be before or after the specified date.Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Austin, TX, USA; San Francisco, CA, USA; Atlanta, GA, USA; Boulder, CO, USA; Addison, TX, USA; Miami, FL, USA; Sunnyvale, CA, USA; Chicago, IL, USA.Bachelor's degree in Computer Science or equivalent practical experience.6 years of experience automating infrastructure provisioning, Developer Operations (DevOps), continuous integration, or delivery utilizing Kubernetes and Linux-based systems. 3 years of experience in project management and technical solution delivery.Experience coding in one or more general purpose languages (e.g., Python, Java, Go, C or C++) including data structures, algorithms, software design, Linux environments and Kubernetes orchestration.Experience working with Cloud Providers such as Google Cloud Platform (GCP).Ability to travel 30% of the time, as needed, for client engagements.
- ...technology firm focused on video intelligence is seeking a Backend / Infrastructure Engineer to enhance its platform. The successful candidate will design and implement cloud-based solutions, optimize GPU workflows, and develop APIs for enterprise-level applications....PlatformCloud
- ...will define architectures spanning InfiniBand interconnects, spine-leaf topologies, and underlay/overlay strategies, working with cloud platform teams and vendors. You will also drive congestion control tuning, document standards, and identify risks with clear roadmaps,...PlatformCloud
- ...innovation by optimizing GPU-accelerated Kubernetes platforms for machine learning. Collaborate... ...projects in advanced infrastructure development. Elevate your... ...clusters across hybrid cloud and bare-metal... ...of experience in platform engineering or DevOps.Proficiency in...PlatformCloud
$84.9k - $209.5k
...computing and AI solutions on Oracle Cloud Infrastructure (OCI). You will engage directly... ...architect and deploy complex HPC and GPU clusters, AI platforms, and intelligent agentic solutions... ...sales technical consulting, solution engineering, and AI transformation strategy....PlatformCloudTemporary workFlexible hours$233k - $300k
...DescriptionThe RoleYou will lead the Cloud Engineering team within Vehicle Autonomy... ...to build an elastic compute platform dedicated to autonomous... ...a product, abstracting away infrastructure complexity so our engineers... ...)Experience with GPU-accelerated workloads and specialized...PlatformCloudFull timeLocal areaWork from homeFlexible hours$184k - $287.5k
...NVIDIA seeks a Senior Software Engineer for our CSP (Cloud Service Provider)... ...-services that expose new GPU capabilities.Drive joint architecture... ...with CSP and internal platform teams; convert findings... ...Prometheus, OpenTelemetry), and infrastructure-as-code.Excellent...PlatformCloudFull timeRemote work$224k - $356.5k
Joining NVIDIA's DGX Cloud AI Efficiency Team means... ...AI researchers and platform teams understand end-to... ...a Senior Performance Engineer to characterize workloads... ...diverse team of infrastructure experts to unlock more... ..., platform teams, and GPU architects to validate...PlatformCloudFull timeRemote work- ...former Navy electrical engineers with a proven track record... ...seeking an experienced Platform Engineer to design, build, and own the infrastructure powering the... ...You will manage a 130+ GPU bare-metal Kubernetes cluster... ...repeatable delivery to cloud and edge targets. Build...PlatformCloudLocal area
$116k - $189.75k
...workloads. We are looking for a Software Engineer focused on bring-up, triage,... ...and inference workloads across NVIDIA GPU platforms at the largest scales we run.In this role... ...validate, and debug large-scale AI clusters, infrastructure, and end-to-end workloads.Bring up,...PlatformCloudFull timeRemote work- ...for talented and passionate software engineers for the Hardware Platform Services team. This team designs and develops Cloud Hardware infrastructure including Compute, Storage, AI servers... ...drivers, developing software for CPU/GPU data centers, maintaining libraries, and...PlatformCloud
$184k - $287.5k
...We are looking for a Senior Software Engineer to lead the bring-up, triage, benchmarking... ...and inference workloads across NVIDIA GPU platforms at the largest scales we run.In this... ...debugging of large-scale AI clusters, infrastructure, and end-to-end workloads, setting the...PlatformCloudFull timeRemote work- ...former Navy electrical engineers with a proven track... ...an experienced CV/ML Platform Engineer with specialization... ..., model, and compute infrastructure powering ACS CV/ML... ...will help manage a 130+ GPU bare-metal Kubernetes... ...artifact management for both cloud and edge targets....PlatformCloudLocal area
- ...supercomputing, high-performance computing, cloud, and AI. Whether you’re designing next-... ...world forward.THE ROLE: The Datacenter Platform Engineering Group (DPEG) organization is looking... ...is a must, understand the flow of a GPU through the different layers of a system...PlatformCloud
$200k - $290k
...seeking a Staff Network Engineer (WAN) to design and... ...multiple data centers and GPU clusters. This role... ...with compute, storage, infrastructure, and facility teams to... ...coherent optics) and platforms like Ciena for large‑scale... ...direct connects with cloud providers and...PlatformCloudFull timeContract work$184k - $287.5k
Joining NVIDIA's DGX Cloud AI Efficiency Team means contributing to the infrastructure that powers our innovative... ...software engineer to join our team. You'... ...underpinning NVIDIA's AI platforms.Define meaningful and... ...and Visualization. The GPU, our invention, serves...PlatformCloudFull timeRemote work$183k - $247.6k
...the world. As a member of the Cloud-Scale Machine Learning... ...reliable, scalable, low-cost infrastructure platform in the cloud that powers hundreds... ...Design Verification Engineers to build the next generation... ...with verifying complex CPU, GPU, or ML accelerator designsAmazon...PlatformCloudLocal areaWork from homeFlexible hours- We Are:The Global AI Infrastructure team is at the center of enabling infrastructure... ...technical expertise across cloud, on-premises, and hybrid... ...that powers AI platforms, GPU-accelerated workloads, large... ...CUDA along with LLM inference engines (TensorRT-LLM), production serving...PlatformCloudFull timeWork experience placementLive inWork at officeLocal area
$170k - $230k
Job DescriptionWe are hiring a Senior Platform Engineer to join the Autonomous Vehicle (AV) Cloud Engineering team within AV Core Infrastructure. AV Cloud Engineering owns core IaaS and... ...offs between hardware-level performance (GPU passthrough) and clean cloud...PlatformCloudFull timeWork experience placementWork at officeLocal areaWork from homeFlexible hours$200k - $300k
...Computing (HPC) Network Engineering team designs and... ...latency communications infrastructure that underpins our incredibly large GPU and CPU compute clusters... ...needed for the future of cloud-scale networking. You’ll... ...interest in exploring new platforms Strong design skills...PlatformCloudWork at officeLocal areaImmediate startWorldwide$135.2k - $306.4k
...Principal Network Development Engineer (IC5) to lead backend NIC... ...-generation networking platforms supporting GPU- and accelerator-based clusters... ...efforts in data center or cloud environmentsStrong... ...initiatives that impact OCI AI infrastructure at fleet levelEnd-to-End...PlatformCloudTemporary workFlexible hours- ...shareholders.How You Will Make an Impact:The Sr. Infrastructure Engineer is responsible for designing,... ...reliable, secure, and scalable platforms across manufacturing sites, remote locations... ..., Windows server platforms, cloud infrastructure, identity services, monitoring...PlatformCloudWork at officeLocal areaImmediate startRemote work
- ...Product Manager, AI GPUs and Platforms, to define and drive product... ...business.This role sits within the Cloud & Enterprise AI (CEAI)... ...responsible for AMD’s CPU and GPU AI roadmap and business results... ...work cross functionally with engineering, architecture, sales,...PlatformCloud
- ...demanding industries. As a DevOps Engineer, you are part of a mission-... ...and highly skilled Senior Infrastructure Engineer tosupport and... ...locations, and multiple public cloud environments.This is an onsite... ...networking, identity platforms, endpoint management systems...PlatformCloudFull timeWork at officeRemote workWorldwide
- ...rapidly evolving state of the art to engineer the platforms that power the innovative, research-driven... ...financial lives. As a member of the Infrastructure Platform Team you will be a hands-on... ...involved with Windows, Citrix and Cloud Services to Dimensional’s Infrastructure...PlatformCloudFull timeWork experience placementRemote work
$79.2k - $209.5k
...performance tests; leverages data plane platforms and distributed state tools for high-... ...standards are met.Oracle Cloud Infrastructure Workflow is a Tier 0 service that is critical... ...step work in a fault tolerant manner. An engineer on this team is responsible for the development...PlatformCloudTemporary workFlexible hours$150k - $170k
Engineering Manager, InfrastructureLocation: Austin, TX - ON SITE ROLESalary... ...Engineering Manager, Infrastructure to lead a team responsible... ...release pipelines, and supporting cloud and on-premises... ...desktop, console, and mobile platforms. The successful candidate will...PlatformCloudRelocation- .... THE TEAM:AMD's Data Center GPU organization is transforming... ...opportunity to lead the Systems/Platform team as they collaborate with... ...companies developing AI Cloud Services, on-prem AI Clusters... ...require interlock with other engineering and business. The role covers...PlatformCloudWorldwide
- ...AI and Bitcoin mining infrastructure. Bitdeer is committed... ...also offers advanced cloud capabilities to customers... ...and cloud platform development architecture... ...architecture, covering GPU cluster interconnects... ...with the cloud platform engineering team on Underlay network...PlatformCloudLocal area
- ...is the Kubernetes-native AI infrastructure company, enabling organizations... ..., Mirantis empowers platform engineering teams to deliver composable,... ...environment—on-premises, in the cloud, at the edge, or in sovereign... ...Mirantis delivers the automation, GPU orchestration, and policy-...PlatformCloudRemote workShift work
$100k - $258k
...highly motivated, and focused on engineering excellence. This... ...seeking a talented and motivated Infrastructure Security Engineer to join our... ...implement, and maintain secure cloud and hybrid infrastructure,... ...environments, with work that may span platforms such as SpaceXAI, X Social,...PlatformCloudTemporary work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cloud Platforms and Infrastructure Engineer, TPU/GPU. Be the first to apply!
- senior cloud security engineer Austin, TX
- azure cloud solution architect Austin, TX
- senior cloud solutions architect Austin, TX
- senior cloud data engineer Austin, TX
- cloud operations engineer Austin, TX
- cloud engineering manager Austin, TX
- informatica cloud developer Austin, TX
- full stack cloud developer Austin, TX
- senior cloud engineer Austin, TX
- cloud architect Austin, TX


