Member of Technical Staff - AI Cloud Infrastructure
emerald-ai
About Emerald AI We’re at a pivotal moment for AI and energy. Demand for compute is skyrocketing, but power constraints are becoming a critical bottleneck. Emerald AI sits at the intersection of these two worlds, enabling AI data centers to scale without overwhelming the grid. Our Emerald Conductor software platform makes data centers flexible and responsive, allowing them to adjust power usage dynamically. This unlocks massive AI growth without major new infrastructure, while also strengthening the grid and supporting the expansion of renewable energy. We’re a team of experts across AI, cloud, software, and energy—on a mission to scale AI sustainably. We’re backed by leading investors and partners including Radical Ventures and NVIDIA. Learn more about our vision, team, and backers at About the Role Emerald AI is building the world's first power flexible managed cloud infrastructure. We are hiring a senior infrastructure engineer to architect and stand up our managed cloud services from end to end. The work covers the platform, the control plane, and the customer experience that together make up a managed AI cloud. The right person has done this before. They have built or served as a core early engineer on a managed cloud or AI platform, whether at a GPU cloud, an internal machine learning platform run at scale, a hyperscaler AI service, or a HPC research computing center operated as a service. This is a role for an architect who still builds. You will make the major design decisions and then implement them yourself. Key Responsibilities Architect our managed services from 0→1. Define the productization of GPU capacity, encompassing isolation boundaries, tenant models, provisioning flows, and service catalogs that scale across diverse providers. Engineer the platform core. Build robust control-plane services, self-service customer interfaces, and automated lifecycle systems, including usage metering integrated with billing infrastructure. Onboard and vet infrastructure partners. Conduct deep technical assessments of bare-metal GPU vendors, evaluating fabric quality, network isolation, and economics to automate the path from handoff to active tenant. Design end-to-end multi-tenancy. Implement rigorous isolation across compute, storage, and networking (InfiniBand/VLANs), ensuring secure boundaries, QoS, and encryption even when customers possess root access. Drive workload orchestration. Manage Kubernetes and Slurm environments for large-scale training and inference, overseeing node health, driver fleets, and kernel management across heterogeneous clouds. Lead high-performance storage strategy. Deploy and integrate parallel storage solutions like Lustre, VAST, or Weka, leveraging your deep experience with these systems to ensure they fold cleanly into our provisioning model. Ensure operational excellence. Define SLOs, observability standards, and incident response protocols that bridge our internal standards with underlying provider SLAs to deliver a reliable, sellable product. Minimum requirements At least 7+ years of experience in infrastructure or platform engineering, including the architecture and launch of a managed cloud or AI platform that reached production users. Strong experience with Kubernetes and Slurm and offering them as managed service Production experience deploying or operating Lustre or a comparable parallel filesystem such as GPFS, Weka, VAST, or BeeGFS, with a solid understanding of parallel filesystem architecture, tuning, and failure modes. A strong grasp of cloud service fundamentals, including control planes, tenancy and isolation models, APIs, quota and metering systems, and the operational discipline of running a service that customers pay for. Deep Linux systems knowledge, mature infrastructure as code practice with tools such as Terraform and Ansible, and solid programming ability in Python or Go. Familiarity with GPU infrastructure, including high performance networking with InfiniBand, RoCE, and RDMA, and the GPU software stack. Preferred requirements Prior time at a GPU cloud, a hyperscaler AI service, or an HPC center that delivers compute and storage as a service, especially one built on rented or colocated capacity. Familiarity with NVIDIA reference architectures such as SuperPOD, along with GPUDirect Storage, NCCL debugging, and DCGM. Experience with Lustre multitenancy features such as nodemap, fileset mounts, and Kerberos, or with service provider deployments of VAST or Weka. Experience negotiating with and integrating multiple infrastructure vendors, together with a practice of designing for portability between them. Experience running object storage at scale with systems such as S3, Ceph, or MinIO, including the design of data tiering. Experience building billing, metering, or FinOps pipelines for services that charge by usage. What We Offer Make an impact. Solve the AI power bottleneck and shape how data centers scale sustainably. Join a world-class team of AI, cloud, software, and energy experts in a collaborative, low-ego environment. Build from 0→1. Influence strategy, GTM, org design, and customer/investor engagement from day one. Competitive pay + equity. Stock options let you share in the value you help create. Comprehensive benefits, including medical, dental, vision, and 401(k) matching. Flexible location. Work from D.C., Boston, or the Bay Area, with 2 WFH days/week. Backed by top investors, including Radical Ventures and NVIDIA. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, disability, and any other protected ground of discrimination under applicable human rights legislation. Emerald AI strives to respect the dignity and independence of people with disabilities and is committed to giving them the same opportunity to succeed as all other employees. Inclusiveness is core to our culture at Emerald AI, and we strive to ensure you get the most from your interview experience. Emerald AI makes reasonable accommodations for applicants with disabilities. If a reasonable accommodation is needed to participate in the job application or interview process, please reach out to the Talent team. #J-18808-Ljbffr emerald-ai
- ...Perimeter Institute building AI systems to discover new... ...engineers to build platform infrastructure at the intersection of computational... ...scale. We are seeking a Member of Technical Staff, Infrastructure &... ...artifact that runs. Run the cloud estate as code across GCP and...CloudRemote work
- ...Perimeter Institute building AI systems to discover new physics... ...next one. We're creating the infrastructure to industrialize scientific... ...infrastructure stack end‑to‑end, from cloud foundations through CI/CD... ...stage equity. We evaluate on technical breadth, systems thinking,...CloudRemote work
$114.6k - $234.6k
...systems, virtualized infrastructure, and highly... ...have a significant technical and business impact. As a member of the software engineering... .... At Oracle Cloud Infrastructure (OCI... ...Member of Technical Staff, you will own the software... ...care. And with AI embedded across our...CloudTemporary workFlexible hours- ...that loop. As an AI robotics company that... ...; at Tutor, every member of our team is... ...characterized by both technical excellence and... ...Member of Technical Staff , you will lead design... ..., and robot infrastructure. You will influence... ...from embedded to cloud, and champion best...CloudWork at officeShift work
- About Emerald AI We're at a pivotal moment for AI and energy... ...growth without major new infrastructure, while also strengthening... ...team of experts across AI, cloud, software, and energy---on... ...at About the Role As a Member of Technical Staff, you will help invent and build...CloudWork from homeFlexible hours2 days per week
- About Emerald AI We’re at a pivotal moment for AI and energy... ...growth without major new infrastructure, while also strengthening... ...team of experts across AI, cloud, software, and energy—on a... ...to the electric grid. As a Member of Technical Staff on the Grid Services team,...CloudWork from homeFlexible hours2 days per week
- ...that loop. As an AI robotics company that... ...; at Tutor, every member of our team is... ...characterized by both technical excellence and... ...are looking for a Staff Software Engineer... ...services, data and ML infrastructure, internal tools,... ...infrastructure Experience with cloud infrastructure,...CloudWork at officeShift work
- ...About Emerald AI We’re at a pivotal moment for... ...growth without major new infrastructure, while also... ...of experts across AI, cloud, software, and energy... ...roadmap, and long-term technical direction. We hire across... ...from early-career to staff+, with expectations and...CloudImmediate startWork from homeFlexible hours1 day per week
$135.2k - $306.4k
...Description As a Senior Staff Engineer, you will play... ...the future of OCI's infrastructure. Your expertise will be... ..., and influencing technical strategies across teams... ...observability, and responsible AI delivery.... ...operating distributed systems, cloud services, SaaS platforms...CloudTemporary workFlexible hours- ...fetch economics across storage tiers and clouds Build versioned, rights-tracked,... ...evaluation corpora, and captured traces for AI Design systems so results remain traceable... ...Five or more years building data infrastructure in production at companies known for engineering...CloudRemote workRelocation
- ...interested in building the future of AI-enabled drug development,... ..., and computational safety infrastructure. If you are excited about... ...backgrounds who are excited to solve technically challenging problems in AI,... ...-stack product engineering Cloud infrastructure and MLOps...CloudRemote work
- ...onboard services, tooling, and infrastructure that tie perception, control,... ...engineers across controls, AI, teleoperation, and hardware,... ...frameworks, Docker, CI/CD systems, or cloud services (AWS). Exposure to... ...who combine world-class technical skills with creative vision,...CloudInternshipWork from homeFlexible hours
- ...Specialist, Production Services Infrastructure Support AnalystAt BNY, our... ...our teams harness cutting-edge AI and breakthrough technologies... ....We’re seeking a future team member for the role of Vice... ...with Azure / VMware / Oracle Cloud.Exposure / Experience with Cisco...CloudWorldwideFlexible hours
- ...Perimeter Institute building AI systems to discover new physics... ...engineers to build platform infrastructure at the intersection of... ...platforms, product engineering, cloud infrastructure, and security.... ...You can walk us through a hard technical decision you made, what you traded...CloudRemote work
$180k - $225k
...Blitzy Blitzy is a Cambridge, MA based AI software development platform on a... ...ups. The Role We are hiring a Senior Member of Technical Staff that has two responsibilities that sharpen... ...stack: backend systems, distributed infrastructure, data architecture, LLM/AI systems,...Immediate start- ...Product Manager - GPU Products & AI InfrastructureWhy This... ...accelerated computing in the cloud. You will play a pivotal role... ...intersection of product strategy, AI infrastructure, cloud computing, and... ...Accelerated Computing: Strong technical understanding of GPU architectures...CloudRemote work
- ...Management, IncSenior Product Manager - Next-Gen Cloud Infrastructure & GPU PlatformsWe are seeking a highly technical Senior Product Manager to lead the strategy and roadmap... ...scalable infrastructure products supporting AI, HPC, enterprise, and cloud workloads.To land...CloudRemote work
- Tutor is seeking a highly skilled Member of Technical Staff to drive robotics software across motion planning, real-time control, perception, and... ..., and own critical subsystems spanning embedded to cloud. As a MoTS, you will shape direction, collaborate cross-functionally...CloudFlexible hours
- Staff Robotics Software Engineer As a Staff Robotics Software Engineer, you... ...optimization systems, tooling, and robot infrastructure. You will influence technical direction, mentor other engineers,... ...across the stack from embedded to cloud, and champion best practices in...Cloud
$131k - $250k
...Overview: Lead product strategy and roadmaps for cloud infrastructure, compute, and GPU platforms supporting AI, HPC, enterprise, and cloud workloads. Key... ...~ Strong product strategy skills. ~ Strong technical stakeholder-management skills. ~ Bachelor's degree...CloudRelocationVisa sponsorship- ...Job Title: Cloud Infrastructure Product Manager Job Location: Cambridge, MA Job Type: Full... ...market needs into product requirements and technical specifications. Partner with... ...Experience with GPU architectures, CUDA, AI/HPC workloads, and accelerated computing...CloudFull time
$116.2k - $229.1k
...Summary Join Deloitte’s AI & Engineering practice and... ...organizations modernize enterprise cloud and data platforms on Azure. As an Azure/Databricks Infrastructure Engineer, you will design,... ...through architecture reviews, technical workshops, and delivery collaborationParticipate...CloudLocal areaVisa sponsorship- ...an operator uses to drive a robot, the fleet infrastructure that connects and updates machines in the field, the cloud services that hold robot and customer data, and... ...and build the security program from a strong technical foundation, partnering closely with our IT and...CloudWork from homeFlexible hours
- ...Cambridge, and the Perimeter Institute building AI systems to discover new physics at scale... ...are seeking engineers to build platform infrastructure at the intersection of computational... .... How We Work We hold a high technical bar and give people full ownership of...Remote work
- ...operations of enterprise IT infrastructure and security. Oversee the Microsoft... ...and scale infrastructure and cloud computing environments to... ...DLP, sensitivity labels, and AI security policies. Advance... ...willingness to engage in hands‑on technical support Excellent...CloudHome office
$120k - $310k
...teams leverage modern cloud platforms, automation,... ...emerging technologies.As a member of the Desktop... ...initiatives, and emerging AI technologies. You will work closely with infrastructure, security, and business... ...solutions. Create and maintain technical documentation,...CloudFull timeLocal areaWorldwide- ...the Role Premier Global Links LLC is seeking a highly technical Senior Product Manager - Cloud Infrastructure & GPU to lead product strategy, roadmap, and... ...to build scalable infrastructure products supporting AI, HPC, enterprise, and cloud workloads . Key Responsibilities...Cloud
- ...topology. Cross-Functional Collaboration: Partner with AI and software infrastructure teams to define interfaces and ensure hardware capabilities... ...of exceptional professionals who combine world-class technical skills with creative vision, grounded in humility and collaboration...Work from homeFlexible hours
- ...in advanced embedded systems, physical AI, and deeply technical product development. We help ambitious... ...employ a flexible hybrid model, team members must be available to work on-site as necessary... ...for a Senior Member of Technical Staff to join our tight-knit team of...Flexible hours
- ...experienced Senior ASIC Designers. This role demands proven technical expertise in advanced ASIC design flows, and... ...Impact: We are tackling a fundamental challenge at the infrastructure layer: unlocking greater AI capability while dramatically improving efficiency. The...Visa sponsorshipRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff - AI Cloud Infrastructure. Be the first to apply!
- sharepoint support analyst Boston, MA
- operations support technician Boston, MA
- product support technician Boston, MA
- senior technical analyst Boston, MA
- systems support technician Boston, MA
- user support analyst Boston, MA
- junior IT service management analyst Boston, MA
- technical support specialist Boston, MA
- remote support technician Boston, MA
- help desk assistant Boston, MA


