Customer Success Engineer (CSE), GPU Cluster
$260k - $290kTogether AI
About the role As a Customer Support Engineer at Together AI, you will serve as the named technical owner for one of our most strategic customer relationships. You will be the primary technical point of contact across all infrastructure domains — compute, networking, storage, and facilities — ensuring flawless delivery and operational health of large-scale GPU deployments. This role sits at the intersection of deep infrastructure expertise and high‑stakes customer partnership, making you a critical driver of both customer success and company growth. Responsibilities Serve as the named technical point of contact for a dedicated strategic customer, owning the end‑to‑end technical relationship across compute, networking, storage, and facilities Drive structured engagement through regular cadences including status reporting, technical steering meetings, and executive business reviews Translate customer operational feedback into actionable input for Engineering, Product, and Infrastructure roadmaps Lead issue lifecycle management, escalation, and RCA authorship across all infrastructure domains in partnership with Support, SRE, DC Ops, and Engineering teams Own end‑to‑end RMA coordination and hardware lifecycle management, including acceptance testing, spare inventory management, and hardware health reporting for large‑scale GPU deployments Maintain deep technical expertise across the customer's infrastructure stack — GPU compute, high‑speed fabric, and large‑scale storage systems — advising on configuration, operational best practices, and incident resolution Own the observability strategy for the customer estate, including alert policy definition, dashboard development, and proactive health management across all infrastructure layers Coordinate DC operations and facilities events in partnership with internal teams and hosting providers, ensuring SLA compliance and cluster availability Act as project manager for all capacity expansions, owning the full node deployment lifecycle from freight receipt through production acceptance Qualifications 5+ years in a customer‑facing technical role, with 2+ years in dedicated technical account management or solutions architecture for large‑scale AI or HPC infrastructure Deep expertise in GPU infrastructure — GPU health diagnostics, RMA workflows, and hardware acceptance testing Hands‑on experience with large‑scale Ethernet and InfiniBand fabric architecture Working knowledge of enterprise storage systems, including high‑density NVMe, parallel file systems, and metadata infrastructure Experience with DC operations, facilities coordination, and hosting provider SLA management Strong ownership mindset for incident management, RCA authorship, and executive‑level customer communication Proficiency in infrastructure monitoring and observability tooling (Prometheus, Grafana, or equivalent) Proven ability to manage multiple concurrent workstreams with hyperscaler‑level rigor and communication standards Proficiency in Python, Bash, or infrastructure automation tools preferred About Together AI Together AI is a research‑driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co‑designing software, hardware, algorithms, and models. We have contributed to leading open‑source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers on our journey in building the next generation of AI infrastructure. Compensation We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full‑time position is: $260-290K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job‑related knowledge. Location San Francisco, CA (Hybrid) or New York, NY (Hybrid) Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at #J-18808-Ljbffr Together AI
- ...and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering... ...the physical health and foundational infrastructure of our GPU clusters. You will be the primary custodian of our compute hardware, responsible...SuggestedWork at officeFlexible hours
- Linuxcareers in San Francisco is building AI research infrastructure. You will design, deploy, and operate large-scale GPU clusters powering training, evaluation, and serving for the research team. The role emphasizes extending orchestration with Kubernetes/Slurm, building...Suggested
$250k
...infrastructure provider building a next-generation GPU platform designed for AI training,... ...for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and... ...and monitoring frameworks for GPU compute clusters Collaborate with ML, data, and platform...SuggestedFull timeRemote work- Together AI seeks a Customer Support Engineer to own technical relationships with strategic customers, focusing on infrastructure health and GPU deployments. You'll engage with client needs, manage incidents, and ensure operational excellence across compute, networking...Suggested
$167.2k - $209k
...DigitalOcean is seeking a Senior Engineer 2 to play a key technical... ...at the inference engine and GPU kernel layers, ensuring our infrastructure... ...across multi-node GPU clusters. Technological Innovation:... ...sense of responsibility for customers, products, employees, and...SuggestedLocal areaRemote workWorldwideFlexible hours- ...us and help build the platform engineers turn to to ship AI products.... ...foundational engineers to lead our GPU Networking efforts, making... ...performance on bleeding-edge clusters (H100/H200, B200/B300, GB200/3... ...NVSHMEM) and potentially write custom communication kernels to overlap...Full timeFlexible hours
$5,000 per month
San Francisco / AustinCustomer Success Group - Customer Success /Full-time /HybridIndustry leader? Well,... ...assets to empower digital transformation.The CSE will combine advising, implementation experience, customer success engineering, enablement, and proactive engagement to...Full timeWork experience placementWork at officeImmediate startFlexible hours- ...boundaries of what's possible in video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and... ...solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure...
$150k - $200k
...our vision at Postman.About the RoleWe're building out the Customer Success Engineering team at Postman, and we need business-minded engineers who... ...solutions that benefit our entire enterprise customer base.As a CSE, you consult on the technical implementation path. When a...Interim roleWork at officeFlexible hours3 days per week$170k - $250k
...This company is looking for an engineer to work directly with the CTO on complex GPU virtualisation challenges. The role... ...significantly more efficient for cloud customers You will be doing performance... ..., and distributed GPU cluster management You will be working...Full timeVisa sponsorshipFlexible hours$179k - $218k
...meaningful work of your career, help our customers and partners advance their AI... ...a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive... ...predictive strategies needed to maintain peak cluster health.The Strategic BridgeFor DC...Temporary work- ...Customer Success Engineer San Francisco, CA & New York, NY Who We Are At Pave, we're building the industry's leading compensation platform... ..., Bessemer Venture Partners, and Craft Ventures. The CSE Team @ Pave: You'll become a Pave product and compensation...Immediate startFlexible hours3 days per week
$300k
...GPU Optimisation Engineer — Real-Time Inference Want to push GPU performance to its limits — not in theory, but in production systems handling... ...kernel fusion, quantisation, and scheduling Writing and tuning custom CUDA / Triton kernels for performance-critical paths...RelocationVisa sponsorshipFree visa- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one... ...Function of PositionAs a Senior Systems GPU Engineer - AI & Robotics, you will be... ..., please note that Intuitive and/or your customer(s) may require that you show current...Local areaWorldwideFlexible hours
- MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models. The ideal candidate has over 4 years of experience with GPU kernels...
$100k - $120k
...cheaper and faster. Responsibilities Lead a team of kernel and system engineers focused on performance-critical code Design, implement, and optimize custom compute kernels for CPU (AVX/ARM NEON), GPU (CUDA/ROCm), and hardware accelerators Find bottlenecks in memory hierarchy...- San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep...Work at officeRelocation package
$285k - $315k
About The Role We're looking for a Founding GPU Kernel Engineer who lives right at the boundary between hardware and software. Someone who thinks... ...parallelism, ZeRO-style optimizations Familiarity with custom accelerators: TPUs, Trainium, Inferentia, or similar...Full timeWork at officeRelocation package$315k
A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams...- A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal...
$285k - $315k
SF Tensor is looking for a Founding GPU Kernel Engineer in San Francisco, specializing in GPU architecture and kernel optimization for machine learning workloads. The ideal candidate has deep expertise, proven capabilities in hand-optimizing performance-critical kernels...Full timeRelocation package- ...transform your dreams into reality. You would collaborate with software engineers, AI researchers, and hardware specialists to develop high-... ...future of mobility. Key Responsibilities Optimize end-to-end GPU performance for real-time autonomous driving workloads, including...
- Mirai Labs in San Francisco seeks engineers to join a senior team building the full on-device stack for real-time local intelligence. You... ...language models work, and experience in writing high-performance GPU kernels or Rust systems programming. We welcome applications from...Local area
$285k - $315k
...AMD. We are partnering with researchers, engineers, and organizations who share our belief... ...About the Role We're hiring a Founding GPU Compiler Engineer to build the core compilation... ...classic compiler optimizations customized for large-scale training (fusion, tiling...Full timeWork at officeRelocation package$220k - $320k
inference.net, a growing company in San Francisco, seeks an experienced engineer to optimize AI inference performance. The ideal candidate will have over 2 years of experience in ML systems and GPU programming. Key responsibilities include implementing optimization techniques...$180k - $280k
...tier investors. Since mid-2024, we've been engineering the foundation for what comes after the... ...done. About the role We're looking for a GPU kernel engineer with deep, low-level CUDA... ...more efficient. You'll write and optimize custom kernels, profile and eliminate...Work at officeVisa sponsorshipShift work- ...Customer Success Engineer San Francisco, CA & New York, NY Who We Are At Pave, we're building the industry's leading compensation platform... ..., Bessemer Venture Partners, and Craft Ventures. The CSE Team @ Pave: You'll become a Pave product and compensation...Immediate startFlexible hours3 days per week
$315k
...committed researchers, engineers, policy experts, and... ...breakthrough innovations in GPU performance and... ...the-art techniques from custom kernel development to... ...infrastructure, fault tolerance, cluster orchestration... ..., we aren't able to successfully sponsor visas for every...Work at officeVisa sponsorshipFlexible hours- ...parallelism across heterogeneous hardware to maximize throughput in production. You will work with state-of-the-art engines, profiling tools, and advanced GPU techniques while collaborating with a hands-on team in a fast-paced SF office environment. #J-18808-Ljbffr...Work at office
- ...sized, and leadership is earned by shipping excellence. We seek engineers with strong intrinsic drive, a true passion for advancing the... ...leverage your knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI. Full-Time On-site at either...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Customer Success Engineer (CSE), GPU Cluster. Be the first to apply!
- airline customer San Francisco, CA
- customer engineer San Francisco, CA
- customer retention San Francisco, CA
- customer project program manager San Francisco, CA
- amazon customer San Francisco, CA
- customer satisfaction San Francisco, CA
- work from home customer San Francisco, CA
- customer liaison San Francisco, CA
- customer San Francisco, CA
- customer marketing San Francisco, CA



