Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Customer Success Engineer (CSE), GPU Cluster

$260k - $290k

Together AI

About the role As a Customer Support Engineer at Together AI, you will serve as the named technical owner for one of our most strategic customer relationships. You will be the primary technical point of contact across all infrastructure domains — compute, networking, storage, and facilities — ensuring flawless delivery and operational health of large-scale GPU deployments. This role sits at the intersection of deep infrastructure expertise and high‑stakes customer partnership, making you a critical driver of both customer success and company growth. Responsibilities Serve as the named technical point of contact for a dedicated strategic customer, owning the end‑to‑end technical relationship across compute, networking, storage, and facilities Drive structured engagement through regular cadences including status reporting, technical steering meetings, and executive business reviews Translate customer operational feedback into actionable input for Engineering, Product, and Infrastructure roadmaps Lead issue lifecycle management, escalation, and RCA authorship across all infrastructure domains in partnership with Support, SRE, DC Ops, and Engineering teams Own end‑to‑end RMA coordination and hardware lifecycle management, including acceptance testing, spare inventory management, and hardware health reporting for large‑scale GPU deployments Maintain deep technical expertise across the customer's infrastructure stack — GPU compute, high‑speed fabric, and large‑scale storage systems — advising on configuration, operational best practices, and incident resolution Own the observability strategy for the customer estate, including alert policy definition, dashboard development, and proactive health management across all infrastructure layers Coordinate DC operations and facilities events in partnership with internal teams and hosting providers, ensuring SLA compliance and cluster availability Act as project manager for all capacity expansions, owning the full node deployment lifecycle from freight receipt through production acceptance Qualifications 5+ years in a customer‑facing technical role, with 2+ years in dedicated technical account management or solutions architecture for large‑scale AI or HPC infrastructure Deep expertise in GPU infrastructure — GPU health diagnostics, RMA workflows, and hardware acceptance testing Hands‑on experience with large‑scale Ethernet and InfiniBand fabric architecture Working knowledge of enterprise storage systems, including high‑density NVMe, parallel file systems, and metadata infrastructure Experience with DC operations, facilities coordination, and hosting provider SLA management Strong ownership mindset for incident management, RCA authorship, and executive‑level customer communication Proficiency in infrastructure monitoring and observability tooling (Prometheus, Grafana, or equivalent) Proven ability to manage multiple concurrent workstreams with hyperscaler‑level rigor and communication standards Proficiency in Python, Bash, or infrastructure automation tools preferred About Together AI Together AI is a research‑driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co‑designing software, hardware, algorithms, and models. We have contributed to leading open‑source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers on our journey in building the next generation of AI infrastructure. Compensation We offer competitive compensation, startup equity, health insurance, and other benefits, as well as flexibility in terms of remote work. The US base salary range for this full‑time position is: $260-290K OTE + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job‑related knowledge. Location San Francisco, CA (Hybrid) or New York, NY (Hybrid) Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our Privacy Policy at #J-18808-Ljbffr Together AI

Vacancy posted 7 hours ago
Similar jobs that could be interesting for youBased on the Customer Success Engineer (CSE), GPU Cluster in San Francisco, CA vacancy
  •  ...and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering...  ...the physical health and foundational infrastructure of our GPU clusters. You will be the primary custodian of our compute hardware, responsible... 
    Suggested
    Work at office
    Flexible hours

    Sciforium

    San Francisco, CA
    4 days ago
  • Linuxcareers in San Francisco is building AI research infrastructure. You will design, deploy, and operate large-scale GPU clusters powering training, evaluation, and serving for the research team. The role emphasizes extending orchestration with Kubernetes/Slurm, building... 
    Suggested

    Linuxcareers

    San Francisco, CA
    4 days ago
  • $250k

     ...infrastructure provider building a next-generation GPU platform designed for AI training,...  ...for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and...  ...and monitoring frameworks for GPU compute clusters Collaborate with ML, data, and platform... 
    Suggested
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • Together AI seeks a Customer Support Engineer to own technical relationships with strategic customers, focusing on infrastructure health and GPU deployments. You'll engage with client needs, manage incidents, and ensure operational excellence across compute, networking... 
    Suggested

    Together AI

    San Francisco, CA
    7 hours ago
  • $167.2k - $209k

     ...DigitalOcean is seeking a Senior Engineer 2 to play a key technical...  ...at the inference engine and GPU kernel layers, ensuring our infrastructure...  ...across multi-node GPU clusters. Technological Innovation:...  ...sense of responsibility for customers, products, employees, and... 
    Suggested
    Local area
    Remote work
    Worldwide
    Flexible hours

    DigitalOcean

    San Francisco, CA
    7 hours ago
  •  ...us and help build the platform engineers turn to to ship AI products....  ...foundational engineers to lead our GPU Networking efforts, making...  ...performance on bleeding-edge clusters (H100/H200, B200/B300, GB200/3...  ...NVSHMEM) and potentially write custom communication kernels to overlap... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    17 hours ago
  • $5,000 per month

    San Francisco / AustinCustomer Success Group - Customer Success /Full-time /HybridIndustry leader? Well,...  ...assets to empower digital transformation.The CSE will combine advising, implementation experience, customer success engineering, enablement, and proactive engagement to... 
    Full time
    Work experience placement
    Work at office
    Immediate start
    Flexible hours

    WalkMe

    San Francisco, CA
    2 days ago
  •  ...boundaries of what's possible in video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and...  ...solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure... 

    Genmo

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...our vision at Postman.About the RoleWe're building out the Customer Success Engineering team at Postman, and we need business-minded engineers who...  ...solutions that benefit our entire enterprise customer base.As a CSE, you consult on the technical implementation path. When a... 
    Interim role
    Work at office
    Flexible hours
    3 days per week

    Postman

    San Francisco, CA
    1 day ago
  • $170k - $250k

     ...This company is looking for an engineer to work directly with the CTO on complex GPU virtualisation challenges. The role...  ...significantly more efficient for cloud customers You will be doing performance...  ..., and distributed GPU cluster management You will be working... 
    Full time
    Visa sponsorship
    Flexible hours
    San Francisco, CA
    1 day ago
  • $179k - $218k

     ...meaningful work of your career, help our customers and partners advance their AI...  ...a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive...  ...predictive strategies needed to maintain peak cluster health.The Strategic BridgeFor DC... 
    Temporary work

    Crusoe

    San Francisco, CA
    3 days ago
  •  ...Customer Success Engineer San Francisco, CA & New York, NY Who We Are At Pave, we're building the industry's leading compensation platform...  ..., Bessemer Venture Partners, and Craft Ventures. The CSE Team @ Pave: You'll become a Pave product and compensation... 
    Immediate start
    Flexible hours
    3 days per week

    P. A.V. E.

    San Francisco, CA
    1 day ago
  • $300k

     ...GPU Optimisation Engineer — Real-Time Inference Want to push GPU performance to its limits — not in theory, but in production systems handling...  ...kernel fusion, quantisation, and scheduling Writing and tuning custom CUDA / Triton kernels for performance-critical paths... 
    Relocation
    Visa sponsorship
    Free visa

    techire ai

    San Francisco, CA
    3 days ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one...  ...Function of PositionAs a Senior Systems GPU Engineer - AI & Robotics, you will be...  ..., please note that Intuitive and/or your customer(s) may require that you show current... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    1 day ago
  • MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models. The ideal candidate has over 4 years of experience with GPU kernels... 

    MakerMaker.AI

    San Francisco, CA
    7 hours ago
  • $100k - $120k

     ...cheaper and faster. Responsibilities Lead a team of kernel and system engineers focused on performance-critical code Design, implement, and optimize custom compute kernels for CPU (AVX/ARM NEON), GPU (CUDA/ROCm), and hardware accelerators Find bottlenecks in memory hierarchy... 

    Coda Robotics

    San Francisco, CA
    17 hours ago
  • San Francisco Tensor Company is seeking a Founding GPU Kernel Engineer to enhance GPU performance for AI applications. You will optimize and write kernels while collaborating with compiler teams to improve efficiencies across architectures. The ideal candidate has deep... 
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    17 hours ago
  • $285k - $315k

    About The Role We're looking for a Founding GPU Kernel Engineer who lives right at the boundary between hardware and software. Someone who thinks...  ...parallelism, ZeRO-style optimizations Familiarity with custom accelerators: TPUs, Trainium, Inferentia, or similar... 
    Full time
    Work at office
    Relocation package

    SF Tensor

    San Francisco, CA
    7 hours ago
  • $315k

    A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams... 

    Anthropic

    San Francisco, CA
    2 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal... 

    Baseten

    San Francisco, CA
    3 days ago
  • $285k - $315k

    SF Tensor is looking for a Founding GPU Kernel Engineer in San Francisco, specializing in GPU architecture and kernel optimization for machine learning workloads. The ideal candidate has deep expertise, proven capabilities in hand-optimizing performance-critical kernels... 
    Full time
    Relocation package

    SF Tensor

    San Francisco, CA
    7 hours ago
  •  ...transform your dreams into reality. You would collaborate with software engineers, AI researchers, and hardware specialists to develop high-...  ...future of mobility. Key Responsibilities Optimize end-to-end GPU performance for real-time autonomous driving workloads, including... 

    Bot Auto

    San Francisco, CA
    2 days ago
  • Mirai Labs in San Francisco seeks engineers to join a senior team building the full on-device stack for real-time local intelligence. You...  ...language models work, and experience in writing high-performance GPU kernels or Rust systems programming. We welcome applications from... 
    Local area

    Mirai Labs

    San Francisco, CA
    2 days ago
  • $285k - $315k

     ...AMD. We are partnering with researchers, engineers, and organizations who share our belief...  ...About the Role We're hiring a Founding GPU Compiler Engineer to build the core compilation...  ...classic compiler optimizations customized for large-scale training (fusion, tiling... 
    Full time
    Work at office
    Relocation package

    San Francisco Tensor Company

    San Francisco, CA
    7 hours ago
  • $220k - $320k

    inference.net, a growing company in San Francisco, seeks an experienced engineer to optimize AI inference performance. The ideal candidate will have over 2 years of experience in ML systems and GPU programming. Key responsibilities include implementing optimization techniques... 

    inference.net

    San Francisco, CA
    3 days ago
  • $180k - $280k

     ...tier investors. Since mid-2024, we've been engineering the foundation for what comes after the...  ...done. About the role We're looking for a GPU kernel engineer with deep, low-level CUDA...  ...more efficient. You'll write and optimize custom kernels, profile and eliminate... 
    Work at office
    Visa sponsorship
    Shift work

    TypeSafe AI

    San Francisco, CA
    7 hours ago
  •  ...Customer Success Engineer San Francisco, CA & New York, NY Who We Are At Pave, we're building the industry's leading compensation platform...  ..., Bessemer Venture Partners, and Craft Ventures. The CSE Team @ Pave: You'll become a Pave product and compensation... 
    Immediate start
    Flexible hours
    3 days per week

    Pave

    San Francisco, CA
    1 day ago
  • $315k

     ...committed researchers, engineers, policy experts, and...  ...breakthrough innovations in GPU performance and...  ...the-art techniques from custom kernel development to...  ...infrastructure, fault tolerance, cluster orchestration...  ..., we aren't able to successfully sponsor visas for every... 
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    more than 2 months ago
  •  ...parallelism across heterogeneous hardware to maximize throughput in production. You will work with state-of-the-art engines, profiling tools, and advanced GPU techniques while collaborating with a hands-on team in a fast-paced SF office environment. #J-18808-Ljbffr... 
    Work at office

    Theory Ventures

    San Francisco, CA
    1 day ago
  •  ...sized, and leadership is earned by shipping excellence. We seek engineers with strong intrinsic drive, a true passion for advancing the...  ...leverage your knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI. Full-Time On-site at either... 
    Full time
    Work at office

    Vast.ai Inc.

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Customer Success Engineer (CSE), GPU Cluster. Be the first to apply!