Senior Production Engineer, Core PE
$300 per monthCrusoe
Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.About This Role:Crusoe is building the most reliable, energy-efficient, AI-optimized cloud platform — and Production Engineering sits at the heart of that mission. As a Production Engineer focused on Operational Excellence, you will help ensure the reliability, scalability, and performance of Crusoe’s GPU cloud that powers next-generation AI workloads.This role is ideal for engineers who enjoy solving complex production problems, improving large-scale distributed systems, and building automation that keeps infrastructure running smoothly. You’ll play a key role in strengthening the operational foundation of Crusoe’s cloud while helping scale infrastructure that supports demanding AI and HPC workloads.You’ll partner closely with Production Engineers, infrastructure teams, and platform engineers to improve system reliability, reduce operational toil, and drive continuous improvements across Crusoe’s rapidly growing GPU cloud.What You’ll Be Working On:Collaborate with cross-functional teams to define and evolve availability metrics for Crusoe’s cloud platform, including establishing, measuring, and improving SLIs and SLOsParticipate in production incident response, diagnosing and resolving service disruptions while contributing to post-incident reviews and root cause analysisBuild, operate, and improve observability across Crusoe’s infrastructure using tools such as Prometheus, Grafana, Alertmanager, and OpenTelemetryIdentify reliability risks, performance bottlenecks, and early indicators of potential production issues across distributed systemsDevelop automation and tooling that reduces operational toil, improves recovery times, and enables self-healing infrastructurePartner with compute, networking, storage, and platform teams to strengthen service resilience and disaster recovery capabilitiesContribute to improving operational processes, knowledge sharing, and reliability best practices across the engineering organizationContinue growing technical depth through mentorship, training, and hands-on work operating large-scale AI infrastructureWhat You’ll Bring to the Team:5+ years of experience in Production Engineering, SRE, or large-scale infrastructure operationsExperience supporting GPU workloads, HPC environments, or latency/throughput-sensitive distributed systemsStrong knowledge of Linux/Unix systems, including debugging complex issues across kernel and user spacePrevious experience in Infrastructure roles building or managing compute, storage or networking platformsUnderstanding of modern cloud infrastructure fundamentals including Kubernetes, distributed systems, virtualization, and cloud platforms (AWS/GCP)Familiarity with incident management practices and reliability frameworks (SRE, ITIL, or similar)Experience with monitoring and observability tools such as Prometheus and Grafana, or a strong desire to deepen expertise in this areaFamiliarity with infrastructure-as-code and configuration management tools such as Terraform or AnsibleScripting or programming experience with languages such as Go, Python, C, or C++Strong communication skills and the ability to collaborate across engineering teamsAbility to remain calm and effective while troubleshooting complex issues in high-impact production environmentsA growth mindset and strong interest in reliability engineering, automation, and operational excellenceBachelor's degree in Computer Science, Electrical Engineering, or a related technical field (or equivalent combination of education and experience)Bonus Points:Experience working with Kubernetes or container orchestration platforms at scaleExposure to change management processes, operational readiness reviews, or structured root cause analysisExperience designing self-healing systems, automated remediation, or event-driven operational toolingInterest in scaling AI or HPC infrastructure and solving reliability challenges in GPU-heavy environmentsPassion for mentorship, learning, and developing deeper expertise in Production EngineeringBenefits:Industry competitive payRestricted Stock Units in a fast growing, well-funded technology companyHealth insurance package options that include HDHP and PPO, vision, and dental for you and your dependentsEmployer contributions to HSA accountsPaid Parental LeavePaid life insurance, short-term and long-term disabilityTeladoc401(k) with a 100% match up to 4% of salaryGenerous paid time off and holiday scheduleCell phone reimbursementTuition reimbursementSubscription to the Calm appMetLife LegalCompany paid commuter benefit; $300 per monthCompensation:Compensation will be paid in the range of $172,000 – $209,000 + Bonus. Restricted Stock Units are included in all offers. Compensation will be determined by the applicant’s education, experience, knowledge, skills, and abilities, as well as internal equity and alignment with market data.Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.LocationSan Francisco, CA - US; Sunnyvale, CA - USEmployment TypeFull timeLocation TypeOn-siteDepartmentCloud Engineering
$170k - $210k
...scalability, and resilience. We deliver products to meet the demands of today while building... ...you. ABOUT THE ROLE: Software Engineers at Bright Machines are responsible for... ...biggest names in the industry. As a Senior Software Engineer, you will build scalable...SeniorFull timeWork at officeRemote workFlexible hours- ...it, drafting emails, running LinkedIn outreach, researching prospects, queuing follow-ups. The Core Product team owns this product surface, and we're hiring senior engineers who want a say in what it becomes. What you'll do Build AI features that drive revenue,...SeniorFull timeWork at office3 days per week
$140k - $200k
...exponential growth. Overview We're looking for a Senior Software Engineer to join our Core Experiences Team. This team builds and maintains the foundational services and SDKs that power Speechify’s product experience across platforms. It's a critical role for someone...SeniorFull timeRemote work$200k - $250k
...the world. StubHub is seeking Senior Software Engineers to design and develop next-generation... ...developing the team's commercial and product strategy. You will be expected to be equally... ..., high-throughput systems that power core seller operations and handle large,...SeniorFull timeWork at officeRemote workWorldwideFlexible hours- ...Startup Employers 2022 List . The Core Product Experience (CPX) team owns and elevates... ...Platform, AI, and other pods across Engineering to deliver cohesive, high-impact experiences... ...strategy. We're looking for a Senior Fullstack Software Engineer with strong...SeniorFull timeWork at officeImmediate startRemote workWork from homeMonday to FridayFlexible hours
- ...scalable science. About the Role The Core Infrastructure team makes the rest of Nooks fast. We're going from one product to many - Sequencing, Dialer, Agent Studio... ...hours a day in the product. We're hiring senior engineers to own the systems everything else is...SeniorFull timeWork at office3 days per week
$151k
.... is on a mission to advance human understanding. Our four products — Scribd®, Slideshare®, Everand™, and Fable — help billions... ...open position within the organization. We’re hiring a Senior Frontend Engineer to help drive Scribd’s Web Modernization program and build...SeniorFull timeFor contractorsLocal areaRemote workHome officeFlexible hours$180k - $290k
...building, and implementing changes to Stellar Core - the primary distributed system that is... ..., working alongside our CTO, our team of engineers, and our community of open source... ...Stellar Core, and help the team hit critical product milestones. Collaborate with the team...SeniorFull timeTemporary workWork at officeLocal areaWorldwideFlexible hours$165k - $240k
...processed over $500 billion in spend. Zip’s team includes product leaders from Apple, Airbnb, and Meta, as well as former procurement... ...Your Role We are looking for an experienced software engineer to join the Core Infrastructure team in San Francisco, CA. This role will...SeniorFull timeHome officeFlexible hours$165k - $247k
...Square, and Under Armour—build better products and digital experiences. With powerful... ...Team We’re looking for a Fullstack Engineer to join the Core Analytics team, the group behind our flagship... ...and scale the platform. As a Senior Engineer, you will: Work with...SeniorFull timeWork at officeWorldwideHome officeFlexible hours$200k - $250k
...1000+ customers in 60+ countries, strong product-market fit, and world-class investor support... ...to our customers — from leadership to engineers — and work together to solve real problems... ...As a Software Engineer on the Core Infrastructure team at Harvey, you'll play...SeniorFull timeInternshipRelocation package- ...Willkie Farr & Gallagher LLP is seeking a Senior Innovation Engineer to support the AI & Innovation team across Palo Alto, San Francisco, or Los Angeles offices. The role focuses on designing secure, scalable AI and low-code solutions, shipping Harvey workflows, and enabling...SeniorWork at office3 days per week
$161.3k - $241.9k
...generational company at a true inflection point. We have strong product-market fit and world-class investor support. We’re... ...workload. We’re looking for a Production Engineer to help build and operate Harvey’s core compute and networking infrastructure, Kubernetes platform...SeniorFull time$160k - $180k
...executive team of former founders and senior leaders from companies... ...looking for a Senior or Staff engineer to join engineering at Tatari,... ...is an hour not spent shipping product. The arrival of AI coding tools... ...occasionally contribute to core reliability and infrastructure...SeniorFull timeWork at officeLocal area2 days per week$200k - $250k
...the world. About the Role: We are seeking a Senior Software Engineer to join our Core Platform - Streaming & Storage team. In this role, you... .... About the Team: The team owns foundational products including message queues, DQL, retry infrastructure, message...SeniorFull timeWork at officeRemote workWorldwideFlexible hours- ...context." We're now building that context infrastructure for production-grade agents on the same open foundation, as agents become the... ..., everywhere. That everyone now includes AI agents. Engineering Hiring Sprint: We're growing our engineering team and are...SeniorFull timeWork at officeLocal areaImmediate startFlexible hours
$190.8k - $267.1k
...sources of information. For more information, visit . The Core Platform team facilitates millions of API requests running the... ...data models that are the sources of truth for Reddit. As a Senior Engineer on this team, you’ll be a technical coach and mentor. You’ll draw...SeniorFull timeFor contractorsWork experience placementRemote workFlexible hours- ...day is currently Tuesday. About the Role As a Senior Software Engineer on Lambda’s Core Cloud Platform team, you will build the control-plane... ...business-critical control-plane services. Debug complex production issues across distributed services, infrastructure...SeniorFull timeWork experience placementWork at officeLocal areaWork from homeFlexible hours
$124k - $171k
...but building what comes next.Job Title:Senior AI Automation EngineerCompany:PrologisA day in the lifeThe Senior AI Automation Engineer accelerates the delivery and adoption of... ...value opportunities and translate them into production capabilities. The engineer leads work...SeniorFull timeContract workFor contractors- Purple DriveSENIOR TEST AUTOMATION ENGINEERShare Contractual San Francisco, CA PDTOverview:Strong understanding of Software quality testing and Software development cycle. Able to review test plan and test design, and lead a group of testers Knowledgeable on tools like ...Senior
$300 per month
...other, come build with us at Crusoe.About the Role:At Crusoe, our Production Engineering team ensures the reliability and scalability of Crusoe’s AI-optimized cloud platform. We’re looking for a Senior Production Engineer with a strong background in distributed systems...SeniorTemporary work$300 per month
...Role:At Crusoe Energy Systems, we are building the most sustainable, AI-first cloud infrastructure, and our Compute-focused Production Engineers are the backbone of that mission. This role is centered on supporting virtualization, hypervisor, and kernel-level performance...SeniorTemporary work- ...About the Role We’re looking for a Senior Backend or Full-Stack Software Engineer (5+ years experience) who wants to build the core systems that power last-mile healthcare delivery. At Sprinter, you’ll work on products that blend logistics, patient experience, safety...SeniorFull timeTemporary workWork at officeMonday to FridayMonday to ThursdayFlexible hours
- ...Senior Software Engineer, Core Platform San Francisco, CA & New York, NY Who We Are At Pave, we're building the industry's leading compensation... ...team sits at the center of Pave Engineering, bridging product and platform engineering. We build the foundational...SeniorFlexible hours3 days per week
$180k - $250k
...Senior Software Engineer, Core Platform London, England, United Kingdom; New York, New York, United States; San Francisco, California, United... ...deploying AI systems—designed to take ideas from research to production with less friction. Through our merger with Voltage...SeniorWork at officeWork from homeFlexible hours2 days per week$10k - $20k
...streamline operations, improve customer experience, and stay competitive through innovative products and services. The company is looking for a Network Automation Engineer to design and build high-speed, low-latency network fabrics for large-scale GPU-accelerated computing...SeniorFull time- ...Horowitz to Blackrock and Fidelity, and employs a team of 450 engineers and entrepreneurs. Astranis designs, builds, and operates... ...153,000 sq. ft. headquarters in Northern California, USA.Senior Hardware/Production Test Software Engineer We are seeking a highly skilled and...SeniorPermanent employmentFlexible hours
- ...compensation platform company in San Francisco is seeking experienced backend engineers to build and maintain robust, data-rich systems for compensation processes. Ideal candidates have strong product intuition and hands-on experience with technologies like Node.js and...SeniorFlexible hours
$250k - $300k
..., come build with us at Crusoe.About the Role:As a Senior Staff/Principal Deployment Automation Engineer for the Compute Team, you will be responsible for deployment... ...to coordinate canary deployments on live production systems, run Blue/Green testing, and perform automatic...SeniorTemporary work$170k - $205k
...at Crusoe. About This Role: We are seeking a Senior Backend Software Engineer to join the Core Backend team within Cloud Customer Experience (CCX).... ...team built and owns the quota model for all Crusoe products and its enforcement path, along with the request and...SeniorTemporary work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Production Engineer, Core PE. Be the first to apply!
- senior security operations engineer San Francisco, CA
- post production engineer San Francisco, CA
- network operations center engineer San Francisco, CA
- production network engineer San Francisco, CA
- remote operation drilling engineer San Francisco, CA
- production operations engineer San Francisco, CA
- data center operations engineer San Francisco, CA
- application operations engineer San Francisco, CA
- senior production engineer San Francisco, CA
- operations engineer San Francisco, CA



