Principal Systems Software Engineer
$258k - $275kSan Diego Stealth Startup
Job Description
Job Description
Principal Systems Software Engineer
Location: San Diego, CA
Job Type: Full-Time
Salary Range: $258,000 - $275,000
Position Overview
We are looking for a Principal Engineer to architect, build, and own the end-to-end data pipeline that drives our high-throughput diagnostic instrument platform — from real-time image acquisition on the instrument, through GPU-accelerated signal processing, to offloading for secondary and tertiary analysis on local HPC clusters and cloud infrastructure.
This is a technical leadership role for an engineer who can design and deliver industrial-grade data processing infrastructure that operates reliably at sustained high throughput. You will be responsible for the full data path: acquiring raw image data from sensors, processing it through GPU pipelines, orchestrating job distribution across local HPC and cloud compute, and ensuring the entire system handles errors, backpressure, and recovery gracefully. The scope spans instrument- embedded software, on-premises Linux HPC infrastructure, and cloud- based compute and storage.
The central challenge of this role is not raw compute optimization — GPU and CPU resources will have adequate headroom. The challenge is building a pipeline architecture that is robust, scalable, and evolvable as instrument throughput increases with each generation, the number of instruments grows, and data volumes scale accordingly. You will design systems that keep a complex multi-stage pipeline running continuously and reliably in a production lab environment, and that can be evolved without wholesale re-architecture as requirements intensify.
Key Responsibilities
End-to-End Data Pipeline Architecture
- Own the architecture of the complete data path from image acquisition to final processed output
- Design pipeline stages with clear interfaces, flow control, and backpressure mechanisms
- Ensure the pipeline sustains continuous high-throughput operation across extended instrument runs
- Define data formats, handoff protocols, and buffering strategies between pipeline stages
- Architect for graceful degradation — the system must handle transient failures without data loss or pipeline stalls
- Establish performance budgets and SLAs for each pipeline stage and monitor adherence
Image Acquisition & On-Instrument Processing
- Develop and optimize real-time image acquisition from high-speed sensors on the instrument
- Implement low-latency, high-bandwidth data capture with minimal frame loss
- Design on-instrument preprocessing stages that reduce data volume before offload
- Manage memory and storage constraints within the instrument compute environment
- Ensure deterministic, repeatable performance under sustained acquisition loads
GPU-Accelerated Signal & Image Processing
- Develop and maintain GPU compute pipelines using CUDA for signal and image processing
- Implement DSP algorithms including frequency-domain analysis, deconvolution, filtering, and detection
- Manage host-to-GPU data transfers and ensure efficient use of GPU resources
- Profile GPU workloads to identify issues and validate performance headroom
- Balance numerical accuracy against throughput requirements
Job Orchestration & Distributed Processing
- Design and implement job queuing, scheduling, and orchestration across instrument, local HPC, and cloud compute
- Build robust work distribution that maximizes resource utilization across heterogeneous compute
- Implement backpressure handling so upstream stages throttle gracefully when downstream is saturated
- Design comprehensive error handling, retry logic, and dead-letter strategies for failed jobs
- Ensure jobs are idempotent and recoverable — partial failures must not corrupt the pipeline
- Implement priority scheduling to balance real-time instrument processing with batch reprocessing
- Monitor queue depths, processing latencies, and resource utilization with actionable alerting
Linux Systems & Performance
- Configure and tune Linux systems for reliable, high-throughput operation across instrument and HPC nodes
- Tune kernel parameters (scheduler, NUMA, IRQs, huge pages) as needed for stable pipeline performance
- Understand and manage DMA paths, PCIe topology, and device-to- memory data movement
- Profile and diagnose system-level issues using perf, ftrace, eBPF, and similar tools
- Ensure system configurations are reproducible and documented across instrument and HPC environments
HPC Compute Platform & Algorithm Infrastructure (co- owned with DevOps)
- Co-design the HPC compute platform architecture with DevOps — define computational requirements, job flow, and data access patterns while DevOps provisions and manages the infrastructure
- Define how algorithms are deployed, versioned, and rolled into production on the HPC platform — support safe side-by-side execution of new and existing algorithm versions
- Design compute allocation strategies that balance real-time instrument processing, batch algorithm development/validation, and historical data reprocessing
- Design the data handoff between instrument-side processing and
- HPC/cloud compute — formats, staging, transfer protocols
- Define storage tiering requirements for the processing pipeline — what data stays hot for active processing, what moves to warm for algorithm development access, and what archives to cold
- Specify when and how workloads should burst from local HPC to cloud (AWS) based on pipeline load and priority
- Optimize data movement across high-speed networks (RDMA, InfiniBand, high-speed Ethernet) between instrument, HPC, and storage
- Design for scalability — the architecture must accommodate increasing instrument throughput, additional instruments, and growing algorithm complexity
Reliability & Observability
- Instrument every pipeline stage with metrics, logging, and tracing
- Build real-time dashboards showing pipeline health, throughput, latency, and queue state
- Design automated recovery mechanisms for common failure modes
- Implement data integrity checks and validation at pipeline stage boundaries
- Support root-cause analysis and post-mortem investigation for pipeline incidents
- Establish runbooks and operational procedures for pipeline operations
Qualifications
Education:
- PhD – 10 yrs, MS – 14 yrs and BS – 17 yrs of experience in Computer Science, Electrical Engineering, or related field.
Experience & Technical Leadership
- 15+ years of professional software engineering experience in performance-critical systems
- Track record of architecting and delivering complex, multi-stage data processing pipelines
- Demonstrated technical leadership — ability to drive architecture decisions and mentor engineers
- Experience operating systems at industrial-grade reliability and throughput requirements
Systems Programming & GPU Computing
- Expert-level C/C++ and systems programming on Linux
- Solid experience with CUDA programming and GPU pipeline development (required)
- Strong understanding of computer architecture: CPU caches,
- NUMA, memory hierarchies, PCIe, DMA
- Experience with Python for tooling, orchestration, and pipeline glue
- Experience with performance profiling and diagnostics tools (perf, ftrace, Nsight, or similar)
Pipeline & Orchestration
- Experience designing multi-stage data pipelines with flow control, buffering, and backpressure management
- Strong understanding of error handling, retry strategies, and fault recovery in performance-critical systems
- Experience with job scheduling and work distribution across heterogeneous compute resources
- Familiarity with workflow orchestration frameworks (Airflow, Celery, custom solutions, or similar) is a plus
Signal Processing & Algorithms
- Practical experience implementing DSP or image processing algorithms in production systems
- Familiarity with frequency-domain analysis, filtering, and detection algorithms
- Ability to reason about numerical accuracy and throughput tradeoffs
Data Movement, Storage & Networking
- Experience optimizing data transfer across high-speed networks (RDMA, InfiniBand, high-speed Ethernet)
- Understanding of shared storage architectures, tiered storagestrategies, and high- throughput data staging
- Experience defining compute platform requirements and collaborating effectively with infrastructure teams
- Familiarity with algorithm deployment and versioning in production computing environments
Preferred:
- Experience with high-throughput diagnostic instrument, imaging, or scientific instrument data pipelines
- Experience scaling a data pipeline through multiple hardware or throughput generations
- Experience with GPUDirect RDMA or other hardware offload technologies
- Familiarity with real-time or low-latency Linux variants
- Background in scientific computing, computational physics, or bioinformatics
- Experience designing systems that span embedded instrument software and datacenter infrastructure
What Success Looks Like
- The end-to-end pipeline from image acquisition to processed output runs continuously and reliably at target throughput
- Backpressure and error handling work transparently — operators are not firefighting pipeline stalls
- Job orchestration seamlessly distributes work across local and cloud compute based on load and priority
- Pipeline performance is predictable, measurable, and well understood with clear per-stage metrics
- New instrument generations with higher data rates can be accommodated through evolution, not redesign
- Adding instruments to the lab scales the pipeline without disproportionate complexity or operational burden
- Algorithm developers can deploy, test, and validate new algorithms on the HPC platform without disrupting production processing
- Storage tiering keeps the right data accessible at the right cost as volumes grow
We are an equal opportunity employer. We thrive on diversity and collaboration.
$211k - $297k
...Debt Solutions - Principal Software Engineer (React/C#/AWS) Job Description Overview CoStar Group (NASDAQ: CSGP) is a leading global provider... ...to own the architecture and design of our software systems, from full stack web products to high-volume, secure data...SuggestedWork at officeWork from homeMonday to Thursday$114k - $171k
...have incredible opportunities to work on revolutionary systems that impact people's lives around the world today,... ...today.We are looking for you to join our team as a Principal/Sr. Principal Embedded Engineer Software based out of San Diego, CA. What You'll Get to Do:In...SuggestedFull timeRelocation packageShift work$114k - $171k
...incredible opportunities to work on revolutionary systems that impact people's lives around the world today,... ...Northrop Grumman Aeronautics Systems is looking to add a Principal or a Sr. Principal Embedded Software Engineer to join our team of qualified and diverse...SuggestedFull timeRelocation packageShift work$114k - $171k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is looking to add a Principal Software Engineer - User Experience Applications to join our team of...SuggestedFull timeRemote workRelocation packageShift work$91.8k - $137.6k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ..., they're making history.Northrop Grumman is seeking a Software Engineer/Principal Software Engineer - Front End to support our Mission...SuggestedFull timeRelocation packageFlexible hoursShift work$91.8k - $137.6k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...history.Northrop Grumman Aerospace Systems is seeking a Software Engineer/Principal Software Engineer - DevSecOps to join the Ground Segment...Full timeInterim roleRelocationFlexible hoursShift work$116.3k - $213.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is seeking a Senior Principal Software Engineer to join our team on site in San Diego, CA or Oklahoma City...Full timeRemote workRelocation packageShift work$150k - $220k
Job description: Senior or Principal Software Engineer (Full-time, On-site, San Diego, CA)We are seeking an experienced Principal orSenior Embedded... ...to design and develop embedded software for space-based systems, including computer boards. This is a hands-on, on-site...Full time$142.2k - $213.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is seeking a Senior Principal Software Engineer - AI Software to join our onsite team on site in San Diego...Full timeRemote workRelocation packageShift work$91.8k - $137.6k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is looking to add a Software Engineer / Principal Software Engineer to join our team of qualified diverse...Full timeRemote workRelocation packageShift work$142.2k - $213.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...Northrop Grumman Aeronautics Systems is looking to add Sr Principal Engineer Software - Integration & Test engineer to join our team of...Full timeInterim roleRelocation packageShift work$156.4k - $211.6k
...We are looking for a software engineering leader who is passionate about creating next-generation... ...regulated industry. Additionally, the Senior Principal Software Engineer will bring deep... ...with modern version control systems (e.g., Git) and tools (e.g., Bitbucket...Full timeTemporary workLocal areaFlexible hours$116.3k - $213.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...Grumman Aeronautics Systems is looking to add a Sr Principal Software Engineer- Thread Lead to join our team on site in San Diego, CA, El...Full timeImmediate startRemote workRelocation packageShift work$91.8k - $137.6k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is looking to add a Software Engineer/Principal Software Engineer - Java Developer to join our team on...Full timeRemote workRelocation packageShift work$91.8k - $137.6k
...have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...Grumman Aeronautics Systems is looking to add a Software Engineer - Java or Principal Software Engineer - Java to join our team on site in San...Full timeRemote workRelocation packageShift work$149.3k - $234.6k
...opportunities to work on revolutionary systems that impact people's lives around the world... ...Grumman brings informed insights and software-secure technology to enable strategic planning... ...team as a Sr. PrincipalCyber Software Engineer. Places of performance for this position...Full timeRelocationShift work$114k - $171k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and... ...Experience organization is seeking a highly motivated Principal AI Software Engineer to join our team onsite in El Segundo or Rancho Bernardo,...Full timeRelocation packageFlexible hoursShift work$204k - $284k
...Principal Software Reverse Engineer San Diego, CA STR is hiring a Principal Software Reverse Engineer who has a passion for research and analysis of vulnerabilities in cyber physical systems. This opportunity will be part of a multidisciplinary team of researchers...Full timeWork experience placementImmediate startNight shift$114k - $171k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is looking to add a Principal Software Engineer - Java Middle Layer to join our team of qualified, diverse...Full timeRemote workRelocation packageShift work- Designs and implements software features and applies special techniques necessary to perform geolocation-related operations, including... ...'s Degree (or foreign academic equivalent) in Electrical Engineering, Computer Engineering, Computer Science or related degree field...Full timeRemote work
$142.2k - $213.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...Grumman Aeronautics Systems has an opening for a Sr. Principal Software Engineer. This role is located on site in El Segundo, CA or San...Full timeRelocation packageShift work- Principal Software Engineer A VC-backed IoT security startup is seeking a Principal Software Engineer to join its growing team. In this role, you’ll report directly to the SVP of Engineering and gain broad visibility across the organization, working on impactful projects...Full time
$144.5k - $195.5k
We are looking for a software engineering leader who is passionate about creating next-generation... ...regulated industry. Additionally, the Principal Software Engineer will bring deep expertise... ...with modern version control systems (e.g., Git) and tools (e.g., Bitbucket...Full timeTemporary workLocal areaFlexible hours- ...solutions that enhance maritime operations through autonomous and intelligent platforms. Job Overview: We are seeking a System Software Engineer to design, implement, and optimize software systems for our autonomous surface vessels. This role will involve working...Permanent employmentTemporary workWork at office
$91.8k - $137.6k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ....Northrop Grumman Aeronautics Systems is looking to add a Software Engineer/Principal Software Engineer- Data Analyst to join our team on site...Full timeRemote workRelocation packageShift work$111.3k - $166.9k
...Company: Qualcomm Technologies, Inc. Job Area: Engineering Group, Engineering Group Software Engineering General Summary: The Snapdragon... ...Snapdragon? Join Qualcomm’s Thermal and Limits Management Systems Software team to work on bleeding edge Windows on...Full timeWork experience placementWork from home- ...thermal performance of various compute use cases. Perform system level software optimizations using hardware and software power/thermal management... ...power, performance and thermal domain to achieve critical engineering goals. Lead the ideation, prototyping, and development of...Full time
$87.1k - $157.45k
...division currently has an exciting opportunity for an Embedded Software Engineer to perform design, development, and hardware/software... ...state-of-the-art processing algorithms into real-time software systems. Projects involve small multi-disciplinary teams of engineers...Full time$155k - $210k
San Diego, CASoftware - Software Platforms and Product /Full-time /HybridZoox is looking for an embedded software engineer to join our Firmware Platforms team. In this role, you will be responsible for developing, extending, and maintaining support for multiple embedded...Full timeTemporary workRelocation package$155k - $185k
...For more information, visit . Job summaryThe Senior Embedded Software Engineer is responsible for the design, development, verification, maintenance... ..., drug delivery, pulse generation, motor control, UI/display systems, battery powered systems, cybersecurity, hardware bring-up,...Full timeTemporary work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Systems Software Engineer. Be the first to apply!
- principal software engineer San Diego, CA
- system programmer San Diego, CA
- IT system engineer San Diego, CA
- systems software developer San Diego, CA
- principal San Diego, CA
- senior principal cloud computing engineer San Diego, CA
- senior principal scientist San Diego, CA
- principal applied scientist San Diego, CA
- principal cloud computing engineer San Diego, CA
- principal scientist San Diego, CA



