Principal AI SoC Runtime Software Architect
$200k - $300kVelaura
Role Overview
We are looking for a Principal AI SoC Runtime Software Architect to own the software architecture that turns Velaura's heterogeneous AI SoC into a coherent, high-performance execution platform.
This role will define and lead development of the end-to-end runtime spanning sensor ingest, preprocessing, AI inference, postprocessing, and delivery of results to robotics and other physical and edge AI applications. The runtime will coordinate execution and data movement across the SoC's CPU cores, AI accelerator, vision and multimedia engines, and other embedded processors. The complete runtime must coordinate heterogeneous workloads, manage ownership, synchronization, and safe reuse of shared data buffers, minimize data movement, provide predictable low-latency execution, recover from failures, and expose cohesive APIs and observability to applications and SDK components.
The ideal candidate combines deep runtime and systems-software expertise with a strong understanding of heterogeneous compute, Linux kernel and driver interfaces, DMA and shared-memory architectures, and performance-sensitive AI or multimedia pipelines. This will be a hands-on principal architect and technical lead who directs engineers across runtime, kernel, driver, firmware, multimedia, and SDK development.
Responsibilities
Define the SoC-wide execution model for coordinating workloads across heterogeneous compute and media engines, including dependency management, resource arbitration, priority and QoS, and concurrent pipeline behavior.
Set the architecture and technical direction for the AI inference runtime, including compiled-model execution, integration with industry-standard AI execution frameworks, and interfaces to applications and the broader Velaura SDK.
Own the end-to-end dataflow architecture for sensor-to-application pipelines, ensuring that camera, media, preprocessing, inference, and postprocessing components operate as an efficient and coherent system.
Define the SoC-wide memory and buffer-sharing architecture across user space, the kernel, and heterogeneous hardware engines, establishing clear ownership, coherency, isolation, synchronization, and lifecycle semantics.
Define the division of responsibility and interface contracts among the runtime, kernel drivers, firmware, and hardware engines, including execution, completion, telemetry, fault management, and recovery semantics.
Partner with the compiler team to define the compiler-runtime contract, ensuring compiled artifacts contain the metadata and execution information needed for the runtime to load, validate, execute, profile, and maintain compatibility across releases.
Establish a system-wide observability and performance architecture that correlates behavior across software and hardware layers and enables optimization against latency, throughput, bandwidth, power, utilization, and predictability goals.
Define the runtime resilience and validation architecture, including fault-containment and recovery policies, architecture-level acceptance criteria, and qualification across correctness, concurrency, compatibility, performance, and sustained workloads.
Required Qualifications
Extensive experience designing and building production runtime systems, embedded middleware, multimedia frameworks, or other performance-critical systems software.
Strong C/C++ programming skills and demonstrated ability to architect and contribute hands-on to production runtime software spanning application-facing APIs, user-space libraries, and low-level driver, firmware, and hardware interfaces.
Strong understanding of heterogeneous and asynchronous execution, including command submission, queues, events, dependencies, synchronization, concurrency, scheduling, and resource management.
Strong understanding of device memory, DMA, IOMMU/SMMU, cache coherency, memory mapping, shared buffers, buffer lifetimes, and kernel/user-space memory interfaces.
Experience optimizing end-to-end data movement and execution across multiple hardware engines rather than focusing solely on individual kernels or accelerator performance.
Experience designing stable runtime APIs with well-defined compatibility, versioning, error handling, diagnostics, and recovery behavior.
Demonstrated ability to debug complex cross-layer correctness and performance problems using disciplined, data-driven methods, profiling, and tracing.
Demonstrated technical leadership across component and organizational boundaries, including translating system requirements into clear architectures, interfaces, implementation guidance, and validation strategies.
Preferred Qualifications
Direct experience developing or extending AI inference runtimes or execution providers, such as ONNX Runtime, TensorRT-like runtimes, OpenVINO, TensorFlow Lite delegates, Qualcomm QNN/SNPE, TVM runtimes, or comparable systems for NPUs, GPUs, DSPs, or other accelerators.
Linux kernel development or upstream contribution experience involving device, accelerator, media, or shared-memory subsystems.
Experience with robotics, autonomous systems, edge AI, ROS 2, camera pipelines, ISP integration, V4L2/media, GStreamer, or other sensor-driven workloads.
Familiarity with compiled-model artifacts, quantized execution, tensor layouts, graph partitioning, memory planning, and compiler/runtime integration.
Experience with embedded Linux SDKs, production deployment, long-term runtime/API compatibility, or the security and isolation requirements of multi-process accelerator systems.
$200,000 - $300,000 a year
Compensation & Benefits
At Velaura, we believe exceptional talent deserves exceptional rewards. Compensation for this role includes a base salary and very competitive equity compensation, allowing team members to share in the company's long-term success.
The base pay listed represents a good faith estimate that the Company reasonably expects to pay for this position at the time of hire. Actual compensation will depend on multiple factors, including the candidate's skills, qualifications, relevant experience, and geographic location.
In addition to base salary and equity compensation, Velaura offers a comprehensive benefits package that may include medical, dental, and vision coverage; paid time off; flexible work arrangements; professional development opportunities; and other benefits designed to support the well-being and growth of our team.
Velaura is committed to pay equity and transparency and regularly benchmarks compensation to ensure we remain competitive in the market.
Why Velaura?
Velaura is building next-generation compute technology for cloud, edge, and PhysicalAI. Our solutions will enable robots, autonomous systems, drones, and other intelligentmachines to operate efficiently in the physical world.
This is an opportunity to help build foundational technology at a time when the industry is undergoing fundamental change. You will work alongside experienced leaders, architects, engineers, and operators who have delivered industry-defining products across mobile, cloud, semiconductor, and AI platforms. If you enjoy solving difficult problems, working across disciplines, and helping shape thefuture of Physical AI.
Equal Employment Opportunity and Accommodations
Velaura is an Equal Opportunity Employer that is committed to inclusion and diversity.Qualified applicants will receive consideration for employment without regard to race,color, religion, national origin, gender, sexual orientation, gender identity, disability orprotected veteran status. We also take affirmative action
#J-18808-Ljbffr- ...paths, cache policies, runtime integrations, performance... ...architecture, runtime software, operating-system... ...technical owner for the AI-HLC architecture. Primary... ...responsibilities ~ Architect the model-aware residency... ...expert storage, SoC-side decode, pinned residency...PrincipalLocal areaRemote work
- ...We are seeking a Software Architect to define and lead the technical vision... ...and building a multi-tenant, AI-native platform that powers... ...a platform that reasons. As Principal Software Architect, you will... ...harness engineering: agent runtimes, tool/function calling, structured...PrincipalLocal area
$262.5k - $394k
...Principal Software Architect, Finance Technology Platform Sunnyvale, California, United States Corporate... ...and building a multi-tenant, AI-native platform that powers financial... ...including harness engineering: agent runtimes, tool/function calling, structured output...PrincipalContract workTemporary workLocal areaRelocation- ...Software Engineer Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon... ...on production-grade embedded runtime environments. You'll work... ...with ML accelerators, GPU, CPU, SoC architecture and micro-architecture...SuggestedFull timeFor contractorsFor subcontractorCasual workWork at officeRemote workDay shift
$195.2k - $361.2k
...explored architectures, algorithms, and software inspired by the brain's... ...power the coming era of physical AI systems-beyond the reach of GPUs... ...systems. In this role, you will architect and lead the development of the firmware, runtime, and performance infrastructure,...SuggestedFull timeWork at officeLocal areaImmediate startShift work$170k - $277k
...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do... ...are looking for a visionary Senior Principal Engineer/Architect to serve as the technical authority... ...with a deep empathy for internal software developers as your primary customers....PrincipalFull timeWork at officeVisa sponsorshipWork visaFlexible hours- ...demand for new data centers and AI compute is rapidly... ...Overview We are seeking a Principal Network Architect to own the datacenter scale... ...that both digital design and software teams implement against. This... ...semantics exposed to the runtime, in partnership with the software...PrincipalFull time
- ...Learning and HPC. We're seeking a Senior Software Architect to help co‑design next‑gen data center... ...communication technologies to accelerate AI and HPC workloads. Explore innovative... ..., SHMEM) and at least one communication runtime (MPI, NCCL, NVSHMEM, OpenSHMEM, UCX, UCC...
$272k - $431.25k
...We are now looking for a Principal Software Engineer for LPX System Software! NVIDIA’s LPX System... ...layers, core system libraries, drivers, and runtime components that workloads enter the... ...An established habit: building with AI coding agents — not as a novelty, but as...PrincipalShift work- ...the potential of generative AI to power the transformation of... .... We are at the forefront of software and hardware innovation, pushing... ...per week. The role: Principal System Software Engineer, AI... ...Experience with deep learning runtimes (such as ONNX Runtime, TensorRT...PrincipalWork experience placement3 days per week
- ...NVIDIA seeks a Principal Architect for SoC and System performance modelling to lead modeling, analysis and validation of cutting-edge architectures. You will collaborate across multiple modeling teams to design, implement and test models, mentor engineers, and drive feature...Principal
$220k - $280k
...Credo is looking for a Principal AI System Architect to join our team in San Jose, CA, reporting to AVP, XPU system AI interface. This role needs... ...skills, daily experience working with compiler, software, SOC & backend team. Preferred Qualifications Master's...PrincipalWork experience placement- ...products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded... ...intelligence, we are a powerhouse - a cutting-edge 'AI Software Solutions Team'. Specialized in AI optimization, fine-tuning large...Principal
$272k - $431.25k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...legacy of innovation! We are seeking a Principal SW Architect, Networking to build the future of AI... ...development of AI networks alongside hardware, software, and applicationsGuaranteeing...PrincipalFull time$272k - $431.25k
...Principal System Software EngineerNVIDIA is seeking a highly motivated Principal System Software Engineer... ...with hardware, architecture, kernel, AI, middleware, and platform teams to... ...including kernel, drivers, middleware, runtime frameworks, and platform services.Collaborate...Principal- ...Lexar Enterprise seeks a senior systems architect to own the end-to-end technical architecture for... ...cache policies, and performance models across runtimes and hardware. You will bridge model architecture, runtime software, OS memory management, and accelerator execution...
$206.4k - $379.1k
...custom multimedia generative AI - deep-tuned image, video, and... ...Premiere. We are hiring a Principal Machine Learning Engineer to... ...how our generative models are architected, optimized , and served at... ...MLOps technologies - serving runtimes, quantization , GPU scheduling...PrincipalTemporary workLocal areaWorldwide- ...from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP... ...under tapeout pressure and propose solutions that balance accuracy, runtime, and methodology constraints ~ Strong understanding of static...Principal
$278.1k - $417.1k
...next generation of AI-driven game... ...modern, browser-native runtime (built on technologies... ...runtime. As our Principal Engineer for On-... ...the renderer. Architect inference systems... ...~8+ years in software/ML engineering, with... ...hardware: mobile SoCs (Apple Neural Engine...PrincipalWork at officeWorldwideRelocation package$152k - $241.5k
...amazing people. NVIDIA is looking for an experienced Senior Software and System Architect to join our Networking Software Architecture group. Nvidia'... ...This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to...Remote work- ...generation computing experiences—from AI and data centers, to PCs,... ...We are seeking a Robotics AI Architect to define and scale next‑... ...the roadmap for our AI SDKs, runtime, and reference architectures to... ...systems Influence compute‑software co‑design across CPU, GPU, and...Principal
- ...performance silicon chips and software content. Join us to transform... ...can work directly with customer architects to close real system-level... ...You are comfortable discussing SoC integration, IP configuration,... ...technical partner for next-generation AI, HPC, storage, networking, and...Principal
$180k - $225k
..., a leading global semiconductor company, is rapidly scaling its AI-focused power delivery business to capture a dominant share of the... ...-world implementation challenges ~ Experience with embedded SoC’s including cross-functional collaboration with HW/FW/SW teams...PrincipalFull time$175.8k - $293k
...competitive advantage. We're looking for a Principal AI Engineer to architect, build, and harden the agentic AI... ...context management, orchestration runtimes, and inference serving. Evaluate and... ...scalable, production-grade software, with significant recent depth in AI...Principal$184k - $287.5k
...As a Senior Software Architect in the GPU Networking Architecture team, you will define Software Defined Networking (SDN) architectural solutions... ...fields related to the modern data center, such as distributed AI and deep learning systems, Networking Operating Systems,...$174k - $252k
...propose new hardware features.Collaborate across Hardware, Driver, Runtime, and Performance Analysis teams and many other stakeholders.... ...or equivalent practical experience.5 years of experience with software development in one or more programming languages.3 years of experience...- ...research team within NVIDIA’s Networking Systems & Software Architecture group is solving some of AI’s hardest infrastructure problems. The team builds systems... ...quantum computing interconnects. The Senior Architect role is to own modules and projects end-to-end—from...
$140k - $190k
...Principal Embedded Software Engineer WiFi team is looking for a Principal Embedded Software Engineer with C programming and networking knowledge to join our team. This is a great opportunity to immerse yourself in all phases of the software development cycle to reach...PrincipalFull timeWorldwide- ...Principal DevOps EngineerTENEX is an AI-native, automation-first, built-for-scale Managed... ...You'll work closely with Software Engineering, AI/ML, and Security... ...and compliant (e.g., SOC 2, ISO 27001)... ...).Proven track record of architecting, building, and operating...PrincipalLive inWork at officeWork from homeRelocationMonday to Thursday
- ...alert: Select how often (in days) to receive an alert: Principal Hardware Systems Architect-Compute Platforms Req ID: 140082 Region: Americas... ...for the architecture and development of next-generation AI, accelerated computing, and high-performance data center...PrincipalLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal AI SoC Runtime Software Architect. Be the first to apply!
- application architect Santa Clara, CA
- senior software architect Santa Clara, CA
- software architect Santa Clara, CA
- .net software architects (remote) Santa Clara, CA
- senior principal cloud computing engineer Santa Clara, CA
- principal Santa Clara, CA
- principal cloud computing engineer Santa Clara, CA
- senior principal scientist Santa Clara, CA
- principal architect Santa Clara, CA
- internship software Santa Clara, CA



