Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal AI SoC Runtime Software Architect

$200k - $300k

Velaura

Role Overview

We are looking for a Principal AI SoC Runtime Software Architect to own the software architecture that turns Velaura's heterogeneous AI SoC into a coherent, high-performance execution platform.

This role will define and lead development of the end-to-end runtime spanning sensor ingest, preprocessing, AI inference, postprocessing, and delivery of results to robotics and other physical and edge AI applications. The runtime will coordinate execution and data movement across the SoC's CPU cores, AI accelerator, vision and multimedia engines, and other embedded processors. The complete runtime must coordinate heterogeneous workloads, manage ownership, synchronization, and safe reuse of shared data buffers, minimize data movement, provide predictable low-latency execution, recover from failures, and expose cohesive APIs and observability to applications and SDK components.

The ideal candidate combines deep runtime and systems-software expertise with a strong understanding of heterogeneous compute, Linux kernel and driver interfaces, DMA and shared-memory architectures, and performance-sensitive AI or multimedia pipelines. This will be a hands-on principal architect and technical lead who directs engineers across runtime, kernel, driver, firmware, multimedia, and SDK development.

Responsibilities

  • Define the SoC-wide execution model for coordinating workloads across heterogeneous compute and media engines, including dependency management, resource arbitration, priority and QoS, and concurrent pipeline behavior.

  • Set the architecture and technical direction for the AI inference runtime, including compiled-model execution, integration with industry-standard AI execution frameworks, and interfaces to applications and the broader Velaura SDK.

  • Own the end-to-end dataflow architecture for sensor-to-application pipelines, ensuring that camera, media, preprocessing, inference, and postprocessing components operate as an efficient and coherent system.

  • Define the SoC-wide memory and buffer-sharing architecture across user space, the kernel, and heterogeneous hardware engines, establishing clear ownership, coherency, isolation, synchronization, and lifecycle semantics.

  • Define the division of responsibility and interface contracts among the runtime, kernel drivers, firmware, and hardware engines, including execution, completion, telemetry, fault management, and recovery semantics.

  • Partner with the compiler team to define the compiler-runtime contract, ensuring compiled artifacts contain the metadata and execution information needed for the runtime to load, validate, execute, profile, and maintain compatibility across releases.

  • Establish a system-wide observability and performance architecture that correlates behavior across software and hardware layers and enables optimization against latency, throughput, bandwidth, power, utilization, and predictability goals.

  • Define the runtime resilience and validation architecture, including fault-containment and recovery policies, architecture-level acceptance criteria, and qualification across correctness, concurrency, compatibility, performance, and sustained workloads.

Required Qualifications

  • Extensive experience designing and building production runtime systems, embedded middleware, multimedia frameworks, or other performance-critical systems software.

  • Strong C/C++ programming skills and demonstrated ability to architect and contribute hands-on to production runtime software spanning application-facing APIs, user-space libraries, and low-level driver, firmware, and hardware interfaces.

  • Strong understanding of heterogeneous and asynchronous execution, including command submission, queues, events, dependencies, synchronization, concurrency, scheduling, and resource management.

  • Strong understanding of device memory, DMA, IOMMU/SMMU, cache coherency, memory mapping, shared buffers, buffer lifetimes, and kernel/user-space memory interfaces.

  • Experience optimizing end-to-end data movement and execution across multiple hardware engines rather than focusing solely on individual kernels or accelerator performance.

  • Experience designing stable runtime APIs with well-defined compatibility, versioning, error handling, diagnostics, and recovery behavior.

  • Demonstrated ability to debug complex cross-layer correctness and performance problems using disciplined, data-driven methods, profiling, and tracing.

  • Demonstrated technical leadership across component and organizational boundaries, including translating system requirements into clear architectures, interfaces, implementation guidance, and validation strategies.

Preferred Qualifications

  • Direct experience developing or extending AI inference runtimes or execution providers, such as ONNX Runtime, TensorRT-like runtimes, OpenVINO, TensorFlow Lite delegates, Qualcomm QNN/SNPE, TVM runtimes, or comparable systems for NPUs, GPUs, DSPs, or other accelerators.

  • Linux kernel development or upstream contribution experience involving device, accelerator, media, or shared-memory subsystems.

  • Experience with robotics, autonomous systems, edge AI, ROS 2, camera pipelines, ISP integration, V4L2/media, GStreamer, or other sensor-driven workloads.

  • Familiarity with compiled-model artifacts, quantized execution, tensor layouts, graph partitioning, memory planning, and compiler/runtime integration.

  • Experience with embedded Linux SDKs, production deployment, long-term runtime/API compatibility, or the security and isolation requirements of multi-process accelerator systems.

$200,000 - $300,000 a year

Compensation & Benefits

At Velaura, we believe exceptional talent deserves exceptional rewards. Compensation for this role includes a base salary and very competitive equity compensation, allowing team members to share in the company's long-term success.

The base pay listed represents a good faith estimate that the Company reasonably expects to pay for this position at the time of hire. Actual compensation will depend on multiple factors, including the candidate's skills, qualifications, relevant experience, and geographic location.

In addition to base salary and equity compensation, Velaura offers a comprehensive benefits package that may include medical, dental, and vision coverage; paid time off; flexible work arrangements; professional development opportunities; and other benefits designed to support the well-being and growth of our team.

Velaura is committed to pay equity and transparency and regularly benchmarks compensation to ensure we remain competitive in the market.

Why Velaura?

Velaura is building next-generation compute technology for cloud, edge, and PhysicalAI. Our solutions will enable robots, autonomous systems, drones, and other intelligentmachines to operate efficiently in the physical world.
This is an opportunity to help build foundational technology at a time when the industry is undergoing fundamental change. You will work alongside experienced leaders, architects, engineers, and operators who have delivered industry-defining products across mobile, cloud, semiconductor, and AI platforms. If you enjoy solving difficult problems, working across disciplines, and helping shape thefuture of Physical AI.

Equal Employment Opportunity and Accommodations

Velaura is an Equal Opportunity Employer that is committed to inclusion and diversity.Qualified applicants will receive consideration for employment without regard to race,color, religion, national origin, gender, sexual orientation, gender identity, disability orprotected veteran status. We also take affirmative action

#J-18808-Ljbffr

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Principal AI SoC Runtime Software Architect in Santa Clara, CA vacancy
  •  ...paths, cache policies, runtime integrations, performance...  ...architecture, runtime software, operating-system...  ...technical owner for the AI-HLC architecture. Primary...  ...responsibilities ~ Architect the model-aware residency...  ...expert storage, SoC-side decode, pinned residency... 
    Principal
    Local area
    Remote work

    Lexarenterprise

    San Jose, CA
    2 days ago
  •  ...We are seeking a Software Architect to define and lead the technical vision...  ...and building a multi-tenant, AI-native platform that powers...  ...a platform that reasons. As Principal Software Architect, you will...  ...harness engineering: agent runtimes, tool/function calling, structured... 
    Principal
    Local area

    Socket.dev

    Sunnyvale, CA
    23 hours ago
  • $262.5k - $394k

     ...Principal Software Architect, Finance Technology Platform Sunnyvale, California, United States Corporate...  ...and building a multi-tenant, AI-native platform that powers financial...  ...including harness engineering: agent runtimes, tool/function calling, structured output... 
    Principal
    Contract work
    Temporary work
    Local area
    Relocation

    Apple Inc.

    Sunnyvale, CA
    23 hours ago
  •  ...Software Engineer Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon...  ...on production-grade embedded runtime environments. You'll work...  ...with ML accelerators, GPU, CPU, SoC architecture and micro-architecture... 
    Suggested
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    4 days ago
  • $195.2k - $361.2k

     ...explored architectures, algorithms, and software inspired by the brain's...  ...power the coming era of physical AI systems-beyond the reach of GPUs...  ...systems. In this role, you will architect and lead the development of the firmware, runtime, and performance infrastructure,... 
    Suggested
    Full time
    Work at office
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    5 days ago
  • $170k - $277k

     ...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do...  ...are looking for a visionary Senior Principal Engineer/Architect to serve as the technical authority...  ...with a deep empathy for internal software developers as your primary customers.... 
    Principal
    Full time
    Work at office
    Visa sponsorship
    Work visa
    Flexible hours

    Palo Alto Networks, Inc.

    Santa Clara, CA
    3 days ago
  •  ...demand for new data centers and AI compute is rapidly...  ...Overview We are seeking a Principal Network Architect to own the datacenter scale...  ...that both digital design and software teams implement against. This...  ...semantics exposed to the runtime, in partnership with the software... 
    Principal
    Full time

    Socket.dev

    Sunnyvale, CA
    4 days ago
  •  ...Learning and HPC. We're seeking a Senior Software Architect to help co‑design next‑gen data center...  ...communication technologies to accelerate AI and HPC workloads. Explore innovative...  ..., SHMEM) and at least one communication runtime (MPI, NCCL, NVSHMEM, OpenSHMEM, UCX, UCC... 

    NVIDIA Gruppe

    Santa Clara, CA
    3 days ago
  • $272k - $431.25k

     ...We are now looking for a Principal Software Engineer for LPX System Software! NVIDIA’s LPX System...  ...layers, core system libraries, drivers, and runtime components that workloads enter the...  ...An established habit: building with AI coding agents — not as a novelty, but as... 
    Principal
    Shift work

    NVIDIA Gruppe

    Santa Clara, CA
    4 days ago
  •  ...the potential of generative AI to power the transformation of...  .... We are at the forefront of software and hardware innovation, pushing...  ...per week. The role: Principal System Software Engineer, AI...  ...Experience with deep learning runtimes (such as ONNX Runtime, TensorRT... 
    Principal
    Work experience placement
    3 days per week

    Entrada Ventures

    Santa Clara, CA
    1 day ago
  •  ...NVIDIA seeks a Principal Architect for SoC and System performance modelling to lead modeling, analysis and validation of cutting-edge architectures. You will collaborate across multiple modeling teams to design, implement and test models, mentor engineers, and drive feature... 
    Principal

    Jobleads-US

    Santa Clara, CA
    5 days ago
  • $220k - $280k

     ...Credo is looking for a Principal AI System Architect to join our team in San Jose, CA, reporting to AVP, XPU system AI interface. This role needs...  ...skills, daily experience working with compiler, software, SOC & backend team. Preferred Qualifications Master's... 
    Principal
    Work experience placement

    Socket.dev

    San Jose, CA
    4 days ago
  •  ...products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded...  ...intelligence, we are a powerhouse - a cutting-edge 'AI Software Solutions Team'. Specialized in AI optimization, fine-tuning large... 
    Principal

    AMD

    Santa Clara, CA
    3 days ago
  • $272k - $431.25k

     ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...legacy of innovation! We are seeking a Principal SW Architect, Networking to build the future of AI...  ...development of AI networks alongside hardware, software, and applicationsGuaranteeing... 
    Principal
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $272k - $431.25k

     ...Principal System Software EngineerNVIDIA is seeking a highly motivated Principal System Software Engineer...  ...with hardware, architecture, kernel, AI, middleware, and platform teams to...  ...including kernel, drivers, middleware, runtime frameworks, and platform services.Collaborate... 
    Principal

    NVIDIA

    Santa Clara, CA
    23 hours ago
  •  ...Lexar Enterprise seeks a senior systems architect to own the end-to-end technical architecture for...  ...cache policies, and performance models across runtimes and hardware. You will bridge model architecture, runtime software, OS memory management, and accelerator execution... 

    Jobleads-US

    San Jose, CA
    3 days ago
  • $206.4k - $379.1k

     ...custom multimedia generative AI - deep-tuned image, video, and...  ...Premiere. We are hiring a Principal Machine Learning Engineer to...  ...how our generative models are architected, optimized , and served at...  ...MLOps technologies - serving runtimes, quantization , GPU scheduling... 
    Principal
    Temporary work
    Local area
    Worldwide

    Adobe

    San Jose, CA
    16 hours ago
  •  ...from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP...  ...under tapeout pressure and propose solutions that balance accuracy, runtime, and methodology constraints ~ Strong understanding of static... 
    Principal

    Synopsys Inc

    Sunnyvale, CA
    a month ago
  • $278.1k - $417.1k

     ...next generation of AI-driven game...  ...modern, browser-native runtime (built on technologies...  ...runtime. As our Principal Engineer for On-...  ...the renderer. Architect inference systems...  ...~8+ years in software/ML engineering, with...  ...hardware: mobile SoCs (Apple Neural Engine... 
    Principal
    Work at office
    Worldwide
    Relocation package

    Unity

    Mountain View, CA
    1 day ago
  • $152k - $241.5k

     ...amazing people. NVIDIA is looking for an experienced Senior Software and System Architect to join our Networking Software Architecture group. Nvidia'...  ...This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to... 
    Remote work

    NVIDIA Gruppe

    Santa Clara, CA
    1 day ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ...We are seeking a Robotics AI Architect to define and scale next‑...  ...the roadmap for our AI SDKs, runtime, and reference architectures to...  ...systems Influence compute‑software co‑design across CPU, GPU, and... 
    Principal

    Advanced Micro Devices

    San Jose, CA
    2 days ago
  •  ...performance silicon chips and software content. Join us to transform...  ...can work directly with customer architects to close real system-level...  ...You are comfortable discussing SoC integration, IP configuration,...  ...technical partner for next-generation AI, HPC, storage, networking, and... 
    Principal

    Synopsys Inc

    Sunnyvale, CA
    a month ago
  • $180k - $225k

     ..., a leading global semiconductor company, is rapidly scaling its AI-focused power delivery business to capture a dominant share of the...  ...-world implementation challenges ~ Experience with embedded SoC’s including cross-functional collaboration with HW/FW/SW teams... 
    Principal
    Full time

    Renesas Electronics

    San Jose, CA
    11 days ago
  • $175.8k - $293k

     ...competitive advantage. We're looking for a Principal AI Engineer to architect, build, and harden the agentic AI...  ...context management, orchestration runtimes, and inference serving. Evaluate and...  ...scalable, production-grade software, with significant recent depth in AI... 
    Principal

    BMC Software

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...As a Senior Software Architect in the GPU Networking Architecture team, you will define Software Defined Networking (SDN) architectural solutions...  ...fields related to the modern data center, such as distributed AI and deep learning systems, Networking Operating Systems,... 

    Facade Today

    Santa Clara, CA
    2 days ago
  • $174k - $252k

     ...propose new hardware features.Collaborate across Hardware, Driver, Runtime, and Performance Analysis teams and many other stakeholders....  ...or equivalent practical experience.5 years of experience with software development in one or more programming languages.3 years of experience... 

    Google

    Sunnyvale, CA
    4 days ago
  •  ...research team within NVIDIA’s Networking Systems & Software Architecture group is solving some of AI’s hardest infrastructure problems. The team builds systems...  ...quantum computing interconnects. The Senior Architect role is to own modules and projects end-to-end—from... 

    NVIDIA Gruppe

    Santa Clara, CA
    3 days ago
  • $140k - $190k

     ...Principal Embedded Software Engineer WiFi team is looking for a Principal Embedded Software Engineer with C programming and networking knowledge to join our team. This is a great opportunity to immerse yourself in all phases of the software development cycle to reach... 
    Principal
    Full time
    Worldwide

    Virtue AI

    Sunnyvale, CA
    1 day ago
  •  ...Principal DevOps EngineerTENEX is an AI-native, automation-first, built-for-scale Managed...  ...You'll work closely with Software Engineering, AI/ML, and Security...  ...and compliant (e.g., SOC 2, ISO 27001)...  ...).Proven track record of architecting, building, and operating... 
    Principal
    Live in
    Work at office
    Work from home
    Relocation
    Monday to Thursday

    TenEx

    San Jose, CA
    23 hours ago
  •  ...alert: Select how often (in days) to receive an alert: Principal Hardware Systems Architect-Compute Platforms Req ID: 140082 Region: Americas...  ...for the architecture and development of next-generation AI, accelerated computing, and high-performance data center... 
    Principal
    Local area

    Celestica Inc.

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal AI SoC Runtime Software Architect. Be the first to apply!