Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Engineer

Accellor

Technical Architect — Ai Systems & Platform InternalsAccellor is looking for a Technical Architect — AI Systems, Inference & Platform Internals to help design, scale, and optimize the systems that power AI systems and internal research workloads.This role is focused on the internal AI systems stack, including inference runtime, model serving, GPU infrastructure, distributed systems, context engineering, cost optimization, evaluation gates, observability, release safety, and production reliability.The ideal candidate is a senior hands-on architect who can reason across the full AI platform — from GPU-level performance and distributed inference to product-scale reliability, model deployment, safety, and cost-efficient operations.Key responsibilities include:1. AI Systems ArchitectureDesign and evolve large-scale AI systems that support AI systems and internal research workloads.Define architecture across inference runtime, model serving, request routing, batching, KV-cache handling, GPU scheduling, distributed execution, observability, release gates, and production rollout.Own technical trade-offs across latency, throughput, reliability, correctness, safety, scalability, cost, and infrastructure efficiency.2. Inference Runtime & Model ServingArchitect high-throughput, low-latency inference systems across large-scale GPU clusters.Work across inference engines, serving layers, scheduling systems, caching, streaming, deployment pipelines, and runtime optimization.Partner with engineering teams to improve model-serving efficiency, tail latency, GPU utilization, memory efficiency, correctness under load, and cost per request.Guide architecture decisions involving PyTorch, JAX, Triton, vLLM-style serving, CUDA/Triton kernels, distributed inference, tensor parallelism, pipeline parallelism, model sharding, and long-context serving.3. GPU, Kernel & Distributed PerformanceAnalyze and improve performance across GPU kernels, memory movement, collective communication, orchestration, and runtime scheduling.Guide engineering decisions involving CUDA, Triton, NCCL/RCCL, GPU profiling, memory pressure, compute utilization, tensor layouts, interconnect behavior, and distributed execution.Identify system-level bottlenecks across compute, memory, networking, scheduling, model execution, and data movement.4. Context EngineeringDesign and guide context engineering frameworks that determine what information should be passed to the model, how it should be structured, how much context should be used, and how context quality should be measured.Own architecture patterns for prompt structure, dynamic context assembly, retrieval-augmented generation, long-context management, conversation memory, tool context, agent state, multimodal context, source grounding, permission-aware retrieval, context compression, and context auditability.Ensure AI systems use the right context, from the right source, with the right permissions, at the right cost, and with measurable quality.5. Cost Optimization FrameworksDesign and build cost optimization frameworks for large-scale LLM and GenAI workloads.Create architecture patterns that reduce unnecessary token usage, redundant retrieval, repeated model calls, inefficient inference paths, and avoidable infrastructure spend.Drive model routing, token budgeting, prompt compression, context pruning, semantic caching, response caching, batch inference, async execution, fallback strategies, and cost telemetry across AI workflows.Ensure cost optimization does not compromise quality, safety, grounding, reliability, or user experience.6. Training & Research InfrastructureCollaborate with research and training infrastructure teams to support large-scale model training and post-training workflows.Contribute to architecture around distributed training, checkpointing, orchestration, fault tolerance, observability, data movement, evaluation infrastructure, and experiment velocity.Support frontier model workflows across pre-training, post-training, reinforcement learning, agent training, evaluation harnesses, and large-scale experiment execution.7. Release Safety, Validation & Evaluation GatesArchitect validation and release systems that ensure model updates, inference engine changes, runtime images, prompt changes, context changes, and platform releases are correct, safe, performant, and regression-free.Define release gates across correctness, numerical stability, latency, throughput, token usage, cost regression, context quality, retrieval quality, safety behavior, reliability, and model output quality.Ensure platform optimizations do not reduce safety, grounding, quality, or user trust.8. Reliability, Observability & Production OperationsDesign systems that make AI infrastructure observable, debuggable, reliable, and operationally safe.Define telemetry, tracing, dashboards, alerts, logs, profiling views, runbooks, SLOs, and post-incident learning loops.Provide visibility into prompts, context payloads, retrieved sources, token consumption, model selection, cache behavior, inference latency, GPU utilization, evaluation scores, safety events, cost, and failures.Turn production issues into stronger platform abstractions, safer rollout mechanisms, better automation, and more reliable infrastructure.9. Agentic & Multimodal Platform InternalsSupport architecture for AI agents, tool use, memory, function calling, multimodal interaction, long-running workflows, and internal or external agent deployment.Work across agent harnesses, evaluation pipelines, workflow orchestration, safety controls, state management, tool execution, memory systems, and product-facing runtime constraints.Ensure agentic and multimodal systems are reliable, observable, secure, cost-aware, and safe under real workloads.10. Technical LeadershipWork closely with Research, Inference, Runtime, Infrastructure, Product, Safety, Security, Technical Success, and Deployment teams.Act as a senior technical authority who can cut across layers, resolve ambiguity, identify systemic risks, and drive architecture decisions.Mentor engineers and technical leads on distributed systems, performance engineering, context engineering, cost optimization, production readiness, AI platform design, and architecture trade-offs.Represent architecture decisions through design docs, RFCs, diagrams, technical reviews, operational plans, and leadership-level summaries.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Principal Engineer in San Francisco, CA vacancy
  •  ...criminal justice partner agencies, supporting mission-critical operations that run 24 hours a day, 7 days a week.\nThis Principal System Integration Engineer role is a key technical position on the JUSTIS development team, with a primary focus on building and supporting the... 
    Suggested
    Permanent employment
    Full time
    Work experience placement
    Work at office
    Immediate start
    Remote work
    Night shift
    2 days per week

    City and County of San Francisco

    San Francisco, CA
    9 days ago
  •  ...may be able to make a hybrid/remote exception for someone in LA or Seattle. About the Role We are looking for a Principal RF Systems & Hardware Engineer to lead the definition and execution of our communication payloads. You will bridge the gap between high-level... 
    Suggested
    Work at office
    Remote work
    Shift work

    AdAstra

    San Francisco, CA
    17 days ago
  • $162k - $243.1k

     ...the world’s leading integrated design practice. Our architects, engineers, interior designers, consultants, sustainability specialists,...  ...us and design your place with Stantec.Your OpportunityAs a Principal, Healthcare Mechanical Engineering, one must bring deep knowledge... 
    Suggested
    Full time
    Temporary work
    Part time
    For contractors
    Casual work
    Work at office
    Local area

    Stantec

    San Francisco, CA
    3 days ago
  • $180k - $230k

     ...from a centralized manual QA team to a model where delivery teams own quality end-to-end, supported by embedded Quality Engineers. We need a Principal-level technical leader to drive that transformation as a hands-on individual contributor.This is not a people-... 
    Suggested
    Full time
    Work at office
    3 days per week

    Dealpath

    San Francisco, CA
    4 days ago
  • $227.9k - $341.9k

    The opportunityWe are looking for a Principal Mobile App Developer to architect and lead the development of a standalone B2C mobile application...  ...and exploit loops.Technical Leadership: Mentor senior mobile engineers, set the standard for code quality, and define the CI/CD and... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    5 days ago
  • $280k - $385k

     ...data and AI infrastructure platform, so our customers can focus on the high-value challenges that are central to their missions.Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data... 
    Local area
    Remote work
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  •  ...person access patterns, cross-organizational platform convergence, and data integrity at financial-services scale. You'll build the engineering processes and culture that let a lean team operate Tier-0 infrastructure with confidence. And you'll push the boundaries of how... 
    Remote work
    Sleeping nights

    SoFi

    San Francisco, CA
    1 day ago
  • $220.4k - $297.4k

     ...data and AI infrastructure platform, so our customers can focus on the high-value challenges that are central to their missions.Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data... 
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  •  ...entity. Responsibilities Dev Infra builds the tools, platforms, and infrastructure that every engineer at Atlassian uses to ship. We are hiring a Senior Principal Engineer for our AI Foundations team, the group responsible for making AI work well across our... 
    Work at office
    Local area

    Atlassian

    San Francisco, CA
    3 days ago
  • $163.1k - $218.7k

    Job Posting Title:Principal Media Streaming EngineerReq ID:10144180Job Description:Technology is at the heart of Disney’s past, present...  ...and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data scientists... 
    Full time
    Work experience placement

    Hulu

    San Francisco, CA
    4 days ago
  • $212.1k - $342.65k

     ...solutions created by the #1 company in e-signature and contract lifecycle management (CLM). What you'll do Join Docusign as a Principal Engineer in the Enterprise Application Technology Engineering team, you will serve as the highest-level individual contributor... 
    Permanent employment
    Full time
    Contract work
    Work at office
    Local area
    Remote work
    Flexible hours
    2 days per week

    DocuSign

    San Francisco, CA
    1 day ago
  • $285.45k

     ...but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here.As a principal engineer on the Online Systems team, you’ll join a team that powers Pinterest’s most business-critical online systems at massive scale... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    4 days ago
  • Our Sacramento/ Davis office has an opening for a self-motivated engineer with 8-12 years of experience to work on projects related to groundwater wells, drinking water, water reuse and stormwater. When you join Brown and Caldwell, you will find that we offer a non-hierarchical... 
    Full time
    Contract work
    Temporary work
    Work experience placement
    Live in

    Brown and Caldwell

    San Francisco, CA
    2 days ago
  • $206.4k - $379.1k

     ...Adobe Firefly’s Generative AI Services team is seeking a Principal Service Engineer to serve as the technical lead for our GenAI Services domain. In this high-impact role, you will lead a team of talented engineers in building scalable, high-performance generative AI... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    1 day ago
  • $162k - $243.1k

     ...the world’s leading integrated design practice. Our architects, engineers, interior designers, consultants, sustainability specialists,...  ....Your OpportunityStantec is seeking a highly accomplished Principal Structural Engineer to lead work on medium and large projects,... 
    Full time
    Temporary work
    Part time
    Casual work
    Work at office
    Local area

    Stantec

    San Francisco, CA
    2 days ago
  • $300 per month

     ...spanning power generation, purpose-built data centers, and the cloud platform that frontier AI runs on. We are looking for a Principal Engineer on our Production Engineering team. Someone who will own the reliability, scalability, and operational excellence of the cloud... 
    Full time
    Temporary work
    Immediate start

    Crusoe

    San Francisco, CA
    1 day ago
  • $240k

     ...their product surfaces.   What you'll do Own outcomes for both Full Service and Institutional teams Manage a team of 7+ engineers across both segments, adapting your approach to each team's stage Build toward one unified Funds experience over time, even while... 
    Work at office
    2 days per week

    AngelList

    San Francisco, CA
    23 days ago
  •  ...response capability across all three; not as a generalist, but as a Principal who can go deep on each.Also note that this position will...  ...:10+ years of experience in detection, response, or security engineering.3+ years of experience commanding security incidents, especially... 
    Contract work
    Work experience placement
    Flexible hours
    Shift work
    Night shift

    Circle

    San Francisco, CA
    4 days ago
  • $240k - $310k

    The RoleYou will be the foundational technical pillar for security at Candid Health. As our first Principal Security Engineer, you won't just be managing a compliance checklist—you will architect, build, and scale the technical systems that protect our customers and their... 

    Candid Health

    San Francisco, CA
    1 day ago
  • $275k - $300k

     ...augmented adversary emulation, and offensive AI security research at Postman's scale.The OpportunityWe are looking for a Principal Offensive Security Engineer who is as much a strategist as they are a hacker. You will own the strategic direction of Postman's offensive... 
    Work at office
    Flexible hours
    3 days per week

    Postman

    San Francisco, CA
    1 day ago
  • $269k - $369k

     ...multiply the people around you, and you are the person product and engineering leadership call when an agent identity problem has no...  ...Credible from the IDE to the boardroom, trusted by CISOs and principal engineers alike, and steady when account politics get sharp.High... 
    Local area
    Remote work
    Worldwide
    Flexible hours
    Shift work

    Okta

    San Francisco, CA
    1 day ago
  • $180.5k - $285k

     ...Principle Communications Engineer One team. Global challenges. Infinite opportunities. At Viasat, we're on a mission to deliver connections with the capacity to change the world. For more than 35 years, Viasat has helped shape how consumers, businesses, governments... 

    ViaSat

    San Francisco, CA
    2 days ago
  •  ...thinking, and a commitment to innovation to help clear the way for millions of Americans to achieve more.About the RoleThe Principal Identity Security Engineer will lead the design and evolution of Happen Bank's enterprise identity security capabilities, helping ensure secure... 
    Full time
    For contractors
    Work at office
    Local area
    Remote work
    Relocation
    Flexible hours

    Lending Club

    San Francisco, CA
    1 day ago
  •  ...Principal EngineerAt Health Universe, we're on a mission to revolutionize science and medicine. We're seeking an accomplished Principal Engineer to join our team to enhance our platform. This role offers the exciting opportunity to work at the intersection of healthcare... 

    HireIdeal

    San Francisco, CA
    2 days ago
  • $108.89k - $176.24k

     ...Principal Engineer-AutonomyJLG began in 1969, when our founder, John L. Grove set out to resolve growing safety concerns in the construction industry. Since then we have been committed to understanding the challenges and delivering innovative solutions to the access market... 
    Permanent employment
    Immediate start
    Shift work
    Weekend work

    Oshkosh Corporation

    San Francisco, CA
    2 days ago
  • $250k - $400k

     ...Office Scripts and Power Automate integration at Microsoft, and did deep learning and RL research before that.I'm hiring a principal-level engineer to own the hardest systems in our care platform, whole. Not tickets, not a lane inside someone else's architecture: entire... 
    Work at office
    Immediate start
    Remote work
    Visa sponsorship
    Flexible hours
    Shift work

    Legion Health

    San Francisco, CA
    1 day ago
  • $207k - $276k

     ...learning and keeping abreast of new technologies and industry best practices and finding ways to bring those practices into the engineering organization.Essential FunctionsPartners with product management to craft product strategy, create product descriptions and ensure... 
    Hourly pay
    Work at office
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning Services

    San Francisco, CA
    2 days ago
  •  ...with software that reasons through the problem, recommends, and eventually executes the best next actions.You'll own the decision engine behind those recommendations, turning messy operational reality into software customers trust with some of the most important decisions... 
    Work at office
    Relocation

    Vitalize Care

    San Francisco, CA
    2 days ago
  •  ...What you’ll do # Partner with medical image reconstruction scientists / engineers to build ML components that improve reconstruction quality, speed, robustness, or quantitative accuracy. # Define training/evaluation pipelines, datasets, and metrics that map to user... 
    Full time

    Midjourney

    San Francisco, CA
    18 hours ago
  • $252k - $374k

     ...hardest problems in medicine. Within Life Science AI (LSAI), ML engineers build and operate the systems that turn generative models and...  ...across Lila's life science domains. We are seeking a Principal ML Engineer to design, build, and scale the ML infrastructure... 
    Full time
    Work at office
    Local area
    Flexible hours

    Lila Sciences

    San Francisco, CA
    18 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Engineer. Be the first to apply!