Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI/ML Infra / Systems Engineer (Inference)

$204k - $216k

Sapience AI Corporation

Job Description

Job Description

About Sapience AI

Sapience AI is the collective intelligence platform for professional communities. We sit above the CRMs, AMS platforms, and knowledge bases that organizations already run, and we turn the expertise scattered across them into something every member can search, act on, and share.

The intelligence a community needs is already inside it. Most organizations just cannot reach it. Knowledge lives in silos, in legacy systems, in the heads of a few experts, and in fragmented records no one can connect. We change that.

Our work is grounded in four commitments: technology elevates people and never replaces them, the best expertise is already inside the community, everything is built on trust, and every deployment is purpose-driven for the organization it serves.

Let's achieve more, together.

Where this role sits

This role owns how Sapience AI serves intelligence at scale. You build and run the inference infrastructure that turns the models and reasoning behind MINERVA and COGENT into fast, reliable, affordable answers for members.

You work where models meet production: serving, scaling, latency, and cost. Every time a member asks the platform something, your systems are what make the answer arrive quickly and hold up under load.

You sit close to applied AI, research, and platform engineering, and you make the difference between a model that works in a notebook and one that serves a community reliably.

Why this role exists

Collective intelligence is only useful if it is fast and dependable. Members will not wait, and communities cannot rely on a platform that buckles under load or costs too much to run.

Serving modern AI at scale is a hard systems problem: large models, tight latency budgets, expensive hardware, and demand that spikes. It takes real infrastructure engineering to get right.

The AI/ML Infrastructure and Systems Engineer owns that problem. You make inference fast, reliable, and affordable, so the intelligence the platform promises actually reaches members.

What you will own (Areas of Responsibility)

You hold seven areas of responsibility across inference infrastructure. Each one is yours to set direction on, build, and measure.

1. Inference serving and systems
  • Build and operate the inference systems that serve models and reasoning behind MINERVA and COGENT.
  • Own the serving path end to end, from request to response, under real load.
  • Make serving robust to failure, traffic spikes, and change.
2. Latency and performance
  • Drive down latency so members get fast answers, and keep it low as the platform grows.
  • Optimize the full path, including model execution, batching, caching, and retrieval.
  • Profile relentlessly and remove the bottlenecks that matter.
3. Scale and reliability
  • Scale inference to more members, more communities, and heavier reasoning without losing reliability.
  • Build autoscaling, load management, and graceful degradation.
  • Own the reliability of the serving layer as a first-order responsibility.
4. Cost and efficiency
  • Own the economics of inference, including GPU and hardware efficiency at scale.
  • Improve utilization and cut waste without hurting quality or speed.
  • Make cost a designed property, not a surprise.
5. Model deployment and lifecycle
  • Build the paths that take models and reasoning components from research to production safely.
  • Support rollout, versioning, and rollback with confidence.
  • Give applied AI and research a fast, safe route to ship.
6. Observability and operations
  • Instrument the serving layer so it can be measured, debugged, and trusted.
  • Build the alerts, dashboards, and tooling that keep inference healthy.
  • Reduce the operational burden of running AI at scale.
7. Hardware, accelerators, and platform choices
  • Make sound choices about accelerators, runtimes, and serving frameworks.
  • Balance performance, cost, and maintainability in platform decisions.
  • Keep the stack current as inference technology moves.
AI-augmented ways of working

You build the infrastructure that serves AI, and you use AI in building it, to generate tooling, reason about performance, and move faster, while you own correctness, reliability, and cost.

The standard is human in partnership: AI accelerates the work, you own the judgment, the interpretation, and the call. The people who create the most value here are not the ones producing the most output. They are the ones turning evidence into clear, durable decisions.

What this role is not

To keep the boundary clear:

  • This is not a model research role. You serve and scale models; you do not develop new model architectures.
  • This is not a general backend role. Your center of gravity is inference systems, performance, and hardware efficiency.
  • This is not a data engineering role. You partner with data teams, but you own serving, not the data platform.
  • This is not a best-effort prototype role. You are accountable for production inference members depend on.
What success looks like

We measure this role on outcomes the team can see:

  • Fast answers. Members get low-latency responses, and latency stays low as the platform grows.
  • Reliable at scale. Inference holds up under load, spikes, and growth.
  • Affordable intelligence. Inference cost per answer improves as usage rises.
  • Safe rollouts. Models and reasoning components ship, version, and roll back with confidence.
  • Operable serving. The serving layer is instrumented, debuggable, and healthy.
  • Sound platform choices. Accelerator and framework decisions age well.
Who you are Required qualifications
  • Five or more years in infrastructure, systems, or ML infrastructure engineering.
  • Hands-on experience serving ML or LLM models in production at scale.
  • Deep understanding of latency, throughput, and performance optimization.
  • Experience with GPUs or accelerators and their efficient use.
  • Strong systems programming and distributed systems fundamentals.
  • A track record of reliable, cost-aware production systems.
  • Fluency with observability and operational excellence.
Preferred qualifications
  • Experience with inference-serving frameworks and model runtimes.
  • Experience optimizing LLM inference, including batching, quantization, and caching.
  • Familiarity with retrieval systems and their performance characteristics.
  • Experience owning cost and capacity for AI workloads.
  • Exposure to neuro-symbolic or agentic systems in production.
How you work
  • You name the real bottleneck before reaching for a fix.
  • You measure before and after, and you trust evidence over intuition.
  • You treat reliability and cost as first-order, alongside speed.
  • You build systems others can operate.
  • You share tooling and knowledge across the team.
Skills & Competencies
  • Inference serving architecture and optimization.
  • Latency, throughput, and performance engineering.
  • GPU and accelerator efficiency.
  • Distributed systems, scaling, and reliability engineering.
  • Model deployment, versioning, and rollout.
  • Observability and production operations for AI.
  • Cost and capacity management for inference.
Services & Tools Experience
  • Inference-serving frameworks and model runtimes (for example vLLM, TensorRT, Triton-class systems).
  • GPU tooling, CUDA-class ecosystems, and accelerator runtimes.
  • Kubernetes, containers, and cloud platforms (AWS, GCP, or Azure).
  • Observability stacks (metrics, tracing, and logging).
  • Python and a systems language such as Go, Rust, or C++.
  • Caching, queuing, and load-management systems.
  • Serving the models and reasoning behind the MINERVA platform and COGENT architecture.
Prior Experience & Background
  • Prior ML infrastructure, platform, or systems engineering at a software or AI company.
  • Experience serving models in production under real load.
  • A track record of performance and reliability improvements at scale.
  • Experience owning inference cost is a plus.
Cross-functional partners

You work most closely with Applied AI, Neuro-Symbolic AI, Platform Engineering, and Research. You serve the models and reasoning that power MINERVA and the COGENT architecture, and you own how they run in production.

How we hire

We review every application, and we encourage you to apply even if you do not match every line above. Research shows that talented people, especially those from underrepresented communities, often hold back when they do not meet every qualification. If that is the only thing holding you back, apply anyway.

Sapience AI is an equal opportunity employer. We are committed to a workplace where everyone, regardless of background, has a voice in building what comes next.

Compensation

Base Salary:  $204,000 - $216,000 + early stage equity

Generous health and wellness benefits

 

Sapience AI is an equal opportunity employer. We do not discriminate on the basis of gender, race or color, ethnicity or national origin, age, disability, religion, sexual orientation, gender identity or expression, veteran status, or any other protected characteristic. If you need an accommodation to complete our application process, let your recruiter know.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the AI/ML Infra / Systems Engineer (Inference) in Portland, OR vacancy
  • $90k - $100k

     ...AI Systems Engineer – Remote Bright Vision Technologies is a technology consulting and software...  ...that powers large-scale AI training and inference workloads. The role focuses on GPU...  ...performance, and developer experience for ML engineers and researchers, with strong... 
    Suggested
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Beaverton, OR
    4 days ago
  • $124k - $280k

    At PwC, our people in data and analytics engineering focus on leveraging advanced...  ...on developing and implementing advanced AI and ML solutions to drive innovation and enhance...  ...and optimising algorithms, models, and systems to enable intelligent decision-making and... 
    Suggested
    H1b

    PwC

    Portland, OR
    4 days ago
  •  ...accelerate how powered athlete products are built. You’ll report into engineering leadership, partner closely with controls and firmware...  ...functional teams responsible for highly integrated mechatronic systems that bring innovative athlete experiences to life. WHO WE... 
    Suggested
    Full time
    Casual work

    Nike

    Beaverton, OR
    22 days ago
  •  ...passionate individuals, OnPoint is looking for our next Senior Systems Engineer. We invite you to explore and grow your career with us! JOB...  ...of the organization. Ability to leverage and implement AI tooling to support job responsibilities and functions. MIN... 
    Suggested
    Hourly pay
    Work experience placement

    OnPoint Community CU

    Portland, OR
    4 days ago
  •  ...dedicated to excellence in the design and engineering ofLam's etch and deposition products. We...  ...industry. The impact you’ll make As a Systems Engineer 4 at Lam, you will lead system...  ..., including responsible adoption of AI-enabled tools to enhance analysis, knowledge... 
    Suggested
    Local area
    Remote work
    Flexible hours
    2 days per week
    3 days per week
    1 day per week

    Lam Research Salzburg GmbH

    Tualatin, OR
    3 days ago
  •  ...dedicated to excellence in the design and engineering of Lam's etch and deposition products. We...  .... The impact you’ll make As a Systems Engineer at Lam, you'll design complex frameworks...  ..., and plotting. Openness to using AI in your day-to-day work Preferred Qualifications... 
    Live in
    Local area
    Remote work
    Flexible hours
    2 days per week
    3 days per week
    1 day per week

    Lam Research

    Tualatin, OR
    1 day ago
  • $141.4k - $234.5k

     ...companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and...  ...space, is seeking a talented and motivated Senior Pre-Sales Systems Engineer within a Pod model for an emerging business positioned for high... 
    Remote work
    Flexible hours

    Pure Storage

    Portland, OR
    1 day ago
  •  ...Vice President of Systems Engineering About the Company Innovative digital banking provider for credit unions & community banks Industry...  ...51-200 Specialties banking api conversational ai online banking online account opening online loan origination... 
    Shift work

    Confidential

    Portland, OR
    3 days ago
  • $95k - $134k

     ...DAT is looking for a Site Reliability Engineer to join our SRE platform team. This position...  ..., performance, and scalability of our systems, implementing robust monitoring solutions...  ...and support future growth. Leverage new AI tools to assist with coding and... 
    Temporary work
    For contractors
    Work experience placement
    Work at office
    Local area
    Immediate start
    Flexible hours

    Dat Services Inc

    Portland, OR
    3 days ago
  • $57 per hour

     ...Systems EngineerJob Duration: 3+ Months Job Location: Kalama, Portland and Tacoma. Zip code 98625Travel- all within a couple hours...  ...experience in Information Technology and/or IT Architecture and Engineering as a Systems Engineer to include:Windows Server administration... 
    Local area

    Saxon Global

    Portland, OR
    1 day ago
  •  ...Sr Systems Engineer USA - Portland, OR Sr. Systems Engineer (Onsite in Wilsonville, OR) We are seeking a Senior Systems Engineer to join our team in Portland, OR (PDX). In this role, you will lead the development, validation, and integration of cutting-edge automated... 

    Twist Bioscience

    Portland, OR
    1 day ago
  •  ...Products Group, we are dedicated to excellence in the design and engineering of Lam's etch and deposition products. We drive innovation to...  ...the semiconductor industry. The Impact You'll Make As a Systems Engineer at Lam, you will design complex frameworks, systems,... 
    Local area
    Remote work
    Flexible hours
    2 days per week
    3 days per week
    1 day per week

    Lam Research

    Tualatin, OR
    4 days ago
  •  ...Distributed Systems Engineer Work closely with solution architects and development stakeholders to resolve ambiguity and conflicts Design, document, and implement scalable and high-performance algorithms in scalable languages such as Scala Support the testing... 
    Permanent employment
    Full time

    TecTammina

    Portland, OR
    2 days ago
  •  ...supporting high quality, reliable and stable infrastructure technologies and security capabilities to Johnstone's enterprise and e-business systems. Support • Documents and maintains all support procedures and provides training to level 1 and 2 support resources as... 
    Temporary work
    For contractors
    Local area
    Immediate start
    Work from home
    Flexible hours

    Johnstone Supply

    Portland, OR
    3 days ago
  •  ...Job Title Responsibilities: Work closely with systems engineers to develop metrology solutions for complex multi-layer, optical critical dimension and overlay measurements for state-of-the-art logic/memory devices. Duties will include application development and... 
    Overseas

    Critical Fit Recruiting

    Portland, OR
    1 day ago
  • $78.03k - $97.53k

     ...than just a job, you'll find a career built on trust, teamwork, and opportunity at OD. Support and maintain the organization’s systems infrastructure, including the implementation of hardware and software. Review previous documents and prepare up-to-date documents,... 
    Full time
    Temporary work
    Work experience placement
    Local area
    Immediate start
    Shift work
    Day shift

    Old Dominion Freight

    Portland, OR
    2 days ago
  •  ...Working for Micro Systems Engineering, Inc. (MSEI) means joining an elite team to work on some of the most exciting challenges in medical technology today. We are a pioneer in developing innovative implantable medical device technologies and devices that save and enhance... 
    Full time
    Work experience placement
    Worldwide

    Biotronik

    Lake Oswego, OR
    22 days ago
  • $130k - $170k

     ...Job Description Job Description Position Description Job Title: Senior Systems and Acoustic Engineer  Job Location: Lake Oswego, OR (On-Site)  Position Status: Full-time  Reports to: EVP of Product Innovation & Quality  Department: Engineering  FLSA Status... 
    Full time
    Contract work
    Relocation
    Visa sponsorship
    Flexible hours

    LIGHTSPEED AVIATION INC

    Lake Oswego, OR
    a month ago
  • $48 - $75 per hour

     ...Job Description We are seeking an experienced Life Safety Systems Engineer to design and manage fire alarm, VESDA, security, detection, and life safety systems for large-scale, mission-critical facilities. This is a long-term role with exciting opportunities to... 
    Hourly pay
    Monday to Friday

    Selectek

    Portland, OR
    1 day ago
  • $37.98 - $59.5 per hour

     ...Senior Life Safety Systems (LSS) Drafter / BIM Designer Pay: $37.98 - $59.50 per hour DOE Qualifications: Strong Revit skills...  ...delivery of complex packages in coordination with multi-disciple engineers and design leads that form the core of our Life Safety Systems... 
    Hourly pay
    Contract work

    HKAA

    Portland, OR
    1 day ago
  •  ...Life Safety Systems Engineer Austin, Texas, United States Qualifications Life Safety Systems Engineer Bachelor's degree in Engineering with the ability to attain PE licensure At least 4 years of AutoCAD design experience Knowledge of IFC, IBC, NFPA 1... 
    Contract work
    Local area

    SolveNow

    Portland, OR
    4 days ago
  •  ...employers in the Life Sciences, IT, and Financial Services sectors, feel free to check us out at Job Description Job Title: Sr Systems Engineer Duration: 6 Months (Possible Contract to Hire) Location: Portland, OR Duties/Responsibilities include but are not limited... 
    Contract work

    Mindlance

    Portland, OR
    2 days ago
  •  ...Job Description Job Description SENIOR SYSTEMS ENGINEER – HARDWARE TEAM SawStop, the world leader in power tool safety, is looking for driven, passionate engineers to join our engineering team. Our engineers develop world-class products that improve the lives of... 
    Temporary work
    Work at office
    Remote work
    Flexible hours
    1 day per week

    Sawstop

    Tualatin, OR
    17 days ago
  •  ...Job Description Job Description Security Systems Engineer About Cache Valley Electric: Founded in 1915, Cache Valley Electric (CVE) is one of the nation's leading electrical contractors and the largest electrical contractor headquartered in Utah. As a family... 
    For contractors

    Cache Valley Electric

    Portland, OR
    12 days ago
  • $18 - $50 per hour

     ...graduation and completion of the program Position Overview: Siemens Digital Industries Software is seeking a Digital Thread & Systems Engineering Intern to support the Capital PreSales organization. This role offers an opportunity to work on innovative digital... 
    Remote job
    Hourly pay
    Full time
    Internship
    Local area

    Siemens

    Gresham, OR
    8 hours ago
  • $101.33k - $180.92k

     ...Yost is a water resource management and engineering firm focused exclusively on water. Since...  ...Principal Engineer - Wastewater Collection System Planning The Senior - Principal Engineer...  .... We may use artificial intelligence (AI) tools to support parts of the hiring... 
    Temporary work
    Work experience placement
    Work at office
    Local area
    Remote work
    Flexible hours
    Night shift
    Afternoon shift

    West Yost

    Lake Oswego, OR
    24 days ago
  •  ...Company Description Hydra-Power Systems, Inc. is a leading distributor of hydraulic and pneumatic components, specializing in integrated...  ...to industry leaders globally. Our dedicated team of in-house engineers and manufacturing professionals provides creative solutions,... 
    Full time

    HYDRA-POWER SYSTEMS INC

    Portland, OR
    more than 2 months ago
  •  ...evolving game. WHO YOU’LL WORK WITH Nike Innovation Underfoot Systems are a team focused on inventing future cushioning technologies and systems.  The team consists of a collection of engineers, scientists, and innovators with a passion to create new technologies... 
    Full time
    Summer work
    Casual work
    Internship
    Relocation package

    Nike

    Beaverton, OR
    4 days ago
  • $45 - $60 per hour

     ...Job Description Job Description Job Title: Distribution Engineer (PE) The position falls within the Distribution Department of the...  ...penalties and civil liability. Use of Artificial Intelligence (AI): We may use Artificial Intelligence (AI) to support parts of... 
    Permanent employment
    Temporary work
    Work at office
    Remote work

    Actalent

    Happy Valley, OR
    1 day ago
  • $83.3k - $95k

     ...Quality Systems Manager Central Kitchen Oregon - Portland, OR 97214 End Date 10/05/2026 Overview Salary Range $83,300.00 - $95,000.00 Salary Position Type Full Time Description Portland, Oregon | Onsite Full-Time | Salaried (Exempt) | Bonus Eligible... 
    Full time
    Temporary work
    Local area
    Flexible hours

    Salt & Straw

    Portland, OR
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI/ML Infra / Systems Engineer (Inference). Be the first to apply!