Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Engineer, Model Quality and Performance

Cerebras Systems

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About The RoleYou'll own model quality and performance for Cerebras' inference offerings. You will define what "good" looks like across the models we serve, building AI-driven systems to measure it at scale, and translating those signals into artifacts our customers and product team actually use.You'll use AI agents to spin up custom eval suites per customer use case, mine trajectories for representative test data, automate the repetitive parts of release qual, and help build performance datasets and benchmarking workflows for customer use cases. We want someone whose first instinct is "how do I get an AI agent to do this on a loop."You'll sit between engineering, product, and customer-facing teams. What You'll DoDesign eval suites with AI agents in the loop. For every model release, curate a thoughtful mix of advanced, basic, long-context, and customer-use-case-specific evals. Use Claude to generate, validate, and prune candidate test cases at speed.Build custom evals for target customers by orchestrating AI agents to mine trajectories from their workloads and synthesize representative eval sets.Automate eval execution end-to-end with AI-driven pipelines on top of standard tooling (Docker, Git, CI). The goal is a system that runs itself between releases, not a script you re-run by hand.Build automations to forecast and benchmark model performance on Cerebras for our top customers, including modeling how fast customer-specific workloads will run in production.Build product-quality tooling that synthesizes quality + performance data into a single, easy-to-use view. Skills & QualificationsExperience building AI agents. You ship real systems with Claude (or equivalent) as a force multiplier. You've built things that would have been infeasible solo without AI agents in the loop.Strong math/stats background..Comfort with Docker, Git, and the standard automation stackA taste for tooling design. You've shipped something that a non-engineer used without complaining. Bonus if AI helped you ship it.AssetsPerformance-tuning experience on custom silicon, GPUs, or FPGAs. Experience designing evals for agentic / coding / long-context / multimodal use cases.Familiarity with open-source eval frameworks (EvalScope, lm-eval-harness, etc.).Experience building AI agents.Why Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CAEmployment TypeFull timeLocation TypeOn-siteDepartmentSoftware Engineering

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the AI Engineer, Model Quality and Performance in Sunnyvale, CA vacancy
  • $117.7k - $221.4k

     ...efficient for embodied AI systems. We believe the...  ...depends not only on stronger models, but also on better...  ...artifacts, and high-quality evaluation loops. That...  ...aware approach that first performs the cheapest reusable...  ...model reflects how Cola engineers think: build durable intermediate... 
    Quality
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...worldwide.We’re a team of engineers, clinicians, and...  ...work helps care teams perform with greater precision...  ...platforms. As a Senior AI/ML Research Engineer, you...  ...fine-tune the foundation models—VFMs, VLMs, and VLA models...  ...label taxonomies, quality control, and the data pipeline... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    a month ago
  • $190k - $260k

     ...an artificial intelligence (AI) powered technology stack purpose...  ...it. Every improvement to our models – from GigaFusionNet to large...  .... We are looking for engineers who make model training fast:...  ...Experience building high-performance data pipelines for large-scale... 
    Performance
    Temporary work
    Work at office
    Visa sponsorship
    Flexible hours

    Kodiak

    Mountain View, CA
    14 days ago
  • $272k - $431.25k

     ...generation of interactive world-model systems. With this...  ..., real-time inference performance in world models. And,...  ...source inferencing engine. We are on a mission to...  ...software: it means high-quality solutions, trusted releases...  ...vacancy. NVIDIA uses AI tools in its recruiting... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  •  ...Meet Our Software Engineers! Meet some of our software...  ...We are looking for an AI engineer to join a small...  ...shipping production-quality code quickly ~ Comfort...  ...literacy – you understand model evaluation, data...  ...requirements, interview performance, and the level and scope... 
    Quality
    Performance
    For contractors
    For subcontractor

    Applied Intuition

    Sunnyvale, CA
    2 days ago
  • $192.2k - $260k

    Join the next science and engineering revolution at Amazon's Delivery Foundation Model team, where you'll work...  ...through advanced AI and foundation models.We...  ...initiatives, ensuring robust performance in production...  ...to improve the safety, quality, and efficiency of Amazon... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    11 days ago
  • $171.6k - $222.2k

     ...that combine innovative AI, sophisticated control...  ...and large language models.We leverage advanced robotics...  ..., safety, and performance needs. You will invent...  ...reproducible results and solid engineering practices, closing the...  .../curation, dataset quality/provenance, and... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    Amazon

    Sunnyvale, CA
    4 days ago
  • $190k - $250k

     ...developed an artificial intelligence (AI) powered technology stack...  ...large-scale generative world models that learn to predict...  ...frameworks that measure world model quality beyond pixel-level metrics, including...  ..., education, skill level and performance during interview. Total... 
    Quality
    Performance
    Temporary work
    Work at office
    Visa sponsorship
    Flexible hours

    Kodiak

    Mountain View, CA
    18 days ago
  • $182k - $242k

     ...Senior Software and AI Engineer Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco...  ...combines superior infrastructure performance with deep technical expertise to accelerate...  ...or experience match. Here are a few qualities we've found compatible with our team.... 
    Quality
    Performance
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    4 days ago
  • $189.3k - $290.7k

     ...Motors is bringing multimodal AI into the vehicle, and we are looking...  ...for a Staff AI/ML Software Engineer to lead the adaptation, fine-...  ...distillation of foundation models for the automotive edge. You will...  ...continuous data loops, and perform reliably after edge quantization... 
    Performance
    Full time
    Work at office
    Local area
    Work from home
    Relocation package
    3 days per week

    General Motors

    Mountain View, CA
    5 days ago
  • $193.3k - $261.5k

     ...Join us to optimize the latest models to run really fast on the...  ...As a Sr. Software Development Engineer on the Inference Model Enablement...  ...inference usability and quality through inference features, infrastructure...  ...* Deliver high-performance models using distributed inference... 
    Quality
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $150k - $210k

     ...the Position – Embodied AI EngineerWe are seeking...  ...exceptional Embodied AI Engineer to build the foundation...  ..., debug, and optimize model architectures (VLA, diffusion...  ...and vacation time. Performance based Short-Term...  ...loved ones improve your quality of life. Personal fitness... 
    Quality
    Performance
    Full time
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

     ...generating it! Our world model team is pushing the...  ...boundaries of multimodal AI, robotics, and world...  ...coherence, SDG quality, and WAM usefulness.Develop...  ...Computer Science, Electrical Engineering, Robotics, Machine...  ...Intelligence, High-Performance Computing and Visualization... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  •  ...Institute of Foundation Models  We are a dedicated...  ...the next generation of AI builders, and drive transformative...  ..., data scientists, and engineers, tackling the most...  ...a global hub for high-performance computing in deep...  ...or similar UIs for data quality assurance  ~ Build and... 
    Quality
    Performance
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    24 days ago
  •  ...builds the world's largest AI chip, 56 times larger...  ...works with the leading model labs, global...  ...Cerebras' Wafer-Scale Engine (WSE). We build the compiler...  ...optimization infrastructure, high-performance kernel enablement, and...  ...focused on execution, quality, and innovation.Scale... 
    Quality
    Performance

    Cerebras Systems

    Sunnyvale, CA
    7 days ago
  • $174.72k - $295.68k

     ...innovation, integrating advanced AI and autonomous driving...  ...Machine Learning Engineers with strong expertise in generative modeling and large-scale deep learning...  .... Develop high-quality multi-view future...  ...Language-Action (VLA) driving performance. Extend prediction... 
    Quality
    Performance
    Full time

    XPENG Motors

    Santa Clara, CA
    24 days ago
  • $104.9k - $218.55k

     ....We are currently seeking a AI Application Engineer to join our team in Santa Clara...  ...engineering, RAG quality, prompt engineering, AI safety...  ...automated evaluation pipelines for model quality, hallucination...  ...testing, and release gating.Perform latency profiling and optimization... 
    Quality
    Performance
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Santa Clara, CA
    1 day ago
  • $300k - $333k

     ...and behavior with the model specifications and taxonomy...  ...and accurate Gemini performance.Build consensus and...  ...dive into the unique engineering challenges we face daily...  ...tone, personality and quality, we call this area...  ...ambiguity inherent in AI research and operates... 
    Quality
    Performance

    Google

    Mountain View, CA
    7 days ago
  •  ...computing experiences—from AI and data centers, to...  ...ROLE:We are hiring AI Engineers to build recursive self...  ...intersection of AI systems, performance engineering, hardware-...  ..., speed, efficiency, quality, reproducibility, and...  ...horizon optimization, and model improvement loops.... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    a month ago
  •  ...Description Who We Are AI is the new electricity:...  ...We are seeking an AI engineer with architecture...  ...composable building blocks: model selection, prompting,...  ...constraints, and measured performance. You will collaborate...  ...task success, retrieval quality, factuality, tool-call... 
    Quality
    Performance

    career

    Mountain View, CA
    26 days ago
  • $152k - $241.5k

     ...computer graphics, high-performance computing, and...  ...everything from generative AI to autonomous systems,...  ...enable researchers and engineers to develop the next generation...  ...developing production-quality softwareHands-on...  ..., time-series data modeling, and real-time performance... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $212.7k - $287.7k

     ...an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize state-...  ...advancing inference usability and quality through inference features,...  ...LLM model architectures, model performance optimizations, and inference... 
    Quality
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $120k - $220k

     ...information powered by advanced AI, recommendation systems...  ...cycle.We're hiring the engineer who owns this agent end...  ...our LLM critic + performance-feedback regeneration...  ...iteration.Generative model router — Pick the right...  ...minimizing $/asset under a quality floor. Prove your... 
    Quality
    Performance
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    a month ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ...a Forward Deployed Research Engineer to build, evaluate, and...  ...and data pipelines to measure model quality and business impact.Build and...  ...Diagnose and optimize AI system performance across models, data,... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    a month ago
  •  ...Description Palona’s AI agents operate in real...  ...than selecting the newest model. It requires disciplined evaluation, high-quality data, modeling judgment...  ...an applied AI Modeling Engineer to improve the...  ...model failures, slice performance by scenario, detect regressions... 
    Quality
    Performance
    Temporary work
    Immediate start

    Palona AI

    Los Altos, CA
    26 days ago
  •  ...Palona is building a category-defining AI platform for restaurants. This role builds...  ...We are looking for an AI Full-Stack Engineer who treats growth as an engineering and...  ...production engineer with a high bar for code quality, performance, measurement, and user experience.... 
    Quality
    Performance
    Temporary work

    Palona AI

    Los Altos, CA
    26 days ago
  • $184k - $287.5k

     ...NVIDIA is the industry leader in high performance computing, gaming and AI. Our GPUs and SOCs give outstanding...  ...alone! Now we're hiring the engineer who will lead the rebuild of that toolchain...  ...regression gates that protect product quality.Help set the team's AI direction.... 
    Quality
    Performance
    Full time
    Immediate start

    Nvidia

    Santa Clara, CA
    a month ago
  • $129.3k - $193.9k

     ...Technologies, Inc. Job Area: Engineering Group, Engineering Group >...  ...for a skilled and motivated AI Model Training Engineer to join our...  ..., scalable models that meet performance, efficiency, and ethical...  ...engineering teams to ensure high-quality, well-labeled, and balanced... 
    Quality
    Performance
    Work experience placement
    Work from home

    Qualcomm

    Santa Clara, CA
    1 day ago
  •  ...computing experiences—from AI and data centers, to...  ...are hiring Applied AI Engineers to work directly with hardware...  ...correctness, measure quality, and help engineers...  ...simulation, firmware, performance debugging, routing, issue...  ...fit as much as model capability.KEY RESPONSIBILITIES... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    a month ago
  •  ...Description Accellor is an AI-native services firm...  ...AI, data, and engineering capabilities. Our mission...  ...including inference runtime, model serving, GPU...  ...platform — from GPU-level performance and distributed inference...  ...used, and how context quality should be measured.... 
    Quality
    Performance

    Accellor

    Mountain View, CA
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Engineer, Model Quality and Performance. Be the first to apply!