Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Engineer, Model Quality and Performance

Cerebras Systems

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About The RoleYou'll own model quality and performance for Cerebras' inference offerings. You will define what "good" looks like across the models we serve, building AI-driven systems to measure it at scale, and translating those signals into artifacts our customers and product team actually use.You'll use AI agents to spin up custom eval suites per customer use case, mine trajectories for representative test data, automate the repetitive parts of release qual, and help build performance datasets and benchmarking workflows for customer use cases. We want someone whose first instinct is "how do I get an AI agent to do this on a loop."You'll sit between engineering, product, and customer-facing teams. What You'll DoDesign eval suites with AI agents in the loop. For every model release, curate a thoughtful mix of advanced, basic, long-context, and customer-use-case-specific evals. Use Claude to generate, validate, and prune candidate test cases at speed.Build custom evals for target customers by orchestrating AI agents to mine trajectories from their workloads and synthesize representative eval sets.Automate eval execution end-to-end with AI-driven pipelines on top of standard tooling (Docker, Git, CI). The goal is a system that runs itself between releases, not a script you re-run by hand.Build automations to forecast and benchmark model performance on Cerebras for our top customers, including modeling how fast customer-specific workloads will run in production.Build product-quality tooling that synthesizes quality + performance data into a single, easy-to-use view. Skills & QualificationsExperience building AI agents. You ship real systems with Claude (or equivalent) as a force multiplier. You've built things that would have been infeasible solo without AI agents in the loop.Strong math/stats background..Comfort with Docker, Git, and the standard automation stackA taste for tooling design. You've shipped something that a non-engineer used without complaining. Bonus if AI helped you ship it.AssetsPerformance-tuning experience on custom silicon, GPUs, or FPGAs. Experience designing evals for agentic / coding / long-context / multimodal use cases.Familiarity with open-source eval frameworks (EvalScope, lm-eval-harness, etc.).Experience building AI agents.Why Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationHeadquarters/Sunnyvale OfficeEmployment TypeFull timeLocation TypeOn-siteDepartmentSoftware

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Engineer, Model Quality and Performance in Sunnyvale, CA vacancy
  • $117.7k - $221.4k

     ...efficient for embodied AI systems. We believe the...  ...depends not only on stronger models, but also on better...  ...artifacts, and high-quality evaluation loops. That...  ...aware approach that first performs the cheapest reusable...  ...model reflects how Cola engineers think: build durable intermediate... 
    Quality
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...worldwide.We’re a team of engineers, clinicians, and...  ...work helps care teams perform with greater precision...  ...platforms. As a Senior AI/ML Research Engineer, you...  ...fine-tune the foundation models—VFMs, VLMs, and VLA models...  ...label taxonomies, quality control, and the data pipeline... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    11 hours ago
  • $190k - $260k

     ...an artificial intelligence (AI) powered technology stack purpose...  ...it. Every improvement to our models - from GigaFusionNet to large...  .... We are looking for engineers who make model training fast:...  ...offsExperience building high-performance data pipelines for large-scale... 
    Performance
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    2 days ago
  • $184k - $287.5k

     ...aren't just powering the AI revolution—we're...  ...cutting-edge deep learning models on every NVIDIA GPU....  ...highly skilled and driven Engineering Manager to take the...  ...standard for AI performance.What You’ll Be Doing:Lead...  ...delivering production-quality software libraries.Demonstrated... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $272k - $431.25k

     ...generation of interactive world-model systems. With this...  ..., real-time inference performance in world models. And,...  ...source inferencing engine. We are on a mission to...  ...software: it means high-quality solutions, trusted releases...  ...vacancy. NVIDIA uses AI tools in its recruiting... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $192.2k - $260k

    Join the next science and engineering revolution at Amazon's Delivery Foundation Model team, where you'll work...  ...through advanced AI and foundation models.We...  ...initiatives, ensuring robust performance in production...  ...to improve the safety, quality, and efficiency of Amazon... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    2 days ago
  • $190k - $250k

     ...developed an artificial intelligence (AI) powered technology stack...  ...large-scale generative world models that learn to predict...  ...frameworks that measure world model quality beyond pixel-level metrics, including...  ..., education, skill level and performance during interview. Total... 
    Quality
    Performance
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    1 day ago
  • $193.3k - $261.5k

     ...Join us to optimize the latest models to run really fast on the...  ...hardware.As a Software Development Engineer on the Inference Model...  ...advancing inference usability and quality through inference features,...  ...responsibilities* Deliver high-performance models using distributed... 
    Quality
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $224k - $356.5k

     ...into the unlimited potential of AI to define the next era of...  ...the forefront of AI and high-performance computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will...  ...strong appreciation for evaluation quality, including correctness,... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $272k - $431.25k

     ...generating it! Our world model team is pushing the...  ...boundaries of multimodal AI, robotics, and world...  ...coherence, SDG quality, and WAM usefulness.Develop...  ...Computer Science, Electrical Engineering, Robotics, Machine...  ...Intelligence, High-Performance Computing and Visualization... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...at the core of modern AI infrastructure, from training large-scale models to running inference in...  ...hardware, and compiler engineering is a big part of what makes...  ..., and target-specific performance signals.Apply RL...  ...benchmarks to assess code quality, correctness, and generation... 
    Quality
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...computing experiences—from AI and data centers, to...  ...ROLEWe are hiring AI Engineers to build recursive self...  ...intersection of AI systems, performance engineering, hardware-...  ..., speed, efficiency, quality, reproducibility, and...  ...horizon optimization, and model improvement loops.... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...computer graphics, high-performance computing, and...  ...everything from generative AI to autonomous systems,...  ...enable researchers and engineers to develop the next generation...  ...developing production-quality softwareHands-on...  ..., time-series data modeling, and real-time performance... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $176.6k - $239k

     ...some of the world’s leading AI model providers companies, AWS Sales...  ....Here are some other qualities we are looking for:Enjoy working...  ...re continuously raising our performance bar as we strive to become Earth...  ..., cloud computing, systems engineering, infrastructure, security,... 
    Quality
    Performance
    Local area
    Worldwide
    Flexible hours

    AmazonWebServices

    Mountain View, CA
    3 days ago
  • $120k - $220k

     ...information powered by advanced AI, recommendation systems...  ...cycle.We're hiring the engineer who owns this agent end...  ...our LLM critic + performance-feedback regeneration...  ...iteration.Generative model router — Pick the right...  ...minimizing $/asset under a quality floor. Prove your... 
    Quality
    Performance
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    2 days ago
  • $192k - $278k

    Lead model releases for Search, evaluating DeepMind...  ...against strict quality bars to determine launch...  ...evolves into a fully AI-enabled product.Design...  ...analyze model applicant performance and refine output quality...  ..., and key enablers of engineering velocity.In this role,... 
    Quality
    Performance
    Shift work

    Google

    Mountain View, CA
    1 day ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ...a Forward Deployed Research Engineer to build, evaluate, and...  ...and data pipelines to measure model quality and business impact.Build and...  ...Diagnose and optimize AI system performance across models, data,... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    3 days ago
  • $149.75k - $275.58k

     ...market. Our diverse team of engineers and researchers have...  ...coming era of physical AI systems-beyond the...  ...customers to build high-performance physical AI applications...  ...customers to improve software quality, performance, and...  ...the hiring process.Work Model for this RoleThis role... 
    Quality
    Performance
    Full time
    Internship
    Work at office
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    11 hours ago
  • $139k - $229k

     ...hybrid, meaning it will be performed both from home and from...  ...'s Machine Learning Engineers are both data/research...  ...implement machine learning models and algorithms. Unlike...  ...production quality code and influence the...  ...newsfeedBuild scalable AI innovations with foundation... 
    Quality
    Performance
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Sunnyvale, CA
    11 hours ago
  • $195.2k - $275.58k

     ...Description: The Software and AI (SAI) organization is...  ...Software Development Engineer to contribute to the...  ...‑platform, open‑source performance library for deep learning...  ..., and maintain high-quality coding and documentation...  ...the hiring process.Work Model for this RoleThis role... 
    Quality
    Performance
    Full time
    Local area
    Immediate start
    Remote work
    Worldwide
    Flexible hours
    Shift work

    Intel

    Santa Clara, CA
    3 days ago
  • $152k - $241.5k

     ...As part of Nvidia's applied AI team for chip design, you will...  ...the intersection of research, engineering, and product development,...  ...their seamless and efficient performance. If you're passionate about the...  ...design, testing, CI/CD, code quality, observability, security, databases... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $150k - $210k

     ...the Position - Embodied AI EngineerWe are seeking...  ...exceptional Embodied AI Engineer to build the foundation...  ..., debug, and optimize model architectures (VLA, diffusion...  ...and vacation time. Performance based Short-Term...  ...loved ones improve your quality of life. Personal fitness... 
    Quality
    Performance
    Full time
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    NVIDIA is the industry leader in high performance computing, gaming and AI. Our GPUs and SOCs give outstanding...  ...alone! Now we're hiring the engineer who will lead the rebuild of that toolchain...  ...regression gates that protect product quality.Help set the team's AI direction.... 
    Quality
    Performance
    Full time
    Immediate start

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...platform upon which every new AI-powered application is built....  ...a senior vision language model engineer to design and build agentic data...  ...search offerings are ease to use, performant and scalable. Your work will...  ..., curate, and maintain high‑quality multimodal datasets (e.g.,... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...built in the age of Generative AI? Join NVIDIA’s TensorRT team...  ...AI agents to produce high-performance, high-quality, modern C++ software at an...  ...are a systems-thinking C++ engineer who wants to help scale out...  ...experience with lightning-fast model onboarding, we want to hear... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $244.14k - $413.16k

     ..., integrating advanced AI and autonomous driving...  ...hands-on Senior Staff AI Engineer to build and scale...  ...reliability, evaluation, and performance.Job Responsibilities:...  ...online metrics to ensure quality, safety, and SLOs....  ...systems using language models, retrieval/grounding (RAG... 
    Quality
    Performance
    Full time

    XPENG Motors

    Santa Clara, CA
    3 days ago
  •  ...computing experiences—from AI and data centers, to...  ...are hiring Applied AI Engineers to work directly with hardware...  ...correctness, measure quality, and help engineers...  ...simulation, firmware, performance debugging, routing, issue...  ...fit as much as model capability.KEY RESPONSIBILITIES... 
    Quality
    Performance

    AMD

    Santa Clara, CA
    11 hours ago
  • $212.7k - $287.7k

     ...an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize state-...  ...advancing inference usability and quality through inference features,...  ...LLM model architectures, model performance optimizations, and inference... 
    Quality
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $200k - $322k

     ...learning ignited modern AI — the next era of...  ...us today.Design-for-X Engineering at NVIDIA works on groundbreaking...  ..., while monitoring performance, automating...  ...offs including cost and quality.What we need to see:BSEE...  ...with SQL, ETL, and data modeling Hands-on experience with... 
    Quality
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $191k - $315k

     ...hybrid, meaning it will be performed both from home and from...  ...LinkedIn’s Enterprise AI team creates meaningful...  ...LLM-as-a-judge, prompt engineering, embedding-based...  ...qualification, improve signal quality, and reduce time-to-...  ...as whether to fine tune models or leverage off the... 
    Quality
    Performance
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    11 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Engineer, Model Quality and Performance. Be the first to apply!