Manager - AI Inference Engineer
$125.5k - $230.2kErnst & Young
Location: Anywhere in Country
At EY, we’re all in to shape your future with confidence.
We’ll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world.
The opportunity
We are seeking an experienced AI Inference Engineer to help design, deploy, and operate private AI inference infrastructure at rack scale . This role is focused on the engineering of a production-grade inference platform: architecture, serving stack design, benchmarking, integration, reliability, and operational readiness.
The ideal candidate has deep hands-on experience building AI inference systems and the platform layers that support them, including multi-GPU inference, Kubernetes-based deployment, autoscaling, observability, routing, access control, and production support.
Your key responsibilities
Private AI inference architecture
- Design and deploy secure private inference architectures for enterprise AI inference.
- Translate workload requirements into technical architecture decisions across compute, caching, interconnect, storage & networking.
- Define reference architectures that support high-throughput, low-latency, enterprise inference at scale.
Inference platform engineering
- Design the model-serving stack, including inference runtime, orchestration, model registry, artifact management, API layer, routing, observability, deployment and fallback.
- Evaluate and recommend inference technologies and platform components for production use.
- Configure and harden the target environment for secure, reliable inference workloads.
Benchmarking and performance evaluation
- Build and run benchmark suites for enterprise workloads.
- Measure latency, throughput, concurrency, utilization, reliability, and cost efficiency.
- Tune serving configurations and system parameters to improve production performance.
- Produce evaluation results and recommendations based on objective testing.
Enterprise integration
- Support integration of inference platforms with client IT systems.
- Work with infrastructure and security stakeholders to validate technical requirements and remediate issues.
Model onboarding and operationalization
- Onboard models into the target inference platform.
- Configure serving endpoints, routing behavior, monitoring, and operational controls.
- Support deployment readiness, runbooks, and operational handoff for production use.
Skills and attributes for success
- Ability to turn business and workload needs into concrete infrastructure and software design decisions.
- Clear technical communication and strong architecture documentation skills.
To qualify you must have
- Bachelor’s Degree in a relevant field
- Strong experience designing and operating AI inference systems in production.
- Deep familiarity with multi-GPU serving and high-performance deployment patterns.
- Hands-on experience with Kubernetes and cloud-native platform technologies.
- Strong understanding of scaling, routing, autoscaling, observability, and reliability for production AI workloads.
- Experience integrating AI platforms into enterprise security and access environments.
Ideally, you'll also have
- Experience with private, sovereign, or enterprise AI infrastructure.
- Experience benchmarking large-model inference on GPU clusters.
- Familiarity with service mesh, policy enforcement, secrets management, and RBAC.
- Experience working across infrastructure, security, and application teams in a regulated environment.
- Experience supporting pilot-to-production AI platform rollouts.
What we offer you
At EY, we’ll develop you with future-focused skills and equip you with world-class experiences. We’ll empower you in a flexible environment, and fuel you and your extraordinary talents in a diverse and inclusive culture of globally connected teams. Learn more .
- We offer a comprehensive compensation and benefits package where you’ll be rewarded based on your performance and recognized for the value you bring to the business. The base salary range for this job in all geographic locations in the US is $125,500 to $230,200. The base salary range for New York City Metro Area, Washington State and California (excluding Sacramento) is $150,700 to $261,600. Individual salaries within those ranges are determined through a wide variety of factors including but not limited to education, experience, knowledge, skills and geography. In addition, our Total Rewards package includes medical and dental coverage, pension and 401(k) plans, and a wide range of paid time off options.
- Join us in our team-led and leader-enabled hybrid model. Our expectation is for most people in external, client serving roles to work together in person 40-60% of the time over the course of an engagement, project or year.
- Under our flexible vacation policy, you’ll decide how much vacation time you need based on your own personal circumstances. You’ll also be granted time off for designated EY Paid Holidays, Winter/Summer breaks, Personal/Family Care, and other leaves of absence when needed to support your physical, financial, and emotional well-being.
Are you ready to shape your future with confidence? Apply today.
EY accepts applications for this position on an on-going basis.
For those living in California, please click here for additional information.
EY focuses on high-ethical standards and integrity among its employees and expects all candidates to demonstrate these qualities.
EY | Building a better working world
EY is building a better working world by creating new value for clients, people, society and the planet, while building trust in capital markets.
Enabled by data, AI and advanced technology, EY teams help clients shape the future with confidence and develop answers for the most pressing issues of today and tomorrow.
EY teams work across a full spectrum of services in assurance, consulting, tax, strategy and transactions. Fueled by sector insights, a globally connected, multi-disciplinary network and diverse ecosystem partners, EY teams can provide services in more than 150 countries and territories.
EY provides equal employment opportunities to applicants and employees without regard to race, color, religion, age, sex, sexual orientation, gender identity/expression, pregnancy, genetic information, national origin, protected veteran status, disability status, or any other legally protected basis, including arrest and conviction records, in accordance with applicable law.
EY is committed to providing reasonable accommodation to qualified individuals with disabilities including veterans with disabilities. If you have a disability and either need assistance applying online or need to request an accommodation during any part of the application process, please call 1-800-EY-HELP3, select Option 2 for candidate related inquiries, then select Option 1 for candidate queries and finally select Option 2 for candidates with an inquiry which will route you to EY’s Talent Shared Services Team (TSS) or email the TSS at View email address on careers.ey.com .
- ...find new areas of inspiration and expand your capabilities, then consider a career in Advisory.KPMG is currently seeking a Manager, AI Engineer to join our Advisory Services practice.Responsibilities:End-to-end design and development of AI/ML solutions, leveraging cloud...SuggestedH1bLocal area
- ...cyber defense, application security, and managed service solutions to rethink the entire security... .... A Cybersecurity Forward Deployed Engineer is a production engineer who works... ...their security and engineering teams—to make AI systems secure, governed, and resilient in...SuggestedFull timeWork experience placementLive inWork at officeLocal area
$128.03k - $261.63k
Position Summary What You’ll Do As a Deloitte Tax, AI Engineer Manager, you will oversee the design, development, deployment, and support of custom AI applications and modules to address key business needs. You will lead a team of engineers, drive project execution...SuggestedWork at officeLocal areaVisa sponsorship2 days per week3 days per week- ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’... ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...systems.Mentor client teams on Snowflake: data management, analytics, AI/ML, and BI integration....SuggestedFull timeWork experience placementLive inWork at officeLocal area
- ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’... ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...with APIs, model monitoring, and prompt management.Operate MLOps / LLMOps pipelines with CI/...SuggestedFull timeWork experience placementLive inWork at officeLocal area
$73.5k - $212.28k
...At PwC, our people in data and analytics engineering focus on leveraging advanced technologies... ...leveraging team member's unique strengths, and managing performance to deliver on client... ...will lead the development of innovative AI solutions that drive remarkable client outcomes...H1b- ...uphold these hallmarks.Job OverviewThe AI Productization Engineer turns successful AI and machine... ...support processes.Define AI incident management processes, including classification,... ...traceability for AI decisions, including inference logging, input/output capture, and end...Local areaFlexible hours
- As a Senior Advanced AI Engineer, you will design, develop, and deploy AI-driven solutions... ...Knowledge of converting models for production inference (TorchScript, ONNX).Experience with... ...platforms, SCADA systems, or energy management solutions.Demonstrated success delivering...Permanent employmentTemporary workFlexible hours
- ...Description We are hiring a Senior AI Engineer to design, build, and operate enterprise... ...end-to-end across the AI stack — from inference engines and platform infrastructure (vLLM... ...inference performance through KV cache management, paged attention, batching strategies,...Full timeWork experience placementLocal area
$93.2k - $155.4k
Position Summary Our Deloitte AI & Engineering team to transform technology platforms, drive innovation, and help make a significant... ...through innovation. Work You’ll Do As an Agentic Engineering Manager, you will operate as a value-stream leader who ensures business...Local area$153k - $297k
...our team.KPMG is currently seeking an Associate Director, AI Security Frontier Engineering to join our Enterprise Security Services organization.... ...to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit...H1bLocal area- We Are:The Global AI Infrastructure team is at the center of enabling infrastructure... ...governance needsDeploy, configure, and manage XPU-based clusters (GPU, DPU, LPU, CPU)... ..., NCCL, NVLink, and CUDA along with LLM inference engines (TensorRT-LLM), production serving frameworks...Full timeWork experience placementLive inWork at officeLocal area
- ...About us New Co is a new AI-native product organization within... ...small, senior team. Our engineering model is agentic: engineers author... ...services around it (case management and work queues, integration... ...modern AI infrastructure: LLM inference and serving, model gateways,...Full timeLocal area
- ...communications by integrating AI-driven solutions into the... ...a full-time AI/CAD Software Engineer to join our team. You will lead... ...Automation: Automate the creation and management of large-scale datasets... ...: Seamlessly integrate ML inference engines into existing RF EDA...Permanent employmentFull time
- We are seeking a Full Stack AI Platform Engineer to join our Data Engineering, AI & ML Platform team... ...insights to plant engineers, facility managers, and OT security analysts.You will... .../CD stacks to automate the testing of inference logic and the deployment of API services...Permanent employmentTemporary workLocal areaFlexible hours
$91.1k - $179.5k
...our clients’ success. We are hiring an AI Engineer to build and operate the data, features,... ...that support model training, real-time inference, and LLM applications using Claude-, GPT... ...lead projects or workstreamsAbility to manage and prioritize multiple tasks in a fast-...Local area$130k - $170k
...following job description:Senior Engineer/Platform Leader accountable... ...operating secure, scalable AI/ML and Generative AI (GenAI)... ...environments, feature/model/prompt management, retrieval and knowledge-... ...workspaces, training/inference patterns, model registry, feature...Full timePart timeShift workDay shift- ...uphold these hallmarks.Job OverviewThe AI/ML Platform Engineer is responsible for the foundational... ...productization initiatives.Design and manage multi-tenancy patterns, resource isolation... ....Experience with model serving, inference optimization, and scalable AI infrastructure...Local areaFlexible hours
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsPython (AI/ML Engineer)Location - Atlanta GA, Irving (TX), Irvine (CA), NJ, or Tampa... ...in agile development methodologies and use project management tools to manage and track project progress.Ensure compliance...
- ...impact, we want to hear from you.We seek a dynamic and driven individual with a strong technical foundation to serve as a Generative AI Engineer. You will leverage our data, technology, and analytics to build LLM pipelines that generate actionable insights for our business...Full timeWork at officeImmediate start
$128k - $252.5k
...span from account executives and data scientists to AI strategists, machine learning specialists, and data engineers. SFL Scientific, a Deloitte Business, is looking... ...scientists, machine learning engineers, project managers, and industry experts to develop robust AI...Local areaVisa sponsorship- ...expertise; deep experience in cloud change management; and cloud-ready operating models with a... ...the elite technical and product engine of the Accenture Google Business Group (AGBG... ...shift in technology: the move to Agentic AI and Product-Led Operating Models. As a Google...Full timeWork experience placementLive inWork at officeLocal areaShift work
- ...grow. What You’ll Bring to The Team:We are looking for a Senior AI Engineer to design, build, and deploy high-quality AI-powered features.... ...and influence AI adoption across teamsPartner with product managers and designers to scope AI featuresContribute to shared patterns...Work experience placementWork at office
$110.7k - $372.9k
...afford the care they need. Deloitte has a new AI-first effort, backed by $1B in committed... ...months, not into a lab. As an Agentic AI Engineer, you will design, build, and... ...authorization to claims integrity and care management. Retrieval, grounding & context engineering...Local areaVisa sponsorship$116k - $175k
...continuous professional development. The Prompt + Skills Engineer is the hands-on builder in Cherry Bekaert’s AI Center of Excellence — the person who writes the... ...across Tax, Assurance, Advisory and Wealth Management use to deliver client work. Working under the direction...Full timeWork experience placementLocal area$131k - $218.3k
Position Summary Our Deloitte AI & Engineering team to transform technology platforms, drive innovation, and help make a significant... ...relationshipsAbility to lead projects or workstreamsAbility to manage and prioritize multiple tasks in a fast-paced and dynamic...Local area$131k - $218.3k
Position Summary Our Deloitte AI & Engineering team works to transform technology platforms, drive innovation, and help make a significant... ...to lead projects or workstreamsAbility to manage and prioritize multiple tasks in a fast-paced and dynamic environmentStrong...Local area$96k - $163k
As an Investigative Data Scientist/AI Engineer, you'll play a critical role in protecting market integrity by combining investigative analysis... ..., and contribute to broader market integrity efforts. Manage multiple concurrent matters with precision, maintaining organized...Full timeTemporary workWork at officeHome officeFlexible hours3 days per week- ...days a week in the Atlanta or Charlotte Office***The Senior AI Security Engineer helps design, implement, test, and operate the controls that... ...testing, and penetration testing.Experience implementing and managing complex information security technologies.Technical Skills &...Full timePart timeWork at officeShift workDay shift
- ...and redefine markets, you’ve come to the right place.Principal AI Engineer is a senior technical leader responsible for designing,... ...development (Airflow, Spark, ETL systems)MLOps, model lifecycle management, and production-grade deploymentCloud platforms (AWS/GCP/Azure...Full timeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager - AI Inference Engineer. Be the first to apply!


