Staff AI Model Optimization Architect (LLM & Multimodal)
Jobleads-US
Qualcomm Technologies, Inc. in Austin, TX, is seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end model transformation and optimization for LLMs, VLMs, diffusion, and multimodal models on Qualcomm accelerators.
You will collaborate with compiler, performance, and accuracy teams to translate models into accelerator-efficient execution, balancing throughput, latency, memory, and quality across Day0 to production deployment.
#J-18808-Ljbffr Jobleads-US$162k - $243k
...compute, connectivity, and AI acceleration to play a... ...-scale foundation models.We are seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end... ...VLMs, diffusion, and multimodal models on Qualcomm inference... ...Profile and optimize LLM/VLM/diffusion...SuggestedWork experience placementWork from home- ...related to Large Language Models (LLMs), from model fine-tune... ...data, machine learning, AI, and Industrial data architects making our team an exciting... ...-of-concept, experiment, optimize, and deploy your models into... ...the full spectrum of LLM tuning including prompt engineering...Suggested
$169.4k - $279.6k
...As a member of AI and Emerging Technology... .... As a Lead Architect, you will collaborate... ..., Crew AI), Model Context Protocol (... ...groups — spanning LLM selection and routing... ...architectures, and optimization techniques Strong... ...Experience with multimodal AI systems Experience...SuggestedTemporary workWork experience placementLocal area$129.4k - $207k
...energetic, visionary, and passionate Agentic AI Architect tospearhead the design, deployment, and optimization of cutting-edge autonomous multi-agent... ...design phases. You will leveragestate-of-the-art LLM orchestration, Model Context Protocol (MCP) integrations, and enterprise...SuggestedFull timeLocal area- ...healthcare and scientific discovery to powering AI and the technologies people rely on... ...: We are seeking highly motivated AI Model Optimization & Software Engineer Interns/Co-op to... ...with software engineers, researchers, architects, platform teams, and project stakeholders...SuggestedFull timeSummer workInternshipSummer internshipWorldwide
$272k - $431.25k
Principal Architect, AI Networking page is loaded## Principal Architect,... ...spanning low-level transport optimization, hardware-software co-design,... ...disaggregated prefill/decode, model parallelism).* Integrating... ...as vLLM, SGLang, and TensorRT-LLM.* Publishing findings, representing...Remote work- Commerce seeks a Staff Software Engineer to lead architecture and evolution of large-scale distributed systems powering our commerce platform... ...will guide cross-domain decisions, mentor engineers, and drive AI integrations including GenAI into current and future products....Worldwide
- ...products that accelerate next‑generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded... .... THE ROLE AMD is looking for a Server and AI Product architect to join our Datacenter System Architecture and Engineering team...
- ...seeking a hands-on Snowflake Architect / Lead with strong expertise... ...Snowflake, Snowflake Cortex AI, Generative AI, and LLM-based solutions . The ideal candidate... ...architectures . Lead data modeling, integration, performance optimization, and Snowflake platform best...Contract work
- Product Manager - AI Inference & Model Serving Full-time Job Summary Mirantis is looking for a... ...observability, and full-stack performance optimization. This person will define how customers... ...runtimes (vLLM, SGLang, TensorRT‑LLM, Dynamo, Triton) and the optimization...Full time
- ...collective is seeking freelance Generative Engine Optimization (GEO) consultants in Austin, Texas, to help clients enhance their visibility in AI-powered search. The role requires... ...optimization, as well as the ability to measure LLM impact. Projects range from 10-40 hours...FreelanceFlexible hours
$129.4k - $207k
...energetic, visionary, and passionate Agentic AI Architect to spearhead the design, deployment, and optimization of cutting‑edge autonomous multi‑agent AI... ...phases. You will leverage state‑of‑the‑art LLM orchestration, Model Context Protocol (MCP) integrations, and enterprise...Local area$129.4k - $207k
...energetic, visionary, and passionate Agentic AI Architect to spearhead the design, deployment, and optimization of cutting-edge autonomous multi-agent AI... ...phases. You will leverage state-of-the-art LLM orchestration, Model Context Protocol (MCP) integrations, and enterprise...Temporary work- ...AMD is seeking a Server and AI Product Architect to join our Datacenter System Architecture & Engineering team to define EPYC and Enterprise AI/HPC GPU product architectures for scalable datacenters. You will collaborate with customers, internal architects, and software...
- ...discovery to powering AI and the technologies people... ...FIRMWARE VALIDATION ARCHITECT THE ROLE: Join our firmware... ...SIMICs like simulation models) or post-silicon... ...execution efficiency, and optimize reporting/triage... ...Copilot, Microsoft Copilot, LLM-based coding assistants...
$224k - $308k
...Agentic AI Architect (Technical Staff) The Software Engineering team delivers next-generation software application... ...across complex systems Design governed data models, vector stores, and metadata layers optimized for LLM context, feedback loops, and inference-time...Full time- Jabil is seeking an FPGA Lead Engineer in Austin to lead FPGA design, integration, and debugging for server, storage, and AI infrastructure platforms. You will own FPGA architecture, RTL implementation, and verification readiness while coordinating with systems, firmware...
- Broadcom Inc. in Austin, Texas, seeks an Agentic AI Architect to spearhead the design, deployment, and optimization of autonomous multi-agent AI architectures within... ...semiconductor design projects. The role emphasizes LLM orchestration, MCP integrations, and enterprise-...
$163k - $413.6k
...than 1,500 certified AWS architects across the company. Join our... ...here: You Are: An AI/ML Sr Architect delivering... ...Generative AI, Foundation Models, and Knowledge & Data Engineering... ...pipelines for ML and LLM lifecycle. Troubleshoot, optimize, and tune AI systems for performance...Work experience placementLive inWork at officeLocal area- ...Opportunity Join a team building AI capabilities that power next-... ...Based in Austin, Texas, the AI Model Developer (JS5/JS6) works... ...product teams to design, evaluate, optimize, and deploy advanced machine learning... ...mechanisms, embeddings, and LLM fundamentals Proficiency in...Temporary workWork at officeRelocation3 days per week
$139.4k - $230k
...Growing Team! As a member of AI and Emerging Technology... ...class solutions. As a Senior Architect, you will partner with Technology... ...integration patterns such as Model Context Protocol (MCP) — to architect... ...platform designs — spanning LLM integration, RAG pipelines,...Work experience placementLocal area- ...exciting innovations across AI infrastructure, cloud... ...and visionary AI Architect to define, drive, and evolve... ..., inference pipelines, model-deployment... ...Autonomous decision making Multimodal AI systems Drive hardware... ...platforms. Evaluate and optimize AI models for accuracy,...Worldwide
$130.7k - $205.2k
[Agentic AI] AI Runtime / ML Model Systems EngineerDescription -This role is responsible for developing runtime abstraction layers and intelligent... ...model placement policies.• Develop xPU routing logic.• Optimize inference execution.• Leads the building of various...Full timeTemporary workWork experience placementWork at officeLocal areaRelocationFlexible hoursShift work$150k - $175k
...software development company delivering cloud, AI, data, and enterprise solutions across... ...career growth potential. Job Title: Model Optimization Engineer Location: 100% Remote (U.S.)... ...Qualifications Experience optimizing LLM inference at production scale....Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...Role: Principal AI/GenAI Platform Architect Location: Austin, TX | San Francisco, CA | Los Angeles,... ...experience. Architect and build model gateways supporting multiple AI providers... ..., Amazon Bedrock, Vertex AI and other LLM providers. Implement model routing...Full time
- ...Job Description Mythic builds AI compute platforms around large... ...ones. Serving frontier-scale models on that kind of architecture pushes... .... We are looking for a System Architect to own the rack-level... ...the software team of the full LLM solution — partitioning strategy...
$91k - $130k
...'re Looking For: We\'re looking for an AI Success Architect to join our Solutions Consulting team and... ...the connector: how a customer uses an LLM platform to turn Meltwater data into... ...fastest-moving part of the AI landscape (models, agent frameworks, MCP tooling) and translate...Local areaFlexible hours- AI-Powered GTM Transformation Role AI has created a significant... ..., workflows, operating models and talent requirements that... ...connect strategy with execution by architecting AI-enabled processes,... ...strengthen decision-making, and optimize Buyer Experiences. What You'll...
- ...experienced System Performance Architect in Austin, Texas, to define and... ...performance requirements, testing models, and collaborating with different teams to ensure optimal system performance. The position... ...the opportunity to shape future AI compute products. #J-18808-Ljbffr...
$240k - $300k
Principal AI Architect Hybrid (Austin, TX) SecurityScorecard is the global... ...Science team, who bring the model and ML depth. Your focus is the... ...native workflow Mentor senior and staff engineers on agentic system... ..., and architecting agentic or LLM-powered systems in production,...Shift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff AI Model Optimization Architect (LLM & Multimodal). Be the first to apply!





