On-Prem AI Engineer: LLaMA/Mistral, GPUs & RAG
Jobleads-US
Texas Integrated Services seeks experts to deploy and maintain on-prem AI infrastructure for healthcare-adjacent clients. You will work across GPU deployment, model quantization, vector databases like Qdrant, and RAG pipelines, while safeguarding data and staying within client networks.
The role requires hands-on server hardware setup and a thoughtful approach to data leaving the building. Travel within Texas for install days is required; familiarity with HIPAA and client privacy is important to
#J-18808-Ljbffr Jobleads-US- ...Texas Integrated Services builds on-prem AI servers for small Texas businesses so they can run LLaMA, Mistral, and similar models on hardware they own. You'll deploy, tune... ...model quantization, vector databases (Qdrant), RAG pipelines, and hands-on server hardware setup. The...Suggested
- OverviewInfosys Topaz is an AI-first suite of services... ...Chain, Lang Graph, Llama Index, or similar... ...frameworks; comfortable with RAG, vector databases,... ...multi-agent), and prompt engineering.• Working with Azure OpenAI... ...-source LLMs (Llama, Mistral, Gemma) and fine-tuning...SuggestedFull timeTemporary workRelocation
$125k - $250k
...OverviewEast West Bank is seeking an experienced Senior AI Engineering to design, build, and operationalize enterprise... ..., AWS Bedrock, and open-source models such as Llama or Mistral.Practical experience with prompt engineering, RAG, embeddings, vector databases, LLM...SuggestedFull time- We Are:The Global AI Infrastructure team is at the center of enabling... ...CUDA along with LLM inference engines (TensorRT-LLM), production... ...computing platforms, including GPUs, DPUs, LPUs, CPUs, high-speed interconnects... ...-augmented generation (RAG), agent orchestration, secure...SuggestedFull timeWork experience placementLive inWork at officeLocal area
- ...That work changes the operating model an engineering organization runs on, the ways of working... ...sets the target. What we design from it is AI-native by construction. We redesign... ...formatting, and retrieval-augmented generation (RAG) patternsAbility to design prompts that...SuggestedFull timeWork experience placementLive inWork at officeLocal area
$70.35k - $235.1k
...Advanced Technology Centers (ATCs) is the engine for reinvention in our clients’ transformation... ...industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...agents, orchestration, context engineering, RAG, workflows) in production environments....Work experience placementLive inWork at officeLocal area3 days per week- ...Position Overview A large grocery retailer is seeking a Lead AI Engineer / Agentic Commerce Lead to drive the architecture, development,... ...platforms. Knowledge of Retrieval-Augmented Generation (RAG), vector databases, and agent frameworks. Experience implementing...Local area
- ...Forward Deployed AI Engineer Tekfortune is a fast-growing consulting firm specialized in permanent, contract & project-based staffing... ...Generative and Agentic AI platforms featuring high-performance RAG (Retrieval-Augmented Generation) pipelines and vector database...Permanent employmentContract workRemote work
- ...AgentsJoin the elite technical and product engine of the Accenture Google Business Group (... ...shift in technology: the move to Agentic AI and Product-Led Operating Models. As a Google... ...use, and Retrieval-Augmented Generation (RAG). Hands-on expertise with Google's Gemini...Full timeWork experience placementLive inWork at officeLocal areaShift work
- ...Exponent Exponent is the only premium engineering and scientific consulting firm with the depth... ...We are currently seeking a Senior AI Engineer for our Data Sciences Practice in... ...including retrieval-augmented generation (RAG) systems, LLM-powered assistants, and domain...Work at officeFlexible hours
$105.5k - $243k
HPC and AI Performance EngineerThis role has been designed as 'Hybrid' with a requirement... ...such as Computer Science, Mathematics, Engineering, Physics, Chemistry, Environmental... ...:Processor technologies, including CPUs, GPUs, and accelerators, from both user and programmer...Full timeWork experience placementWork at officeLocal areaImmediate start2 days per week$109k - $251k
HPC & AI Performance EngineerThis role has been designed as 'Hybrid' with a requirement... ...and AI industry. We are looking for driven engineers who thrive on complex technical challenges... ...AI system architecture, including CPUs, GPUs and other accelerators, memory, networking...Permanent employmentFull timeWork experience placementInternshipWork at officeLocal areaImmediate startWorldwide2 days per week$75 - $120 per hour
...Job Description Job Description Position: Forward Deployed AI Engineer Location: Houston, TX - Fully Onsite Compensation: $75.0... ...patterns, including: Retrieval-augmented generation (RAG) AI agents and tool calling Semantic or vector search Prompt...Hourly payLong term contractLocal areaShift work$91.1k - $179.5k
Position Summary Join our AI & Engineering team in transforming technology platforms, driving innovation, and helping make a significant... ...LLM applications using LLM APIs, prompt engineering, RAG, or vector retrieval2+ years of hands-on experience with Python...Local areaVisa sponsorship- •4–6 years software engineering experience, including 2+ hands-on experience building LLM-based... ...experience, open-source contributions to AI tooling, or published work on agents.... ...implement integration designs. Build RAG pipelines, vector search, and context-management...Full time
$203.5k
General Information Job Title Lead, AI Engineering Job ID 102641 Work Areas Analytics, Data & Research, Management Consulting... ...AI pipelines, including:Retrieval-Augmented Generation (RAG) Fine-tuning and parameter-efficient tuningEmbedding generation...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week- ...ideas into reality.We are Secure, Responsible AI & Data Protection professionals who enable... ...You AreManagers are the hands-on delivery engine of the Secure AI practice. They lead day-... ...fundamentals: LLMs, agentic systems, A2A, RAG, MLOps pipelines, MCP, AI Gateway, model...Full timeWork experience placementLive inWork at officeLocal area
- ...engagement. A Cybersecurity Forward Deployed Engineer is a production engineer who works... ...their security and engineering teams—to make AI systems secure, governed, and resilient in... ...environments—LLM systems, multi-agent pipelines, RAG architectures, and MLOps infrastructure—...Full timeWork experience placementLive inWork at officeLocal area
- ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’... ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...pipelines and retrieval-augmented generation (RAG) systems.Design knowledge structures that...Full timeWork experience placementLive inWork at officeLocal area3 days per week
$55k - $151.47k
...SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies... ...team you will develop and implement AI solutions that enhance product offerings.... ...production- Designing and optimizing RAG pipelines- Leading technical discovery in...Full timeH1b- ...Overview: We are seeking a highly skilled Senior Agentic AI Engineer to design, develop, and deploy modern Agentic AI solutions. This role... ..., multi-agent workflows, and Retrieval-Augmented Generation (RAG) pipelines for enterprise use on Microsoft Azure. Collaborate with...
$59.35 - $77.03 per hour
...sponsorship of an employment Visa at this time. Senior Agentic AI Engineer Location: Houston, TX Onsite Flexibility: Hybrid — 3... ...applications, multi-agent workflows, and Retrieval-Augmented Generation (RAG) pipelines for enterprise use on Microsoft Azure. You will...Contract workWork visa3 days per week- ...thinking services company at the forefront of AI-native innovation. We partner with... ...next-generation, agent-powered workflows engineered to scale in real-world settings. Our engineers... ...agents, orchestration, context engineering, RAG, workflows) in production environments....Full timeWork experience placementLive inWork at officeLocal area
$293.5k
General Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work Areas Analytics, Data & Research,... ...component pipelines, including:Retrieval-Augmented Generation (RAG)Fine-tuning and parameter-efficient tuningEmbedding generation...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week- ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’... ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...and Unity Catalog governance. You've built RAG and multi-agent systems on the Databricks...Full timeWork experience placementLive inWork at officeLocal area
- ...Advanced Technology Centers (ATCs) are the engine for reinvention in our clients’... ...deepest industry knowledge, the latest in Gen AI solutions, and tech expertise from around... ...Analyst, Cortex LLM functions. You've built RAG and multi-agent systems on the Snowflake AI...Full timeWork experience placementLive inWork at officeLocal area
$110k - $145k
...Job Description Summary The Senior Data Engineer designs and builds the AWS-native data foundation behind our enterprise AI applications - knowledge graphs, semantic layers... ...Experience supporting AI/ML or LLM systems - RAG pipelines, embeddings, eval datasets,...Permanent employmentContract workRemote workVisa sponsorshipWork visaRelocation package- ...Overview GSI Environmental Inc. (GSI) seeks a highly innovative AI Engineer to join our growing team and help how artificial intelligence,... ...is a Microsoft 365 environment. Experience building LLM/RAG applications, document automation, or AI assistants for...Remote work
$135k - $162k
General InformationJob TitleSenior AI/ML Engineering ManagerJob ID108808Work AreasTechnology & EngineeringEmployment TypePermanent Full-TimeLocation... ..., model serving and LLMOps infrastructure, and production RAG and retrieval systems. You manage and develop a team of Data...Full timeWork at officeLocal areaImmediate start$109k - $251k
...HPC & AI Performance Engineer This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from... ...Knowledge of HPC and AI system architecture , including CPUs , GPUs and other accelerators, memory, networking, storage, and...Permanent employmentWork experience placementInternshipWork at officeWorldwide2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to On-Prem AI Engineer: LLaMA/Mistral, GPUs & RAG. Be the first to apply!



