Principal AI Platform Engineer
Kutir Technologies
Position Title: Principal AI Platform Engineer
Location: New York NY (Hybrid 3 days onsite)
Duration: 12+ months
Job Description:
The Client is seeking a Principal AI Platform Engineer / Agentic AI Architect to lead the architecture, development and production deployment of an enterprise-grade AI platform supporting financial services, insurance, payments and healthcare operations.
This position requires a deeply technical architect who can design the complete AI platform layer, remain hands-on with engineering and guide multiple product teams. The selected consultant will own LLM infrastructure, agent orchestration, retrieval systems, model evaluation, enterprise integrations, observability, security and Responsible AI controls.
The ideal candidate will have successfully taken multiple Generative AI and Agentic AI solutions from proof of concept into highly regulated production environments.
Key Responsibilities:
Define the end-to-end architecture for a multi-tenant enterprise Agentic AI platform.
Design and implement multi-agent systems containing planner, supervisor, retriever, executor, validator and human-in-the-loop components.
Build reusable agent frameworks using LangGraph, LangChain, Semantic Kernel, AutoGen, CrewAI or comparable orchestration technologies.
Architect a model gateway supporting Azure OpenAI, AWS Bedrock, Anthropic Claude, Google Gemini, Meta Llama and internally hosted models.
Design Retrieval-Augmented Generation and GraphRAG solutions across structured and unstructured enterprise data.
Implement hybrid retrieval combining vector search, semantic search, keyword search, reranking and knowledge graphs.
Build enterprise memory layers supporting episodic, semantic and procedural agent memory.
Design MCP-native tools, connectors and secure integration patterns for enterprise applications.
Integrate AI services with SAP, Oracle Fusion, Salesforce, ServiceNow, payment platforms and internal data products.
Design and operate LLM inference and model-serving infrastructure across cloud and on-premises environments.
Establish LLMOps and MLOps pipelines covering model registration, deployment, evaluation, monitoring, rollback and continuous improvement.
Develop automated evaluation frameworks for accuracy, groundedness, hallucination, safety, latency and cost.
Lead fine-tuning, instruction tuning, RLHF, RLAIF, model distillation and post-training initiatives.
Implement AI guardrails, content filtering, PII masking, access controls and prompt-injection defenses.
Establish model-risk documentation, lineage, explainability and audit evidence for regulated environments.
Design Kubernetes-based deployment architectures using Helm, Argo CD, Terraform and GitOps.
Implement distributed tracing and AI observability using OpenTelemetry, Grafana, Prometheus and specialized LLM monitoring platforms.
Define platform standards, reusable reference architectures and engineering best practices.
Lead architecture reviews and provide technical direction to AI engineers, platform engineers and product teams.
Partner with security, legal, risk, data governance and executive stakeholders.
Own technical decisions concerning build-versus-buy, model selection, infrastructure and platform scalability.
Optimize GPU utilization, token consumption, inference latency and overall LLM operating costs.
Maintain hands-on involvement through prototyping, code reviews, troubleshooting and production support.
Required Qualifications:
15+ years of overall software engineering, platform engineering or enterprise architecture experience.
8+ years designing and operating cloud-native platforms in production.
5+ years of hands-on AI/ML engineering or machine-learning platform experience.
3+ years building production Generative AI solutions using commercial or open-source LLMs.
Demonstrated experience architecting and deploying at least two enterprise Agentic AI implementations.
Expert-level Python development experience, including FastAPI, Flask, REST, gRPC and event-driven microservices.
Advanced experience with LangGraph and at least two additional agent-orchestration frameworks.
Hands-on experience with Azure OpenAI, AWS Bedrock and at least one additional foundation-model platform.
Deep knowledge of RAG, GraphRAG, embeddings, semantic search, reranking and context engineering.
Production experience with at least three vector technologies, including pgvector, Pinecone, Weaviate, Milvus, OpenSearch or Azure AI Search.
Advanced experience with Neo4j or another enterprise knowledge-graph platform.
Hands-on experience designing and implementing Model Context Protocol servers and clients.
Experience with LangSmith, MLflow and at least one evaluation framework such as Ragas, TruLens, Phoenix or Arize.
Strong Kubernetes engineering experience, including Helm, service meshes, autoscaling and production troubleshooting.
Advanced Infrastructure-as-Code experience using Terraform and strong GitOps experience using Argo CD.
Experience implementing model gateways, LLM routing, fallback strategies and multi-model architectures.
Hands-on experience deploying self-hosted LLMs using vLLM, NVIDIA Triton, Ray Serve or Hugging Face TGI.
Experience with GPU infrastructure, inference optimization, batching, quantization and distributed model serving.
Strong working knowledge of PostgreSQL, ClickHouse, object storage and distributed caching technologies.
Experience implementing OpenTelemetry-based tracing across agents, models, tools and microservices.
Strong understanding of OAuth 2.0, OIDC, workload identity, secrets management, encryption and zero-trust security.
Experience delivering AI platforms within banking, insurance, payments or another highly regulated industry.
Strong knowledge of Responsible AI, model governance, data privacy and model-risk management.
Experience presenting complex architectural decisions to executive and nontechnical stakeholders.
Proven ability to lead globally distributed engineering teams while remaining hands-on.
Mandatory Domain Experience:
Recent production experience in at least two of the following: banking and payment operations; insurance claims; fraud detection or financial-crime operations; healthcare operations; financial regulatory compliance; enterprise case management and workflow automation.
Highly Preferred Qualifications:
Experience building agentic workflows for claims adjudication, payment exceptions, fraud investigation or healthcare case management.
Experience integrating AI agents with Oracle Fusion, SAP S/4HANA, Salesforce or ServiceNow.
Hands-on experience with reinforcement learning, RLHF, RLAIF or Direct Preference Optimization.
Experience taking a fine-tuned or distilled LLM/SLM into production.
Experience implementing confidential computing or private AI environments.
Experience with NVIDIA NIM, NeMo, CUDA or GPU scheduling.
Experience designing AI platforms capable of operating across AWS, Azure, GCP and on-premises infrastructure.
FinOps experience specifically focused on GPU and LLM workloads.
Experience with multi-tenancy, usage metering, chargeback and token-level cost allocation.
Published research, patents or recognized open-source contributions related to agent infrastructure, LLM evaluation or AI security.
Microsoft Azure AI Engineer, AWS Machine Learning Specialty, Google Professional Machine Learning Engineer or equivalent certification.
TOGAF, Kubernetes CKA/CKAD or advanced cloud architecture certification.
Master's degree or Ph.D. in Computer Science, Artificial Intelligence, Machine Learning or a related discipline.
Candidate Validation Requirements:
Candidates must be prepared to describe at least two production implementations, including the business problem and measurable outcome; complete architecture; models and orchestration frameworks; RAG or GraphRAG design; agent planning, memory and tool use; evaluation and hallucination controls; security and Responsible AI; cloud deployment; monitoring; and latency, accuracy and cost results.
Candidates whose experience is limited to prototypes, demonstrations, copilots or proof-of-concept implementations will not be considered.
$230k - $290k
...New York( Digital Solutions Group ) - Platform Engineering /Full Time /RemoteAHEAD helps large enterprises... ...enterprises design, build, and run AI agent platforms on top of it. This role... ...depth toward AI agent platforms. As a Principal Technical Consultant, you own that work...PrincipalFull timeWork at officeShift work- ...appgate.com. About the Role We're looking for a AI/ML Engineer (Senior/Staff/Principal) - Threat Detection who will design, build, and... ...aggregation systems that power our autonomous threat detection platform. You'll work at the intersection of identity...PrincipalFull timeFor contractorsWorldwide
$265k - $285k
...Principal AI Engineer DriveWealth is the pioneer of fractional equities trading and embedded investing. We are a visionary technology company... ...the world. Our mission is realized through an API-based platform, empowering our partners to offer seamless investing and...PrincipalWorldwide- As an AI Platform Engineer for AI & Emerging Tech, you will drive AI platform enablement across the enterprise. This role sits at the intersection of engineering, governance, and user enablement: you will partner with Information Security, Risk, Legal and Compliance teams...Suggested
- Freshworks is seeking a Principal AI Solutions Engineer to serve as the AI technical leader for our North America Employee Experience business. This is a highly strategic, customer-facing role focused on accelerating AI revenue growth, helping customers realize measurable...Principal
- ...Principal AI Engineer Midtown Manhattan, New York City At our consulting company, we solve the most complex and critical challenges by... ...at the intersection of software engineering, data science, platform architecture, and AI governance. You will be the technical...PrincipalFull timeWork at officeRemote workRelocation package
- ...capital and venture buyout strategies, each powered by an AI operating system and team of leading technologists, entrepreneurs... ...We're hiring an ambitious, self-initiated, product-minded Principal Applied AI Engineer to build the agentic workflows that power Redesign's most...Principal
- ...Job Title: AI Platform Engineer Location: NYC, NY (Hybrid Model) - 10003 Zip code Energy & Utility domain with experience in Google solutions . 1 level in-person interview Only Locals AI Platform Engineer: Design and implement AI/ML infrastructure...Local area
$206.6k - $284.3k
...today about how it will leverage AI over the next decade. Few have... ...its viability through hands-on engineering. This role requires both. We build the platform that transforms millions of clinical... ...outcomes for members. As a Principal AI Applied Engineer, you will define...PrincipalFull timeTemporary workApprenticeshipWork at officeRemote workRelocation3 days per week$185k - $264.5k
...Within the Tech, Data, AI, Ventures (TDAV) organization, our work is guided by a... ...built to last. Role Overview As a Principal AI Engineer within the Artificial Intelligence &... ...solutions on New York Lifes enterprise AI platform. Rather than developing the underlying...Principal- Capital One is seeking a Lead AI Engineer for Gen AI Platform Services in New York, focusing on Agentic AI, guardrails, memory, and evaluation. You will collaborate with engineers, researchers, and PMs to build scalable AI-powered products for customers and associates....
$130k - $147k
...institutions to put customers at the center of every decision. Our AI-first platform transforms proprietary data, advanced analytics and deep... ...Job Description We are seeking a Senior ML/AI Platform Engineer to help build and operate Curinos' Databricks-native AI...Part timeWork at officeRemote workWork from homeFlexible hours- About the job Senior AI Platform Engineer Job Title: Senior AI Platform Engineer Location: New York (Hybrid - 4 days/week onsite) Interview Process: Onsite interview required Company Overview Glint Tech Solutions is a women-owned global IT staffing and recruiting firm...
- Title: Senior AI Platform Engineer Client: CPPIB Location: Hybrid - New York City, NY Visa: USC/GC Only Duration: 6+ Months Interview: Video and In-person Rate: $65-70/hr. C2C Note: Candidate must have to include a summary of why he/she is a good fit for the role...
- Senior AI Platform Engineer Join Hire Hangar and work with fast-growing global companies while building a long-term career. Location: Remote Time Zone: Flexible / Aligned to HQ Time Zone Role Overview: This role sits at the intersection of AI engineering and platform development...Work experience placementRemote workFlexible hours
- Owning the administration and performance of critical data and AI platforms, the full-time Principal Data Platform Engineer will manage Power BI, AtScale, and Databricks while ensuring reliability, cost optimization, and governance in a remote environment. Key Responsibilities...PrincipalFull timeRemote work
$155k - $215k
...mission is to develop a firmwide Artificial Intelligence (AI) Development Platform that aligns with the firm's Technology principles and drives... ...AI across our businesses. This role is for a platform engineering specialist who will help build a firmwide AI Development...Full timeTemporary work$162k - $215k
...experience building and operating production AI systems, including LLM-based workflows.... ...of modern AI methods such as prompt engineering, retrieval-augmented generation, fine-tuning... ...designing and delivering shared platforms or services at scale. We need deep technical...Full timeApprenticeshipWork at officeWork from homeFlexible hours$141k - $177k
...runtime insights, open innovation, and agentic AI. Creators of technology trusted by over 6... ...manage the MCP registry. Own the agent platform other teams build on — the deploy pattern... .... Qualifications 6+ years in software engineering, platform engineering, or DevOps,...Full time- Job Title: AI Platform Engineer Job Summary We are looking for an AI Platform Engineer to design, build, and maintain scalable AI infrastructure and platforms that enable data scientists and machine learning engineers to develop, deploy, monitor, and manage AI/ML models...Full timeRemote work
$100k - $150k
...consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This... ...organization offering tremendous career growth potential. Job Title: AI Platform Engineer Location: 100% Remote (Continental United States) Position...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship$125k - $165k
AI Platform Engineer TELCOR Inc, a leading innovator in laboratory software, is looking for an AI Platform Engineer to join our TELCOR AI Systems team! This role is ideal for a product-minded engineer who enjoys shipping fast, owning features end-to-end, and building high...Work at officeRemote work$180k - $200k
...securities, mortgage servicing rights, and other credit-related assets. POSITION SUMMARY Bayview Asset Management is seeking an AI Platform Engineer to help build the firm's Enterprise AI Platform that powers intelligent applications, AI agents, workflow automation, and...Work at officeLocal area- AI Platform Engineer Location-Type: Remote (US Based) Start Date Is: ASAP Duration: 1-2 Month Contract Compensation Range: $60-80/hour W2 Benefits: Eligible for Health, Dental, Vision, 401K Must be authorized to work in the U.S. This position is not eligible for sponsorship...Contract workImmediate startRemote work
- HireNow Staffing is seeking a Real-Time Voice AI Agent Systems Engineer for our NYC client. You will own production-grade voice AI infrastructure... ...KPI-driven improvements. Candidates must have 3+ years in platform/infra engineering, 2+ years in voice AI, and strong Node....Visa sponsorship
- AI Platform Engineer The AI Platform Engineer II is an experienced, hands-on engineer who helps build, automate, deploy, and operate enterprise AI applications and platform capabilities. This is an engineering-first AI role. Strong software or platform engineering fundamentals...
$250k
Castleton Tower Consulting, LLC is seeking a Technical Leader to design and build data platforms for investment firms in New York City. The role focuses on hands-on engineering while engaging with clients to ensure robust, automated tools. Ideal candidates will have 7+...- To support a growing data platform, the full-time remote Principal AI/ML Engineer will contribute to building and scaling production-grade AI/ML systems, focusing on data ingestion, DevOps reliability, and security compliance. Key responsibilities: Contribute to the platform...PrincipalFull timeRemote work
- Designing and building the technical layer of the AI platform, the full-time AI Platform Engineer will manage integration surfaces, internal tooling, and governed integration gateways while working remotely. Key responsibilities Design and build the technical layer of...Full timeRemote work
$150k - $210k
...communities. We have multiple Software Engineering positions open at Associate, Director... ...help modernize core accounting and P&L platforms used daily by Finance and Treasury. We are... ...and apply new technologies, including AI/ML and prompt‑based approaches , where...PrincipalFull timeTemporary workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal AI Platform Engineer. Be the first to apply!
- senior chief engineer New York, NY
- associate director engineering New York, NY
- general engineer New York, NY
- project engineer assistant project manager New York, NY
- chief design engineer New York, NY
- principal infrastructure engineer New York, NY
- principal cloud engineer New York, NY
- assistant chief engineer New York, NY
- chief engineer New York, NY
- principal developer New York, NY



