Staff Engineer, Platform
$160k - $290kShield AI
Shield AI is a venture-backed defense-tech company with the mission of protecting service members and civilians with intelligent systems. Its products include Hivemind autonomy software, V-BAT and X-BATaircraft, and Aechelon simulation and synthetic reality technologies. With offices and facilities across the U.S., Europe, the Middle East, and Asia-Pacific, Shield AI’s technology actively supports operations worldwide. For more information, visit Follow Shield AI on LinkedIn, X, Instagram, and YouTube.
Job Description:
We are looking for a Staff Platform Engineer tocontribute to the Platform solutions and infrastructure that power Forge, the AI Factory.These componentsprovidedistributed runtime capabilities that enable teams to reliably orchestrate workloads,process data,and underpin the workflows of building an AI Pilot.
TheForgePlatform Engineering teamprovidesKubernetes-nativecapabilitiesthatsupport autonomy development, simulation, testing, training, evaluation, deployment, and operational workflows across commercial cloud, on-premises infrastructure, sovereign deployments, edge environments, and fully air-gapped systems.
This is a hands-on technical leadership role. You will define platform architecture, implement production software,establishreusable operational patterns, and partner with teams across autonomy, ML, simulation, test, infrastructure, and product engineering. Success requires balancing developer productivity, reliability,extensibility,performance, portability, and long-term operational maintainability.
What you'll do:
- Build Kubernetes-native platform services: Develop and operate Kubernetes-based services, controllers, operators, deployment patterns, and runtime integrations that support distributed workloads across multiple environments.
- Develop distributed orchestration capabilities: Design and build reusable primitives for authoring, scheduling, and scaling pipeline work.
- Build reliable data-processing infrastructure: Develop platform capabilities for data storage, ingestion, validation, transformation, and governance.
- Develop highly extensible platform components: The Forge Platform base provides standardized tooling around authentication, authorization, observation, networking, routing, secret management, and more to the services that are integrated on top of the ecosystem.
- Create reference architectures: Establish recommended deployment patterns, operating profiles, capacity guidance, benchmarks, reliability practices, and distribution approaches across cloud providers, on-prem, edge, and air-gapped environments.
- Advance observability and operability: Establish end-to-end metrics, logs, traces, structured events, dashboards, alerting, service-level objectives, operational diagnostics, and runbooks for workflows, pipelines, event streams, and platform services.
- Partner with downstream teams: Work directly with autonomy, ML Ops, simulation, test, infrastructure, product, and customer-facing teams to turn recurring distributed-systems problems into reusable platform capabilities.
Key Outcomes:
- The Forge Platform continues to improve in KPIs around reliability, scalability, operational use cases, and customer adoption.
- New services from downstream teams are guided to successful platform integration. Shared distributed services have clear ownership, repeatable deployment patterns, tested recovery procedures, practical observability, and well-defined operational standards.
- The Platform is demonstrated, evaluated, and benchmarked across a wide variety of operational environments.
- Interfaces are maintained for long periods of time to instill customer confidence and reduce upgrade burdens.
Required qualifications:
- Significant experience designing and operating production distributed systems, cloud-native platforms, backend infrastructure, or data-intensive services.
- Strong software engineering skills and a record of delivering production systems in Go and Python.
- Deep understanding of distributed-systems fundamentals, including failure handling, idempotency, consistency tradeoffs, retries, ordering, delivery semantics, backpressure, partitioning, state management, and fault tolerance.
- Experience designing or operating workflow orchestration, distributed job execution, asynchronous processing, event-driven systems, or long-running service workflows.
- Ability to define architecture and technical standards while remaining hands-on in implementation, production troubleshooting, performance analysis, and reliability improvement.
- Experience working across multiple teams to turn recurring infrastructure needs into reusable, well-documented platform capabilities.
- Clear technical communication and the ability to make complex distributed-systems architecture understandable to both specialists and downstream users.
Preferred qualifications:
Experience in any of the following is beneficial but not required:
- Kubernetes controllers, operators, Custom Resource Definitions, admission control, scheduling extensions, KubeRay, or workload-management systems.
- Distributed execution and workflow technologies such as Ray, Temporal, Argo Workflows, Flyte, Dagster, Airflow, Prefect, Kubernetes Jobs, or comparable systems.
- Durable messaging and event-streaming technologies such as NATS JetStream, Kafka, Redpanda, Pulsar, RabbitMQ, SQS/SNS, or comparable systems.
- ETL/ELT, batch processing, event-driven data pipelines, CDC, schema evolution, data validation, artifact processing, or large-file transfer workflows.
- Service networking technologies and practices such as Envoy, service meshes, Kubernetes networking, and CNI plugins.
- Observability software such as OpenTelemetry, Prometheus, Grafana, Loki, Tempo, Jaeger, distributed tracing, structured logging, SLOs, service-level indicators, alerting, and incident-management practices.
- Terraform, Helm, ArgoCD, GitOps, Kubernetes package management, repeatable platform distribution, and Infrastructure as Code.
Why join us:
Platform Engineering is foundational to how Shield AI develops, tests, evaluates, deploys, and operates autonomy systems. This role offers the opportunity to shape the distributed systems foundation used by engineering teams across the company and delivered into demanding customer environments.
Your work will determine how reliably data moves through the organization, how services coordinate across complex environments, how teams execute and recover long-running workflows, and how operators understand the health of mission-critical systems. You will work at the intersection of Kubernetes, distributed computing, event-driven architecture, data processing, networking, observability, and autonomy.
You will help establish reusable platform capabilities that allow specialized teams—including autonomy, simulation, test, ML Ops, and application engineering—to move faster while building on dependable operational foundations.
$160,000 - $290,000 a year
San Diego, CA pay range: $160,000 - $240,000
San Mateo, CA pay range: $190,000 - $290,000
Full-time regular employee offer package:
Pay within range listed + Bonus + Benefits + Equity
Temporary employee offer package:
Pay within range listed above + temporary benefits package (applicable after 60 days of employment)
Salary compensation is influenced by a wide array of factors including but not limited to skill set, level of experience, licenses and certifications, and specific work location. All offers are contingent on a cleared background and possible reference check. Military fellows and part-time employees are not eligible for benefits. Please speak to your talent acquisition representative for more information.
Shield AI is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed toequal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, marital status, disability, gender identity or Veteran status. If you have a disability or special need that requires accommodation, please let us know.
#J-18808-Ljbffr$205.65k - $342.75k
...What is the opportunity? As a DevOps Engineer at Omnissa, you’ll play a key role in advancing the reliability, scalability, and automation... ...of Horizon Cloud, our cutting‑edge Desktop‑as‑a‑Service (DaaS) platform. You’ll work hands‑on across CI/CD, cloud infrastructure,...SuggestedWork experience placementLocal areaWorldwideVisa sponsorship$194k - $267k
...Staff Site Reliability Engineer - Kubernetes Important: if an employer asks you to log into their system via iCloud or Google, send a code, an... ...will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours$181k - $265k
...wealth management industry by building an AI platform for wealth professionals. We partner... ...we’re excited to hire a high performing Staff SRE to join our growing Platform team.... ...teams. You will collaborate closely with engineering leadership, product managers, and cross-...SuggestedWork at officeImmediate start3 days per week$225k - $300k
...machine learning to automate performance engineering. Today, we help our customers reduce... ...Daniel Gross. About the Role As a Staff Infrastructure Engineer, you will design... ...reliability, and observability across the platform Influence architectural decisions and...Suggested$180k - $250k
...SentiLink is hiring a remote candidate for Staff Infrastructure Engineer. This is a full time position. Work location: USA. The role typically... ...infrastructure that serves as the foundation of the SentiLink platform. You’ll operate as a technical leader across...SuggestedFull timeWork at officeRemote workHome officeFlexible hours- ...Replit is the agentic software creation platform that enables anyone to build applications... ...role: Join our Site Reliability Engineering (SRE) team and help ensure the reliability... ...millions of developers worldwide. As a Staff Site Reliability Engineer, you will bridge...Full timeTemporary workWork at officeWorldwideFlexible hours
$226k - $283k
...scribe. We're building the AI intelligence platform that restores humanity to healthcare and... ...can't be bolted on - it must be engineered into the product. This is a senior technical... ...secure paved paths. Who You Are: Staff-level product security judgment. You...Work at office3 days per week$190k - $235k
...bathymetric and imagery data products to customers through our own platform, and we're scaling toward continuous, around-the-clock data... ...on shore while cutting cost and time. The Role The Staff Robotics Engineer sets the technical direction for autonomy and vehicle...Visa sponsorshipWork visa$180k - $250k
...makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and... ...ecosystem that ambitious teams build on. You are a hands-on engineer who builds the software and processes that keep a large fleet of...Temporary workLocal areaRelocation package$204k - $247k
...Senior Software Engineer Join us in building the next generation of AI infrastructure that will power innovation across our organization. We're seeking a senior software engineer to support our Platform team, which owns the AWS infrastructure, Kubernetes clusters, and...$216k - $270k
...throughout life. As the world adjusts to this new reality, leading platform companies are scrambling to build LLMs at billion scale, while... .... At Scale, our products include the Generative AI Data Engine, SGP, Donovan, and others that power the most advanced LLMs and...Full timeLive in- ...Senior Software Engineer, Platform FullTime Professional Round Rock, TX, US 16 days ago Requisition ID: 1024 Job Purpose/Summary The Senior Platform Engineer will develop and implement the backend and data platform architecture for our next-generation space...Full timeContract work
$208k - $312k
...Engineering New York City, San Francisco Full Time About Vercel: Vercel is the agentic infrastructure company. We free people... ..., extended, and operated by agents. We are building the platform for that future, trusted by companies like OpenAI, PayPal, Ramp...Full timeWork at officeRemote workWork from homeWorldwideMonday to FridayFlexible hours$236k - $300k
...Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how... ...— that separates good products from great ones. As a Staff Design Engineer at Mixpanel, you'll sit within the design organization and...Contract workLive in- ...Position: Staff Engineer, Machine Learning Operations The Staff Engineer, Machine Learning Operations will architect, own,... ...years architecting and deploying production ML systems on cloud platforms (Azure preferred) Proven track record building and scaling...Flexible hours
$185k - $260k
...About Zip Zip is the AI platform for enterprise procurement — built for humans and agents working together. By orchestrating procurement... ...LinkedIn Top Startups. Your Role As a Senior Software Engineer on the Developer Platform team, you will be responsible for...Immediate startRemote workHome officeFlexible hours$195k - $227.5k
...Chaturbate, one of the most heavily trafficked live streaming platforms in the world. We support a global network of independent... ...impact of what they build. The Role We are looking for a Staff Database Engineer to take primary ownership of the reliability, availability,...Temporary workWork at officeRemote work$166.9k - $230.9k
...technology that’s both radically intelligent and deeply human. Our platform runs over one million predictions per borrower using more than... ...hear from you. The Team: As a member of our Servicing Engineering team, you will play a pivotal role in ensuring smooth loan...Summer workCurrently hiringLocal areaRemote workWork from home$164.2k - $240k
...massive positive impact on the hundreds of millions of drivers who carry auto insurance in the US. Root's Engineering team is committed to building a flexible platform on which our product designers and quantitative scientists can quickly test ideas, deploy them into...Flexible hours- ...Chief of Staff for Engineering About Us Parity is one of the world's most experienced companies building the core infrastructure behind blockchain - the system that allows information and value to be shared securely without needing an intermediary. We're laying the foundation...For contractorsRemote workShift work
$217k - $303.9k
...information, visit redditinc.com. Reddit is hiring a Staff Product Security Engineer to make the secure path the easiest path for engineers and... ...gap structurally - through guardrails, automation, and platform-level prevention that scale with the engineering org....For contractorsWork experience placementRemote work$215k - $250k
...profit. Billions of dollars in commerce run through the Upside platform every year, and that value goes directly back to our... ...sustainability initiatives. About the role: We're looking for a Staff Data Engineer to serve as a technical leader on the Data Engineering team...Full timeWork at officeRemote workFlexible hours$175k - $225k
...across our products. That means raising the engineering bar for how we design, build, and... ...together end to end. We're looking for a Staff Frontend Engineer who already works this... ...more time on frontend architecture, platform design, and technical direction across teams...Local areaWorldwideShift work- ...About CrewAI CrewAI is the leading framework and enterprise platform for building and orchestrating multi-agent AI systems,... ...agentic automations in production. This is a full-stack product engineering role with a strong frontend bias. You'll work across Rails, React...
$186.07k - $218.9k
...quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase. As a Senior Software Engineer on the Data Platform team within the Platform group, you will help build the systems that centralize all of Coinbase's internal and third-...Local area- ...engaged community of millions of daily active users who use the platform for many different reasons, but there’s one thing that nearly... ...for building lovable products for Discord users and Discord engineers. We’re building the next generation Data Platform that powers...Worldwide
$229.9k - $262.4k
...Senior Lead AI Engineer (GenAI Platform Services) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized...Full timePart timeLocal area$250.8k - $286.2k
...Senior Lead AI Engineer (GenAI Platform Services) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer...Full timePart timeLocal area$253.9k - $298.7k
...quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase. As a Senior Staff Software Engineer on the Data Platform team within Platform, you'll define and lead the technical strategy for Coinbase's data infrastructure, spanning...Local area$190k - $220k
| About us: Foodsmart is the leading Foodcare platform in the U.S., built to deliver nutrition-driven healthcare at scale. Powered... ...food. | About the role: Foodsmart is seeking a Staff AI Engineer who will play a critical leadership role in shaping the architecture...Local areaRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, Platform. Be the first to apply!
- engineering aide Eastern, KY
- assistant engineer Eastern, KY
- technology administrator Eastern, KY
- staff engineer Eastern, KY
- platform engineer Eastern, KY
- senior platform engineer Eastern, KY
- platform developer Eastern, KY
- platform product manager Eastern, KY
- power platform Eastern, KY
- platform manager Eastern, KY

