Senior Platform Engineer
The Judge Group Inc
Senior Platform Engineer
Irvine, CA
Fulltime We are seeking a Senior Platform Engineer to build, administer, automate, secure, and operate enterprise Databricks and cloud data-platform environments. The role is responsible for Databricks workspace administration, Unity Catalog governance, identity and access management, compute and cluster policies, infrastructure automation, CI/CD enablement, monitoring, reliability, cost management, and production support.
The ideal candidate combines hands-on Databricks administration with strong AWS or Azure platform engineering, Terraform, Python, Kubernetes, CI/CD, security, networking, observability, and Site Reliability Engineering practices.
Key Responsibilities
Databricks Administration and Platform Engineering
• Administer Databricks accounts, workspaces, metastores, catalogs, schemas, external locations, storage credentials, connections, shares, recipients, and platform configurations.
• Provision and manage development, test, staging, and production workspaces using standardized, repeatable patterns.
• Configure workspace settings, repositories, jobs, notebooks, SQL warehouses, instance pools, job clusters, all-purpose compute, and serverless capabilities.
• Define and enforce cluster policies, approved runtime versions, libraries, init scripts, autoscaling, tagging, and compute-usage standards.
• Manage Databricks Runtime and platform upgrades, compatibility testing, release planning, maintenance windows, and rollback procedures.
• Support Databricks Workflows, Delta Live Tables or Lakeflow Declarative Pipelines, Databricks SQL, MLflow, model registry, feature engineering, and structured-streaming services.
• Troubleshoot workspace, permissions, connectivity, compute, storage, job execution, library, runtime, and performance issues.
• Maintain administration standards, platform documentation, knowledge articles, support procedures, and operational runbooks.
Unity Catalog, Identity, Security, and Governance
• Design and manage Unity Catalog metastores, catalogs, schemas, managed and external tables, volumes, external locations, storage credentials, and grants.
• Implement user and group provisioning through single sign-on, SCIM, identity-provider integration, and enterprise directory services.
• Administer account-level and workspace-level users, groups, service principals, permissions, entitlements, and access-control models.
• Implement role-based and attribute-based access controls, least-privilege permissions, separation of duties, and privileged-access procedures.
• Configure secure access patterns for secrets, tokens, credentials, service principals, private endpoints, storage, and external systems.
• Enable audit logging, lineage, system tables, tagging, data classification, row-level security, column masking, and compliance reporting.
• Partner with security and governance teams on encrypti on, key management, network controls, data loss prevention, retention, auditability, and regulatory requirements.
• Review access, monitor privileged activities, remediate policy violations, and support internal and external audits.
Cloud Infrastructure and Networking
• Build and operate Databricks on AWS or Azure, including secure integration with cloud storage, identity, networking, encryption, and monitoring services.
• On AWS, work with S3, IAM, KMS, VPC, PrivateLink, security groups, Route 53, CloudWatch, Secrets Manager, and related services.
• On Azure, work with ADLS Gen2, Microsoft Entra ID, managed identities, Key Vault, virtual networks, private endpoints, network security groups, Azure Monitor, and related services.
• Configure control-plane and data-plane connectivity, private networking, DNS, routing, firewall, proxy, and egress controls.
• Integrate Databricks with cloud data lakes, APIs, databases, message platforms, and enterprise applications.
• Support Kubernetes, Docker, EKS or AKS, and container-based platform services where required.
• Contribute to capacity planning, disaster recovery, high availability, backup, restoration, and business-continuity exercises.
Infrastructure as Code and Automation
• Build and maintain reusable Terraform modules using cloud and Databricks providers.
• Automate workspace, network, Unity Catalog, storage, identity, compute-policy, cluster, job, permission, and monitoring configurations.
• Manage Terraform state, workspaces, variables, modules, versioning, policy checks, drift detection, and controlled promotion across environments.
• Use Python, shell scripting, Databricks CLI, REST APIs, and SDKs to automate administrative and operational tasks.
• Implement self-service workspace, catalog, schema, access, and compute vending with appropriate approval and governance controls.
• Maintain configuration standards and reduce manual administration through repeatable automation.
DevOps, CI/CD, and Release Engineering
• Design and support CI/CD pipelines using GitHub Actions, GitLab CI/CD, Jenkins, Azure DevOps, or Harness.
• Automate deployment of notebooks, jobs, workflows, libraries, policies, infrastructure, and platform configuration.
• Support Databricks Asset Bundles, Git integration, artifact management, environment promotion, testing, approvals, and rollback.
• Integrate security scanning, policy validation, infrastructure testing, and release evidence into delivery pipelines.
• Enable engineering teams through templates, reusable pipelines, documentation, and self-service platform capabilities.
• Partner with application, data-engineering, and DevOps teams to ensure deploy ment standards are consistent and supportable.
Reliability, Monitoring, Operations, and FinOps
• Establish monitoring, alerting, dashboards, logs, metrics, traces, and health checks for Databricks and connected cloud services.
• Use CloudWatch, Azure Monitor, Datadog, Splunk, New Relic, or similar platforms to monitor availability, compute utilization, failures, security events, and cost.
• Define platform service-level indicators, service-level objectives, operational metrics, and error budgets.
• Lead incident response, problem management, root-cause analysis, corrective actions, and post-incident reviews.
• Manage vulnerability remediation, runtime patching, dependency updates, security exceptions, and platform lifecycle activities.
• Optimize cluster sizing, autoscaling, pools, SQL warehouses, job concurrency, serverless usage, storage, and workload scheduling.
• Implement budget controls, chargeback or showback tagging, utilization reporting, anomaly detection, and cost-optimization recommendations.
• Participate in operational support rotations and maintain escalation paths with Databricks and cloud providers.
• Coordinate platform upgrades, disaster-recovery tests, security reviews, and production-readiness assessments.
Collaboration and Technical Leadership
• Partner with architecture, data engineering, security, cloud, network, governance, FinOps, and service-management teams.
• Advise engineering teams on Databricks platform standards, secure patterns, deployment models, performance, and cost.
• Conduct technical reviews and ensure solutions meet enterprise architecture and operational-support requirements.
• Mentor platform engineers and administrators and lead knowledge-transfer sessions.
• Communicate platform health, risks, dependencies, incidents, and improvement roadmaps to technical and business stakeholders.
• Drive continuous improvement in automation, reliability, security, developer experience, and operational efficiency.
Required Qualifications
• Typically 7-10 years of cloud, DevOps, Site Reliability Engineering, infrastructure, or platform-engineering experience.
• At least 3 years of hands-on Databricks platform administration in an enterprise environment.
• Strong experience administering Databricks workspaces, Unity Catalog, compute, cluster policies, jobs, SQL warehouses, permissions, and service principals.
• Strong experience with AWS or Azure infrastructure, identity, storage, networking, encryption, monitoring, and private connectivity.
• Strong proficiency with Terraform and infrastructure-as-code practices.
• Experience with Python, shell script
Irvine, CA
Fulltime We are seeking a Senior Platform Engineer to build, administer, automate, secure, and operate enterprise Databricks and cloud data-platform environments. The role is responsible for Databricks workspace administration, Unity Catalog governance, identity and access management, compute and cluster policies, infrastructure automation, CI/CD enablement, monitoring, reliability, cost management, and production support.
The ideal candidate combines hands-on Databricks administration with strong AWS or Azure platform engineering, Terraform, Python, Kubernetes, CI/CD, security, networking, observability, and Site Reliability Engineering practices.
Key Responsibilities
Databricks Administration and Platform Engineering
• Administer Databricks accounts, workspaces, metastores, catalogs, schemas, external locations, storage credentials, connections, shares, recipients, and platform configurations.
• Provision and manage development, test, staging, and production workspaces using standardized, repeatable patterns.
• Configure workspace settings, repositories, jobs, notebooks, SQL warehouses, instance pools, job clusters, all-purpose compute, and serverless capabilities.
• Define and enforce cluster policies, approved runtime versions, libraries, init scripts, autoscaling, tagging, and compute-usage standards.
• Manage Databricks Runtime and platform upgrades, compatibility testing, release planning, maintenance windows, and rollback procedures.
• Support Databricks Workflows, Delta Live Tables or Lakeflow Declarative Pipelines, Databricks SQL, MLflow, model registry, feature engineering, and structured-streaming services.
• Troubleshoot workspace, permissions, connectivity, compute, storage, job execution, library, runtime, and performance issues.
• Maintain administration standards, platform documentation, knowledge articles, support procedures, and operational runbooks.
Unity Catalog, Identity, Security, and Governance
• Design and manage Unity Catalog metastores, catalogs, schemas, managed and external tables, volumes, external locations, storage credentials, and grants.
• Implement user and group provisioning through single sign-on, SCIM, identity-provider integration, and enterprise directory services.
• Administer account-level and workspace-level users, groups, service principals, permissions, entitlements, and access-control models.
• Implement role-based and attribute-based access controls, least-privilege permissions, separation of duties, and privileged-access procedures.
• Configure secure access patterns for secrets, tokens, credentials, service principals, private endpoints, storage, and external systems.
• Enable audit logging, lineage, system tables, tagging, data classification, row-level security, column masking, and compliance reporting.
• Partner with security and governance teams on encrypti on, key management, network controls, data loss prevention, retention, auditability, and regulatory requirements.
• Review access, monitor privileged activities, remediate policy violations, and support internal and external audits.
Cloud Infrastructure and Networking
• Build and operate Databricks on AWS or Azure, including secure integration with cloud storage, identity, networking, encryption, and monitoring services.
• On AWS, work with S3, IAM, KMS, VPC, PrivateLink, security groups, Route 53, CloudWatch, Secrets Manager, and related services.
• On Azure, work with ADLS Gen2, Microsoft Entra ID, managed identities, Key Vault, virtual networks, private endpoints, network security groups, Azure Monitor, and related services.
• Configure control-plane and data-plane connectivity, private networking, DNS, routing, firewall, proxy, and egress controls.
• Integrate Databricks with cloud data lakes, APIs, databases, message platforms, and enterprise applications.
• Support Kubernetes, Docker, EKS or AKS, and container-based platform services where required.
• Contribute to capacity planning, disaster recovery, high availability, backup, restoration, and business-continuity exercises.
Infrastructure as Code and Automation
• Build and maintain reusable Terraform modules using cloud and Databricks providers.
• Automate workspace, network, Unity Catalog, storage, identity, compute-policy, cluster, job, permission, and monitoring configurations.
• Manage Terraform state, workspaces, variables, modules, versioning, policy checks, drift detection, and controlled promotion across environments.
• Use Python, shell scripting, Databricks CLI, REST APIs, and SDKs to automate administrative and operational tasks.
• Implement self-service workspace, catalog, schema, access, and compute vending with appropriate approval and governance controls.
• Maintain configuration standards and reduce manual administration through repeatable automation.
DevOps, CI/CD, and Release Engineering
• Design and support CI/CD pipelines using GitHub Actions, GitLab CI/CD, Jenkins, Azure DevOps, or Harness.
• Automate deployment of notebooks, jobs, workflows, libraries, policies, infrastructure, and platform configuration.
• Support Databricks Asset Bundles, Git integration, artifact management, environment promotion, testing, approvals, and rollback.
• Integrate security scanning, policy validation, infrastructure testing, and release evidence into delivery pipelines.
• Enable engineering teams through templates, reusable pipelines, documentation, and self-service platform capabilities.
• Partner with application, data-engineering, and DevOps teams to ensure deploy ment standards are consistent and supportable.
Reliability, Monitoring, Operations, and FinOps
• Establish monitoring, alerting, dashboards, logs, metrics, traces, and health checks for Databricks and connected cloud services.
• Use CloudWatch, Azure Monitor, Datadog, Splunk, New Relic, or similar platforms to monitor availability, compute utilization, failures, security events, and cost.
• Define platform service-level indicators, service-level objectives, operational metrics, and error budgets.
• Lead incident response, problem management, root-cause analysis, corrective actions, and post-incident reviews.
• Manage vulnerability remediation, runtime patching, dependency updates, security exceptions, and platform lifecycle activities.
• Optimize cluster sizing, autoscaling, pools, SQL warehouses, job concurrency, serverless usage, storage, and workload scheduling.
• Implement budget controls, chargeback or showback tagging, utilization reporting, anomaly detection, and cost-optimization recommendations.
• Participate in operational support rotations and maintain escalation paths with Databricks and cloud providers.
• Coordinate platform upgrades, disaster-recovery tests, security reviews, and production-readiness assessments.
Collaboration and Technical Leadership
• Partner with architecture, data engineering, security, cloud, network, governance, FinOps, and service-management teams.
• Advise engineering teams on Databricks platform standards, secure patterns, deployment models, performance, and cost.
• Conduct technical reviews and ensure solutions meet enterprise architecture and operational-support requirements.
• Mentor platform engineers and administrators and lead knowledge-transfer sessions.
• Communicate platform health, risks, dependencies, incidents, and improvement roadmaps to technical and business stakeholders.
• Drive continuous improvement in automation, reliability, security, developer experience, and operational efficiency.
Required Qualifications
• Typically 7-10 years of cloud, DevOps, Site Reliability Engineering, infrastructure, or platform-engineering experience.
• At least 3 years of hands-on Databricks platform administration in an enterprise environment.
• Strong experience administering Databricks workspaces, Unity Catalog, compute, cluster policies, jobs, SQL warehouses, permissions, and service principals.
• Strong experience with AWS or Azure infrastructure, identity, storage, networking, encryption, monitoring, and private connectivity.
• Strong proficiency with Terraform and infrastructure-as-code practices.
• Experience with Python, shell script
Vacancy posted 21 hours ago
Similar jobs that could be interesting for youBased on the Senior Platform Engineer in Tustin, CA vacancy
- ...Senior Platform Engineer We are seeking a Senior Platform Engineer to build, administer, automate, secure, and operate enterprise Databricks and cloud data-platform environments. The role is responsible for Databricks workspace administration, Unity Catalog governance...Senior
$151.2k - $204.6k
...interview process.About the RolePlatform Core Engineering (PCE) builds the foundational systems,... ...'s engineers build on every day. As a Senior Product Manager on PCE, your customers... ...on earning their trust. You will own a platform product area and be measured on whether...SeniorFlexible hours- ...Required Skills ~ Bachelor’s degree in computer science, Engineering, or a related field (or equivalent practical experience).... ...years of cloud infrastructure experience ~7+ years of Azure platform engineering expertise ~ Proven experience delivering enterprise...Senior
$120k - $202.5k
Who We Are Looking For State Street's Cyber Data & Analytics (CyberDNA) team is seeking a Sr.Platform Operations Engineer to help shape the next generation of cybersecurity data, analytics, and AI-powered platforms. Partnering closely with Global Cyber Security, Infrastructure...SeniorFull timeTemporary workFlexible hours$166k - $220k
...Unmanned Aerial System (UAS) threats. Working across product, engineering, sales, logistics, operations, and mission success, the... ...environments worldwide.ABOUT THE JOBWe are seeking a Senior Mission Data Platform Engineer to design, build, and operate the foundational data...SeniorFull timeWork experience placementImmediate startWorldwide$136.75k - $218.8k
...demand professional development resources that allow you to hone existing skills and learn new ones "I can succeed as a Senior Platform Engineer at Capital Group." As a Senior Engineeron our enterprise-wide technology team, you won't just maintain infrastructure...SeniorTemporary workLocal areaFlexible hours$150k - $250k
...alongside some of the brightest minds in aerospace, this is your chance to make a real impact. We're seeking a Senior Platform/Real-Time Embedded Firmware Engineer with strong low-level embedded software expertise and a passion for building robust, deterministic firmware...Senior- ...approaches or pure transformer-only architectures, combining rigorous engineering with learning systems proven in globally deployed solutions... .... ~4+ years of industry experience in ML infrastructure or platform engineering. ~ Strong coding skills in Python/TypeScript...SeniorLocal area
- Job Title Match Insights % Know your job match score, get recommended jobs and connect with recruiters looking for talent like you! Log in or create your free ClearanceJobs profile Job Requirements Security Clearance not specified Polygraph not specified ...Senior
- ...every piece of hardware on the board comes up and works when our image boots. We’re hiring a Senior Software Engineer to own that layer.This role sits at the seam between our platform organization, which maintains the underlying NixOS platform and vendor BSP integration,...SeniorFull timeWork experience placementLive inImmediate start
$191k - $253k
A leading defense technology company in California seeks a skilled engineer for its mission simulator platform. Candidates should have strong C++20 skills and experience with performance optimization in simulation environments. Responsibilities include designing APIs and...Senior$237.15k - $300k
...Google, located in Irvine, CA, seeks a senior engineering leader to own outcomes and influence stakeholders while solving ambiguous problems... ...Requires strong domain expertise and people leadership. The Platforms and Devices team builds computing platforms across desktop,...SeniorWork at office- ...compared to 60% at a typical company. Role Summary As a Senior Software Engineer specializing in agentic applications, you will be a key technical leader and influential voice in shaping our GenAI platform's architecture and strategy. You will play a pivotal role in...SeniorFull timeContract work
$60 per hour
...more exclusive features. Direct message the job poster from Yoh, A Day & Zimmermann Company Senior Technical Recruiter at Yoh, A Day & Zimmermann Company Senior Platform Engineer with Large Media Company Location: Hybrid - 1-2 days per week onsite in Aliso Viejo, CA...SeniorHourly payContract work2 days per week1 day per week$191k - $253k
Anduril Industries is looking for a Senior Software Engineer in Costa Mesa, California. The role involves architecting, building, and scaling digital backbone systems for smart factories. Candidates should have 5+ years of experience in software engineering and proficiency...Senior- A leading marketing services company in Newport Beach is looking for a Salesforce Platform Engineer to design and optimize their Salesforce platform. You will work with cross-functional teams to implement solutions, manage data integrity, and support end-users. Ideal candidates...SeniorFull time
$128.6k - $160.8k
...individuals who are excited by emerging technologies and eager to use them responsibly to solve problems and drive impact. The Sr. Platform Engineer will be responsible for designing, building, and maintaining internal tools and platform capabilities that accelerate...SeniorFull timeLocal areaFlexible hours- A global technology firm is seeking an experienced Data Scientist / Senior Consultant to drive the strategy and execution of their search and recommendation systems. The ideal candidate will possess a Master’s or Ph.D. in a quantitative field and over a decade of relevant...Senior
$150k - $165k
...Job Description Job Description 10779 – Manager, Platform Engineering Location: Irvine, CA 92614 (5 days on-site) Company Overview: Hyundai Autoever America (HAEA) is the dynamic IT powerhouse behind Hyundai Motor Corporation, a Fortune 500 global leader in...Work experience placementLocal area- ...make a meaningful impact. Learn more at About the Data Platform Team Every robot we deploy generates a continuous stream of... ...them, and turn them into the ranked problem list that drives our engineering roadmap. About the Job As a Data Platform Engineer,...Work at officeFlexible hours
- ...cloud environments including Google Cloud Platform (GCP), Microsoft Azure, and Amazon Web... ...). This role partners with the CISO and senior leadership to modernize security capabilities... ..., FedRAMP, CMMC) Mentor architects and engineers on cloud security best practices Cloud...SeniorFull time
$191k - $253k
...the military in months, not years.ABOUT THE TEAM The Simulation Platform team owns the core architecture and infrastructure of Anduril'... ...ergonomic UI atop them, working alongside Simulation Modeling engineers.Custom Entity-Component-System architecture that underpins the...SeniorFull timeWork experience placementImmediate start$191k - $253k
...networking technology to the military in months, not years.ABOUT THE TEAMThe Software Platform team at Anduril is dedicated to building the products and systems that empower our engineering teams to develop, test, and deploy cutting-edge defense technology with...SeniorFull timeWork experience placementImmediate start$95.5k - $181.7k
...push the boundaries of known science and find new ways to connect and protect our world.We are currently seeking a Senior Software Engineer to join the Platform and Infrastructure Services (IS) team in the Air C2 (Command and Control) and Battlefield Sensors software...SeniorTemporary workWork experience placementWork at officeRemote workRelocation packageFlexible hours$175k - $240k
....YouAre excited to be part of a vibrant engineering community that values diversity, hard work... ...solutions.We are seeking a VP‑level Senior Software Developer to lead development and... ...production systems, and experience with cloud platforms (AWS preferred).Familiarity with...SeniorFull time$220k - $292k
...develops Lattice for Mission Autonomy, Anduril’s premier software platform that enables masses of Fury, Barracuda, and other first and... ...teams like Perception, Motion Planning, Hardware, and Test Engineering to solve some of the hardest problems facing our customers. This...SeniorFull timeWork experience placementImmediate start- Anduril Industries is a defense technology company driving advanced autonomy, AI, and ERP-like integration across its CorpTech Platform, ArsenalOS, and CorporateOS. This role contributes to high-availability, enterprise-grade software powering production and corporate...Senior
$171.3k - $299.8k
...reach, diverse solutions and services portfolio, and digital platform Ingram Micro Xvantage set us apart. Learn more at Come join our... ...be a fun journey!Ingram Micro is seeking a Director, Platform Engineering to own both the day-to-day operational health of our technology...Full timeTemporary workWorldwideShift work- ...YOU WILL DOThe Sr. Application Developer Associate, Insurance Platform supports auto insurance products and platforms through SQL and... ...Bachelor’s degree in Computer Science, Information Technology, Engineering or a related field.Relevant certifications preferred, but not...SeniorFull timeContract workWork at officeLocal areaImmediate start
- Anduril Industries seeks a Senior Software Engineer in Costa Mesa, California, to architect and build digital solutions for smart factories. This role demands 5+ years of experience and proficiency in technologies like Javascript and React. The ideal candidate excels in...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Platform Engineer. Be the first to apply!
Related searches
- senior network engineer remote Tustin, CA
- senior manager legal Tustin, CA
- senior commercial counsel Tustin, CA
- senior manager tax Tustin, CA
- senior living Tustin, CA
- senior implementation project manager Tustin, CA
- senior level Tustin, CA
- senior cloud network engineer Tustin, CA
- senior activities Tustin, CA
- international tax senior Tustin, CA



