Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal, Platform Engineering

$174.9k - $292.3k

The Options Clearing Corporation

What You'll Do

Summary 2-3 sentences describing the overall responsibilities. This is a people-management role with full supervisory responsibility for a team of engineers and managers. The Director will lead and manage this team to drive reliability, observability, and cloud platform engineering excellence across a large, complex cloud-based computing environment. The ideal candidate is a hands-on, data-driven technical leader, who can personally raise the bar on SRE and observability practices, reduce waste, optimize cloud efficiency, and improve performance, in close partnership with a dedicated SRE/monitoring team and a centralized architecture function.

The individual will also serve as a player-coach, passionate about new technologies, mentoring technical teams through complex initiatives while keeping a steady eye on the extensive regulatory/compliance demands on our company (e.g., CIS, NIST).

Primary Duties and Responsibilities

  • Manage and develop a team of engineers and managers focused on SRE, observability and cloud platform engineering, growing and retaining talent
  • Lead hiring and staffing for the team, including sourcing, interviewing, and selecting engineers and managers to build out the organization
  • Meet with team members regularly to provide coaching and feedback on performance
  • Perform evaluations and deal effectively with staff problems and corrective actions as needed
  • Develop employee career development plans to assist with team member career growth and development
  • Manage and participate in the implementation of production changes during defined maintenance windows and support on-call rotation
  • Serve as a point of escalation within the team for reliability, observability, and platform support issues
  • Foster an atmosphere of trust, respect, and high performance while displaying strong ethics and integrity
  • Manage project and daily work task planning and prioritization, meeting project deadlines while maintaining a high quality of work
  • Ensure team compliance with all appropriate OCC policies and procedures, and institute corrective actions to address audit and other regulatory or compliance findings
  • Operate within budget; establish and assure adherence to schedules, work plans, and performance requirements

Technical & Platform Responsibilities

  • Bring deep, hands-on SRE and observability expertise to Platform Engineering, elevating the maturity, rigor, and technical depth of the organization’s observability, monitoring, and reliability practices in close partnership with the dedicated SRE and monitoring team
  • Serve as a technical authority on reliability engineering, mentoring engineers and managers and raising the bar on SLOs/SLIs, error budgets, incident response, and post-mortem discipline across the platform organization
  • Own the definition, governance, and lifecycle of SLOs and SLAs across Platform Engineering, codified as SLO-as-code (e.g., OpenSLO, Sloth, or Nobl9) and version-controlled in the repo rather than only living in a vendor UI
  • Drive resilience engineering as a discipline across the platform, including chaos engineering (e.g., AWS FIS, Gremlin) and load/performance testing, to proactively validate and improve system resilience ahead of production incidents
  • Deliver golden paths and templates (Terraform modules, Helm charts, pipeline templates) that ship with logging, metrics, tracing, dashboards, alerts and SLO defaults out of the box, reducing toil and accelerating safe onboarding for engineering teams
  • Act as product owner for observability within the platform organization, defining the roadmap and requirements for observability capabilities while the dedicated SRE and monitoring team retains ownership of the underlying tooling and operations
  • Own and drive platform engineering delivery and reliability metrics, including DORA metrics (deployment frequency, lead time for changes, change failure rate, MTTR), cycle time and throughput, holding the organization accountable to continuous improvement
  • Define and report on a rounded set of platform and observability metrics: platform adoption (percentage of services on paved roads, service catalog completeness, time to first deploy, onboarding time), observability coverage (percentage of tier-1 services with SLOs, instrumented traces, and runbooks), observability health (MTTD, MTTA, alert actionability ratio, toil percentage, and paging load per engineer), observability economics (telemetry cost per service, cardinality and ingestion growth, and retention tiers), and developer experience (SPACE/DevEx measures alongside DORA)
  • Champion the Golden Signals, RED, and USE methods as the platform’s common language for monitoring and alerting
  • Ensure audit-grade telemetry practices, including log retention and immutability, PII and sensitive-data scrubbing, and access controls on observability data, mapped to relevant compliance frameworks (e.g., CIS, NIST and SIFMU-specific resilience and reporting expectations)
  • Drive reliability across four core pillars: a data-driven culture ("data junkies") that uses telemetry and metrics, not intuition, to make decisions; waste reduction through elimination of toil and inefficient processes; cloud measures such as utilization, rightsizing, and FinOps metrics; and continuous performance optimization
  • Translate reliability signals (SLOs, error budgets, incident learnings) surfaced by the SRE organization into concrete cloud platform investments and roadmap priorities
  • Drive results in the cloud platform by building a reliable, scalable, secure technology stack in collaboration with engineers and leaders, using leading industry practices
  • Champion infrastructure-as-code (IaC) practices across the cloud platform, ensuring provisioning, configuration, and environment management are automated, version-controlled and consistently applied
  • Assess and plan for capacity needs within the cloud platform and forecast accordingly
  • Collaborate with the centralized architecture function on platform architecture decisions, ensuring cloud platform engineering execution aligns with broader architectural standards and direction
  • Implement and manage initiatives within your assigned area of responsibility with accountability for results and compliance with all controls and security requirements
  • Lead in the development of technology roadmaps and end-of-life technology plans
  • Effectively communicate project and operational service issues to senior management promptly with observations, decisions, and recommendations for corrective measures
  • Other duties as assigned

Supervisory Responsibilities

  • Manage a team of engineers and engineering managers focused on SRE, observability, and cloud platform engineering
  • Full people-management responsibilities, including hiring, performance management, coaching, career development, and corrective action

Qualifications

  • [Required] 5+ years of demonstrated experience leading engineering teams, with an emphasis on developing key talent and cultivating positive, high-performing cultures
  • [Required] 10+ years of progressive, hands-on experience in software engineering with an understanding of large-scale computing solutions (primarily AWS), including software design and development, database architectures, IP networking, security, cloud operations, and performance tuning
  • [Required] Demonstrated, hands-on expertise building and maturing SRE and observability practices at scale, able to personally raise the technical bar for a dedicated SRE/monitoring team, not just consume their output
  • [Required] Demonstrated track record owning the definition and governance of SLOs and SLAs for mission- critical systems, and personally driving resilience engineering practices (chaos engineering, load/performance testing) to validate reliability ahead of incidents
  • [Required] Demonstrated track record driving reliability through a data-driven culture, waste/toil reduction, cloud efficiency measures, and performance optimization, treating metrics as the primary lens for every decision
  • [Required] Experience defining, instrumenting, and acting on software delivery performance metrics (DORA metrics, cycle time, deployment frequency, lead time, etc.) to drive engineering improvement initiatives
  • [Required] Strong consultative, communication, team player, and analytical skills, with the ability to regularly interact between various teams distributed across the US
  • [Required] Strong technical team leadership and technical project management skills
  • [Required] Relevant experience leading highly technical team members through adopting new technologies while maintaining highly available, mission-critical systems, with a proven track record of success
  • [Required] Ability to clearly communicate verbally and in writing to business and technology leaders, architects, developers, and team members
  • [Required] Must be able to collaborate effectively with a group of high-performing, technical individuals
  • [Required] Experience acting as a product owner, defining roadmap, requirements, and priorities, for a platform capability such as observability, ideally in partnership with a separate team that owns the underlying tooling and operations
  • [Required] Experience with architecting, implementing, and maintaining highly available mission-critical environments for 24x7 availability
  • [Required] Demonstrated history of working within deadlines and ability to work well under pressure
  • [Required] Experience managing work tasks using Agile methodology/scrum desired
  • [Preferred] Comfort with ambiguity and demonstrated ability to lead complex programs in a decentralized environment
  • [Preferred] Experience working in an environment with a defined production change control process; experience working with audits and compliance or in a regulated environment a plus
  • [Preferred] Experience in organizations with a mature, centralized SRE function, avoiding role/scope overlap

Technical Skills

  • [Required] Deep expertise in OpenTelemetry, including instrumentation standards, auto-instrumentation, semantic conventions, and the OTel Collector, as the foundation for a vendor-neutral, paved-road instrumentation strategy
  • [Required] Hands-on experience with metrics engines and time-series databases, including Prometheus and PromQL, plus scale-out options such as Mimir, Thanos, VictoriaMetrics, or Amazon Managed Prometheus, including cardinality management
  • [Required] Hands-on experience with tracing and logging backends such as Tempo/Jaeger, Loki/Elastic/Splunk (Splunk is common in financial services), and AWS X-Ray
  • [Required] Experience with Kubernetes and Kafka observability specifically, including EKS metrics, kube-state-metrics, consumer lag, and broker health
  • [Required] Experience with alerting and incident tooling such as PagerDuty or Opsgenie, including ServiceNow integration, alert routing, and noise reduction
  • [Required] Hands-on experience with SLO-as-code frameworks such as OpenSLO, Sloth, or Nobl9, defining and governing SLOs in the repo rather than only in a vendor UI
  • [Required] Experience delivering golden paths and templates (Terraform modules, Helm charts, pipeline templates) that ship with logging, metrics, tracing, dashboards, alerts, and SLO defaults out of the box
  • [Required] Experience with resilience validation practices, including chaos engineering (e.g., AWS FIS, Gremlin) and load/performance testing
  • [Required] Deep, hands-on mastery of observability tooling and practices (metrics, distributed tracing, centralized logging, dashboards, alerting), e.g., Datadog, Prometheus/Grafana, CloudWatch, or equivalent, sufficient to elevate, not just consume, a dedicated SRE team’s capability
  • [Required] Deep understanding of SRE principles including SLOs/SLIs, error budgets, incident management, and post-mortem culture, with a track record of driving adoption and maturity
  • [Required] Fluency in the Golden Signals, RED, and USE methods as applied frameworks for monitoring and alerting design
  • [Required] Hands-on experience with: Terraform, Kubernetes, Jenkins or other CI/CD tooling, Kafka, Github, and configuration management tools such as Puppet, Chef, or Ansible
  • [Required] Deep, hands-on expertise with infrastructure-as-code (IaC) tools and practices (e.g., Terraform, CloudFormation, CDK, Pulumi), with a track record of driving IaC adoption at scale across a cloud platform organization
  • [Required] Relevant experience with configuration and implementation of IaaS, Infrastructure as Code, AWS, Azure, etc.
  • [Required] Expert working knowledge of infrastructure design and components, such as servers, operating systems, networks, and storage
  • [Preferred] Basic understanding of good delivery practices and continual integration and improvement; Agile/Lean background for projects and project delivery
  • [Preferred] Experience with telemetry pipeline tools such as Fluent Bit, Vector, or Cribl for routing, sampling, redaction, and cost control
  • [Preferred] Experience with synthetic monitoring, real user monitoring (RUM), and eBPF-based observability
  • [Preferred] Familiarity with audit-grade telemetry practices (log retention/immutability, PII and sensitive-data scrubbing, access controls on observability data) and mapping observability controls to compliance frameworks such as CIS, NIST, or SIFMU-specific resilience requirements
  • [Preferred] Familiarity with engineering metrics/analytics platforms (e.g., LinearB, Jellyfish, Sleuth, Haystack, or internal equivalents) used to track DORA metrics and delivery performance
  • [Preferred] Competent in all phases of application development and implementation, including SDLC; hands-on scripting/development skills in Python, Ruby, Go, Java, etc. in a corporate environment strongly desired
  • [Preferred] Experience establishing IaC governance and standards (module libraries, policy-as-code, drift detection) across multiple teams or business units
  • [Preferred] Experience building a metrics-driven engineering culture, including scorecards, dashboards, or leadership reporting on delivery performance

Education and/or Experience

  • [Required] Bachelor’s degree, preferably in a technical discipline (Computer Science, Mathematics, etc.), or equivalent combination of education and experience required; Master’s degree and relevant experience also considered
  • [Preferred] 10+ years’ experience in IT systems installation, operations, administration, and maintenance of cloud systems / virtualized servers, including 5+ years in a technical leadership role
  • [Preferred] Experience working in a financial services or highly regulated environment preferred

Certificates or Licenses

  • [Required] AWS Solutions Architect Associate Certification or higher strongly desired
  • [Preferred] Relevant industry certifications such as Microsoft Azure or Google Cloud

About Us

The Options Clearing Corporation (OCC) is the world’s largest equity derivatives clearing organization. Founded in 1973, OCC is dedicated to promoting stability and market integrity by delivering clearing and settlement services for options, futures and securities lending transactions. As a Systemically Important Financial Market Utility (SIFMU), OCC operates under the jurisdiction of the U.S. Securities and Exchange Commission (SEC), the U.S. Commodity Futures Trading Commission (CFTC), and the Board of Governors of the Federal Reserve System.

Benefits

  • A highly collaborative and supportive environment developed to encourage work-life balance and employee wellness
  • A hybrid work environment, up to 2 days per week of remote work
  • Tuition Reimbursement to support your continued education
  • Student Loan Repayment Assistance
  • Technology Stipend allowing you to use the device of your choice to connect to our network while working remotely
  • Generous PTO and Parental leave
  • 401k Employer Match
  • Competitive health benefits including medical, dental and vision

Visit for more information.

Compensation

The salary range listed for any given position is exclusive of fringe benefits and potential bonuses. If hired at OCC, your final base salary compensation will be determined by factors such as skills, experience and/or education. In addition, we believe in the importance of pay equity and consider internal equity of our current team members as part of any final offer. We typically do not hire at the maximum of the range in order to allow for future and continued salary growth. We also offer a substantial benefits package as noted on All employees may be eligible for a discretionary bonus. Discretionary bonuses are based on various factors, including, but not limited to, company and individual performance and are not guaranteed.

Salary Range $174,900.00 - $292,300.00

Incentive Range 23% to 30%

Equal Opportunity Employer

OCC is an Equal Opportunity Employer. OCC is an equal opportunity employer that is committed to diversity, equity, and inclusion. OCC provides equal employment opportunities to all employees and applicants for employment without regard to race, color, national origin, citizenship status, sex, sexual orientation, gender identity or expression, disability, age, marital status, religion, veteran status, or any other characteristics protected by applicable federal, state, or local laws.

If you need a disability related accommodation for any part of the application process, please email View email address on click.appcast.io.

#J-18808-Ljbffr

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Principal, Platform Engineering in Chicago, IL vacancy
  •  ...ll Do:This is a senior individual contributor role responsible for personally driving reliability, observability, and cloud platform engineering excellence across a large, complex cloud-based computing environment. The ideal candidate is a hands-on, data-driven technical... 
    Suggested
    Full time
    Remote work
    2 days per week

    The Options Clearing Corporation

    Chicago, IL
    1 day ago
  •  ...eligible.CCC Intelligent Solutions Inc. (CCC) is a leading cloud platform for the multi-trillion-dollar insurance economy, creating...  ...about CCC at .The RoleWe are seeking a highly skilled Platform Engineer with deep expertise in designing, deploying, and managing scalable... 
    Suggested
    Full time

    CCC Information Services

    Chicago, IL
    1 day ago
  • $119.4k - $204.6k

     ...Essential Responsibilities Define enterprise-wide platform strategy, vision, and target-state architectures for platforms such...  .... Establish standards and guardrails across all platform engineering domains. Drive innovation in cloud, data, and AI platforms... 
    Suggested
    Full time
    Temporary work
    Part time
    Work from home
    3 days per week

    Alliant Credit Union

    Chicago, IL
    1 day ago
  • $184k - $304k

     ...hire. This position is ineligible for employment Visa sponsorship. Job Description Overall Purpose The Principal Platform Security Engineer is a hands-on enterprise technical leader responsible for defining the long-term technical vision, secure target states... 
    Suggested
    Hourly pay
    Full time
    Work at office
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning®

    Chicago, IL
    3 days ago
  •  ...Northern Trust is seeking a Sr Principal Software Engineer to lead the AI Security Platform for code analysis across enterprise codebases. The role emphasizes platform ownership, architectural evolution, and scalable security automation in a cloud-native environment.... 
    Suggested

    Jobleads-US

    Chicago, IL
    4 days ago
  • $119.77k - $140.9k

     ...What You'll Do Act as the technical SME for our enterprise API ecosystem, leading engineering efforts across Apigee OPDK, Apigee Hybrid (Azure), and Apollo GraphQL platforms. Design and deliver secure, scalable, and highly available API gateway solutions that... 
    Full time
    Temporary work
    Work experience placement
    Local area
    3 days per week

    U.S. Bank

    Chicago, IL
    2 days ago
  •  ...Talentify is seeking a Principal Data Engineer in a hybrid role based in Dallas or Chicago. You will own the architecture and engineering for a new analytics platform, partnering with stakeholders to translate business goals into scalable data solutions. The role prioritizes... 

    Jobleads-US

    Chicago, IL
    15 hours ago
  • $211.5k - $235k

     ...Platform Engineering Manager - Infra + DevOps Remote Position Honor Technology’s mission is to change the way society cares for older adults. As a leader in aging care innovation, Honor provides the technology, tools, and services that empower older adults to live... 
    Permanent employment
    Temporary work
    Work at office
    Local area
    Remote work
    Relocation
    Home office

    Honor

    Chicago, IL
    1 hour ago
  • $212.5k - $275k

     ...delivers cutting-edge trading, clearing and investment solutions to market participants around the world. The goal of the Cboe Platform Engineering group focuses around building a scalable and secure foundations platform, enabling the Cboe operations, software engineering... 

    Cedar Cares, Inc

    Chicago, IL
    2 days ago
  • $147.76k - $240.11k

     ...technology, advanced analytics, and AI capabilities to help our customers build a better, more sustainable world.Job SummaryThe PIM Platform Engineering Manager leads engineering strategy and delivery for the PIM platform, owning technical direction, architecture review, and... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Caterpillar

    Chicago, IL
    2 hours ago
  •  ...McDonald’s Corporation is seeking a Director of Software Engineering for the Kiosk and Restaurant Platform within the Global Technology team in Chicago. You will lead a 28-person team plus external resources, establish engineering excellence, and deliver capabilities for... 

    Jobleads-US

    Chicago, IL
    5 days ago
  • Senior DevOps Engineer - Cloud Platform Engineering The RoleUST is searching for a Senior DevOps Engineer to join our Cloud Platform Engineering team. This is a role for an engineer who thinks in code first, infrastructure second — someone who can bridge complex business... 
    Full time
    Temporary work
    Part time
    Work at office
    Local area
    Remote work
    Flexible hours

    UST Global

    Chicago, IL
    1 day ago
  • Reporting to the Director Engineering, you will serve as a technical expert and resource across the Conagra Engineering organization. You...  ...practices.A Taste of Your ResponsibilitiesPartner with Platform Engineering leadership to develop annual and three-year capital... 
    Full time
    Local area
    Remote work
    Flexible hours

    Conagra Brands

    Chicago, IL
    3 days ago
  •  ...ServiceNow in Chicago is seeking an engineer to execute platform administration across ServiceNow and SAP. You will support internal stakeholders, troubleshoot issues, and drive automation, upgrades, and governance of sub-production environments. You will design debt... 

    Jobleads-US

    Chicago, IL
    1 day ago
  •  ...Principal EngineerVouch is the insurance broker that powers ambition. We're a tech-enabled...  ....This role owns that layer. Vouch's engineering organization is becoming agent-native:...  ...governed guardrails. This role owns the platform that governs agent-built software, makes... 
    Work at office
    Flexible hours
    3 days per week

    Vouch

    Chicago, IL
    5 days ago
  •  ...recruiting experience for everyone. With our Atlas SaaS Platform , Hunt Club is bringing its technology into the B2B and B2C...  ...Hunt Club is looking for a strategic and self-motivated Principal Engineer to help drive customer goals and objectives for our Expert team... 
    Full time
    Relocation

    Hunt Club

    Chicago, IL
    2 days ago
  • # Principal EngineerNorth Chicago, IL Full-timeOther / Corporate Functions PrincipalNotify me about similar jobsView Full Job & Apply...  ...LinkedIn, Facebook , Instagram , X and YouTube.Job DescriptionAn engineering professional who, working with little or no supervision,... 

    Biopharma Careers

    Chicago, IL
    3 days ago
  • $131.75k - $170.5k

     ...model. While the internal title for this position is Senior Linux Engineer, this role has been posted externally under a different title...  ...and skill sets.Role Overview We are seeking a Senior Platform Engineer to join our Systems Platform Engineer team. In this role... 
    Full time
    Work at office
    Immediate start

    Cboe Exchange

    Chicago, IL
    1 day ago
  •  ...Principal Engineer - 100% RemoteAn industrial real estate supply chain platform we have one of the largest portfolios of industrial space in the US, with over 400 million square feet of high-quality, well-located industrial assets. We actively construct and manage our... 
    Shift work

    1872 Consulting

    Chicago, IL
    2 days ago
  • $245k - $285k

     ...building. Why This Role Matters Vouch's engineering organization is becoming agent-native:...  .... This role owns that layer. This is a principal-level individual contributor seat...  ...contexts need to become real, and the cloud platform underneath all of it. We are not going... 
    Live in
    Work at office
    Flexible hours
    3 days per week

    Vouch Insurance

    Chicago, IL
    2 hours ago
  • $210k - $314k

     ...legal teams, governments, and enterprises around the world. Our platform powers investigations, litigation, and compliance for some of...  ...with the energy of a startup. What We Do At Relativity, engineers don't just write code—we shape how industries uncover critical... 
    Remote work
    Home office
    Flexible hours

    Relativity

    Chicago, IL
    2 hours ago
  • $209k - $238.5k

     ...Sr. Manager, Software Engineering, Back End (Enterprise Platforms Technlogy) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast‑paced, collaborative, inclusive, and iterative delivery environment? At Capital... 
    Full time
    Part time
    Internship
    Local area

    Capital One National Association

    Chicago, IL
    2 days ago
  •  ...Adyen is seeking an Engineering Manager to lead the Issuing team in Chicago. You will build and scale a best-in-class engineering organization around the issuing product, collaborating with product, operations, and business teams to deliver secure and scalable payment... 

    Jobleads-US

    Chicago, IL
    1 day ago
  •  ...We are seeking a Tonkean Platform Support Engineer to provide application support, workflow administration, integration troubleshooting, and continuous improvement for Tonkean-based solutions. The role will work closely with business teams, enterprise application teams... 
    Permanent employment
    Full time

    Techvilla Solutions

    Chicago, IL
    a month ago
  • $100k - $130k

     ...firm, and a bourgeoning investment adviser. Overview: DV is looking for a highly motivated and customer-focused Client Platforms Engineer to join our dynamic Chicago IT support team. In this role, you will be responsible for providing exceptional technical support... 
    Work at office
    Worldwide
    Flexible hours

    DV Trading

    Chicago, IL
    a month ago
  • $124.36k - $146.3k

     ...excel at—all from Day One.Job DescriptionJob SummaryThe Senior Engineer (Generative AI) is responsible for designing, developing, and deploying...  ...observability practices for production environments3. Cloud, Platform & Scalability EngineeringDevelop and deploy GenAI systems... 
    Work experience placement
    Local area
    3 days per week

    US Bank

    Chicago, IL
    5 days ago
  • $102.85k - $133.1k

     ...that carries traditional finance onto DeFi rails. CCUS Software Engineering is a high trust environment. We treat failure as a source of...  ...CI pipelines in GitHub Actions, and integrating them with the platform team's deployment tooling.Owning the Helm charts and values that... 
    Full time
    Work experience placement
    Live in
    Work at office
    Immediate start

    Cboe Exchange

    Chicago, IL
    20 hours ago
  • $119k - $169.4k

     ...work model in either our Chicago, IL or Overland Park, KS office.While the internal title for this position is Senior Automation Engineer, this role has been posted externally under a different title to better align with market norms and to attract talent with comparable... 
    Full time
    Work at office
    Immediate start

    Cboe Exchange

    Chicago, IL
    2 hours ago
  • $200k - $250k

     ...expectations, integrity, innovation and a willingness to challenge consensus. About the Role We're looking for an AI Inference Platform Engineer to build, operate, and optimize the systems that serve large language, vision, multimodal, and embedding models across DRW.... 
    Temporary work
    Flexible hours

    DRW

    Chicago, IL
    2 days ago
  • $145k

     ...Platform Engineer at Akuna Capital Akuna Capital is an innovative trading firm with a strong focus on collaboration, cutting-edge technology, data driven solutions, and automation. We specialize in providing liquidity as an options market maker – meaning we are committed... 
    Internship
    Work at office
    Remote work

    Akuna Capital

    Chicago, IL
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal, Platform Engineering. Be the first to apply!