Senior Site Reliability Engineer Consultant (AI)
$135k - $160kUniversity Information Technology - Indiana Wesleyan University
IWU is seeking a Senior Site Reliability Engineer Consultant (AI) (Independent Contractor ‑ 1099 / Contract-to-Hire) for a large-scale technology reorganization project.
ABOUT THE ROLE:
- IWU is re-platforming its enterprise on a dual-site, active/active, and on-premises data center foundation and an AI-first application stack. Reliability is a product requirement – not an afterthought bolted on after launch. We are seeking a Senior Site Reliability Engineer Consultant with Sovereign/On-Prem AI Stack Experience – a principal-level individual contributor to define service-level objectives, build the observability platform, and lead incident response for the applications and platform services that serve our 15,000 learners and internal operations.
- This is not a cloud SRE role . IWU runs production workloads on infrastructure we own and operate in commercial colocation. You will be responsible for application and platform reliability – SLOs, monitoring, alerting, on-call, and post-incident learning – while partnering with Infrastructure on hardware and network availability signals. Application-level performance and error budgets are yours; the data center floor is theirs – and the partnership between the two is what makes uptime real.
- You will be a hands-on principal : you will translate industry reliability standards into measurable SLOs and deployed tooling, mentor engineers through runbooks and postmortems, and hold the bar for operational excellence across dual sites.
Reliability Strategy and SLOs
- Define and maintain service-level objectives (SLOs), SLIs, and error budgets for critical platform and application services – availability, latency, throughput, and data freshness – aligned to business impact.
- Partner with product, application, and platform teams to negotiate realistic targets, document dependencies, and prioritize reliability work against feature delivery.
- Contribute to the enterprise observability reference architecture – metrics, logs, traces, and synthetic checks – per organizational standards.
Observability Platform
- Build AI platform observability in partnership with the Enterprise AI Architect – GPU health (DCGM), inference latency and throughput, model endpoint SLOs, queue depth, and cost/utilization dashboards.
- Design, deploy, and operate the observability stack – metrics platform, log aggregation, distributed tracing, and dashboards.
- Implement alerting and escalation – on-call rotations, alert hygiene (actionable alerts only), and runbook linkage.
- Ensure observability services themselves meet availability and retention targets – dual-site redundancy, capacity planning, and lifecycle management.
Incident Response and Operations
- Lead incident response for application and platform outages – triage, communication, mitigation, and resolution – including AI-specific incidents (model degradation, inference failures, data pipeline stalls) where standards apply.
- Drive post-incident reviews – root-cause analysis, corrective actions, and tracking to completion; feed learnings back into SLOs, runbooks, and architecture.
- Maintain and improve runbooks, playbooks, and escalation paths – clear ownership, severity definitions, and executive communication templates.
Capacity, Performance, and Resilience
- Partner with Infrastructure, Database, and Network teams on capacity signals – trend analysis, saturation forecasting, and proactive scaling before user impact.
- Design and exercise failure-mode testing – game days, failover drills, and chaos experiments appropriate for dual-site active/active topologies.
- Support release and deployment reliability – canary patterns, rollback criteria, deployment health checks, and integration with CI/CD pipelines.
- Analyze performance regressions – APM, tracing, and profiling data to isolate bottlenecks across application, database, and infrastructure layers.
Governance and Collaboration
- Align with ITIL incident, problem, and change management – ServiceNow or equivalent ticketing, CAB participation for high-risk changes, and audit-ready incident records.
- Partner with Information Security on security incident coordination and observability data handling – retention, access control, and compliance (FERPA, HIPAA, SOC 2) as applicable.
- Mentor application and platform engineers on reliability patterns – graceful degradation, circuit breakers, idempotency, and building services that are operable by default.
WHAT WE ARE LOOKING FOR:
Required
- According to Indiana Wesleyan University policy, all employees and contract-to-hire contractors must possess a strong Christian commitment and adhere to the standards outlined in the IWU Community Lifestyle Statement .
- 8+ years of progressive experience in site reliability engineering, production operations, or platform engineering, with 3+ years owning SLOs and on-call for business-critical services.
- Demonstrated experience building and operating observability at scale – metrics, logs, and traces – on stacks in production.
- Strong incident management skills – leading Sev-1/Sev-2 response, writing postmortems, and driving permanent fixes rather than repeated firefighting.
- Experience defining and operating SLO/SLI/error-budget programs – not just uptime percentages, but user-centric reliability targets tied to business outcomes.
- Hands-on proficiency with Windows, Linux, containers, and Kubernetes operations – pod health, resource limits, HPA, and debugging production clusters.
- Scripting and automation skills – Python, Go, or Bash – for tooling, alert automation, and toil reduction.
- Bachelor's degree in Computer Science, Information Systems, Engineering, or a related technical field – or equivalent demonstrated experience .
Strongly Preferred
- Experience with on-premises or colocation production environments – dual-site active/active topologies, and partnering with infrastructure teams on shared incident boundaries.
- Experience observing AI/ML platform workloads – GPU monitoring, inference SLOs, batch pipeline latency, or model-serving health checks.
- Familiarity with APM and synthetic monitoring .
- Experience with Infrastructure as Code – for observability-as-code deployments.
- Background in a regulated or compliance-sensitive environment (education/FERPA-Title-IV, financial services/GLBA, healthcare/HIPAA, insurance, public sector).
- Familiarity with ITIL service management and on-call best practices (Google SRE book principles applied pragmatically).
- Relevant certifications or demonstrated depth equivalent to CKA/CKAD, or similar.
How You Work
- You measure before you optimize – SLOs and data drive priorities, not loudest stakeholder or latest outage memory.
- You design for operability – if on-call cannot diagnose it from the dashboard, the service is not done.
- You reduce toil relentlessly – automation and self-service beat heroics every time.
- You mentor by example – your runbooks, SLO docs, and calm incident leadership become the standard others follow.
WHY THIS ROLE?
- Principal scope, hands-on impact - you drive reliability strategy and the hardest incidents without needing a management title.
- Greenfield reliability culture . You are building SLO and observability practice alongside a funded dual-site re-platforming – not inheriting a decade of alert spam and no error budgets.
- AI-native challenge . Inference latency, GPU saturation, and model-serving SLOs are first-class problems here, not edge cases.
- Clear ownership boundary . Application SLOs are yours; infrastructure hardware is Infrastructure's – clean partnership, real uptime.
- Mission and meaning . IWU is a century-old Christ-centered university serving roughly 15,000 learners. The reliability you deliver keeps the systems running that prepare them to change the world.
PROJECT REMUNERATION:
Estimated annual project remuneration is $135,000 - $160,000 for this independent contractor (1099) / Contract-to-Hire role.
Contractor Insurance/Bond Requirements
- General Liability: $1m per occurrence / $2m aggregate
- Cyber Liability: $2m per occurrence
- Workers’ Compensation / Employers’ Liability: $500,000; not applicable for approved sole operators
- Professional Liability / Errors & Omissions: $1m per occurrence / $3m aggregate
Equipment & Travel Expenses
- IWU will provide required equipment and software. Personal devices and software are generally not permitted for assigned work unless approved by IWU.
- IWU will pay for required and approved travel and other expenses.
- Specific policies, rules, and details will be included in the executed contract.
Engagement Process
- Easy-Apply on LinkedIn
- LinkedIn Survey
- Phone Screen
- Video Interviews
- On-Site Interview/Tour
- Contract Engagement Begins
Compliance & Disclaimers
IWU is an equal opportunity employer committed to compliance with Title VII of the Civil Rights Act of 1964, Title IX of the Educational Amendments of 1972 and Section 504 of the Rehabilitation Act of 1973 or other federal, state, or local laws or executive orders except as claimed in a filed religious exemption.
Indiana Wesleyan University is managing this search directly. Third-party recruiters, staffing agencies, search firms, subcontracting firms, and candidate marketers may not submit candidates for this role unless they have a current, written agreement with IWU that specifically authorizes recruitment support for this opening.
Unsolicited candidate submissions will not create any obligation, fee, commission, placement right, or ownership claim . Any candidate submitted without prior written authorization may be contacted directly by IWU and considered without payment of any recruiting or placement fee.
Candidates must apply directly through this LinkedIn posting or the official IWU application process. This role does not authorize third-party representation, bench submissions, or agency-to-agency candidate referrals.
$141.8k - $195k
...company that's building the telemetry infrastructure for the AI era. At Cribl, we partner with IT and Security teams at... ...the herd.Why You'll Love This RoleCribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all...SeniorRemote work$168k - $200k
...patient's request for their medical records to powering the AI revolution in healthcare, Datavanters are building the... ...healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the...Senior$121.4k - $218.6k
...that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and... ...and resource optimization. As a Senior Site Reliability Engineer, you will be... ...biggest moments without a glitch. AI : Enabling our customers to build, secure...SeniorWork experience placementWork at office$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role... ...everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers...SeniorTemporary workImmediate startFlexible hoursShift work$155k - $175k
...IWU is seeking a Senior Network Engineer Consultant (Independent Contractor ‑ 1099 / Contract-to-Hire) for... ...ROLE: IWU is standing up a dual-site, active/active, and on-premises data... ...support enterprise re-platforming and an AI-first technology stack. The network...SeniorContract workFor contractorsLocal areaRemote work$66k - $158.4k
...operational excellence, reliability, and quality are non-... ...transformation into the Digital and AI era. They inspire... ...The Reliability Engineering team is the engineering... ...within the standards the Senior Principal SRE Engineer... ...with meaningful time as a Site Reliability Engineer,...Full timeH1bVisa sponsorshipWork visaFlexible hours$75.7k - $136.3k
...and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications... ...: Scaling the world's biggest moments without a glitch. AI : Enabling our customers to build, secure, and scale AI...Work experience placementWork at office$95k - $171k
...Are you passionate about cutting-edge AI infrastructure? Do you want to build... ...infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$140k - $210k
...the world’s number 1 job site*, our mission is to help people... ...26) Day to Day As an Engineering Manager in Site Reliability Engineering at Indeed, you... ..., from new hires to senior contributors. Create conditions... ...for that opening. AI Notice Indeed is...Work experience placementLocal area- ...operational goals through the use of Generative AI solutions. We are a fast-growing, remote-... ...Our team believes in empowering talented engineers to solve meaningful problems, collaborate... ...Position Overview We are seeking a Senior Software Engineer who thrives in a...SeniorRemote workFlexible hours
- ...the leading edge of emerging technology and rapidly expand your engineering skill set? Join Best Coding Squad (BCS), where we develop full-... ...exploring new and emerging technologies at the frontier of AI.ResponsibilitiesDesign and develop technical solutions as part...SeniorLocal areaRemote workVisa sponsorshipWork visa
$130k - $200k
...CCaaS Technical Lead, Senior ConsultantOur Deloitte Sales... ...advanced analytics, Generative AI, transformative technologies,... ...operators, creatives, designers, engineers, and architects. Our team balances... ...of experience in technology consulting, software engineering, cloud...SeniorLocal area$92.7k - $161.85k
As a Senior DevOps Engineer, you will be a key technical contributor responsible... ..., automated testing, and reliable deployment across environments... ...integration feasibility.Apply Site Reliability Engineering (SRE... ...without the assistance of AI tools or external prompts. Our...SeniorFull timeWork at office- ...learn more about working at Coinbase. Senior Software Engineer (EAA) The CEE Surfaces team, part of... ...and Intelligence layers. You'll ship reliable, measurable features that improve... ...escalation rates. Lead integration of AI/ML and vendor capabilities into customer...SeniorLocal area
- ...Bright Vision Technologies is seeking an experienced AI Systems Performance Specialist to optimize AI training and inference workloads... ...requires 10+ years in AI infrastructure, HPC, or performance engineering, plus hands-on expertise with distributed frameworks and cloud...SeniorRemote job
$93k - $189k
DescriptionSenior Manager, AI Engineering PlatformsBuilding the Google-powered foundation that... ..., and support models.Ensure platform reliability, scalability, security, and compliance.... ...Practice: Visit Huntington's Career Web Site for more details.Note to Agency Recruiters...SeniorFull timeWork at officeRemote workWork from homeFlexible hours$77k - $202k
...The Opportunity As a GenAI Python Systems Engineer – Senior Associate, you will play a pivotal role in transforming raw data into actionable... ..., you will focus on developing and implementing advanced AI and ML solutions to drive innovation and enhance business processes...SeniorH1b- ...writing production code across Ruby on Rails, React, and Java. You think in service boundaries, you raise the engineering bar around you, and you have made AI-assisted development a measurable advantage for your whole team, not just for yourself. WHAT YOU'LL DO Lead...SeniorTemporary workWork at office
$150k - $200k
...As the most widely adopted self-service AI triage platform, weve reached patients in... ...this role This role sits on the Engineering team and reports to the Engineering Manager... .... This role carries a security and reliability bent. Health systems trust us with patient...SeniorRemote workWork from homeFlexible hours- ## Senior Software EngineerApply: Indianapolis: Full time: Posted 3... ...meet you.**Position Summary**The AI and Innovation team builds and... ...looking for a Senior Software Engineer to focus on building the core... ...to turn real problems into reliable systems.**What You’ll Do*** Build...SeniorFull timeContract workWork at office
$79.2k - $209.5k
...design, including tradeoffs around security, reliability, scalability, maintainability,... ...technical debt. Improve the team through engineering practices, operational practices, development... ...to life-saving care. And with AI embedded across our products and services...SeniorTemporary workFlexible hours$88.7k - $139.26k
...and excellence.Your Role at Delta Faucet:Position SummaryThe Senior Analytics Engineer owns the layer between curated enterprise data and business... ...controls used by analysts, self-service users, dashboards, and AI agents. This role replaces a traditional Data Engineer...SeniorFull timeLocal area$186.07k - $218.9k
...builds the systems every Coinbase engineer relies on to write, validate,... ...feedback. We’re hiring Senior Software Engineers across several... ..., deploy safety, and test reliability scale across thousands of engineers... .... ~ Utilizes generative AI responsibly, maintaining...SeniorLocal area$122k - $240.5k
...Summary Agentic AI is moving from... ...'re growing a team of engineers who want to work at the... ...improve the performance, reliability, and usability of AI solutions... ...: Experience in consulting or other client-facing... ...entry-level employees to senior leaders, we believe...SeniorWork at officeLocal areaVisa sponsorshipShift work$110.7k - $218.3k
...advanced analytics, Generative AI and intelligent automation,... ...operators, creatives, designers, engineers, and architects. Our team... ...Required: 5+ years of relevant consulting or industry experience2+... ...From entry-level employees to senior leaders, we believe there’s always...SeniorLocal areaVisa sponsorship- Position Summary Deloitte Global is the engine of the Deloitte network. Our professionals... ...to it, and a real appetite to bring AI into how we build, detect, and analyze. If... ...development From entry-level employees to senior leaders, we believe there’s always room to...SeniorImmediate startWorldwideVisa sponsorship
- ...Senior Vice President, Software Engineering, SaaS About the Company An established global leader in digital learning and education technology. Industry... .... Experience in B2B2C product environments, driving AI adoption, and overseeing data architecture for legacy...Senior
$186.07k - $218.9k
...more about working at Coinbase. As a Senior Software Engineer on the Data Platform team within the... ...Coinbase to interact with the data platform reliably at scale. Deliver self-service... ...monitoring. ~ Utilizes generative AI responsibly, maintaining human oversight...SeniorLocal area- ...EY is seeking an AI Security Engineer to own the security posture of EY’s Agentic AI platform end to end. The role covers security across cloud-native infrastructure, Kubernetes, identity, and supply-chain integrity to defend against agentic threats. The candidate...Senior
$116.2k - $229.1k
...Recruiting for this role ends on 11/1/2026. Work you'll do As a Senior Consultant on the Moveworks team, you will be responsible for: Leading... ...stories to support the design and configuration of the Moveworks AI Assistant, including knowledge management, conversational...SeniorLocal areaVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer Consultant (AI). Be the first to apply!
- senior service associate Indianapolis, IN
- senior safety specialist Indianapolis, IN
- senior vice president of business development Indianapolis, IN
- senior mulesoft developer Indianapolis, IN
- senior business manager Indianapolis, IN
- senior linux systems engineer Indianapolis, IN
- senior mainframe developer Indianapolis, IN
- senior cloud security engineer Indianapolis, IN
- sr tax manager Indianapolis, IN
- senior level Indianapolis, IN



