Senior AI Ops & Incident/Site Reliability Engineer
$93.6k - $170.64kPerficient, Inc.
We currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a hybrid role, 3 days a week in office.Job Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience, and intelligent automation initiatives across enterprise technology environments. This role will partner with Network Operations Center (NOC), Infrastructure Operations, Cloud Engineering, DevOps, and Application Support teams to proactively detect, respond to, and prevent technology incidents. The ideal candidate combines hands-on incident management expertise with experience implementing observability, automation, and AI-driven operational solutions to improve system reliability, reduce operational overhead, and enhance customer experience. The candidate should possess deep expertise in AIOps, ITSM, ITIL, SRE, Incident Management, Cloud Operations, and Enterprise Infrastructure.This is a hybrid role requiring on-site attendance up to three days per week. Candidates must reside within approximately one hour commuting distance of one of the following office locations: Fort Mill, SC, Austin, TX, Boston, MA, New York, NY, Tempe, AZ, or San Diego, CA.Perficient is always looking for the best and brightest talent and we need you! We’re a quickly-growing, global digital consulting leader, and we’re transforming the world’s largest enterprises and biggest brands. You’ll work with the latest technologies, expand your skills, and become a part of our global community of talented, diverse, and knowledgeable colleagues.Perficient is the global AI and technology consulting firm disrupting the traditional consulting model. Powered by our 7,000+ advisors, engineers, and designers, Perficient implements AI-first solutions that break conventions and deliver outcomes that matter. Proudly serving clients that represent the world’s most innovative brands, and in collaboration with our powerful technology partner ecosystem, we bring deep industry expertise and data-driven design to redefine how businesses run and succeed. Perficient is different. For real. Learn more at perficient.com.Bachelor's degree in Computer Science, Information Technology, Engineering, or related field (or equivalent experience).8+ years of experience in IT Operations, Site Reliability Engineering, Infrastructure Operations, Network Operations, or Production Support environments.5+ years of experience leading incident management, operational transformation, or reliability engineering initiatives.Strong experience with: Site Reliability Engineering (SRE)IT Service Management (ITSM)ITIL FrameworkIncident, Problem, Change, and Event ManagementNetwork Operations Center (NOC)Infrastructure OperationsService Desk OperationsApplication Production SupportCloud Platforms (AWS, Azure, or GCP)DevOps Practices and ToolchainsHands-on experience with Dynatrace, monitoring platforms, and observability solutions.Experience using ServiceNow for ticketing, workflow automation, and service management.Strong understanding of infrastructure, networking, cloud architecture, and enterprise application ecosystems.Proven experience conducting root cause analysis and implementing preventive controls.Experience leading enterprise AIOps implementations.Experience building AI-powered operational agents and intelligent automation solutions.Certifications such as: ITIL Foundation or ITIL Managing ProfessionalCertified Site Reliability Engineer (SRE)AWS, Azure, or Google Cloud certificationsServiceNow certificationsExperience with workflow orchestration and enterprise automation platforms.Familiarity with predictive analytics, machine learning operations, and autonomous operations frameworks. ADDITIONAL INFORMATIONPerficient, Inc. proudly provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, gender, sexual orientation, national origin, age, disability, genetic information, marital status, amnesty, or status as a protected veteran in accordance with applicable federal, state and local laws. Perficient, Inc. complies with applicable state and local laws governing nondiscrimination in employment in every location in which the company has facilities. This policy applies to all terms and conditions of employment, including, but not limited to, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. Perficient, Inc. expressly prohibits any form of unlawful employee harassment based on race, color, religion, gender, sexual orientation, national origin, age, genetic information, disability, or covered veterans. Improper interference with the ability of Perficient, Inc. employees to perform their expected job duties is absolutely not tolerated.Disability Accommodations: Perficient is committed to providing a barrier-free employment process with reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or accommodation due to a disability, please contact us.Applications will be accepted until the position is filled or the posting is removed.The salary range for this position takes into consideration a variety of factors, including but not limited to skill sets, level of experience, applicable office location, training, licensure and certifications, and other business and organizational needs. The new hire salary range displays the minimum and maximum salary targets for this position across all US locations, and the range has not been adjusted for any specific state differentials. It is not typical for a candidate to be hired at or near the top of the range for their role, and compensation decisions are dependent on the unique facts and circumstances regarding each candidate. A reasonable estimate of the current salary range for this position is $93,600.00 to $170,640.00. Please note that the salary range posted reflects the base salary only and does not include benefits or any potential equity or variable bonus programs. Information regarding the benefits available for this position are in our benefits overview.Disclaimer: The above statements are not intended to be a complete statement of job content, rather to act as a guide to the essential functions performed by the employee assigned to this classification. Management retains the discretion to add or change the duties of the position at any time. #LI-BV1Incident & Recovery ManagementMonitor, document, and analyze major incident response efforts and service recovery activities.Serve as a senior escalation point for Tier 1 and Tier 2 operational incidents.Conduct incident reviews, root cause analysis, and corrective action planning.Improve Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).Site Reliability EngineeringImplement SRE practices to improve platform reliability, scalability, and resiliency.Define and monitor SLAs, SLOs, and operational KPIs.Develop proactive reliability and availability strategies.AIOps & AutomationImplement AIOps solutions to automate incident detection, diagnosis, remediation, and prevention.Build and optimize AI-powered operational agents and self-healing workflows.Reduce operational effort through intelligent automation.Observability & MonitoringLead enterprise monitoring initiatives using Dynatrace and related observability platforms.Improve visibility across cloud, infrastructure, applications, and user experiences.Enable predictive monitoring and anomaly detection.ITSM & Service OperationsDevelop and enhance incident, problem, change, and event management frameworks aligned with ITIL and ITSM best practices. Leverage ServiceNow workflow automation to improve service delivery.Cross-Functional LeadershipPartner with Infrastructure, DevOps, Cloud, Security, Application Development, and NOC teams.Mentor operational teams and promote an automation-first culture.Full timePosting Date: 2026-07-21
$160k - $200k
...days/per week.Tulip, the leader in AI-native frontline operations, is helping... ...best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining... ...processes as a player / coachPerform incident response and debug production issues...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible... ...that ensure cluster reliability and security (e.g., CoreDNS,... ...manual processes ("allergic to ops work")We are a small team of... ...redefined the data platform for the AI era, enabling builders to...SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$127k - $249k
Platform Engineering is the department within SRE that is responsible for... ...infrastructure, ensuring reliable code deployment from development... ...manual process (“allergic to ops work”). We are a small team of... ...redefined the database for the AI era, enabling innovators to create...SeniorLocal areaWorldwideFlexible hours$166k - $220k
...powered by Lattice OS, an AI-powered operating system that... ...that Anduril services are reliable and maintainable. This means... ....ABOUT THE JOBAs a Site Reliability Engineer on the Observability team,... ...investigation of production incidents — the ultimate mandate of the...SeniorFull timeWork experience placementImmediate start$188k - $282k
...DescriptionPosition SummaryVertex is seeking a Sr. Principal AI Engineer, Agentic AI Platform & Control Tower Ops to define and build the foundational enterprise AI... ...services that accelerate development of safe, reliable, and scalable AI-powered applicationsLead...SeniorFull timeSummer work$134.25k - $214.8k
...where you matter.Your ImpactAre you an engineer who gets excited about the challenge of... ...of the Observability team within Axon's Site Reliability organization — a focused team responsible... ...infrastructure engineeringExperience with agentic AI tooling or building LLM-powered...SeniorWork experience placementWork at officeRemote work$134.3k - $189.9k
...biology easier to engineer. Ginkgo is... ...new organisms. Senior Software Engineer... ...to data APIs and AI-enabled agentic... ...candidate to work on-site Monday - Friday... ...performant and reliable.Data Management... ...and accessible.Ops & Infra TeamThis... ...start.Contribute to incident response, post-...SeniorFull timeH1bCurrently hiringWork at officeRelocationRelocation packageMonday to Friday$123.25k - $175.1k
...does the team work on? The Ad Ops team at Roku plays a crucial... ...the role? Roku is seeking a Senior Solutions Engineer, Ad Data Ingestion to join... ...paid time off. How will I use AI at Roku? At Roku, we are... ...technical point of contact for incidents What experience would help...SeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$138.1k - $198.2k
...that simply works. The SRE Engineering Enablement Team supports our... ...at Cisco. Your Impact As a Site Reliability Engineer, you will be at the... ...stakeholders. Practice sustainable incident responses (on call rotation)... ...organizations in the AI era - and beyond. We’ve been...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$158k - $197.5k
...practice. About the RoleThe Senior AI Defense Engineer is a technical leader... ...evaluation and assessments that reliably finds issues before production... ..., identity platforms). Incident Response & Forensics for AI... ...threat research, or ML/ML Ops engineering.EducationBachelor...SeniorWork experience placementShift work$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary: The Senior... ...in on-call rotations and incident response, including root cause analysis and...SeniorFull timeTemporary workWork at officeFlexible hoursWeekend work$130k - $140k
...technology.Job DescriptionSite Reliability EngineerLocation(s):... ...MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to... ...to and resolving escalated incidents for customer issues or monitoring... ..., AMQ.Familiar with AI tools.What Sets You Apart (...SeniorOngoing contractFull timeTemporary workWork experience placement$51.9 per hour
...is responsible for the reliability, availability, and... ...role blends software engineering, clinical engineering,... ...functionally with AHN site leaders and teams to navigate... ...IT trends, including AI, security patching,... ...participates in post-incident reviews to identify root...For contractorsLocal area- ...systems scale securely and reliably is core to this mission.The AI Devex team exists to... ...acceleration.As a Senior Platform Engineer on the AI Devex team, you... ...practices.Participate in incident response, root cause analysis... ...in DevOps, Platform, Site Reliability,...SeniorWork at officeRelocation
$226k - $307k
....As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system... ...FP16).Proficiency in low-level programming for AI accelerators, specifically developing and optimizing custom ML OPs and TensorRT Plugins with efficient CUDA kernel...SeniorFull timeTemporary workRelocation package- Senior Software Engineer — File SystemEngineering | File System Team... ...that directly affect reliability, performance,... ...operational quality, and using AI-assisted engineering... ..., design reviews, incident follow-up, and technical... ...workspaces Free on-site fitness centers and stocked...SeniorFull timeFor contractorsWork at officeRemote workFlexible hours
$180k - $200k
...This is a hands-on platform engineering role that blends software... ...; employees are on-site in the Boston office 3 days... ...operational workflows to improve reliability, observability and incident response.Contribute to... ...ChatGPT, Cursor and other AI model experience a plus.The...SeniorCasual workWork at officeWorldwideFlexible hours3 days per week$191k - $253k
...is powered by Lattice OS, an AI-powered operating system that... ...and systems that empower our engineering teams to develop, test, and deploy... ...unparalleled efficiency and reliability. We design and implement... ...operations, troubleshooting, and incident response.REQUIRED...SeniorFull timeWork experience placementImmediate start$140.88k - $153.75k
...Information Job Title Senior Platform Engineer Job ID 105799 Work... ...call rotation for platform incidents; drive incident response to... ..., observability, security, reliability, and operational hygiene.... ...ongoing review feedback.Use AI coding assistants to accelerate...SeniorPermanent employmentFull timeContract workWork at officeLocal area$185k - $227k
...read on for more details. ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance... ..., and acting as the final escalation point for critical incidents to ensure theplatform is scalable and efficient. Nutanix...SeniorRemote work- ...We are seeking a product-minded Software Engineer to join our core team. In this role, you’... ...to reimagine it from the ground up. As an AI-native company, we lean in heavily to AI-... ...domain — SIEM, SOAR, EDR, MDR, or incident response workflows. ~ Excellent software...SeniorFull time
$148k - $222k
...own their own destiny.Senior Software Engineer - Agent PlatformAt Klaviyo... ...state-of-the-art AI and machine learning technologies... ...vLLM and Ray.Develop reliable, scalable data... ...listed. AI Ops is a rapidly evolving... ...on our official career site. Please be cautious of...Senior$140k - $160k
...leading travel search engine. With billions of queries... ....KAYAK is seeking a Senior Java Software Engineer... ...messaging at scale, ensuring reliability and performance across... ...including alerting and incident response for team-owned... ...— including AI-assisted tooling and automation...SeniorWork at officeFlexible hours3 days per week$150k - $185.3k
...FabricCollaborate with a team of BI Engineers to identify and implement a path forward in leveraging AI assisted BI development in MS... ...with resolution of escalated incidents and participates in problem... ...Continuous Delivery) in Microsoft Dev Ops, supporting Azure Data Factory...SeniorWork experience placementShift work$165k - $190k
Maven AGI is an enterprise AI platform founded in July 2023... ...The RoleWe’re looking for a Senior DevOps Engineer to own and evolve the infrastructure... ...our platform scales reliably as we onboard enterprise customers... ..., scaling, monitoring, and incident responseBuild and optimize...Senior- HPE Labs - Senior Software Engineer - Integrated HPC & Quantum SolutionsThis role has been designed as... ...computing platforms, GPU computing, AI/ML frameworks (e.g., TensorFlow, PyTorch... ...or make any payments and report the incident to your local authorities immediately....SeniorFull timeWork experience placementWork at officeLocal areaImmediate start2 days per week
$195.5k - $352.1k
...business. The mission of the Ad Engineering Team is to build this platform. We are hiring a Senior Software Engineer for the Advertising... ...paid time off.How will I use AI at Roku? At Roku, we don’t just... ..., QA, and infrastructure/ops Be an evangelist for platform innovation...SeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$75 - $80 per hour
...officesPOSITION OVERVIEWThe Senior Developer role... ...leader, bringing deep engineering expertise while working... ...performance, scalability, and reliability across FP&A workloads.... ..., data vectors, and AI-based methods where... ...point for critical incidents, providing Tier-3 expertise...SeniorHourly payWork at office$190k - $250k
Senior Software Engineer, ApplicationsAcuityMD is a software and data platform that... ...Health. We're a high-growth AI and Data company scaling rapidly... ...VPs, marketers, sales ops, and sales reps. We deliver top... ...opportunities where physicians or sites of care can better serve...SeniorFor contractorsWork at officeRemote workWork from homeHome officeFlexible hours$140k - $170k
...learn more.Perk is growing its engineering footprint in North America, and we're looking for a Senior Software Engineer to help us do... ...practices; defines SLOs, manages incident responses, and conducts... ...and other job-related factors.AI at Perk:AI is embedded in how we...SeniorSummer workWork at officeImmediate startWorldwideRelocationRelocation package3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior AI Ops & Incident/Site Reliability Engineer. Be the first to apply!
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- senior business analyst Boston, MA
- senior risk manager Boston, MA
- senior cost estimator Boston, MA
- senior manager tax Boston, MA
- senior automation engineer Boston, MA
- senior devops Boston, MA
- senior recruiter Boston, MA
- senior property manager Boston, MA


