Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior AI Ops & Incident/Site Reliability Engineer

$93.6k - $170.64k

Perficient, Inc.

We currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a hybrid role, 3 days a week in office.Job Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience, and intelligent automation initiatives across enterprise technology environments. This role will partner with Network Operations Center (NOC), Infrastructure Operations, Cloud Engineering, DevOps, and Application Support teams to proactively detect, respond to, and prevent technology incidents. The ideal candidate combines hands-on incident management expertise with experience implementing observability, automation, and AI-driven operational solutions to improve system reliability, reduce operational overhead, and enhance customer experience. The candidate should possess deep expertise in AIOps, ITSM, ITIL, SRE, Incident Management, Cloud Operations, and Enterprise Infrastructure.This is a hybrid role requiring on-site attendance up to three days per week. Candidates must reside within approximately one hour commuting distance of one of the following office locations: Fort Mill, SC, Austin, TX, Boston, MA, New York, NY, Tempe, AZ, or San Diego, CA.Perficient is always looking for the best and brightest talent and we need you! We’re a quickly-growing, global digital consulting leader, and we’re transforming the world’s largest enterprises and biggest brands. You’ll work with the latest technologies, expand your skills, and become a part of our global community of talented, diverse, and knowledgeable colleagues.Perficient is the global AI and technology consulting firm disrupting the traditional consulting model. Powered by our 7,000+ advisors, engineers, and designers, Perficient implements AI-first solutions that break conventions and deliver outcomes that matter. Proudly serving clients that represent the world’s most innovative brands, and in collaboration with our powerful technology partner ecosystem, we bring deep industry expertise and data-driven design to redefine how businesses run and succeed. Perficient is different. For real. Learn more at perficient.com.Bachelor's degree in Computer Science, Information Technology, Engineering, or related field (or equivalent experience).8+ years of experience in IT Operations, Site Reliability Engineering, Infrastructure Operations, Network Operations, or Production Support environments.5+ years of experience leading incident management, operational transformation, or reliability engineering initiatives.Strong experience with: Site Reliability Engineering (SRE)IT Service Management (ITSM)ITIL FrameworkIncident, Problem, Change, and Event ManagementNetwork Operations Center (NOC)Infrastructure OperationsService Desk OperationsApplication Production SupportCloud Platforms (AWS, Azure, or GCP)DevOps Practices and ToolchainsHands-on experience with Dynatrace, monitoring platforms, and observability solutions.Experience using ServiceNow for ticketing, workflow automation, and service management.Strong understanding of infrastructure, networking, cloud architecture, and enterprise application ecosystems.Proven experience conducting root cause analysis and implementing preventive controls.Experience leading enterprise AIOps implementations.Experience building AI-powered operational agents and intelligent automation solutions.Certifications such as: ITIL Foundation or ITIL Managing ProfessionalCertified Site Reliability Engineer (SRE)AWS, Azure, or Google Cloud certificationsServiceNow certificationsExperience with workflow orchestration and enterprise automation platforms.Familiarity with predictive analytics, machine learning operations, and autonomous operations frameworks. ADDITIONAL INFORMATIONPerficient, Inc. proudly provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, gender, sexual orientation, national origin, age, disability, genetic information, marital status, amnesty, or status as a protected veteran in accordance with applicable federal, state and local laws. Perficient, Inc. complies with applicable state and local laws governing nondiscrimination in employment in every location in which the company has facilities. This policy applies to all terms and conditions of employment, including, but not limited to, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. Perficient, Inc. expressly prohibits any form of unlawful employee harassment based on race, color, religion, gender, sexual orientation, national origin, age, genetic information, disability, or covered veterans. Improper interference with the ability of Perficient, Inc. employees to perform their expected job duties is absolutely not tolerated.Disability Accommodations: Perficient is committed to providing a barrier-free employment process with reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or accommodation due to a disability, please contact us.Applications will be accepted until the position is filled or the posting is removed.The salary range for this position takes into consideration a variety of factors, including but not limited to skill sets, level of experience, applicable office location, training, licensure and certifications, and other business and organizational needs. The new hire salary range displays the minimum and maximum salary targets for this position across all US locations, and the range has not been adjusted for any specific state differentials. It is not typical for a candidate to be hired at or near the top of the range for their role, and compensation decisions are dependent on the unique facts and circumstances regarding each candidate. A reasonable estimate of the current salary range for this position is $93,600.00 to $170,640.00. Please note that the salary range posted reflects the base salary only and does not include benefits or any potential equity or variable bonus programs. Information regarding the benefits available for this position are in our benefits overview.Disclaimer: The above statements are not intended to be a complete statement of job content, rather to act as a guide to the essential functions performed by the employee assigned to this classification. Management retains the discretion to add or change the duties of the position at any time. #LI-BV1Incident & Recovery ManagementMonitor, document, and analyze major incident response efforts and service recovery activities.Serve as a senior escalation point for Tier 1 and Tier 2 operational incidents.Conduct incident reviews, root cause analysis, and corrective action planning.Improve Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).Site Reliability EngineeringImplement SRE practices to improve platform reliability, scalability, and resiliency.Define and monitor SLAs, SLOs, and operational KPIs.Develop proactive reliability and availability strategies.AIOps & AutomationImplement AIOps solutions to automate incident detection, diagnosis, remediation, and prevention.Build and optimize AI-powered operational agents and self-healing workflows.Reduce operational effort through intelligent automation.Observability & MonitoringLead enterprise monitoring initiatives using Dynatrace and related observability platforms.Improve visibility across cloud, infrastructure, applications, and user experiences.Enable predictive monitoring and anomaly detection.ITSM & Service OperationsDevelop and enhance incident, problem, change, and event management frameworks aligned with ITIL and ITSM best practices. Leverage ServiceNow workflow automation to improve service delivery.Cross-Functional LeadershipPartner with Infrastructure, DevOps, Cloud, Security, Application Development, and NOC teams.Mentor operational teams and promote an automation-first culture.Full timePosting Date: 2026-07-21

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior AI Ops & Incident/Site Reliability Engineer in Boston, MA vacancy
  • $160k - $200k

     ...days/per week.Tulip, the leader in AI-native frontline operations, is helping...  ...best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining...  ...processes as a player / coachPerform incident response and debug production issues... 
    Senior
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    3 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible...  ...that ensure cluster reliability and security (e.g., CoreDNS,...  ...manual processes ("allergic to ops work")We are a small team of...  ...redefined the data platform for the AI era, enabling builders to... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    5 days ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for...  ...infrastructure, ensuring reliable code deployment from development...  ...manual process (“allergic to ops work”). We are a small team of...  ...redefined the database for the AI era, enabling innovators to create... 
    Senior
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    14 hours ago
  • $166k - $220k

     ...powered by Lattice OS, an AI-powered operating system that...  ...that Anduril services are reliable and maintainable. This means...  ....ABOUT THE JOBAs a Site Reliability Engineer on the Observability team,...  ...investigation of production incidents — the ultimate mandate of the... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Boston, MA
    4 days ago
  • $188k - $282k

     ...DescriptionPosition SummaryVertex is seeking a Sr. Principal AI Engineer, Agentic AI Platform & Control Tower Ops to define and build the foundational enterprise AI...  ...services that accelerate development of safe, reliable, and scalable AI-powered applicationsLead... 
    Senior
    Full time
    Summer work

    Vertex Pharmaceuticals

    Boston, MA
    5 days ago
  • $134.25k - $214.8k

     ...where you matter.Your ImpactAre you an engineer who gets excited about the challenge of...  ...of the Observability team within Axon's Site Reliability organization — a focused team responsible...  ...infrastructure engineeringExperience with agentic AI tooling or building LLM-powered... 
    Senior
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    1 day ago
  • $134.3k - $189.9k

     ...biology easier to engineer. Ginkgo is...  ...new organisms. Senior Software Engineer...  ...to data APIs and AI-enabled agentic...  ...candidate to work on-site Monday - Friday...  ...performant and reliable.Data Management...  ...and accessible.Ops & Infra TeamThis...  ...start.Contribute to incident response, post-... 
    Senior
    Full time
    H1b
    Currently hiring
    Work at office
    Relocation
    Relocation package
    Monday to Friday

    Ginkgo Bioworks

    Boston, MA
    5 days ago
  • $123.25k - $175.1k

     ...does the team work on? The Ad Ops team at Roku plays a crucial...  ...the role? Roku is seeking a Senior Solutions Engineer, Ad Data Ingestion to join...  ...paid time off. How will I use AI at Roku? At Roku, we are...  ...technical point of contact for incidents What experience would help... 
    Senior
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    Boston, MA
    4 days ago
  • $138.1k - $198.2k

     ...that simply works.  The SRE Engineering Enablement Team supports our...  ...at Cisco. Your Impact As a Site Reliability Engineer, you will be at the...  ...stakeholders.  Practice sustainable incident responses (on call rotation)...  ...organizations in the AI era - and beyond. We’ve been... 
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Boston, MA
    2 days ago
  • $158k - $197.5k

     ...practice. About the RoleThe Senior AI Defense Engineer is a technical leader...  ...evaluation and assessments that reliably finds issues before production...  ..., identity platforms). Incident Response & Forensics for AI...  ...threat research, or ML/ML Ops engineering.EducationBachelor... 
    Senior
    Work experience placement
    Shift work

    Wilmer Hale

    Boston, MA
    3 days ago
  • $140k - $205k

    Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary: The Senior...  ...in on-call rotations and incident response, including root cause analysis and... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Boston, MA
    3 days ago
  • $130k - $140k

     ...technology.Job DescriptionSite Reliability EngineerLocation(s):...  ...MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to...  ...to and resolving escalated incidents for customer issues or monitoring...  ..., AMQ.Familiar with AI tools.What Sets You Apart (... 
    Senior
    Ongoing contract
    Full time
    Temporary work
    Work experience placement

    SS&C Technologies

    Waltham, MA
    2 days ago
  • $51.9 per hour

     ...is responsible for the reliability, availability, and...  ...role blends software engineering, clinical engineering,...  ...functionally with AHN site leaders and teams to navigate...  ...IT trends, including AI, security patching,...  ...participates in post-incident reviews to identify root... 
    For contractors
    Local area

    Highmark Health

    Boston, MA
    2 days ago
  •  ...systems scale securely and reliably is core to this mission.The AI Devex team exists to...  ...acceleration.As a Senior Platform Engineer on the AI Devex team, you...  ...practices.Participate in incident response, root cause analysis...  ...in DevOps, Platform, Site Reliability,... 
    Senior
    Work at office
    Relocation

    WHOOP

    Boston, MA
    5 days ago
  • $226k - $307k

     ....As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system...  ...FP16).Proficiency in low-level programming for AI accelerators, specifically developing and optimizing custom ML OPs and TensorRT Plugins with efficient CUDA kernel... 
    Senior
    Full time
    Temporary work
    Relocation package

    Zoox

    Boston, MA
    1 day ago
  • Senior Software Engineer — File SystemEngineering | File System Team...  ...that directly affect reliability, performance,...  ...operational quality, and using AI-assisted engineering...  ..., design reviews, incident follow-up, and technical...  ...workspaces Free on-site fitness centers and stocked... 
    Senior
    Full time
    For contractors
    Work at office
    Remote work
    Flexible hours

    Nasuni

    Boston, MA
    2 days ago
  • $180k - $200k

     ...This is a hands-on platform engineering role that blends software...  ...; employees are on-site in the Boston office 3 days...  ...operational workflows to improve reliability, observability and incident response.Contribute to...  ...ChatGPT, Cursor and other AI model experience a plus.The... 
    Senior
    Casual work
    Work at office
    Worldwide
    Flexible hours
    3 days per week

    Acadian Asset Management

    Boston, MA
    3 days ago
  • $191k - $253k

     ...is powered by Lattice OS, an AI-powered operating system that...  ...and systems that empower our engineering teams to develop, test, and deploy...  ...unparalleled efficiency and reliability. We design and implement...  ...operations, troubleshooting, and incident response.REQUIRED... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Boston, MA
    4 days ago
  • $140.88k - $153.75k

     ...Information Job Title Senior Platform Engineer Job ID 105799 Work...  ...call rotation for platform incidents; drive incident response to...  ..., observability, security, reliability, and operational hygiene....  ...ongoing review feedback.Use AI coding assistants to accelerate... 
    Senior
    Permanent employment
    Full time
    Contract work
    Work at office
    Local area

    Bain & Company

    Boston, MA
    4 days ago
  • $185k - $227k

     ...read on for more details. ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance...  ..., and acting as the final escalation point for critical incidents to ensure theplatform is scalable and efficient. Nutanix... 
    Senior
    Remote work

    GrabJobs

    Boston, MA
    1 day ago
  •  ...We are seeking a product-minded Software Engineer to join our core team. In this role, you’...  ...to reimagine it from the ground up. As an AI-native company, we lean in heavily to AI-...  ...domain — SIEM, SOAR, EDR, MDR, or incident response workflows. ~ Excellent software... 
    Senior
    Full time

    Seven Ai

    Boston, MA
    14 hours ago
  • $148k - $222k

     ...own their own destiny.Senior Software Engineer - Agent PlatformAt Klaviyo...  ...state-of-the-art AI and machine learning technologies...  ...vLLM and Ray.Develop reliable, scalable data...  ...listed. AI Ops is a rapidly evolving...  ...on our official career site. Please be cautious of... 
    Senior

    Klaviyo

    Boston, MA
    2 days ago
  • $140k - $160k

     ...leading travel search engine. With billions of queries...  ....KAYAK is seeking a Senior Java Software Engineer...  ...messaging at scale, ensuring reliability and performance across...  ...including alerting and incident response for team-owned...  ...— including AI-assisted tooling and automation... 
    Senior
    Work at office
    Flexible hours
    3 days per week

    Kayak Europe

    Cambridge, MA
    1 day ago
  • $150k - $185.3k

     ...FabricCollaborate with a team of BI Engineers to identify and implement a path forward in leveraging AI assisted BI development in MS...  ...with resolution of escalated incidents and participates in problem...  ...Continuous Delivery) in Microsoft Dev Ops, supporting Azure Data Factory... 
    Senior
    Work experience placement
    Shift work

    Wilmer Hale

    Boston, MA
    1 day ago
  • $165k - $190k

    Maven AGI is an enterprise AI platform founded in July 2023...  ...The RoleWe’re looking for a Senior DevOps Engineer to own and evolve the infrastructure...  ...our platform scales reliably as we onboard enterprise customers...  ..., scaling, monitoring, and incident responseBuild and optimize... 
    Senior

    Maven AGI

    Boston, MA
    4 days ago
  • HPE Labs - Senior Software Engineer - Integrated HPC & Quantum SolutionsThis role has been designed as...  ...computing platforms, GPU computing, AI/ML frameworks (e.g., TensorFlow, PyTorch...  ...or make any payments and report the incident to your local authorities immediately.... 
    Senior
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    Boston, MA
    1 day ago
  • $195.5k - $352.1k

     ...business. The mission of the Ad Engineering Team is to build this platform. We are hiring a Senior Software Engineer for the Advertising...  ...paid time off.How will I use AI at Roku? At Roku, we don’t just...  ..., QA, and infrastructure/ops Be an evangelist for platform innovation... 
    Senior
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    Boston, MA
    3 days ago
  • $75 - $80 per hour

     ...officesPOSITION OVERVIEWThe Senior Developer role...  ...leader, bringing deep engineering expertise while working...  ...performance, scalability, and reliability across FP&A workloads....  ..., data vectors, and AI-based methods where...  ...point for critical incidents, providing Tier-3 expertise... 
    Senior
    Hourly pay
    Work at office

    SRI Tech

    Boston, MA
    3 days ago
  • $190k - $250k

    Senior Software Engineer, ApplicationsAcuityMD is a software and data platform that...  ...Health. We're a high-growth AI and Data company scaling rapidly...  ...VPs, marketers, sales ops, and sales reps. We deliver top...  ...opportunities where physicians or sites of care can better serve... 
    Senior
    For contractors
    Work at office
    Remote work
    Work from home
    Home office
    Flexible hours

    AcuityMD

    Boston, MA
    3 days ago
  • $140k - $170k

     ...learn more.Perk is growing its engineering footprint in North America, and we're looking for a Senior Software Engineer to help us do...  ...practices; defines SLOs, manages incident responses, and conducts...  ...and other job-related factors.AI at Perk:AI is embedded in how we... 
    Senior
    Summer work
    Work at office
    Immediate start
    Worldwide
    Relocation
    Relocation package
    3 days per week

    Perk

    Boston, MA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior AI Ops & Incident/Site Reliability Engineer. Be the first to apply!