Observability Engineer
Rimini Street
Observability Engineer
The Observability Engineer builds the monitoring, tracing, and logging infrastructure that provides visibility into Rimini Street's Agentic AI ERP Platform. You will ensure that operators, developers, and customers can understand system behavior, diagnose issues, and optimize performance across the distributed platform.
Reporting to the Sr Director Engineering – India, this role is critical for production operations and customer trust. You will implement OpenTelemetry-based instrumentation, build dashboards and alerting, and create the observability foundation that enables reliable AI agent operations in enterprise ERP environments.
Essential Duties & Responsibilities
Observability Infrastructure
- Design and implement observability architecture using OpenTelemetry as the foundation.
- Deploy and manage distributed tracing infrastructure with Jaeger or similar tools.
- Build metrics collection and storage using Prometheus, Thanos, or Cortex.
- Implement centralized logging with structured log formats and efficient querying.
- Ensure observability infrastructure scales with platform growth and customer deployments.
Instrumentation & Integration
- Define instrumentation standards for Java (Quarkus), Python, and Angular applications.
- Implement automatic and manual instrumentation for MCP servers and agent workflows.
- Create trace context propagation across service boundaries and ERP system calls.
- Build custom metrics for AI agent behavior, LLM performance, and RAG retrieval quality.
- Integrate observability with CI/CD pipelines for deployment tracking.
Dashboards & Visualization
- Create Grafana dashboards for platform health, performance, and business metrics.
- Build customer-facing dashboards showing agent activity and processing status.
- Design SLI/SLO dashboards and error budget tracking.
- Implement trace visualization for debugging complex agent workflows.
- Create documentation and training materials for dashboard usage.
Alerting & Incident Response
- Design alerting strategies that balance signal quality with noise reduction.
- Implement multi-level alerting with appropriate escalation paths.
- Build runbooks linking alerts to diagnostic procedures and remediation steps.
- Support incident response with observability expertise and root cause analysis.
Experience
- 4-6 years in observability, SRE, or platform engineering roles.
- Experience implementing observability for distributed systems or microservices.
- Track record building dashboards and alerting for production systems.
- Experience with OpenTelemetry or similar instrumentation frameworks.
- Background in enterprise software or B2B SaaS preferred.
Technical Skills
Required
- Strong expertise in OpenTelemetry (traces, metrics, logs).
- Experience with Prometheus, Grafana, and alerting systems.
- Knowledge of distributed tracing with Jaeger, Zipkin, or similar tools.
- Experience with log aggregation (ELK stack, Loki, or similar).
- Programming skills in Python, Java, or Go for instrumentation and tooling.
Preferred
- Experience with Kubernetes observability and service mesh telemetry.
- Knowledge of AI/ML observability, including LLM monitoring.
- Familiarity with Quarkus and Python instrumentation.
- Experience with SLI/SLO frameworks and error budgets.
- Understanding of PromQL, LogQL, and TraceQL query languages.
Skills & Competencies
- Strong analytical and troubleshooting skills.
- Good written and verbal English communication skills.
- Ability to work in distributed, multi-timezone teams.
- Data-driven mindset with focus on actionable insights.
- Collaborative approach to working with development and operations teams.
Desired Qualifications
- Bachelor's degree in Computer Science or related field.
- Grafana, Prometheus, or observability platform certifications.
- CKA or cloud platform certifications.
Language: Fluent English required.
Why Rimini Street?
We are looking for talented, passionate people to help us build our future at Rimini Street. We hire only the best, the most extraordinary professionals and provide compensation, bonuses, and benefits to match the skills of our top-performing team members. Do you thrive in a fast-paced environment, enjoy growing together, and get excited about learning new skills? Are you looking for an opportunity to make a true impact as part of a team of extraordinary professionals? This is the place for you.
Our work is challenging and meaningful. We start and end each day with a sense of achievement and purpose guided by our core values, the Four Cs:
- Company: We dream big and innovate boldly.
- Colleagues: We work with extraordinary people who create a culture of mutual respect and collaboration.
- Clients: We relentlessly pursue solutions that help clients achieve their goals. Our unmatched client care is rooted in our passion for exceptional service.
- Community: We believe in leaving the world a better place than we found it. With the Rimini Street Foundation, we've made positive impacts in six continents for over 425 charities.
Accelerating Company Growth
- Nasdaq-listed under ticker symbol RMNI since October 2017
- Over 6,300+ signed contracts to date, including Fortune 500 and Global 100 companies
- Over 2,000 team members in 23 countries
- US and international recognition for industry leadership and philanthropic efforts.
Rimini Street is committed to creating a diverse and inclusive environment and is proud to be an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to age, race, color, religion, national origin, sexual orientation, gender or gender identity, disability, protected veteran status, or any other characteristic protected by law.
$129k - $171k
...networking technology to the military in months, not years. ABOUT THE TEAM We are looking for a Staff Product Quality Engineer (Low Observable Systems) to support our Fury autonomous aircraft program at our Arsenal 1 facility in Ashville, OH. This role will be on...SuggestedFull timeWork experience placementImmediate start$180k - $270k
...the core infrastructure that supports the Control Plane and Observability infrastructure for products in the Hyperscale line-of-business... ...~3+ years of experience in software development and systems engineering, using languages such as Golang, Python, C++, or Java. ~...SuggestedFull timeWork at officeFlexible hours$91.7k - $163.7k
...discover the meaning behind Caring. Connecting. Growing together.The OptumServe Enterprise Monitoring team is seeking a Senior Observability Engineer to lead the design, automation, and operational excellence of our enterprise observability platforms. This role is heavily...SuggestedMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work$107.5k - $204.5k
...the strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today’s mission and stay ahead... ...phases of development, production, and maintenance of Low Observable (LO) weapon systems. You will be working on cutting-edge projects...SuggestedTemporary workWork experience placementWork at officeRemote workRelocationFlexible hours- ...performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-... ...systems. You will lead reliability initiatives, champion observability and automation, drive major incident response, and partner...SuggestedRemote workFlexible hours
$146k - $194k
...military in months, not years. The Air Dominance and Strike Low Observable Team are advancing into the next generation of LO Technology... ...effectiveness of unmanned air systems. Our Low Observables engineering team is responsible for designing, analyzing, testing, and...Full timeWork experience placementFor subcontractorImmediate start- ...: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsProduction Support / Observability Engineering Total experience mandate : 7+ YearsLocation : Plano, TX (Hybird)Required Skill:- 3+ years Development or Architecture experience...Full timeFlexible hoursShift workWeekend workAfternoon shift
$115k - $151.8k
As a Sr. Systems Engineer II, you’ll deploy, operate, and support our observability platform — a containerized monitoring stack (TimescaleDB, OpenSearch, Telegraf, Promethus) running on Linux hosts in an on-premises, air-gapped environment — and to serve as a strong network...Full timeLocal area$159k - $272k
...opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (SREs) focused on the...Full timePrivate practiceLocal areaRemote workWork from home3 days per week$100k - $150k
...Site Observability Engineer- Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- Mass General Brigham is seeking a Senior Observability Systems Engineer to enhance reliability and visibility across enterprise services. You will work hands-on with Dynatrace and ThousandEyes, build automation with Terraform, and collaborate with application, cloud, network...
$138.4k - $173k
...help improve the reliability, quality of services and overall observability patterns. Along with your team, you’ll ensure all aspects of... ...code, and disaster recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability and quality of...Full timeFlexible hours- ...customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability...Full timeWorldwideFlexible hours
- ...systems. We have an exciting opportunity for an RCS Project Engineer to support Airframe and Structures development projects at General... ...solutions for aircraft structures as a Survivability/Low Observable Integration Project Engineer in the Advanced Aircraft Engineering...Full timeWork at officeRemote work
$93.95k - $136.74k
...invite all applicants to join us and experience what it means to be part of Mass General Brigham. Job Summary Senior Observability Systems Engineer Join Mass General Brigham’s Digital Enterprise Observability team as a Senior Observability Systems Engineer. In this role...Full timeLocal areaRemote workMonday to FridayShift work$146k - $194k
Senior Low Observables Engineer - Mission Systems Anduril Industries is a defense technology company dedicated to transforming U.S. and allied military capabilities with advanced technology. By bringing 21st‑century innovation to the defense industry, Anduril designs,...For subcontractorRelocation package$146k - $194k
A defense technology company is seeking a Senior Low Observables Engineer in Costa Mesa, California. This role involves designing antennas for unmanned air vehicles, requiring extensive RF experience and proficiency in various antenna modeling tools. Candidates must have...- ...SS&C for expertise, scale, and technology.Job DescriptionJob Title: Sr. Observability EngineerLocations: [Waltham, MA - Hybrid]About the RoleWe are seeking a skilled Monitoring & Alerting Engineer to design, build, and enhance tools and applications that power SS&C Intralinks...Ongoing contractFull time
- Platform Engineer (DevOps / CI/CD / Automation / Observability) DETAILS Location : Remote Position Type : 16M Contract Hourly / Salary : to $90W2 based on experience) JOB SUMMARY Vaco is currently seeking a Platform Engineer (DevOps / CI/CD / Automation / Observability)...Hourly payContract workWork at officeLocal areaRemote work
- ...Job Description Job Description SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities on this team. Location: Remote | Time Zone: (US, Canada, Brazil, Chile, Colombia,...Remote work
- ...Observability / Site Reliability Engineer (SRE) Ontrac Solutions is a leading technology consulting firm, specializing in cutting-edge solutions that drive business transformation. We partner with organizations to modernize their infrastructure, streamline processes...Remote work
$60 - $68 per hour
DescriptionKforce has a client that is seeking a Senior Observability Engineer in Boston, MA.Responsibilities:* Senior Observability Engineer will design, implement, and support enterprise monitoring and observability solutions for applications, infrastructure, and cloud...- ...customers feel every second of downtime.We're building a new observability function that runs the way we run incident response: the system... ..., customers, and the exceptions. As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold...Permanent employmentWork at officeShift work
- San Francisco, CAMission & Program Engineering - Satellite Systems Engineering /Full time /On-siteWanna join the adventure?Are you passionate... ...per year and is building the next generation of Earth observation capability - from payload design through calibration, processing...Full timeTemporary work
$90k - $140k
Job DescriptionWhat will you do?Design and implement observability solutions using industry-leading platforms, establishing logging standards... ...and prevent failuresCollaborate with cross-functional teams (Engineering, DevOps, Security, production support, product) to...Full timeFlexible hours$146k - $194k
...years. ABOUT THE TEAM The Air Dominance and Strike Low Observable Team are advancing into the next generation of LO... ...and effectiveness of unmanned air systems. Our Low Observables engineering team is responsible for designing, analyzing, testing, and sustaining...Full timeWork experience placementImmediate start- LangChain is looking for a Systems/Database Engineer based in San Francisco to design and optimize a storage layer for AI observability. You will write performant Rust code, optimize database services, and ensure efficient cloud storage integration. The ideal candidate...
- What You'll Be DoingThe Sr Observability Engineer will support the design, implementation, and optimization of enterprise observability solutions. This role will drive the product vision, roadmap, and adoption of monitoring platforms, ensuring alignment with business goals...
- Atlassian in San Francisco is seeking a systems engineer to own parts of an observability layer for AI-assisted software development. You will detect and attribute AI-generated code across various workflows, ensuring seamless integration with developer tooling. The ideal...
$124.36k - $146.3k
...with our customers, our communities, and each other. Job DescriptionResponsibilitiesAs a senior-level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and business stakeholders...Full timeWork experience placementLocal area3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Observability Engineer. Be the first to apply!


