Observability Engineer
Rimini Street
Observability Engineer
The Observability Engineer builds the monitoring, tracing, and logging infrastructure that provides visibility into Rimini Street's Agentic AI ERP Platform. You will ensure that operators, developers, and customers can understand system behavior, diagnose issues, and optimize performance across the distributed platform.
Reporting to the Sr Director Engineering – India, this role is critical for production operations and customer trust. You will implement OpenTelemetry-based instrumentation, build dashboards and alerting, and create the observability foundation that enables reliable AI agent operations in enterprise ERP environments.
Essential Duties & Responsibilities
- Design and implement observability architecture using OpenTelemetry as the foundation.
- Deploy and manage distributed tracing infrastructure with Jaeger or similar tools.
- Build metrics collection and storage using Prometheus, Thanos, or Cortex.
- Implement centralized logging with structured log formats and efficient querying.
- Ensure observability infrastructure scales with platform growth and customer deployments.
- Define instrumentation standards for Java (Quarkus), Python, and Angular applications.
- Implement automatic and manual instrumentation for MCP servers and agent workflows.
- Create trace context propagation across service boundaries and ERP system calls.
- Build custom metrics for AI agent behavior, LLM performance, and RAG retrieval quality.
- Integrate observability with CI/CD pipelines for deployment tracking.
- Create Grafana dashboards for platform health, performance, and business metrics.
- Build customer-facing dashboards showing agent activity and processing status.
- Design SLI/SLO dashboards and error budget tracking.
- Implement trace visualization for debugging complex agent workflows.
- Create documentation and training materials for dashboard usage.
- Design alerting strategies that balance signal quality with noise reduction.
- Implement multi-level alerting with appropriate escalation paths.
- Build runbooks linking alerts to diagnostic procedures and remediation steps.
- Support incident response with observability expertise and root cause analysis.
Experience
- 4-6 years in observability, SRE, or platform engineering roles.
- Experience implementing observability for distributed systems or microservices.
- Track record building dashboards and alerting for production systems.
- Experience with OpenTelemetry or similar instrumentation frameworks.
- Background in enterprise software or B2B SaaS preferred.
Technical Skills
Required
- Strong expertise in OpenTelemetry (traces, metrics, logs).
- Experience with Prometheus, Grafana, and alerting systems.
- Knowledge of distributed tracing with Jaeger, Zipkin, or similar tools.
- Experience with log aggregation (ELK stack, Loki, or similar).
- Programming skills in Python, Java, or Go for instrumentation and tooling.
Preferred
- Experience with Kubernetes observability and service mesh telemetry.
- Knowledge of AI/ML observability, including LLM monitoring.
- Familiarity with Quarkus and Python instrumentation.
- Experience with SLI/SLO frameworks and error budgets.
- Understanding of PromQL, LogQL, and TraceQL query languages.
Skills & Competencies
- Strong analytical and troubleshooting skills.
- Good written and verbal English communication skills.
- Ability to work in distributed, multi-timezone teams.
- Data-driven mindset with focus on actionable insights.
- Collaborative approach to working with development and operations teams.
Desired Qualifications
- Bachelor's degree in Computer Science or related field.
- Grafana, Prometheus, or observability platform certifications.
- CKA or cloud platform certifications.
Language: Fluent English required.
Why Rimini Street?
We are looking for talented, passionate people to help us build our future at Rimini Street. We hire only the best, the most extraordinary professionals and provide compensation, bonuses, and benefits to match the skills of our top-performing team members. Do you thrive in a fast-paced environment, enjoy growing together, and get excited about learning new skills? Are you looking for an opportunity to make a true impact as part of a team of extraordinary professionals? This is the place for you.
Our work is challenging and meaningful. We start and end each day with a sense of achievement and purpose guided by our core values, the Four Cs:
- Company: We dream big and innovate boldly.
- Colleagues: We work with extraordinary people who create a culture of mutual respect and collaboration.
- Clients: We relentlessly pursue solutions that help clients achieve their goals. Our unmatched client care is rooted in our passion for exceptional service.
- Community: We believe in leaving the world a better place than we found it. With the Rimini Street Foundation, we've made positive impacts in six continents for over 425 charities.
Accelerating Company Growth
- Nasdaq-listed under ticker symbol RMNI since October 2017
- Over 6,300+ signed contracts to date, including Fortune 500 and Global 100 companies
- Over 2,000 team members in 23 countries
- US and international recognition for industry leadership and philanthropic efforts.
Rimini Street is committed to creating a diverse and inclusive environment and is proud to be an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to age, race, color, religion, national origin, sexual orientation, gender or gender identity, disability, protected veteran status, or any other characteristic protected by law.
- ...opportunities to connect, grow, and make an impact.Why this role:Most engineering roles hand you a slice. A product owner sets the “what,” a... ...platform — including locally-hosted models — that powers observability and automated issue detection across the organization....SuggestedFull timeWork at officeWork visaNight shift
$91.7k - $163.7k
...discover the meaning behind Caring. Connecting. Growing together.The OptumServe Enterprise Monitoring team is seeking a Senior Observability Engineer to lead the design, automation, and operational excellence of our enterprise observability platforms. This role is heavily...SuggestedMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work- ...performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-... ...systems. You will lead reliability initiatives, champion observability and automation, drive major incident response, and partner...SuggestedRemote workFlexible hours
$86.8k - $165.2k
...the strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today’s mission and stay ahead... ...phases of development, production, and maintenance of Low Observable (LO) weapon systems. You will be working on cutting-edge projects...SuggestedTemporary workWork experience placementWork at officeRemote workRelocationFlexible hours$146k - $194k
...military in months, not years.The Air Dominance and Strike Low Observable Teamare advancing into the next generation of LO Technology... ...and effectiveness of unmanned air systems. Our Low Observables engineering team is responsible for designing, analyzing, testing, and...SuggestedFull timeWork experience placementFor subcontractorImmediate start- ...what it means to be part of Mass General Brigham.Job SummarySenior Observability Systems EngineerJoin Mass General Brigham’s Digital Enterprise Observability team as a Senior Observability Systems Engineer. In this role, you’ll help strengthen the reliability,...Full timeLocal areaRemote workMonday to Friday
- ...: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsProduction Support / Observability Engineering Total experience mandate : 7+ YearsLocation : Plano, TX (Hybird)Required Skill:- 3+ years Development or Architecture experience...Full timeFlexible hoursShift workWeekend workAfternoon shift
$100k - $150k
...Site Observability Engineer- Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...Job Description Job Description The Municipal Engineer - Construction Observer represents existing municipal/governmental clients and performs observation and documentation of construction activities and performs in office technical work assignments on municipal...For contractorsFor subcontractorWork at officeNight shiftWeekend work
$159k - $272k
...opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (SREs) focused on the...Full timePrivate practiceLocal areaRemote workWork from home3 days per week- Mass General Brigham is seeking a Senior Observability Systems Engineer to enhance reliability and visibility across enterprise services. You will work hands-on with Dynatrace and ThousandEyes, build automation with Terraform, and collaborate with application, cloud, network...
$138.4k - $173k
...help improve the reliability, quality of services and overall observability patterns. Along with your team, you’ll ensure all aspects of... ...code, and disaster recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability and quality of...Full timeFlexible hours- ...systems. We have an exciting opportunity for an RCS Project Engineer to support Airframe and Structures development projects at General... ...solutions for aircraft structures as a Survivability/Low Observable Integration Project Engineer in the Advanced Aircraft Engineering...Full timeWork at officeRemote work
$146k - $194k
Senior Low Observables Engineer - Mission Systems Anduril Industries is a defense technology company dedicated to transforming U.S. and allied military capabilities with advanced technology. By bringing 21st‑century innovation to the defense industry, Anduril designs,...For subcontractorRelocation package$146k - $194k
A defense technology company is seeking a Senior Low Observables Engineer in Costa Mesa, California. This role involves designing antennas for unmanned air vehicles, requiring extensive RF experience and proficiency in various antenna modeling tools. Candidates must have...- PNC is seeking a Software Engineering Manager in Site Reliability Engineering to lead a distributed team across multiple technology hubs... ..., drive incident response, and spearhead monitoring, observability, and automation initiatives to #J-18808-Ljbffr FairygodbossWork at office
- ...Job Description Job Description SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities on this team. Location: Remote | Time Zone: (US, Canada, Brazil, Chile, Colombia,...Remote work
- What You'll Be DoingThe Sr Observability Engineer will support the design, implementation, and optimization of enterprise observability solutions. This role will drive the product vision, roadmap, and adoption of monitoring platforms, ensuring alignment with business goals...
$60 - $68 per hour
DescriptionKforce has a client that is seeking a Senior Observability Engineer in Boston, MA.Responsibilities:* Senior Observability Engineer will design, implement, and support enterprise monitoring and observability solutions for applications, infrastructure, and cloud...- ...customers feel every second of downtime.We're building a new observability function that runs the way we run incident response: the system... ..., customers, and the exceptions. As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold...Permanent employmentWork at officeShift work
$121k - $163.4k
...DescriptionAre you a seasoned professional with a passion for observability, telemetry, and system reliability? Do you excel in a dynamic... ...? At CoBank, we are seeking a Senior Observability Engineer to help design, implement, and scale enterprise observability...Work at officeWork visaFlexible hours- San Francisco, CAMission & Program Engineering - Satellite Systems Engineering /Full time /On-siteWanna join the adventure?Are you passionate... ...per year and is building the next generation of Earth observation capability - from payload design through calibration, processing...Full timeTemporary work
$90k - $140k
Job DescriptionWhat will you do?Design and implement observability solutions using industry-leading platforms, establishing logging standards... ...and prevent failuresCollaborate with cross-functional teams (Engineering, DevOps, Security, production support, product) to...Full timeFlexible hours- ...Observability / Site Reliability Engineer (SRE)Ontrac Solutions is a leading technology consulting firm, specializing in cutting-edge solutions that drive business transformation. We partner with organizations to modernize their infrastructure, streamline processes, and...
- Atlassian in San Francisco is seeking a systems engineer to own parts of an observability layer for AI-assisted software development. You will detect and attribute AI-generated code across various workflows, ensuring seamless integration with developer tooling. The ideal...
- LangChain is looking for a Systems/Database Engineer based in San Francisco to design and optimize a storage layer for AI observability. You will write performant Rust code, optimize database services, and ensure efficient cloud storage integration. The ideal candidate...
$98.18k - $115.5k
...career goals: partnering with our customers, our communities, and each other. Job DescriptionResponsibilitiesThe Reliability Observability Engineer 3 is responsible for enabling reliable, measurable, and supportable application operations across a broad portfolio of...Full timeWork experience placementLocal area3 days per week- Job Title : Observability Operations Engineer Location : Phoenix,AZ ( only Local candidate) Client: TCS Rate: $35/hr on W2 Positions: 2 JD: Job Title: Observability Operations Engineer Role Description: "• Administer and optimize enterprise Dynatrace...Local area
$50 per hour
...Senior Observability Operations Engineer Location: Phoenix, Arizona, USA | Employment Type: Full-Time | Job Code: PAN-00000050 Role Overview The ideal candidate has expertise in Dynatrace, Splunk, OpenSearch/Elasticsearch, Kubernetes, Linux, and cloud-native...Full time- ...Mid-Level Observability Engineer – End-User Experience (Aternity) Location: Remote USA Type: Contract Role Summary: The Observability Engineer will support Customer’s enterprise observability program with a focus on end-user and digital experience monitoring...Contract workRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Observability Engineer. Be the first to apply!



