Senior Site Reliability Engineer
$160k - $250kHive
DevOps And Systems Engineer
Hive is the leading provider of cloud-based AI solutions to understand, search, and generate content, and is trusted by hundreds of the world's largest and most innovative organizations. The company empowers developers with a portfolio of best-in-class, pre-trained AI models, serving billions of customer API requests every month. Hive also offers turnkey software applications powered by proprietary AI models and datasets, enabling breakthrough use cases across industries. Together, Hive's solutions are transforming content moderation, brand protection, sponsorship measurement, context-based ad targeting, and more.
Our unique machine learning needs led us to open our own data centers, with an emphasis on distributed high performance computing integrating GPUs. Even with these data centers, we maintain a hybrid infrastructure with public clouds when the right fit. As we continue to commercialize our machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering for our customers. Our ideal candidate is someone who is able to thrive in an unstructured environment and takes automation seriously. You believe there is no task that can't be automated and no server scale too large. You take pride in optimizing performance at scale in every part of the stack and never manually performing the same task twice.
Responsibilities
- Automate manual operational processes
- Improve workflows of developer, data, and machine learning teams
- Manage secure integration and deployment tooling
- Create, maintain, monitor, and audit secure infrastructure
- Manage a diverse array of technology platforms, following best practices and procedures
- Participate in on-call rotation and root cause analysis
- Maintain awareness of industry best practices for data maintenance handling as it relates to your role
- Adhere to policies, guidelines and procedures pertaining to the protection of information assets
- Report actual or suspected security and/or policy violations/breaches to an appropriate authority
Requirements
- Minimum 3 - 5 years of previous experience in development, operations, IT, or a related field
- Comfortable working on Linux infrastructures (Debian) via the CLI
- Able to learn quickly in a fast-paced environment
- Able to debug, optimize, and automate routine tasks
- Able to multitask, prioritize, and manage time efficiently independently
- Able to physically lift equipment at least 30 pounds
- Can communicate effectively across teams and management levels
- Degree in computer science, or similar, is an added plus!
Technology Stack
- Operating Systems - Linux/Debian Family/Ubuntu
- Configuration Management - Chef
- Containerization - Docker
- Container Orchestrators - Mesosphere/Kubernetes
- Scripting Languages - Python/Ruby/Node/Bash
- CI/CD Tools - Jenkins
- Network hardware - Arista/Cisco/Fortinet
- Hardware - HP/SuperMicro
- Storage - Ceph, S3
- Database - Scylla, Postgres, Pivotal GreenPlum
- Message Brokers: RabbitMQ
- Logging/Search - ELK Stack
- AWS: VPC/EC2/IAM/S3
- Networking: TCP / IP, ICMP, SSH, DNS, SSL / TLS, Storage systems, RAID, distributed file systems, NFS / iSCSI / CIFS
We are a group of ambitious individuals who are passionate about creating a revolutionary AI company. At Hive, you will have a steep learning curve and an opportunity to contribute to one of the fastest growing AI start-ups in San Francisco. The work you do here will have a noticeable and direct impact on the development of the company.
Thank you for your interest in Hive and we hope to meet you soon!
The current expected base salary for this position ranges from $160,000 - $250,000. Actual compensation may vary depending on a number of factors, including a candidate's qualifications, skills, competencies and experience, and location. Base pay is one part of the total compensation package that is provided to compensate and recognize employees for their work; stock options may be offered in addition to the range provided here.
Employees are eligible to participate in a number of Company-sponsored benefits, including health, vision and dental insurance. Employees are also eligible to participate in a gym membership as part of our commitment to employee wellness. In addition, employees will be entitled to paid vacation in accordance with the Company's vacation policy.
Hired applicant may receive an equity grant in the form of an option to purchase stock in the future for a specified price.
$170k - $220k
Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating...Senior$147k - $202.4k
...career-defining work. We're all in on this mission. If you are too, let's talk.Our company is seeking a highly skilled Senior Site Reliability Engineer to join our team. We are a SaaS company specializing in securing large-scale systems. This role is a blend of software...SeniorWork at officeLocal areaWorldwideFlexible hoursShift work$166k - $258k
...Seattle office a minimum of 4 days/week in order to be considered for this position.Nordstrom is looking for a Senior Engineer 2 to join our Site Reliability Engineering (SRE) team — and we think that person could be you.You'll help build the scalable, reliable, and resilient...SeniorFull timeWork at office$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed... ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,...SeniorWork experience placementWork at officeRemote work$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...SeniorPermanent employmentLocal areaWorldwideFlexible hours$167.7k - $245.2k
...platform. This team is responsible for architecting, delivering, and maintaining our FedRAMP offering. Your ImpactAs a FedRAMP Site Reliability Engineer(SRE), you will lead the operations and architecture of our mission-critical Federal region, ensuring peak availability,...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours2 days per week$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As a Staff SRE, you will be very hands-on technically while also mentoring a small team of SREs.The InfraSec team collaborates...SeniorLocal areaRemote workWorldwideFlexible hours- ...team is responsible for the reliability, scalability, and efficiency... ...building features, but about engineering the resilience and performance... ...maintain system stability.As a Site Reliability Engineer, you... ...practices while working alongside senior engineers to solve...Senior
$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...enabled with self-serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours- ...Excellent analytical and problem-solving skills with a proactive approach. AI/ML experience or a strong interest in applying AI/ML to reliability, security, or operational efficiency is a plus. Benefits Health, dental, and vision insurance; 401(k); flexible spending...SeniorFull timeWork at officeFlexible hours
- ...of the cloud footprint Evaluate AWS compute, networking, and security architecture for scalability and growth Collaborate across engineering and product teams to scope projects to core business requirements Drive engineering-wide improvements to service management for deployments...Senior
$160k - $210k
...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale...SeniorWork at officeImmediate startRemote workWork from home$134.25k - $214.8k
...real change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability...SeniorWork experience placementWork at officeRemote workFlexible hours$143k - $191k
...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental... ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and...Full timeTemporary workWork experience placementImmediate start- Company DescriptionComtech LLC is a woman-owned small business focused on delivering end-to-end solutions and products. Since 1998, we have successfully serviced enterprises across the public and private sectors, and the Department of Defense. Our services span all aspects...
$120k - $170k
Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager... ...from the office or from a remote space of your choosing. Seniority level Seniority level Mid-Senior level Employment type...SeniorFull timeWork at officeRemote workFlexible hours$152k - $241.5k
...the world.Join the Simulation Software team at NVIDIA as a Senior System Software Engineer! This role offers an outstanding opportunity to work on... ...development by enabling Chips Simulation as a trusted and reliable virtual platform.What you will be doing:Drive early...SeniorFull time$194k - $267k
...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$194k - $267k
...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$204k - $306k
...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from... ...in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$194k - $267k
...career-defining work. We're all in on this mission. If you are too, let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services...Local areaWorldwideFlexible hours- ...future together. We are responsible for the reliability of all the company's major data warehouse products, services, and query engines. We serve business needs across domains... ...practices, and emerging technologies related to site reliability and infrastructure engineering....
- ...data planes. We are hoping to enhance engineering efficiency by concentrating our expertise... ...framework that will ensure the reliability of databases being used by critical tier... ...is an exciting opportunity to work with senior architects & leaders at OCI as well as...SeniorWorldwideFlexible hours
- ...to our people, and the incredible connections we get to make in every community we are in. About this team Site Reliability Engineering We are looking for a motivated engineer to join the Foundations team which is responsibility for observability and...
- Job Title Required Skills: CHEF experience - Must have most critical Azure Cloud – experience - Must have most critical AKS- Azure Kubernetes services - Must have most critical Kubernetes - Must have most critical NoSQL DB – Cassandra / Mongo DB ...
$236k - $339.2k
...the ecosystem that powers our AI-pilled engineers and our agentic organizations and reinvents... ...a bias for measurable impact through reliability, performance, and ease of use.Bonus points... ...job posting on the Snowflake Careers Site for salary and benefits information: careers...Senior$132.23k - $176.31k
...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem....SeniorFull timeTemporary workRemote work$152k - $241.5k
As a Vulkan Performance driver engineer, you will have a hand in everything from the game engine down to bare metal! You will be part of a team whose mission it is to achieve the best possible performance, power-efficiency, and latency for the latest games and creative...SeniorFull time$184k - $287.5k
...workloads worldwide. Our team of skilled engineers is committed to addressing major global... ...lives around the world!We are looking for a Senior Systems Software Engineer with strong... ...depth needed to maintain cluster reliability at frontier AI scale. In this vital role...SeniorFull timeWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Seattle, WA
- site reliability engineer remote Seattle, WA
- site reliability engineer sre Seattle, WA
- senior technical analyst Seattle, WA
- senior associate attorney Seattle, WA
- senior developer Seattle, WA
- sr industrial security specialist Seattle, WA
- senior aws cloud engineer Seattle, WA
- senior manager business development Seattle, WA
- remote senior salesforce administrator Seattle, WA




