Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Cloud Reliability Engineer

Versant Media

Job Description

Job Description

Company Description

VERSANT (Nasdaq: VSNT) is an industry-changing media and entertainment business and home to trusted brands that shape culture, inform audiences, and build lasting connections. It operates across four core markets: political news and opinion, business news and personal finance, golf, and sports and genre entertainment. These markets are served through a powerful portfolio of iconic and innovative brands, including CNBC, MS NOW, USA Network, Golf Channel, Oxygen, E!, SYFY, and Versant's sports division USA Sports, along with complementary digital assets including Fandango, Rotten Tomatoes, GolfNow and GolfPass.

Job Description

The Cloud Reliability Engineer is responsible for ensuring the availability, performance, scalability, and operational excellence of VERSANT’s cloud platforms and services. 

This role works closely with cloud engineering, application development, networking, security, and operations teams to build and maintain highly reliable systems across a large multi-account AWS environment. The engineer will leverage automation, observability, and reliability engineering practices to improve platform resilience, reduce operational risk, and enhance the customer experience. 

As a leading media company, VERSANT operates digital products, streaming platforms, content delivery systems, and media workflows that demand high levels of uptime and performance. The Cloud Reliability Engineer will help ensure these services remain resilient, scalable, and operationally mature. 

The ideal candidate has strong experience with AWS, monitoring and observability platforms, incident management, automation, infrastructure as code, and operational best practices. Experience with AWS Organizations, Control Tower, Identity Center, Terraform, and modern cloud operations tooling is highly desirable. 

Responsibilities 

Reliabiliy Engineering 

  • Design, implement, and maintain reliability practices for cloud infrastructure and platform services. 

  • Define and monitor service-level objectives (SLOs), service-level indicators (SLIs), and operational metrics. 

  • Identify reliability risks and implement solutions that improve availability, scalability, and resilience. 

  • Drive continuous improvement initiatives focused on operational excellence and system stability. 

Monitoring, Observability & Performance 

  • Design and maintain monitoring, logging, alerting, and observability solutions across AWS environments. 

  • Develop dashboards and reporting that provide visibility into platform health and performance. 

  • Analyze system behavior, identify bottlenecks, and implement performance improvements. 

  • Establish proactive monitoring practices that detect issues before they impact customers. 

Incident Response & Operational Excellence 

  • Participate in incident response, troubleshooting, and root cause analysis activities. 

  • Lead post-incident reviews and identify corrective actions to prevent recurrence. 

  • Improve operational processes, runbooks, and recovery procedures. 

  • Support disaster recovery and business continuity initiatives. 

 

AWS Platform Reliability 

  • Support the reliability and operational health of large-scale AWS environments utilizing AWS Organizations, Control Tower, and Identity Center. 

  • Partner with cloud engineering teams to improve platform architecture, resiliency, and operational consistency. 

  • Assist in maintaining secure, scalable, and highly available cloud services. 

  • Automation & Infrastructure as Code 

  • Develop automation that reduces operational toil and improves system reliability. 

  • Support infrastructure-as-code solutions using Terraform, CloudFormation, and related technologies. 

  • Automate operational workflows, monitoring, remediation, and recovery activities. 

  • Contribute to CI/CD pipelines and deployment automation initiatives. 

 

Media & Digital Platform Reliability 

  • Support the reliability of streaming platforms, content delivery systems, media workflows, APIs, and customer-facing applications. 

  • Collaborate with engineering teams to improve application reliability and operational readiness. 

  • Assist in capacity planning and scaling efforts for high-traffic events and media workloads. 

 

Collaboration & Continuous Improvement 

  • Partner with cloud, networking, security, and application teams to identify and address operational risks. 

  • Promote reliability engineering best practices throughout the organization. 

  • Contribute to documentation, standards, and operational procedures. 

  • Evaluate emerging technologies and recommend improvements to platform reliability and observability. 

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, Information Systems, or equivalent practical experience. 

  • 3–7 years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, Infrastructure Engineering, or related roles. 

  • Strong hands-on experience with AWS cloud services and enterprise-scale AWS environments. 

  • Experience with: 

  • Monitoring and observability platforms 

  • Incident management and root cause analysis 

  • Operational troubleshooting and performance tuning 

  • AWS Organizations, Control Tower, and Identity Center 

  • Experience with Infrastructure as Code: 

  • Terraform 

  • CloudFormation 

  • Experience with CI/CD platforms and deployment automation. 

  • Experience with scripting and automation using Python, PowerShell, Bash, or similar languages. 

  • Strong understanding of AWS networking, resiliency, and cloud architecture concepts. 

  • Experience with logging, metrics, tracing, and alerting technologies. 

  • Strong troubleshooting, communication, and collaboration skills. 

 

Additional Information 

Location: New York City, NY or Englewood Cliffs, NJ - (Hybrid – 3 days onsite) 

Employees based in our Englewood Cliffs office have access to a variety of convenient on-site services and amenities, including free employee parking, complimentary electric vehicle charging stations, an on-site fitness center with locker rooms and towel service, and dry-cleaning drop-off and pick-up.

To support an easier commute, shuttle service is also available to and from the Englewood Cliffs campus, with pickup locations across Manhattan’s East Side and West Side, Brooklyn, Hoboken, Jersey City, Newark, and Secaucus.

Additional Information

As part of our selection process, external candidates may be required to attend an in-person interview with a VERSANT Media employee at one of our locations prior to a hiring decision. VERSANT Media's policy is to provide equal employment opportunities to all applicants and employees without regard to race, color, religion, creed, gender, gender identity or expression, age, national origin or ancestry, citizenship, disability, sexual orientation, marital status, pregnancy, veteran status, membership in the uniformed services, genetic information, or any other basis protected by applicable law.

If you are a qualified individual with a disability or a disabled veteran and require support throughout the application and/or recruitment process as a result of your disability, you have the right to request a reasonable accommodation. You can submit your request to View email address on us.fitly.work.

VERSANT Media is committed to fair and equitable compensation practices. We include a good faith pay range for each position to comply with applicable state and local pay transparency laws and to promote equity across our organization. Actual compensation will be based on factors such as the candidate's skills, qualifications, experience, and location and may include additional forms of compensation and benefits such as health insurance, retirement plans, paid time off, etc.

VERSANT Media is not accepting unsolicited assistance from search firms for this employment opportunity. All resumes submitted by search firms to any employee at VERSANT via-email, the Internet, or in any form and/or method without a valid written Statement of Work in place for this position from VERSANT's Talent Acquisition team will be deemed the sole property of VERSANT. No fee will be paid in the event the candidate is hired by VERSANT as a result of the referral or through other means.

Vacancy posted 23 days ago
Similar jobs that could be interesting for youBased on the Cloud Reliability Engineer in Englewood Cliffs, NJ vacancy
  • As a Lead Site Reliability Engineer at JPMorgan Chase within the Public Cloud team, you will blend hands-on engineering with program leadership to promote platform stability, ensure consistent execution across SRE teams, and partner closely with Engineering and Product... 
    Suggested

    JP Morgan Chase

    Jersey City, NJ
    18 hours ago
  •  ...Collaborate with development teams to identify reliability risks and improve system architecture...  ...automation • Experience with cloud platforms such as AWS and/or GCP • Solid...  ...and best practices • Certifications in SRE, DevOps, or Performance Engineering are a plus
    Suggested

    Omni Inclusive

    Englewood Cliffs, NJ
    4 days ago
  • $61k - $101k

     ...Requirements: We need 5+ years of experience in SRE, production engineering, platform reliability, or infrastructure operations at enterprise scale. We...  ...Technologies: AI AWS Azure Bash CI/CD Cloud GCP Support LLM Python Terraform DevOps... 
    Suggested
    Full time

    J.P. Morgan

    Jersey City, NJ
    7 days ago
  • $110k - $125k

     ...Site Reliability Engineer Must Have Technical/Functional Skills • 6-7 years of experience in Site Reliability Engineering, Production Support...  ...and troubleshooting. • Hands-on experience with AWS cloud services (EC2, RDS, IAM, VPC, CloudWatch, S3). • Experience... 
    Suggested

    Tata Consultancy Services

    Englewood Cliffs, NJ
    3 days ago
  •  ...Information Technology group delivers secure, reliable technology solutions that enable DTCC to...  ....As a Principal Site Reliability Engineer (SRE), you will drive operational excellence...  ...advancing SRE best practices through modern cloud, observability, automation, and AI-... 
    Suggested
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Jersey City, NJ
    18 hours ago
  • $30.7 - $46.05 per hour

     ...Engineer I, Reliability Engineering The Engineer I, Reliability Engineering, at Thermo Fisher Scientific will play a crucial role in supporting equipment reliability, maintenance strategy, and manufacturing operations, along with process improvements and engineering... 
    Hourly pay
    Temporary work
    Internship
    Work at office

    Thermo Fisher

    Fair Lawn, NJ
    4 days ago
  • $124k - $186k

     ...where you can make an impact.With interesting opportunities in engineering, marketing, sales, supply chain, operations, HR, finance, and...  ...have something special for you.POSITION SUMMARYThe Test and Reliability Engineer is responsible for ensuring we deliver product... 
    Full time
    Work experience placement
    Local area

    IDEX

    Rutherford, NJ
    4 days ago
  • $99k - $132k

     ...Description Job Description Job Title: Reliability Engineer Location: Bayonne, NJ Department: Engineering Bayonne Experience: 5-7 years FLSA Status: Exempt Safety Sensitive: No We are a global leader in food & beverage ingredients.  Pioneers at heart... 
    Work at office

    OFI

    Bayonne, NJ
    2 days ago
  • $105k - $154k

     ...Reliability Engineer About the Role Our reliability team is responsible to evaluate, develop, design, and implement software and product...  ...the latest technologies that go into building a hyperscale cloud services. What You'll Do The successful candidate will... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    ZT Systems

    Secaucus, NJ
    1 day ago
  • $123.8k - $175k

     ...customers to make the world healthier, cleaner, and safer through reliable manufacturing operations. You'll work with cross-functional...  ...'s Degree plus 8 years of experience in maintenance or engineering in pharmaceutical/biotech manufacturing • Preferred Fields of... 
    Temporary work
    Work at office

    Thermo Fisher Scientific

    Fair Lawn, NJ
    4 days ago
  •  ...Job Title: Reliability Test Engineer Location : Hoboken, NJ  Division: Technology Department:  Technology About Us Quantum Computing Inc. (QCi) (Nasdaq: QUBT) is an innovative, integrated photonics company that provides accessible and affordable quantum machines... 
    Work at office
    Local area
    Remote work

    Quantum Computing Inc.

    Hoboken, NJ
    1 day ago
  • $120k - $140k

     ...person can experience a sense of belonging. Join us on our One Perrigo journey as we evolve to win in self-care. The Reliability Engineer serves as the site reliability leader and technical subject matter expert for asset reliability, maintenance strategy,... 
    For contractors

    Perrigo Company plc

    Bronx, NY
    a month ago
  •  ...you can push the limits of what's possible. As a Lead Software Engineer at JPMorganChase within Corporate Technology, you are an integral...  ...the financial services industry and their IT systemsPractical cloud native experience JPMorganChase, one of the oldest financial... 
    Work at office
    Local area

    JP Morgan Chase

    Jersey City, NJ
    18 hours ago
  •  ...Role:Infrastructure Reliability Engineer Location: Jersey City / Columbus, OH Job Description: We are looking for a hands-...  ...and operability across enterprise infrastructure -spanning cloud and on-prem/hybrid environments . This role focuses on Terraform... 

    Saransh Inc

    Jersey City, NJ
    1 day ago
  •  ...member of our recruitment team will provide more details.Job Summary: MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment in AWS, support the configuration, deployment and... 
    Full time
    Work at office
    Local area
    Remote work

    MUFG

    Jersey City, NJ
    3 days ago
  •  ...world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office...  ...with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and... 
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    18 hours ago
  •  ...Fandango is looking for a SENIOR PLATFORM ENGINEER to build our next big thing in platform...  ...diverse domains, offering expertise in AWS cloud infrastructure and resources, containers...  ...automation solutions that enable rapid, reliable, and secure deployments Design and implement... 
    Local area

    Versant Media

    Englewood Cliffs, NJ
    23 days ago
  •  ...effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Audit Technology, Production...  ...maturity.Hands-on technology experience across cloud (AWS/Azure/GCP), automation and scripting (Python, Shell... 

    JP Morgan Chase

    Jersey City, NJ
    3 days ago
  •  ...complex and mission-critical systems. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Commercial Investment Banking team...  ...with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and... 

    JP Morgan Chase

    Jersey City, NJ
    2 days ago
  •  ...firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate strong knowledge... 

    JP Morgan Chase

    Jersey City, NJ
    4 days ago
  • $137.75k - $185k

     ...most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Consumer and Investment...  ...with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize... 
    Local area

    JPMorgan Chase Bank, N.A.

    Jersey City, NJ
    18 hours ago
  • $60 - $65 per hour

     ...SRE Engineer (W2) Jersey City, NJ (Onsite) 6 Months Contract to Hire Job Description: Proficient in application development...  ...Integration, Continuous Delivery, Test Driven Development, Cloud Development, application resiliency and security.... 
    Full time
    Contract work
    Work experience placement

    Pinnacle Group

    Jersey City, NJ
    18 hours ago
  •  ...Job Title: Site Reliability Engineer (SRE) Location: Wood Ridge, NJ Duration: 6 Months Position Overview We...  ...technologies (Docker, Kubernetes). ~ Hands-on experience with cloud platforms (AWS, Azure, or GCP). ~ Excellent... 
    Permanent employment

    PROLIM Corporation

    Wood Ridge, NJ
    4 days ago
  •  ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities... 

    JP Morgan Chase

    Jersey City, NJ
    4 days ago
  •  ...Gartner® Magic Quadrant™ for Supplier Risk Management. Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S....  ...Linux/Unix or other operating systems. ~ Familiarity with cloud platforms (AWS) and secure system integration. ~ Comfort integrating... 
    Work at office
    Work from home
    Flexible hours

    Exiger

    Jersey City, NJ
    2 days ago
  •  ...BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ...  ...Integration, Continuous Delivery, Test Driven Development, cloud development, application resiliency, and security. Used terraform... 
    Temporary work

    BCForward

    Jersey City, NJ
    a month ago
  • $42.5k - $70.5k

     ...formal training or certification in software engineering concepts, along with 3+ years of applied...  ...plus exposure to modern domains such as cloud, AI/ML, and/or mobile and their...  ...expertise in SRE best practices, including reliability, scalability, performance, security, enterprise... 
    Full time

    J.P. Morgan

    Jersey City, NJ
    10 days ago
  •  ...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing...  ...service health reporting across on-premises and cloud platforms. Develops reusable IaC modules,... 
    Work at office
    Shift work
    Day shift

    Bank of America Corporation

    Jersey City, NJ
    24 days ago
  •  ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Consumer and Investment...  ...with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize... 
    Local area

    J.P. Morgan

    Jersey City, NJ
    7 hours ago
  •  ...Site Reliability Engineer As a Site Reliability Engineer, your role is to provide reliability engineering services through observability...  ...incidents. This role requires a strong background in scripting, cloud platforms, and a passion for optimizing operational... 
    Work experience placement

    Omni Inclusive

    Secaucus, NJ
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Cloud Reliability Engineer. Be the first to apply!