Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

SRE Engineer

AceStack LLC

Role : SRE Engineer

Location : San Jose, CA

(ONSITE)

FULL TIME ONLY

Job Description


Must Have Technical/Functional Skills:
• Exp. in Apache SPARK development, Kubenetes, CI-CD Pipeline, Jenkins, Dockers, Kubernetes, PL SQL, Python
• Writing SQL queries and procedures
• Writing Python code to automate and develop small functionalities
• Creating CI/CD pipelines
• Writing Jenkins jobs
• Managing applications in Kubernetes environments, including deployment, configuration, and triaging
• Hands-on experience with Apache Spark


Roles & Responsibilities:


The candidate will provide technical leadership for the team(s) they are associated with and participate in key technical decisions. They will engage with customers on escalations and ensure that there is continuous improvement in all areas. Participate in technical discussions within the team and with other groups within Business Units associated with specified projects
• You design, develop, and maintain our real time data processing, data Lakehouse infrastructure.
• You have experience with Python writing data pipelines and data processing layers.
• You develop and maintain Ansible playbooks for infrastructure configuration and management
• You develop and maintain Kubernetes manifests, Helm charts, and other deployment artifacts
• You have hands-on experience on Docker and containerization and how to manage/prune the images in private registries.
• You have hands-on experience on access control in K8S cluster
• You have hands-on experience on SPARK and maintaining SPARK CLUSTER
• You monitor and troubleshoot issues related to Kubernetes clusters and containerized applications
• You drive initiatives to containerize standalone apps to be containerized i n Kubernetes.
• You develop and maintain infrastructure as code (IaC) and collaborate with other teams to ensure consistent infrastructure management across the organization
• You use observability tools to do "capacity management" of our services and infrastructure resources.
• You are for guiding the development and testing activities of other engineers that involve several inter-dependencies
• Experience in AWS ECS and EKS is added advantage
• Experience in Dremio is added advantage
• Experience in Dynatrace or any tracing, infrastructure, or real time monitoring tool is added advantage
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the SRE Engineer in San Jose, CA vacancy
  •  ...Position- SRE Engineer Duration-Contract Location- San Jose, C JD Roles & Responsibilities • Extensive experience working with linux flavors like rhel/centos os, shells, filesystems and utilities • Knowledge of distributed computing and experience... 
    Suggested
    Contract work
    Immediate start

    Syntricate Technologies

    San Jose, CA
    4 days ago
  •  ...ability to multi-task in fast-paced environments. Global Collaboration: Comfortable working cross-functionally with product/engineering units across multiple time zones. Documentation: High care in creating detailed design specifications and presenting... 
    Suggested

    VBeyond

    Santa Clara, CA
    4 days ago
  •  ...SRE Engineer St Louis, MO (Onsite from day 1) Client Required Skills: • Bachelor's Degree in Computer Science, Computer Systems, Information Technology or related. Equivalent experience is acceptable. • Experience with web applications and distributed systems... 
    Suggested

    Omega Solutions

    Santa Clara, CA
    4 days ago
  •  ...Enterprise Technologies Inc. is a recognized provider of professional IT Consulting services in the US. We are actively seeking SRE Devops Engineer Fulltime Role for one of our direct client. Role: SRE Devops Engineer Location :- Santa Clara,CA (Remote... 
    Suggested
    Full time
    Local area
    Remote work

    Rootshell Enterprise Technologies

    Santa Clara, CA
    4 days ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before...  ...EngineerAbout the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability... 
    Suggested
    Night shift

    Forward Networks

    Santa Clara, CA
    22 hours ago
  • $152k - $241.5k

     ...next wave of artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and help build the next generation...  ...programming languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,... 
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature...  ...observability and security products. Work with your global SRE team to optimize operations, increase efficiency in our use of cloud... 
    Flexible hours

    Sumo Logic

    San Jose, CA
    3 days ago
  •  ...our San Francisco/San Jose/Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering. Our scope includes the Lambda website, cloud APIs and systems... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $148k - $235.75k

     ...make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw...  ...experience) and 5+ years operating production distributed systems as SRE/DevOps/Platform Ops.Proven ownership of reliability for an... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting...  ...365.Who You AreExperienced Architect: 5+ years of experience in SRE, DevOps, or Systems Engineering, with a proven track record of... 
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    22 hours ago
  • $101k - $161k

     ...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-Life...  ...our growing Arista’s CloudVision-as-a-Service (CVaaS) global SRE team. SREs at Arista combine strong software engineering background... 

    Arista Networks

    Santa Clara, CA
    3 days ago
  • $120k - $180k

     ...and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.Sr. SRE & DevOps EngineerAbout the Role:At CrowdStrike, our engineering organization depends on shared infrastructure platforms that power critical product capabilities at... 
    Full time
    Work experience placement
    Work at office
    Local area

    CrowdStrike

    Sunnyvale, CA
    4 days ago
  •  ...Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots. FII provides customers with... 
    Permanent employment
    Full time
    Work at office
    Local area

    Foxconn Technology Group

    San Jose, CA
    4 days ago
  •  ...Site Reliability Engineer Location – San Jose, CA What You'll Do - Responsibilities Engage in and improve the whole lifecycle of services—from inception and design, through automated deployment, operation and refinement. Work with all relative teams to... 

    Netpace

    San Jose, CA
    4 days ago
  • $187.04k - $359.72k

     ...changes that improve reliability and velocity. Qualifications Minimum Qualifications: BS or MS degree in Computer Science, Electrical Engineering, Computer Engineering or related areas. Experience in one or more programming languages such as Go, Java, C++, Python etc. Good... 
    Temporary work
    Local area
    Overseas
    Shift work

    Tik Tok

    San Jose, CA
    22 hours ago
  • $185k - $227k

     ...opportunity to build your career is compelling, read on for more details. ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance ofJuul’s hybrid cloud infrastructure (Nutanix, AWS/GCP). This involves... 
    Remote work

    GrabJobs

    San Jose, CA
    1 day ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 

    TechDigital Group

    Santa Clara, CA
    22 hours ago
  •  ...Job Title : Site Reliability Engineer Location: San Jose, CA Duration: Contract Job Description: Extensive experience working with linux flavors like rhel/centos os, shells, filesystems and utilities Knowledge of distributed computing... 
    Contract work
    Immediate start

    Syntricate Technologies

    San Jose, CA
    3 days ago
  • $210.6k - $305.1k

     ....S. citizen on U.S. soil. Lead, inspire, and develop a talented SRE team, fostering a culture of innovation, collaboration, and excellence...  ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 days ago
  • $207.4k - $259.2k

     ...and supports and celebrates all of our team members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability, performance, and... 
    Permanent employment
    Local area

    Archer Aviation

    San Jose, CA
    4 days ago
  • $207k - $301k

     ...that improve reliability for multiple teams.Engage in software engineering on services written in Java, C++, and Go (including instrumenting...  ...of ownership and motivation.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large... 

    Google

    San Jose, CA
    1 day ago
  • $186.9k - $267.7k

     ...customers to expertly deploy and manage AI-powered applications with enhanced observability and control.As a Staff Site Reliability Engineer (SRE), you will provide technical leadership for the reliability, scalability, and operational architecture of Splunk Agent... 
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Milpitas, CA
    18 hours ago
  • $122.5k - $175k

     ...of cybersecurity.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose,...  ...in the Cloud Infrastructure & Operations department. You are an SRE with proven experience in Linux/UNIX System Administration and... 
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    1 day ago
  • $262k - $365k

     ...analyzing, and troubleshooting distributed systems.Preferred qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant... 

    Google

    San Jose, CA
    3 days ago
  • $184k - $287.5k

    At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency...  ...operation with consistent reliability and uptime. As an SRE here, you will be part of a welcoming team that values... 
    Full time

    Nvidia

    Santa Clara, CA
    22 hours ago
  •  ...compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base grow.You... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  •  ...Position: Site Reliability Engineering (SRE) Location: Santa Clara, CA (Onsite) Duration: W2 / C2C Contract Experience: 10+ Years Job Description: • WS application and CI/CD pipelines, Microsoft Server admin and workload support (Data Center and AWS... 
    Contract work
    Immediate start

    Syntricate Technologies

    Santa Clara, CA
    4 days ago
  • Coding experience in one or more of Python or Java. Experienced with automating infrastructure with scripting (Shell Script, Python) and tooling (Puppet, Terraform, Ansible, Chef, etc). Experienced with Splunk for investigating or monitoring problems on systems...
    Flexible hours

    VBeyond

    Sunnyvale, CA
    1 day ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure...  ...real problems and fix them?What you'll be doing:Serve as an SRE Architect part of GPU Private Cloud team used by thousands of... 
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    2 days ago
  • $163.43k - $213.97k

     ...Location: Santa Clara, CA Travel: Up to 25% Job ID: 1739 The Role: We are seeking a Staff Site Reliability Engineer. As Staff SRE Engineer, you set the technical direction for reliability across regions and services. You own the reliability strategy,... 
    Permanent employment
    Contract work
    Work at office

    IonQ Inc.

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to SRE Engineer. Be the first to apply!