Site Reliability Engineer (Azure + Data)
VDart Inc
Job Title : Site Reliability Engineer (Azure + Data)
Work Mode : Bellevue, WA - Hybrid
Job type ; Contract
Job Description :
6 years of experience in infrastructure support site reliability engineering cloud operations or platform engineering including strong hands-on ownership of production Azure environments
Demonstrated expertise in reliability performance capacity cost optimization and high severity incident leadership
Deep knowledge of incident problem change security privacy compliance audit and governance processes
Strong experience with Azure access control certificates secrets monitoring logging CICD Infrastructure as Code and scripting
Solid understanding of data platform architecture and dependencies across ADF ADLS Gen2 Synapse Cosmos DB Azure Data Explorer SQL Server and Microsoft Fabric
Strong production ownership customer facing communication and the ability to drive operational discipline across onsite and offshore teams
Key Responsibilities
Infrastructure reliability and optimization Own the availability performance capacity and cost of production Azure infrastructure supporting high volume data platforms target availability SLA adherence and QoS while proactively addressing bottlenecks saturation and scaling risks
Azure architecture security and governance Design and operate landing zones subscriptions resource groups VNets peering ExpressRoute Azure Firewall Bastion DDoS protection Azure Policy Entra ID RBAC Managed Identities PIM and Conditional Access
Incident problem and change management Lead triage mitigation stakeholder communication root cause analysis corrective actions risk assessment approvals validation and rollback planning improve MTTR and prevent repeat incidents
Compliance and operational readiness Maintain S360 security privacy audit and governance compliance sustain accurate runbooks SOPs CENs and operational playbooks
Certificates secrets and dependencies Manage the lifecycle of certificates keys secrets identities and service dependencies track expirations automate renewals and secure service to service communication
Data platform infrastructure Support and optimize Azure Data Factory ADLS Gen2 Synapse Analytics Cosmos DB Azure Data Explorer SQL Server and Microsoft Fabric troubleshoot throughput and dependency issues and guide platform modernization and Fabric migration
Monitoring and observability Use Azure Monitor Log Analytics Application Insights and KQL platform metrics ing and cost dashboards to detect risks analyze trends and drive evidence based decisions
Automation and engineering practices Standardize infrastructure and operational workflows through Azure DevOps GitHub Actions ARM Bicep Terraform YAML PowerShell Azure CLI Python and Power Automate apply SRE practices including error budgets automated recovery and self-healing
Cost and performance management Right size resources optimize compute storage networking SQL and Cosmos capacity and implement budgets and reservation planning without compromising reliability
Customer and team collaboration Serve as the infrastructure SRE contact for Redmond customers communicate risks and optimization opportunities clearly and coordinate consistent execution across onsite offshore infrastructure SRE and data engineering teams
Skills
Mandatory Skills : Azure Infra Services, Azure Monitor, Azure Storage
Good to Have Skills : Azure Data Factory
$204k - $306k
...mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco,... ...Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and... ...and/or have access to protected federal data. As a condition of employment for this...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ...Engineering: Optimize the collection, processing, and storage of log data to ensure high reliability and low latency of our Splunk...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours$194k - $267k
...self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing... ...federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours- ...This is an engineering-first Senior SRE role. We’re looking for senior engineers who have... ...production (design → launch → on-call → reliability improvements) Led incident response and... ...infrastructure (AWS preferred, GCP/Azure acceptable). Strong Signal You’ve...Suggested
- ...Job Title Technical/Functional Skills: Windows Servers, Digital: Microsoft Azure Windows Powershell, Digital: DevOps Roles & Responsibilities: Windows Server 2012 -2019 Administration Microsoft Azure Azure AAD DFSR, DHCP DNS, KMS, WSUS TCP/IP Hyper...Suggested
$232k - $319k
...scale the service with great people and reliable, cost-effective, and efficient infrastructure... ...the velocity of SRE and product engineering by developing robust platforms, powerful... ...and/or have access to protected federal data. As a condition of employment for this position...Permanent employmentLocal areaWorldwideFlexible hours- The Data Infrastructure SRE team is responsible for the reliability, scalability, and efficiency of the core data services that... ...building features, but about engineering the resilience and performance... ...maintain system stability.As a Site Reliability Engineer, you will...
- The Data Infrastructure SRE team is responsible for the reliability, scalability, and efficiency of the core data services that power our products. We manage a... ...work is not about building features, but about engineering the resilience and performance of the underlying...
$133.2k - $219.6k
...development through technological innovation and engineering practices. Our team focuses on the... ...high performance, and enterprise-level reliability. By doing so, we aim to empower numerous... ...a passionate and technically skilled Site Reliability Engineer (SRE) to join the our...Full time- ...Data EngineerSkills Required:Cosmos SQL experience Azure Power BIKnowledge of Azure Data Factory ADLSIndividuals with strong technical background in SQL SQL Server data engineering BI Kusto Data Analysisand strong knowledge of Data Warehousing ConceptsCustomer facing...
- ...Azure Data EngineerWe are seeking a highly skilled professional to join our team as an Azure Data Engineer. This role involves designing, developing, and maintaining scalable data pipelines using Azure Data Factory and related services. The ideal candidate will have a...Work experience placement
$165k - $230k
...goal of enabling human life on Mars.SR. KUBERNETES PLATFORM SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in... ...improvement techniquesUnderstanding of distributed databases and data modelingExperience with automatically managing dozens,...Permanent employmentTemporary workWork at officeWorldwideMonday to FridayWeekend work- ...Platform SRE team supports all Big Data services and products across... .... We are responsible for the reliability of all the company's major... ...products, services, and query engines. We serve business needs across... ...emerging technologies related to site reliability and infrastructure...
- ...day. We're hiring a senior, hands-on engineer to own the reliability, availability, security, and performance... ...of hands-on Cloud Operations and Site Reliability Engineering, operating production... ...~ A second cloud (Google Cloud or Azure) is a plus, not a substitute. We...Full time
- ...Senior Site Reliability Engineer (SRE) Location: Seattle, hybrid - 2 times a week in the office Job Type: Full-time, direct hire Industry... ...Experience with multi-cloud environments (AWS, GCP, Azure). Chaos engineering experience (Gremlin, Chaos Mesh, or...Full timeWork at office
$127k - $249k
Platform Engineering is the department within SRE that is responsible... ...components that ensure cluster reliability and security (e.g., CoreDNS,... ...market. We have redefined the data platform for the AI era,... ...Google Cloud, and Microsoft Azure.With offices worldwide and over...Work at officeLocal areaRemote workWorldwideFlexible hours$165k - $227k
...on this mission. If you are too, let's talk.About the TeamThe Data Platform team is responsible for the foundational data services... ...flexible. We encourage ownership. We expect great things from our engineers and reward them with stimulating new projects, new technologies...Local areaWorldwideFlexible hours- Company DescriptionComtech LLC is a woman-owned small business focused on delivering end-to-end solutions and products. Since 1998, we have successfully serviced enterprises across the public and private sectors, and the Department of Defense. Our services span all aspects...
$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering... ...production accessExperience building platforms on either Azure or AWS (preferably both)Building, operating, and innovating...Work experience placementWork at officeRemote workFlexible hours$151.2k - $204.6k
Would you like to be an engineer who builds the systems that power advertising at scale,... ...advertising queries every day, where latency, reliability, and quality translate directly into... ...Development Engineer, operating as a Site Reliability Engineer, to raise the reliability...Flexible hours$143k - $194k
...operating system that turns thousands of data streams into a realtime, 3D command and... ...to our customers. System Deployment Engineers work in complex environments with shared... ...demonstrations and exercisesWork with site reliability engineers to provide and refine requirements...Full timeTemporary workWork experience placementImmediate start$165k - $270k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building... ...systems to become sharded and geo-redundant in multiple data centers Advance existing deployment, monitoring, and alerting...Permanent employmentTemporary workWorldwideWeekend work$165k - $230k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging... ...distributed systems to become sharded and geo-redundant in multiple data centers Advance existing deployment, monitoring, and...Permanent employmentTemporary workWorldwideWeekend work$112.8k - $153.7k
...your ideas. This is a senior engineering role. You will architect, operate... ...continuously improve cloud data and analytics platforms (e.g.... ...the standards for platform reliability, CI/CD, environment design, security... ....Relevant certifications (Azure, Snowflake, Databricks, etc.)...Full timeContract workLocal areaFlexible hours$168.1k - $227.4k
...keep the cloud running. We support all AWS data centers and all of the servers, storage,... ...team of software, hardware, and network engineers, supply chain specialists, security... ...design or architecture (design patterns, reliability and scaling) of new and existing systems...InternshipFlexible hoursDay shift- ...Job Description: The successful Support Engineer has the drive and intellectual horsepower... ...or .NET, C# with SQL Server Having Azure domain experience is a big asset Bigdata... ...Skills Expertise in Spark ecosystems, Data frame, API, Data set API, RDD, APIs,...Full timeFlexible hoursShift work
$157.7k - $213.8k
...are passionate about enabling data teams to solve the world's toughest... ...their business. Founded by engineers — and customer obsessed — we... ...data.Data Plane Storage: Provide reliable and high performance services... ...backends, e.g., AWS S3, Azure Blob Store.Delta Lake: A storage...Local areaWorldwide$149.8k - $262.2k
...DescriptionIt all started when engineer Fred Luddy wrote code that... ...brings together any AI, any data, and any workflow— helping 85... ...standards.You will contribute to the reliability, scalability, and operability... ...one major hyperscaler (AWS, Azure, GCP), including its core...Permanent employmentWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week$149.8k - $262.2k
...DescriptionIt all started when engineer Fred Luddy wrote code that... ...brings together any AI, any data, and any workflow— helping 85... ...resolution, with a strong focus on reliability, scalability, and operability... ...one major hyperscaler (AWS, Azure, GCP), including its core...Permanent employmentWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week$55k - $151.47k
...Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a... ...performance of networks, servers, and data centers, minimizing downtime and confirming... ...Google Cloud Platform, and Microsoft Azure to optimize system performance-...Full timeH1b
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (Azure + Data). Be the first to apply!
- data engineer machine learning Bellevue, WA
- aws data engineer Bellevue, WA
- sr data engineer Bellevue, WA
- big data cloud engineer Bellevue, WA
- finance data engineer Bellevue, WA
- sr information security engineer Bellevue, WA
- data center engineer Bellevue, WA
- senior data integration developer Bellevue, WA
- senior cloud data engineer Bellevue, WA
- data engineer Bellevue, WA




