SRE
Keylent Inc
SRE MAHIN-JOB-31492
Location: Plano TX
Skill: Web application designing basics-1
8+ years of professional experience
processing a culture of learning through the
development and sharing of skills,
knowledge, process and tools
2. A driving passion for finding solutions to
hard problems at scale and operationalizing
them
3. Exceptional critical thinking and
communication skills, with a passion for
leveraging documentation as a tool for
constant improvement
4. Experienced with designing, building, and
optimizing automated pipelines with
automated testing and automated security
controls
5. Experience in performing Root Cause
Analysis and Problem Management
6. Experience with working in Agile Scrum
teams with demonstrated success leading
improvements
Understand the system and application architecture
2. Guide the architecture and development teams on how
to make applications highly available, reliable, and
performant at global scale
3. Partner with architecture team to ensure operability,
measurability, and manageability are accounted for in
business features and enablers
4. Collaborate with product owners and managers to
establish service level objectives for applications and
agreed consequences if the objectives are not being
met
5. Collaborate with development team members to
swarm, troubleshoot, and resolve problems
6. Drive the Root Cause Analysis of production issues
and other failures within the application and system
7. Design, build, and champion automated solutions to
optimize application/service/platform uptime with
minimal human intervention
8. Create and implement standards and best practices,
driving adoption across development teams and
external vendors as applicable
Location: Plano TX
Skill: Web application designing basics-1
8+ years of professional experience
processing a culture of learning through the
development and sharing of skills,
knowledge, process and tools
2. A driving passion for finding solutions to
hard problems at scale and operationalizing
them
3. Exceptional critical thinking and
communication skills, with a passion for
leveraging documentation as a tool for
constant improvement
4. Experienced with designing, building, and
optimizing automated pipelines with
automated testing and automated security
controls
5. Experience in performing Root Cause
Analysis and Problem Management
6. Experience with working in Agile Scrum
teams with demonstrated success leading
improvements
Understand the system and application architecture
2. Guide the architecture and development teams on how
to make applications highly available, reliable, and
performant at global scale
3. Partner with architecture team to ensure operability,
measurability, and manageability are accounted for in
business features and enablers
4. Collaborate with product owners and managers to
establish service level objectives for applications and
agreed consequences if the objectives are not being
met
5. Collaborate with development team members to
swarm, troubleshoot, and resolve problems
6. Drive the Root Cause Analysis of production issues
and other failures within the application and system
7. Design, build, and champion automated solutions to
optimize application/service/platform uptime with
minimal human intervention
8. Create and implement standards and best practices,
driving adoption across development teams and
external vendors as applicable
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the SRE in Plano, TX vacancy
- Job Title SRE Lead Location Plano, TX - Onsite Job Description We are seeking an experienced 13 to 18 years of experience to join our team. The ideal candidate will have expertise in AWS, SRE, and Datadog, and a background in the automotive industry is a plus. This hybrid...SuggestedDay shift
- We are seeking a Delivery SRE leader who will ensure security applications are delivered with strong SDLC discipline and measurable reliability. This role partners closely with Product Owners and engineering leadership to challenge assumptions, sharpen the Definition of...Suggested
- JPMorgan Chase & Co. seeks a Delivery SRE leader to ensure security applications are delivered with strong SDLC discipline and measurable reliability. This role partners closely with Product Owners and engineering leadership to challenge assumptions, sharpen the Definition...Suggested
- JPMorganChase & Co. is seeking a Delivery SRE leader to ensure security applications are delivered with strong SDLC discipline and measurable reliability. The role partners with Product Owners and engineering leadership to bake SRE requirements into design and build phases...Suggested
- ...SRE Production Support Engineer Location: Plano, TX Duration: 6 Months (Contract to hire) Interview Process: 1st round - Zoom 2nd round – In Person Role Overview: Position is part of the Central Site Reliability Engineering (SRE) Team. Looking for...SuggestedContract workShift work
$87.5k - $125k
Observability Engineer DISH is transforming the future of connectivity. We're doing it by building the country's first virtualized, standalone 5G wireless network from scratch. The foundation of a connected world, it's a network free of the limitations of the past, ...Flexible hoursNight shift- ...Role : SRE Engineer Location : Plano, Texas Job Summary: We are looking for a highly motivated Site Reliability Engineer (SRE) to improve system reliability, scalability, and performance of mission-critical applications. The ideal candidate should have strong...
$229.9k - $262.4k
...Overview Full-Stack Engineer 5 (Java, Python, SRE, AWS, AI) (Cloud Operations Resilience Engineering) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive and iterative...Full timePart timeInternshipLocal area- ...workflows Experience with cloud-native technologies in AWS, Azure, or GCP Experience with CI/CD, Infrastructure as Code (IaC), and/or SRE operations What we'll bring During your interview process, our team can fill you in on all the details of our industry-leading...H1bRelocation package
- ...environment. Operational Frameworks: Advanced understanding and application of ITIL v4 principles, FinOps, and Site Reliability Engineering (SRE) concepts. Automation/IaC: Strong practical experience with: IaC: Terraform, GCP instance Templates, Azure ARM/Bicep, CloudFormation....
- ...Architectures Cloud Engineering & Infrastructure as Code (IaC) DevOps, CI/CD & Engineering Productivity Site Reliability Engineering (SRE) & Operational Excellence Monitoring, Observability & Platform Reliability Security, Governance & Engineering Controls The Senior...Contract workWork experience placement2 days per week
- ...practicesCreating of APIs and Dashboards to determine and report data on environment health and application quality.Partnership with SRE team to enable and improve automation with tools and processes in non-prod environments.Interact and communicate with technical and non...Contract workWork experience placementRemote work3 days per week
- ...Cloud-Native Architecture & Platform Modernization DevOps, CI/CD & Engineering Productivity Platforms Site Reliability Engineering (SRE) & Operational Excellence Monitoring, Observability & Platform Intelligence Security, Governance, Compliance & Engineering Controls LLM...Contract workTemporary workWork experience placement2 days per week1 day per week
$229.9k - $262.4k
...critical incident response efforts and ensure root-cause analysis drives durable reliability improvements Partner with Cloud, Security, and SRE teams to define cross-platform observability, telemetry, and self-healing automation Mentor and upskill engineers across the...Full timePart timeH1bLocal area- ...limits, N+1 mitigation (DataLoader). Proven delivery of API/schema governance, versioning/deprecation, and CI policy gates. Strong SRE practices: SLIs/SLOs, error budgets, OpenTelemetry, data-driven post-incident improvements. Developer productivity: time-to-first-hello...Contract work
$8,896.99 per month
...GitOps pipelines using Cloud Build/GitHub Actions/Artifact Registry and integrate IaC with policy-as-code. Establish observability and SRE practices: Cloud Monitoring, Logging, Trace, Error Reporting, SLOs/SLIs, incident runbooks, and game days. Define cost governance and...Work experience placement- ...Architecture Cloud Engineering & Cloud-Native Architecture Infrastructure as Code (IaC) DevOps & CI/CD Site Reliability Engineering (SRE) Monitoring & Observability Platform Reliability & Operational Excellence Developer Experience (DevEx) Self-Service Platforms APIs, SDKs...Contract workTemporary work2 days per week1 day per week
- ...Own alerting strategy and on-call runbooks; participate in incident response and post-incident reviews Partner with engineering and SRE teams on reliability improvements, deployment safety, and change management practices Qualifications Required 5+ years in DevOps, platform...Full timeWork at officeRemote work3 days per week
- ...CloudWatch - ServiceNow - Kafka - API design/integration and troubleshooting - Incident, problem, and change management (ITIL-aligned) - SRE/operations metrics (availability, reliability, MTTR, SLA/SLO reporting) - Root cause analysis and post-incident governance - SQL and...
- ...recommendations before use, escalating when uncertain and following data handling expectations. Must have a background in development or SRE Advanced expertise in stakeholder management, with the ability to establish productive working relationships and influence decision-...
$140.1k - $234.85k
...and new operations outcomes. This high velocity Digital Transformation necessitates Effective, Modern & Resilient operations in an SRE construct for all the programs under TS& EP Portfolio, per the main purpose to drive higher order outcomes to our customers who use our...Shift work- ...Splunk certifications (Power User, Admin, Architect). Experience supporting collaboration platforms at an enterprise scale. Background in SRE, NOC, or production support environments. Experience building executive dashboards and operational scorecards. Everforth Apex is a...Contract workNight shift
- ...and operational stability Proficient in coding in Java, GoLang, or equivalent modern programming languages. Experience with DevOps and SRE principles. Strong experience in Private or Public cloud Good understanding of Cloud Infrastructure and API Demonstrated experience...
- ...deployment patterns, and lead incident response to improve reliability. The role emphasizes collaboration with Cloud, Security, and SRE teams, plus mentoring engineers in networking, reliability, and automation practices. Strong communication skills are essential. #J-...
- ...multiple technical areas within various business functions in support of the firm's business objectives. Vice President Engineering/SRE leadership role on the Platform Identity Engineering team, building and running the always-on identity systems the whole firm relies...Work at officeImmediate start
- ...drive autoscaling strategy and right-sizing. Qualifications Senior experience in performance engineering for distributed systems and/or SRE-style reliability engineering in production. Strong cloud/container background (AWS + Kubernetes/EKS; ECS exposure beneficial)....
- ...(schema evolution, backfills, error handling, contract-driven interoperability). Establish production-grade operations and controls (SRE practices, monitoring/on-call, incident response/RCA, auditability, least-privilege, disciplined change management) and deliver governed...Contract work
- ...and familiarity with Infrastructure as Code (IaC) tools like Terraform or CloudFormation. 5+ years in Site Reliability Engineering (SRE), Disaster Recovery Planning, or Distributed Systems Engineering. Demonstrated experience leading effective use of approved AI-assisted...
- ...Description Job Description Responsibilities: \t3-4 years of experience in production engineering and site reliability engineering (SRE) to design, implement, and maintain highly available, scalable, and resilient systems. \tOwn end-to-end operational...
- ...and networking. Design highly available, scalable, and disaster-resilient solutions. Troubleshoot production issues and apply SRE/reliability practices. Participate in code reviews, technical design discussions, and engineering best practices. Provide technical...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!


