SRE
InterSources Inc
SRE Production Support Engineer
Location: Plano, TX
Duration: 6 Months (Contract to hire)
Interview Process:
1st round - Zoom
2nd round – In Person
Role Overview:
- Position is part of the Central Site Reliability Engineering (SRE) Team.
- Looking for a strong individual contributor, not someone who simply follows instructions.
- Candidate should have a Software Engineering background, with a strong focus on production environments and application reliability.
- This is not a traditional support or operations role.
Required Skills:
- Solid understanding of Site Reliability Engineering (SRE) principles.
- Experience managing and maintaining production applications ("care and feeding" of applications).
- Ability to review and manage production backlogs.
- Production-focused mindset with experience in reliability, stability, and operational excellence.
- Understanding of how applications are deployed, monitored, and maintained.
- Strong proficiency in Java coding and scripting – Python, Shell Scripting
- Good understanding of DevOps concepts and practices.
- Familiarity with .NET applications is beneficial.
- Experience with any major cloud platform is acceptable, including:
- AWS
- Azure
- Google Cloud Platform (GCP)
- IBM Cloud
- Dell Technologies Cloud
- The specific cloud platform is less important than understanding how applications are deployed, managed, and operated in a cloud environment.
Role Expectations:
- This is not a shift-based position.
- The role does not involve continuously monitoring dashboards or providing traditional production support.
- The focus is on improving application reliability through engineering and automation.
- Candidate does not need to know the exact source code of every application but should understand how applications function in production and how to improve their reliability.
Job Description:
- Expert in at least one technology and design technique as well as experience working across large environments with multiple operating systems/infrastructure for large-scale programs (e.g., Expert Engineers) starting to be firm-wide resources working on projects across Client
- Is multi-skilled with expertise across software development lifecycle and toolset
- May be recognized as a leader in Agile and cultivating teams working in Agile frameworks
- Sought out as coach for at least one technical skill
- Strong understanding of techniques such as Continuous Integration, Continuous Delivery, Test Driven Development, Cloud Development, resiliency, security
- Stays abreast of cutting-edge technologies/trends and uses experience to influence application of those technologies/trends to support the business; may give speeches and outside the firm, writes articles
Roles and Responsibilities:
- Executes standard software solutions, design, development, and technical troubleshooting
- Writes secure and high-quality code using the syntax of at least one programming language with limited guidance
- Designs, develops, codes, and troubleshoots with consideration of upstream and downstream systems and technical implications
- Applies knowledge of tools within the Software Development Life Cycle toolchain to improve the value realized by automation
- Applies technical troubleshooting to break down solutions and solve technical problems of basic complexity
- Gathers, analyzes, and draws conclusions from large, diverse data sets to identify problems and contribute to decision-making in service of secure, stable application development
- Learns and applies system processes, methodologies, and skills for the development of secure, stable code and systems
- Adds to team culture of diversity, equity, inclusion, and respect
Additional Skills:
- Formal training or certification in software engineering /Site Reliability Engineering concepts and 6 plus years of applied experience.
- Hands-on practical experience in system design, application development, testing, and operational stability
- Exposure to product engineering or production/Platform support activities with a good understanding on scalability, security, and reliability.
- Experience in developing, debugging, and maintaining code in a large corporate environment with one or more modern programming languages and database querying languages
- Demonstrable ability to code in one or more languages
- Experience across the whole Software Development Life Cycle
- Exposure to agile methodologies such as CI/CD, Application Resiliency, and Security
- Emerging knowledge of software applications and technical processes within a technical discipline (e.g., cloud, artificial intelligence, machine learning, mobile, etc.)
About Us:
InterSources Inc, a Certified Diverse Supplier, was founded in 2007 and offers innovative solutions to help clients with Digital Transformations across various domains and industries. Our history spans over 16 years and today we are an Award-Winning Global Software Consultancy solving complex problems with technology. We recognize that our employees and our clients are our strengths as the diverse talents and opportunities they bring to the table enable us to grow as a global platform and they are causally linked with our success. We provide strategic and technical advice, and we have expertise in areas covering
Artificial Intelligence, Cloud Migration, Custom Software Development, Data Analytics Infrastructure & Cloud Solutions, Cyber Security Services, etc. We make reasonable accommodations for clients and employees and we do not discriminate based on any protected attribute including race, religion, color, national origin, gender sexual orientation, gender identity, age, or marital status. We also are a Google Cloud partner company. We align strategy with execution and provide secure service solutions by developing and using the latest technologies that thrive our resources to deliver industry-leading capabilities to our clients and customers, making it convenient for our clients to do business with InterSources Inc. Our teams also drive growth by refining technology-driven client experiences that put the users first, providing an unparalleled experience. This results in strengthening the core technologies of clients, enabling them to scale with flexibility, create seamless digital experiences and build lifelong relationships.
- JPMorgan Chase & Co. is recruiting a Lead Site Reliability Engineer to define the future of reliability on the AI ML and Data platform. You will lead a team, drive resiliency reviews, and mentor engineers while advocating for secure, observable, scalable systems across...Suggested
- ...collaborating with J.P. Morgan to connect them with exceptional professionals for this role. JOB DESCRIPTION We are seeking a Delivery SRE leader who will ensure security applications are delivered with strong SDLC discipline and measurable reliability. This role partners...Suggested
- JPMorgan Chase & Co. is seeking a Lead Software Engineer within the Consumer and Community Banking technology -Deposits Platform to lead resiliency and DR efforts across the product portfolio. The role involves cross-team collaboration, development of failover frameworks...Suggested
- ...Role description We are seeking an experienced Site Reliability Engineer SRE Lead to drive platform reliability observability and operational excellence across the API Services ecosystem Description Role description We are seeking an experienced Site Reliability Engineer...SuggestedTemporary workLocal area
- ...SRE MAHIN-JOB-31492 Location: Plano TX Skill: Web application designing basics-1 8+ years of professional experience processing a culture of learning through the development and sharing of skills, knowledge, process and tools 2. A driving passion for...Suggested
$229.9k - $262.4k
...Overview Full-Stack Engineer 5 (Java, Python, SRE, AWS, AI) (Cloud Operations Resilience Engineering) Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative, inclusive and iterative...Full timePart timeInternshipLocal area- ...Role : SRE Engineer Location : Plano, Texas Job Summary: We are looking for a highly motivated Site Reliability Engineer (SRE) to improve system reliability, scalability, and performance of mission-critical applications. The ideal candidate should have strong...
$87.5k - $125k
Observability Engineer DISH is transforming the future of connectivity. We're doing it by building the country's first virtualized, standalone 5G wireless network from scratch. The foundation of a connected world, it's a network free of the limitations of the past, ...Flexible hoursNight shift- ...integration, and end-to-end tests for control plane components; participate in code reviews Collaborate with infrastructure, platform, and SRE teams to define scheduling policies, resource quotas, and placement constraints Document architecture decisions, APIs, and...
$147.25k - $190k
...storage, load balancing, DNS, certificates/TLS), managing scope, dependencies, risk, and status. Partner with security, IAM, network, SRE/operations, and vendors to architect and implement scalable, resilient solutions and platform modernization. Automate and...Full timeWork at office$229.9k - $262.4k
...incident response efforts and ensure root‑cause analysis drives durable reliability improvements Partner with Cloud, Security, and SRE teams to define cross‑platform observability, telemetry, and self‑healing automation Mentor and upskill engineers across the...Full timePart timeH1bLocal area- ...Cloud-Native Architecture & Platform Modernization* DevOps, CI/CD & Engineering Productivity Platforms* Site Reliability Engineering (SRE) & Operational Excellence* Security, Governance, Compliance & Engineering Controls* LLM Workflows, Prompt Engineering & Context...Hourly payContract workWork experience placement
$54.68 - $64.68 per hour
...OpenSSH,Oracle Enterprise Linux,performance tuning,performance optimization,postfix,Python,system reliability,scripting,sendmail,SMTP,SRE Practices,vulnerability remediation,high availability,TCP/IP,analytical,communication,organizational skills,Leadership,Troubleshoot,...Hourly payContract workTemporary workWork experience placement- ...limits, N+1 mitigation (DataLoader). Proven delivery of API/schema governance, versioning/deprecation, and CI policy gates. Strong SRE practices: SLIs/SLOs, error budgets, OpenTelemetry, data-driven post-incident improvements. Developer productivity: time-to-first-hello...Contract work
- ...reliability, and scalability; follow Agile practices such as Scrum and Continuous Delivery. Support Site Reliability Engineering (SRE) practices to ensure excellent user experience and system performance. Required qualifications, capabilities, and skills: Formal...
- ...ServiceNow - Kafka - API design/integration and troubleshooting - Incident, problem, and change management (ITIL-aligned) - SRE/operations metrics (availability, reliability, MTTR, SLA/SLO reporting) - Root cause analysis and post-incident governance -...
- ...recommendations before use, escalating when uncertain and following data handling expectations. Must have a background in development or SRE Advanced expertise in stakeholder management, with the ability to establish productive working relationships and influence decision-...
- ...Toyota Financial Services Technology Operations Center is looking for a passionate and highly motivated Senior Site Reliability Engineer (SRE) - Backup Infrastructure. In this role, you will apply software engineering principles to ensure the reliability, availability, and...H1bRelocation package
- ...cloud, and colocation environments. You will own end‑to‑end connectivity, security, and automation, partnering with Cloud, Security, and SRE teams to improve reliability and performance. The role requires 8+ years in network engineering, hands‑on experience with BGP, L2/...
- ...MongoDB, and Redis. Experience building and maintaining event streaming solutions using Kafka or RabbitMQ. Well-versed in DevOps and SRE practices, including CI/CD with Jenkins and observability using Splunk and Grafana. Hands-on experience using enterprise-...
- ...and reliable technology environment across banking operations. The ideal candidate brings 10+ years of IT operations leadership in regulated financial services, expertise in DevOps/SRE, and deep experience with hybrid cloud (Azure/AWS). #J-18808-Ljbffr Jobleads-US
- ...latency, scalability, and disaster recovery requirements. ~ Monitor system performance, troubleshoot production issues, and apply SRE and reliability engineering practices. ~ Identify opportunities for performance optimization and cloud cost efficiency. ~ Participate...Local area
- ...scalability, and security for mission-critical workloads. Champion the shift toward next-generation operating models. Drive the adoption of SRE, DevOps, and intelligent automation to reduce toil, increase speed-to-market, and optimize unit economics. Move beyond basic SLA...Local areaFlexible hoursShift work
- ...Toyota Financial Services is seeking a Senior Site Reliability Engineer (SRE) focused on Backup Infrastructure in Plano, TX. You will ensure reliability, availability, and performance of enterprise backup ecosystems, and automate backup workflows for efficiency. You...
- ...incident management, disaster recovery, and business continuity ~ Drive automation, observability, and platform stability through DevOps/SRE principles ~ Ensure robust cybersecurity posture in alignment with FFIEC and NIST frameworks ~ Manage vendor relationships and...Full timeLocal areaRelocation
- ...Description Job Description Responsibilities: \t3-4 years of experience in production engineering and site reliability engineering (SRE) to design, implement, and maintain highly available, scalable, and resilient systems. \tOwn end-to-end operational...
- ...to $66.00/hr. w2 Responsibilities: Build and maintain CI/CD pipelines using Jenkins. Implement and refine observability for SRE, including metrics, logs, and traces. Participate in on-call rotations and ensure service readiness. Drive incident management...Hourly payFor contractorsLocal area
- ...and familiarity with Infrastructure as Code (IaC) tools like Terraform or CloudFormation. ~5+ years in Site Reliability Engineering (SRE), Disaster Recovery Planning, or Distributed Systems Engineering. ~ Demonstrated experience leading effective use of approved AI-...
- ...assisted practices and ensuring security and auditability. You will manage cross-functional workstreams, collaborate with security, IAM, and SRE teams, and deliver scalable, resilient infrastructure solutions with IaC and robust runbooks. #J-18808-Ljbffr Jobleads-US
- ...and networking. Design highly available, scalable, and disaster-resilient solutions. Troubleshoot production issues and apply SRE/reliability practices. Participate in code reviews, technical design discussions, and engineering best practices. Provide technical...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!


