NEW JOB OPENING
SENIOR SITE RELIABILITY ENGINEER AND BACKUP ENGINEER
IN CHICAGO, IL, USA!

 

Date Posted: 09/03/2026
Hiring Organization: Rose International
Position Number: 506940
Industry: Financial Services
Job Title: Senior Site Reliability Engineer and Backup Engineer
Job Location: Chicago, IL, USA, 60604
Work Model: Hybrid
Work Model Details: 3 times a week is mandatory
Shift: M-F, 8am - 5pm
Employment Type: Temp to Hire
FT/PT: Full-Time
Estimated Duration (In months): 10
Min Hourly Rate($): 80.00
Max Hourly Rate($): 90.00
Must Have Skills/Attributes: AWS, GitHub, Grafana, IAM, Infrastructure, Linux, PowerShell, Python, Site Reliability Engineering (SRE), Splunk, Terraform
Experience Desired: Enterprise Backup Engineering (10 yrs); Cyber Recovery Architecture (10 yrs); Site Reliability Engineering (SRE) (10 yrs); Automation & Infrastructure as Code (IaC) (10 yrs)
Required Minimum Education: Bachelor’s Degree

 

Job Description
Required Education
• Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent work experience

Required Experience
• 7+ years in Backup Engineering, Infrastructure Engineering, or Site Reliability Engineering
• 5+ years designing enterprise backup solutions
• 3+ years supporting cyber recovery architectures
• Experience implementing SRE principles within enterprise infrastructure environments
• Strong understanding of distributed systems and high availability architectures
• Cohesity
• Dell PowerProtect Data Manager
• Dell Data Domain
• Dell Cyber Recovery
• Rubrik
• Commvault
• Veritas NetBackup
• Veeam
• Air-gapped vaults
• Immutable backups
• Clean Rooms
• Isolated Recovery Environments (IRE)
• Recovery orchestration
• Cyber resilience testing
• Ransomware recovery
• Recovery validation
• Microsoft Azure
• AWS
• Google Cloud Platform
• Cloud-native backup
• Cross-region recovery
• Hybrid cloud resiliency
• VMware
• Hyper-V
• Kubernetes
• OpenShift
• Linux
• Windows Server
• Active Directory
• Enterprise storage platforms
• Ansible
• Terraform
• Python
• PowerShell
• Bash
• GitHub
• GitHub Actions
• CI/CD pipelines
• Dynatrace
• Grafana
• Prometheus
• Splunk
• ELK Stack
• ServiceNow
• Zero Trust architecture
• NIST Cybersecurity Framework
• CIS Controls
• Encryption and key management
• Identity and Access Management (IAM)
• Multi-factor authentication (MFA)
• Secure recovery processes

Preferred Qualifications
• Experience in financial services or another highly regulated industry
• Experience supporting GSIB cyber resiliency programs
• Knowledge of regulatory expectations from agencies such as the Federal Reserve, OCC, or FFIEC
• Experience with chaos engineering and resilience testing
• Familiarity with SRE tooling and reliability metrics
• Experience implementing AI-assisted operations (AIOps) and predictive analytics
• Strong systems thinking and engineering mindset
• Excellent troubleshooting and root cause analysis skills
• Ability to lead cross-functional technical recovery efforts
• Strong communication and executive presentation skills
• Proven ability to influence engineering standards and drive operational excellence
• Commitment to continuous improvement through automation and reliability engineering

Job Description – Project Overview
• Seeking a highly technical Senior Site Reliability Engineer (SRE) with deep expertise in enterprise backup engineering, cyber recovery, and platform resiliency
• Responsible for engineering highly available, secure, and automated recovery capabilities that protect against operational failures, ransomware, and other cyber threats
• Combines traditional SRE principles (automation, observability, reliability engineering, and resilience) with experience designing and operating enterprise backup platforms, immutable storage, air-gapped cyber vaults, isolated recovery environments (IREs), and recovery orchestration
• Partners closely with Infrastructure, Cyber Security, Cloud Engineering, Application Development, and Disaster Recovery teams to ensure critical services remain recoverable, resilient, and continuously validated
• Engineer and maintain highly available, resilient enterprise platforms using SRE principles
• Define and measure Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets for backup and recovery services
• Develop automation to reduce operational toil and improve reliability
• Perform root cause analysis (RCA) and implement permanent corrective actions
• Continuously improve platform reliability, scalability, performance, and recoverability
• Establish proactive monitoring, alerting, and observability for backup and cyber recovery platforms
• Participate in incident response and major incident recovery activities
• Design, implement, and administer enterprise backup and recovery solutions across on-premises, cloud, and SaaS platforms
• Engineer immutable backup architectures that support ransomware resilience
• Design backup strategies for virtual environments, physical servers, databases, Kubernetes/OpenShift, cloud-native workloads, NAS/Object Storage, and enterprise applications
• Optimize backup performance, retention, replication, encryption, and recovery objectives
• Implement policy-based backup automation and lifecycle management
• Ensure compliance with enterprise RPO and RTO requirements
• Design and implement enterprise cyber recovery solutions including air-gapped recovery vaults, clean rooms, Isolated Recovery Environments (IRE), and immutable storage architectures
• Develop secure recovery workflows following cyberattack scenarios
• Engineer automated malware scanning and recovery validation processes
• Design and test recovery orchestration for severe-but-plausible cyber events
• Support recovery point validation and promotion into production recovery environments
• Collaborate with Cyber Security teams on ransomware resilience strategies
• Develop Infrastructure as Code (IaC) and Recovery as Code automation
• Build automated recovery runbooks using Ansible, Terraform, PowerShell, Python, and GitHub Actions
• Automate recovery validation, reporting, and compliance evidence generation
• Eliminate manual recovery processes wherever possible
• Implement monitoring for backup success rates, replication health, recovery readiness, storage utilization, cyber vault health, and infrastructure dependencies
• Build dashboards for executive and operational visibility
• Integrate with enterprise observability platforms (Dynatrace, Grafana, Splunk, Prometheus)
• Plan and execute cyber recovery exercises, clean room validation, air-gap recovery testing, full isolated recovery environment exercises, Bare Metal Recovery (BMR) testing, and Disaster Recovery testing
• Validate application recoverability against defined RTO/RPO objectives
• Produce executive reporting on recovery readiness and testing outcomes


 

Benefits:
For information and details on employment benefits offered with this position, please visit here. Should you have any questions/concerns, please contact our HR Department via our secure website.

California Pay Equity:
For information and details on pay equity laws in California, please visit the State of California Department of Industrial Relations' website here.

Rose International is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, age, sex, sexual orientation, gender (expression or identity), national origin, arrest and conviction records, disability, veteran status or any other characteristic protected by law. Positions located in San Francisco and Los Angeles, California will be administered in accordance with their respective Fair Chance Ordinances.

If you need assistance in completing this application, or during any phase of the application, interview, hiring, or employment process, whether due to a disability or otherwise, please contact our HR Department.

Rose International has an official agreement (ID #132522), effective June 30, 2008, with the U.S. Department of Homeland Security, U.S. Citizenship and Immigration Services, Employment Verification Program (E-Verify). (Posting required by OCGA 13/10-91.).

 

Apply Now

 

About Rose

  • Founded in 1993
  • Office Locations Across the U.S.
  • 150+ Clients: Corporations and Government Agencies
  • Employee Oriented Company
  • Challenging Assignments Across the U.S.
  • Continuous Professional Development

The interactions that I have had with your representatives have always been prompt and very professional. I am very pleased and impressed with your company and services.

Sioe, Consultant

Rose International maintained good communication during assignments and are very informative through email and phone calls.

Sade, Consultant

It was great working for Rose International. Everyone was extremely helpful.

Rosann, Consultant

I had a very positive experience working for Rose. The entire process is very efficient and easy.

Joanne, Consultant

I have been very pleased with my experience with Rose International. Everyone that I encountered was very helpful and courteous.

Stephanie, Consultant

EMPLOYEE COMMENTS

  • We want you to work with us, but don't take our word for it. Take a look at this sampling of employee comments. They speak for themselves.