Site Reliability Engineer

LeidosBaltimore, MarylandOn-siteFull-timeStaff, 8–12 yearsListed 2 hours ago

Apply now

About this role

The Digital Sector at Leidos currently has an opening for a Site Reliability Engineer (SRE) / Senior Cloud Engineer to work in our Baltimore, Maryland office.  This is an exciting opportunity to use your experience helping the Center for Medicare and Medicaid Services (CMS) modernize its legacy Contact Center CRM platform within the CMS AWS Enclave and Pega Cloud for Government.

Primary Responsibilities

The SRE/Senior Cloud Engineer shall design, build, and operate the highly available cloud infrastructure supporting the CRM modernization effort.

Responsibilities shall include, but are not limited to:

- Design, build, and operate highly available AWS infrastructure within the CMS AWS Enclave (FedRAMP Moderate), applying AWS Well-Architected Framework best practices.
- Architect secure, scalable multi-account VPC interconnectivity between the CMS AWS Enclave and Pega Cloud for Government (PCFG) in AWS GovCloud US-West, including PrivateLink, API Gateway, and Direct Connect.
- Support container and serverless architectures (e.g., AWS Lambda, Glue) for data integration, batch processing, and API layers supporting the modernized CRM.
- Apply Site Reliability Engineering (SRE) practices, including defining and tracking Service Level Objectives (SLOs), error budgets, and reliability metrics aligned to contract SLAs (e.g., ≥99.9% availability).
- Build and maintain observability across the AWS and Pega environments, including centralized logging, monitoring, and alerting, to enable proactive detection of performance and availability issues.
- Lead incident response and root-cause analysis for production issues, and long-term reliability improvements.
- Automate infrastructure provisioning, configuration, and environment build-out using Infrastructure as Code (e.g., Terraform, CloudFormation, Ansible).
- Design and test Disaster Recovery capability for cloud-based workloads, including backup, failover, and Multi-AZ/Multi-Region resilience.
- Support performance testing and capacity planning to validate the platform's ability to scale to 20,000 concurrent CSR sessions and peak Open Enrollment Period (OEP) volumes.
- Support continuous security monitoring, vulnerability remediation, and Zero Trust alignment across the AWS and Pega environments.
- Partner with the DevOps Lead/Configuration Manager to build and maintain CI/CD pipelines, ensuring automated testing, security scanning, and deployment across all SDLC environments.
- Support cloud connectivity and data movement for the AWS-based data migration pipeline (e.g., AWS Glue, S3, RDS/Aurora PostgreSQL) between legacy Siebel and the modernized CRM.
- Coordinate with the CMS Hybrid Cloud Team on cloud environment provisioning, patching, and lifecycle management activities.
- Manage release coordination and change windows in support of OEP blackout periods and other critical operational periods, minimizing risk of service disruption.
- Collaborate with the Release Train Engineer, Solution Architect, and Agile delivery teams to align infrastructure readiness with sprint and PI planning.
- Support integration of Genesys Cloud CX infrastructure and telephony/chat channels with the modernized CRM environment.
- Continuously identify opportunities to reduce operational toil through automation of repetitive tasks, log analysis, and routine operational activities.
- Document cloud architecture, operational runbooks, and disaster recovery procedures to support the Transition-Out Plan and audit readiness.
- Provide technical mentoring and knowledge-sharing to other engineers on cloud architecture, automation, and reliability engineering best practices.
- Communicate technical status, risks, and dependencies to CMS leadership, and Leidos management.
- Support requirements traceability and technical documentation related to infrastructure and integration architecture.
- Actively participate in planning sessions, requirements gathering activities, design sessions, Agile sessions, and other events supporting the CRM modernization effort.

Required Qualifications:

- Bachelor’s degree and a minimum of 6-8 years of relevant experience in cloud engineering, site reliability engineering, or infrastructure, or an equivalent combination of education and experience
- Experience building highly available AWS infrastructure based on industry best practices and the AWS Well-Architected Framework
- Experience with Infrastructure as Code, automation, and configuration management of cloud-based resources
- Experience designing Disaster Recovery for cloud-based workloads, including Multi-AZ/Multi-Region resilience
- Ability to obtain Public Trust

Preferred Qualifications:

- AWS Solutions Architect or SysOps Administrator Certification
- Experience with AWS GovCloud and FedRAMP-authorized cloud environments
- Experience supporting Pega Cloud for Government (PCFG) or similar SaaS platform connectivity (e.g., AWS PrivateLink, VPC peering)
- Familiarity with container (Docker, ECS, EKS) and serverless (Lambda) architectures
- Experience with observability/monitoring tooling (e.g., Splunk, New Relic, CloudWatch) and incident response practices
- Experience supporting federal contact center or other 24x7 mission-critical government systems
- Agile delivery experience

If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.

##

##

## Original Posting:
September 24, 2026

For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

## Pay Range:
Pay Range $107,900.00 - $195,050.00

The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.