Senior AWS Site Reliability Engineer

🏢 VIATEQ Corporation · all 4 jobs
📍 United States
💰 USD 145,000 - 185,000 / annual
📅 Posted Sep 16, 2026 · via Himalayas
🏷 Site Reliability Engineering, AWS Cloud Engineering, Platform Engineering, DevOps Engineer, Cloud Infrastructure, Senior Site Reliability Engineer +6 more
Apply on original site ↗

VIATEQ is looking for a Senior AWS Site Reliability Engineer (SRE) to support our prime contractor and government customer in Washington, DC. The position is remote and requires the candidate to be able to obtain a Public Trust.

The Senior AWS SRE will support the CRM, Integration, and Platform Services workforce. This role will support AWS cloud platform engineering, Kubernetes operations, HashiCorp platform automation, CI/CD, observability, reliability, and AI-enabled delivery practices while helping the customer deliver secure, reliable, accessible, and maintainable digital services. The successful candidate will improve reliability, observability, incident response, performance, capacity, and operational readiness and will work closely with government stakeholders, architects, engineers, product owners, cybersecurity, operations, and business users to convert requirements into mission-aligned platform capabilities.

The Senior AWS SRE will provide site reliability, production operations, and cloud platform engineering support for AWS capabilities within the CRM, Integration, and Platform Services environment.
Responsibilities:

- Design, configure, develop, integrate, test, document, and sustain AWS capabilities using AWS, EKS, ECS, Terraform, Packer, Vault, Consul, GitHub Actions, CI/CD pipelines, CloudWatch, Open Telemetry, container platforms, and security automation.

- Implement and improve CI/CD pipelines, infrastructure as code, container platform operations, monitoring, alerting, and secure deployment automation.

- Translate business, mission, security, accessibility, and operational requirements into practical technical solutions.

- Support platform architecture, backlog refinement, implementation planning, release readiness, and production transition activities.

- Develop reusable patterns, configuration standards, automation, documentation, and support procedures that reduce delivery and sustainment risk.

- Troubleshoot complex issues across platform configuration, code, data, APIs, identity, security, performance, and user experience.

- Collaborate with cybersecurity, privacy, data, infrastructure, QA, and change management teams to align delivery with federal operating expectations.

- Maintain clear technical documentation, design decisions, implementation notes, test evidence, and operational runbooks.

Required Education and Experience:

- Bachelor's degree in Computer Science, Information Systems, Software Engineering, Data Analytics, Cybersecurity, or a related discipline, or equivalent work experience.

- 7+ years of experience in site reliability, production operations, and cloud platform engineering, preferably in complex enterprise or government environments.

- Hands-on experience with AWS implementation, configuration, development, integration, testing, or operations.

- Experience working with Agile delivery teams and translating stakeholder needs into maintainable technical outcomes.

Preferred Education and Experience:

- AWS Solutions Architect, AWS DevOps Engineer Professional, AWS Security Specialty, Kubernetes certification, Terraform certification, or comparable cloud credential.

- Experience supporting federal government IT environments, Public Trust programs, or regulated enterprise delivery.

- Experience with Section 508, cybersecurity, auditability, data protection, and operational documentation expectations.

Required Skills and Competencies:

- Strong hands-on knowledge of AWS capabilities, implementation patterns, administration, development, integration, and lifecycle management.

- Ability to design and implement secure, supportable, upgrade-aware solutions that avoid unnecessary customization and reduce technical debt.

- Experience with APIs, identity and access controls, data management, testing, monitoring, troubleshooting, and release coordination.

- Ability to document technical designs, configuration decisions, operational procedures, test results, and risks in c

Flights + hotels

This role requires you to be in the United States. If that means relocating or flying in, it is worth checking fares before you commit to a start date.

Compare flights and hotels →

← All remote jobs

Comparing Site Reliability Engineer pay and openings — the live median is $128k?All remote Site Reliability Engineer jobs →Site Reliability Engineer salary data →
Want more like this? Browse every live remote developer role.All remote developer jobs →
Get new developer jobs by email
Daily email, only when there's something new. One click to stop.

Get remote developer jobs like this by email

10 hand-picked jobs, one email a day. No spam, unsubscribe anytime.

Similar for you