Senior DevOps Engineer

๐Ÿข Snapsheet ยท all Snapsheet jobs
๐Ÿ“ United States
๐Ÿ’ฐ USD 120,000 - 135,000 / annual
๐Ÿ“… Posted 2026-07-11 ยท via Himalayas
๐Ÿท Cloud-Infrastructure-Engineering,Platform-Engineering,AWS-Engineering,Site-Reliability-Engineering,Senior-DevOps,Senior-Staff-DevOps-Engineer,Senior-DevOps-Platform-Engineer,Senior-AWS-DevOps-Engineer,Senior-DevOps-Manager,Senior-DevOps-Architect,DevOps-Engineer,Cloud-DevOps-Engineer
Apply on original site โ†—

Job Title: Senior DevOps Engineer
Company: Snapsheet

Job Location: Remote
Job Type: Full-time

About Snapsheet : Snapsheet is claims technology the way it should be: purposeful, precise, and designed to deliver outcomes. Where others bolt things on, we engineer them into our core systems and processes across cloud-based claims management, virtual vehicle appraisals, and elite loss and recovery services. Trusted by over 170+ P&C Carriers, MGAs, MGUs, TPAs, and logistics companies, our open architecture is built to fit how our companies work, not the other way around.
What you'll get:

- Remote working environment - your new commute is however long it takes to walk to your desk!

- Flexibility - empathy is ingrained in who we are and we are happy to offer a flexible PTO policy, casual dress code, and more!

- Development - Mentorship programs, 1-on-1 management, promote when ready culture, quarterly internal promotion opportunities, and goal setting sessions.

- Fun - Celebrations just because, yearly in-person and remote events, Snapsheet Swag, Employee Resource Groups, and more!

Job Overview:

Snapsheet is building out its DevOps function, and this is a key hire in that journey. As a Senior DevOps Engineer reporting to the Director of IT, you will own the infrastructure layer underpinning our AWS cloud platform โ€” including cloud networking, database infrastructure, availability architecture, security hardening, and operational automation. You'll serve as the infrastructure platform partner for our Full Stack engineers and a close collaborator with SRE, which owns CI/CD, developer tooling, and observability. Success in this role means bringing clarity, reliability, and rigor to how our AWS environment is managed and evolved.
Responsibilities:

- Own infrastructure architecture and standards โ€” naming, tagging, resource sizing, lifecycle management, and documentation โ€” while auditing and closing gaps across environments

- Own internal and customer-facing cloud networking (VPCs, subnets, security groups, Transit Gateway, VPN/Direct Connect) and define standards for secure inter-service connectivity

- Manage database infrastructure including clustering, failover, encryption, backups, and performance tuning, with multi-AZ high-availability strategies to meet recovery objectives

- Design and maintain high-availability and disaster recovery strategies (RTO/RPO targets, runbooks, tested recovery processes), plus build cost/FinOps reporting to drive AWS spend visibility

- Administer AWS IAM (roles, policies, least-privilege, cross-account access) and enforce infrastructure security baselines in partnership with IT Security

- Co-own Terraform HCL infrastructure with SRE, writing modular, production-grade IaC and engaging with evolving workflow tooling (Terragrunt, Spacelift, etc.)

- Manage IT-layer AWS footprint (RPA server infrastructure, VDI) and co-own GitHub administration, and partner with Full Stack engineers on AWS service selection (Lambda, Fargate, App Runner, ElastiCache)

Qualifications:

- 5+ years of experience in DevOps, infrastructure, or platform engineering roles preferred

- Deep hands-on experience with core AWS services: EC2, ECS/Fargate, Lambda, RDS/Aurora (MySQL & PostgreSQL), DynamoDB, ElastiCache, OpenSearch, CloudFront, VPC, IAM, Secrets Manager

- Expert-level Terraform (HCL) experience, including environment orchestration and workflow automation/state management best practices

- Hands-on Aurora MySQL cluster management (clustering, failover, storage scaling, backups) with working knowledge of Aurora PostgreSQL and standalone RDS

- Strong networking background: VPCs, routing, security groups, NACLs, Transit Gateway, VPN, and Route 53

- Security expertise across IAM policies, SCPs, VPC isolation, GuardDuty, Security Hub, and encryption standards

- Experience with Docker image hardening, ECS/Fargate deployment patterns, and scripting proficiency in Python, Bash, or Ruby

- Working proficiency

โ† All remote jobs