Senior DevOps Engineer
Job Title: Senior DevOps Engineer
Company: Snapsheet
Job Location: Remote
Job Type: Full-time
About Snapsheet : Snapsheet is claims technology the way it should be: purposeful, precise, and designed to deliver outcomes. Where others bolt things on, we engineer them into our core systems and processes across cloud-based claims management, virtual vehicle appraisals, and elite loss and recovery services. Trusted by over 170+ P&C Carriers, MGAs, MGUs, TPAs, and logistics companies, our open architecture is built to fit how our companies work, not the other way around.
What you'll get:
- Remote working environment - your new commute is however long it takes to walk to your desk!
- Flexibility - empathy is ingrained in who we are and we are happy to offer a flexible PTO policy, casual dress code, and more!
- Development - Mentorship programs, 1-on-1 management, promote when ready culture, quarterly internal promotion opportunities, and goal setting sessions.
- Fun - Celebrations just because, yearly in-person and remote events, Snapsheet Swag, Employee Resource Groups, and more!
Job Overview:
Snapsheet is building out its DevOps function, and this is a key hire in that journey. As a Senior DevOps Engineer reporting to the Director of IT, you will own the infrastructure layer underpinning our AWS cloud platform โ including cloud networking, database infrastructure, availability architecture, security hardening, and operational automation. You'll serve as the infrastructure platform partner for our Full Stack engineers and a close collaborator with SRE, which owns CI/CD, developer tooling, and observability. Success in this role means bringing clarity, reliability, and rigor to how our AWS environment is managed and evolved.
Responsibilities:
- Own infrastructure architecture and standards โ naming, tagging, resource sizing, lifecycle management, and documentation โ while auditing and closing gaps across environments
- Own internal and customer-facing cloud networking (VPCs, subnets, security groups, Transit Gateway, VPN/Direct Connect) and define standards for secure inter-service connectivity
- Manage database infrastructure including clustering, failover, encryption, backups, and performance tuning, with multi-AZ high-availability strategies to meet recovery objectives
- Design and maintain high-availability and disaster recovery strategies (RTO/RPO targets, runbooks, tested recovery processes), plus build cost/FinOps reporting to drive AWS spend visibility
- Administer AWS IAM (roles, policies, least-privilege, cross-account access) and enforce infrastructure security baselines in partnership with IT Security
- Co-own Terraform HCL infrastructure with SRE, writing modular, production-grade IaC and engaging with evolving workflow tooling (Terragrunt, Spacelift, etc.)
- Manage IT-layer AWS footprint (RPA server infrastructure, VDI) and co-own GitHub administration, and partner with Full Stack engineers on AWS service selection (Lambda, Fargate, App Runner, ElastiCache)
Qualifications:
- 5+ years of experience in DevOps, infrastructure, or platform engineering roles preferred
- Deep hands-on experience with core AWS services: EC2, ECS/Fargate, Lambda, RDS/Aurora (MySQL & PostgreSQL), DynamoDB, ElastiCache, OpenSearch, CloudFront, VPC, IAM, Secrets Manager
- Expert-level Terraform (HCL) experience, including environment orchestration and workflow automation/state management best practices
- Hands-on Aurora MySQL cluster management (clustering, failover, storage scaling, backups) with working knowledge of Aurora PostgreSQL and standalone RDS
- Strong networking background: VPCs, routing, security groups, NACLs, Transit Gateway, VPN, and Route 53
- Security expertise across IAM policies, SCPs, VPC isolation, GuardDuty, Security Hub, and encryption standards
- Experience with Docker image hardening, ECS/Fargate deployment patterns, and scripting proficiency in Python, Bash, or Ruby
- Working proficiency