Senior Software Engineer- Site Reliability Engineering (SRE)
location_onChestnut Street, Franklin, Virginia, Saint Louis County, Minnesota, 55792, United States
Job Description
About the Team
The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations.
Work location
Work model: Remote
Chestnut Street, Franklin, Virginia, Saint Louis County, Minnesota, 55792, United States
Virginia, Minnesota
Key Responsibilities
- check_circleDefine and adopt SLIs and SLOs across multiple services or platforms
- check_circleDesign and architect Infrastructure as Code solutions for large-scale environments
- check_circleImplement and manage containerized and serverless architectures using Docker and Kubernetes
- check_circleBuild and maintain reliable and self-healing CI/CD pipelines
- check_circleImplement comprehensive monitoring, alerting, and logging systems
- check_circleLead high-severity incident response and coordinate blameless postmortems
- check_circleImplement cloud security best practices including IAM and encryption
- check_circleServe as primary SRE liaison for development teams on architecture and design
- check_circleAct as subject matter expert for clients on cloud infrastructure and reliability strategy
Requirements
- verified5+ years of experience in site reliability engineering, cloud engineering, or related fields
- verifiedUS citizens eligible to obtain and maintain an active Secret security clearance or above
- verifiedStrong software engineering skills with emphasis on clean, modular, and maintainable code
- verifiedDeep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation
- verifiedDeep experience with containerization and orchestration tools like Docker and Kubernetes
- verifiedDeep knowledge of networking concepts, cloud security best practices, and identity management
- verifiedExperience with programming or scripting languages such as Python, Bash, or Go
- verifiedExperience with CI/CD pipelines and DevOps methodologies
- verifiedBachelor's or advanced degree in Computer Science or a related field (preferred)
- verifiedGoogle Cloud Professional Cloud Architect or DevOps Engineer certification (preferred)
- verifiedAWS Certified Solutions Architect, Developer, or SysOps Administrator certification (preferred)
- verifiedCompTIA Security+ or equivalent DoD 8140/8570 IAT Level II baseline certification (preferred)
Skills, education and keywords
Skills: site reliability engineering, infrastructure as code, terraform, cloudformation, docker, kubernetes, python, bash, go, ci/cd.
Education: Bachelor's or advanced degree in Computer Science or a related field.