Senior Software Engineer - Site Reliability Engineering
Job
Noctua Technology, Inc
Reston, VA (In Person)
$175,200 Salary, Full-Time
Review key factors to help you decide if the role fits your goals.
Pay Growth
?
out of 5
Not enough data
Not enough info to score pay or growth
Job Security
?
out of 5
Not enough data
Calculating job security score...
Total Score
99
out of 100
Average of individual scores
Skill Insights
Compare your current skills to what this opportunity needs—we'll show you what you already have and what could strengthen your application.
Job Description
Job Requirements Reston, VA Chantilly, VA Fort Belvoir, VA San Diego, CA Secret Polygraph not specified Mid Level Career (5+ yrs experience) $148,920 - $201,480 Job Description Company Overview Noctua Technology, Inc. is a software engineering and consulting corporation focused on data engineering, machine learning, and cloud technologies. We specialize in delivering premier quality software engineering solutions to Public Sector and Commercial customers across the US. Department Overview The Site Reliability Engineering discipline at Noctua Technology, Inc is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don't just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations. Job Summary We are seeking a highly experienced and autonomous Senior Site Reliability Engineer (SRE) to join our dynamic team. As a technical leader, you will define the strategy and apply advanced software engineering principles to operations, focusing on the architecture, reliability, and long-term performance of large-scale production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure.
Security Clearance Requirement:
Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above. Key Responsibilities- Site Reliability Engineering ○ Drive the definition and adoption of SLIs and SLOs across multiple services or entire platforms, ensuring alignment with business goals.
- Toil Reduction and Incident Management ○ Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
- Testing and Service Resiliency ○ Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls. ○ Proactively identify and address system weaknesses and ensure performance under stress. ○ Support disaster recovery and high availability strategies through backup and failover planning.
- Collaboration and Knowledge Sharing ○ Serve as a primary SRE liaison for development teams, influencing application architecture and design to meet reliability and scalability targets from inception.
- Stakeholder Communication ○ Act as a subject matter expert and trusted advisor to clients and internal leadership on cloud infrastructure, reliability strategy, and Service Level Agreement (SLA) negotiations. ○ Act on client feedback to refine and enhance cloud solutions. ○ Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively.
- Continuous Learning and Innovation ○ Stay updated on the latest developments in cloud infrastructure and technology trends. ○ Drive innovation by proposing and implementing new techniques and technologies. Qualifications
- 5+ years of experience in site reliability engineering, cloud engineering, or related fields.
- Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management.
- Deep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
- Deep experience with containerization and orchestration tools like Docker and Kubernetes.
- Deep knowledge of networking concepts, cloud security best practices, and identity management.
- Experience with programming or scripting languages such as Python, Bash, or Go.
- Experience with CI/CD pipelines and DevOps methodologies.
- Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
- Demonstrated ability to influence technical decision-making across organizational boundaries Preferred qualifications:
- Bachelor's or advanced degree in Computer Science or a related field.
- Any of the below cloud certifications: ○ Google Cloud Professional Cloud Architect ○ Google Cloud Professional Cloud DevOps Engineer ○ AWS Certified Solutions Architect ○ AWS Certified Developer ○ AWS Certified SysOps Administrator group id: 91166237 N Name Hidden Recruiter Apply now
Similar remote jobs
Similar jobs in Reston, VA
M&T Bank
Reston, VA
Posted2 days ago
Updated18 hours ago
Similar jobs in Virginia
Amazon
Herndon, VA
Posted2 days ago
Updated18 hours ago
Accountable Healthcare Staffing
Arlington, VA
Posted2 days ago
Updated18 hours ago