A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.
$0 Cost Medical Plan options, Dental & Vision, Discounted Childcare, Pet Insurance, 3.5 weeks PTO, 401k match, etc.
Location:
East Lansing, MI (Onsite)
Job Summary:
a { text-decoration: none; color: #464feb; } tr th, tr td { border: 1px solid #e6e6e6; } tr th { background-color: #f5f5f5; } a { text-decoration: none; color: #464feb; } tr th, tr td { border: 1px solid #e6e6e6; } tr th { background-color: #f5f5f5; } We are seeking a Site Reliability Engineer (SRE) with strong networking expertise to help support and modernize our hybrid infrastructure. This role will focus on network reliability, automation, cloud networking, and Infrastructure as Code (IaC) as we continue migrating workloads from on-premises environments to Azure. The ideal candidate combines deep networking knowledge with experience managing infrastructure through automation, CI/CD pipelines, and Git-based workflows. Key Responsibilities of the
Site Reliability Engineer:
Manage and support enterprise networking infrastructure, including routing, switching, wireless, and firewalls Design and maintain secure network connectivity across on-premises and Azure environments Support Azure networking initiatives, including VNets, NSGs, gateways, and hybrid connectivity Implement and manage network infrastructure using Infrastructure as Code (Terraform, Ansible) Develop and maintain CI/CD pipelines and Git-based deployment workflows Administer network security and remote access solutions, including Zero Trust technologies Troubleshoot complex infrastructure and network-related incidents Participate in infrastructure automation, monitoring, performance optimization, and continuous improvement efforts Contribute to on-call support and incident response activities Preferred Skills of the
Site Reliability Engineer:
Strong network management background including management & configuration of firewalls, switches, etc. Knowledgeable of routing & switching (BGP, OSPF, VLANs, WAN) Solid Infrastructure as Code (IaC) in large-scale environments Deep experience with Git, GitHub Actions , and Git-based deployment models Experience with CI/CD management and infrastructure delivery through automated pipelines Experience with Ansible or Terraform Experience with Azure networking (VNets, Network Security Groups, Gateways, hybrid connectivity) Knowledge of Microsoft Global Secure Access (GSA) Network automation and configuration management Bonus Skills of the
Site Reliability Engineer:
Zero Trust Network Access (ZTNA) solutions Microsoft Global Secure Access (GSA) Advanced Terraform or Ansible experience Knowledgeable of Cloudflare Cloud migration projects involving network transformation Financial services or regulated industry experience #LI-NB5 #