Find Jobs
Find Jobs Near You – Available Work in Your Location
Site Reliability Engineer
Career Insights for Site Reliability Engineer
See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.
Scorecard
Based on Texas data
Review key factors to help you decide if this role fits your goals. How is this calculated?
What they do
A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.
$123,674 / year median in Texas
Job Description
Site Reliability Engineer:
On behalf of our technology client, Procom is searching for a Site Reliability Engineer for a 12-month contract role. This position is a hybrid position with 4 days onsite at our client's Southlake, Texas 76092 office. Site Reliability Engineer - We are seeking a motivated Site Reliability Engineer (Contractor) with a focus on improving reliability, observability, operational efficiency, and automation across on-premises and cloud platforms. The project involves enhancing automation, cloud infrastructure, and production operations. Site Reliability Engineer -
Responsibilities:
- Develop Python-based automation solutions to reduce manual operational effort.
- Automate infrastructure management across Linux, Windows, Kubernetes, GCP, and cloud-native environments.
- Integrate tools and platforms through APIs and client libraries.
- Support CI/CD automation and deployment reliability initiatives.
- Monitor and maintain production systems to meet reliability and availability objectives.
- Participate in incident response, troubleshooting, and root cause analysis activities.
- Build and maintain dashboards, alerts, and monitoring solutions using Splunk, Grafana, Prometheus, GCP Operations Suite, or similar tools.
Site Reliability Engineer -
Mandatory Skills:
- 3 to 5 years of experience in Site Reliability Engineering, DevOps, Systems Engineering, or Platform Engineering.
- Strong programming skills in Python for automation and tooling development.
- Experience supporting Kubernetes and cloud platforms (GCP, AWS, or Azure).
- Familiarity with infrastructure automation and configuration management tools.
- Experience with monitoring and observability platforms such as Splunk, Grafana, Prometheus, Datadog, or similar.
- Understanding of Linux systems, networking, and distributed applications.
- Strong analytical, troubleshooting, and problem-solving skills.
Site Reliability Engineer -
Nice-to-Have Skills:
- Experience with Terraform, Ansible, or Infrastructure as Code solutions.
- Exposure to OpenTelemetry and modern observability practices.
- Experience with CI/CD pipelines and deployment automation.
- Knowledge of AI/ML, AIOps, or intelligent operational tooling.
- Experience supporting highly available production systems in regulated or enterprise environments.
Site Reliability Engineer -
Assignment Length:
This is a 12-month contract position. Site Reliability Engineer -
Start Date:
ASAP. Site Reliability Engineer -
Assignment Location:
Southlake, Texas, 76092 United States. This is a hybrid role requiring 4 days onsite.