Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details

Back to Results

Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

Paradigm

Site Reliability Engineer

Choose a Location

This role is available in multiple locations. Pick one to apply.

Career Insights for Site Reliability Engineer

See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.

Scorecard

Based on Washington data

Review key factors to help you decide if this role fits your goals. How is this calculated?

Were these scores useful?

What they do

A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.

$133,805 / year median in Washington

Explore Career

Job Description

Paradigm is a software company transforming the way that the residential, construction & building product industries operate across the globe. We are looking for a Site Reliability Engineer to be part of revolutionizing these industries. We are building the future of how homes are designed, estimated, and built by revolutionizing construction workflows with modern software engineering, agent-assisted systems, and mobile-first experiences. We are powered by our parent company, Builders FirstSource (
NYSE:
BLDR): a Fortune 300 company with over $23 billion in revenue and more than 29,000 employees across 550+ locations, BFS is redefining construction through data, digital infrastructure, and AI-powered innovation. We're seeking a Site Reliability Engineer to support the reliability, scalability, and performance of our cloud platforms. This role blends software engineering and systems engineering to build and maintain highly available, scalable, and resilient systems. The engineer will apply SLO‑driven practices, contribute to automation‑focused operations, and support ongoing reliability improvements. These efforts will contribute to a strong culture of reliability, automation, and engineering excellence to build systems that scale confidently and recover gracefully.
What You Will Do:
Contribute to defining and refining SLOs, SLIs, and error budgets. Implement infrastructure automation and reduce operational toil using IaC, scripting, and CI/CD pipelines. Build and maintain observability tooling (metrics, logs, tracing, alerting). Identify issues and implement changes that enhance system scalability, performance, and resilience. Contribute to incident response and blameless postmortems. Support the optimization of AWS and/or Azure cloud environments for reliability and cost efficiency. What You Need to
Succeed:
Bachelor's degree in Computer Science, Software Engineering, or related field or equivalent experience. 0-4 years of production experience in Azure and/or any major cloud platform is preferred. Familiarity with SLOs, incident management, and reliability engineering practices. Proficient with Infrastructure as Code tools (Terraform/Ansible) and scripting languages (Python/Bash/PowerShell). Experience with monitoring/observability platforms (e.g., Datadog or similar).
Beneficial Skills and Knowledge:
Experience with Docker and Kubernetes. Knowledge of CI/CD pipelines (Jenkins, Git-based workflows). Linux/Windows administration skills. Experience with cloud cost optimization and performance testing.