Find Jobs
Find Jobs Near You – Available Work in Your Location
Skip to job details
UT
US Tech Solutions
Site Reliability Engineer
Career Insights for Site Reliability Engineer
See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.
Scorecard
Based on New York data
Review key factors to help you decide if this role fits your goals. How is this calculated?
What they do
A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.
$128,318 / year median in New York
Job Description
$70-$75 per hour Corning, NY Contract
Duration:
06 months with possible extensionJob Description:
- Looking for an experienced Site Reliability Engineer who can strengthen our team's platform engineering and operational capabilities. You will play a key role in supporting Kubernetes infrastructure managed through Rancher, improving system reliability and automation, and advancing infrastructure-as-code and GitOps practices across our environment.
Key Responsibilities:
- + •
Platform Operations:
- Maintain and enhance Kubernetes platforms across on-premises and cloud environments, ensuring reliability, scalability, and operational efficiency. +
Cluster Management:
- Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher. +
Linux Systems Administration:
- Provide deep technical expertise in Linux-based systems, including performance tuning, troubleshooting, automation, and operational support. +
- Infrastructure as
Code:
- Develop and maintain infrastructure-as-code solutions to standardize and automate platform deployment and management, with a preference for Cluster API (CAPI)-based approaches. +
GitOps and Deployment Automation:
- Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration in a consistent, auditable manner. +
Collaboration:
- Work closely with developers, scientists, and infrastructure teams to deliver reliable platform services and translate operational needs into sustainable engineering solutions. +
Continuous Improvement:
- Identify opportunities to improve platform resilience, observability, security, and maintainability through automation and modern SRE practices.
- Required Skills
- + 5+ years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering roles.
Technical Skills:
- +
Kubernetes:
Cluster operations, upgrades, networking, storage, troubleshooting, and workload support. +Platform Management:
Rancher or similar Kubernetes management platforms. +Linux:
Advanced administration of Linux/Unix systems. + Infrastructure asCode:
Strong IaC experience; Cluster API (CAPI) preferred. +GitOps/CI-CD:
ArgoCD, Git version control, and deployment automation practices. +Scripting/Automation:
Bash, Python, or similar scripting languages for automation and operational tooling.Preferred:
- + Experience with hybrid infrastructure spanning on-premises and public cloud platforms (AWS, Azure, GCP).
Education:
- + BS in Computer Science, Software Engineering, Information Technology, or related field preferred; or equivalent professional experience.
About US Tech Solutions:
- US Tech Solutions is a global staff augmentation firm providing a wide range of talent on-demand and total workforce solutions.
AI Statement:
- _By applying, you acknowledge that AI-assisted tools may be used during hiring.