Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details
Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

Rose International

SRE (Site Reliability Engineer)

Career Insights for Site Reliability Engineer

See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.

Scorecard

Based on New York data

Review key factors to help you decide if this role fits your goals. How is this calculated?

Were these scores useful?

What they do

A Site Reliability Engineer is responsible for designing, implementing, and maintaining highly reliable and scalable software systems and infrastructure. They emphasize automation, code-driven infrastructure, and the use of software tools to manage systems efficiently. Monitors performance within production environments, identifies causes of incidents, and implements preventative measures to ensure software reliability.

$128,318 / year median in New York

Explore Career

Job Description

Required Education//Certifications:
  • Bachelor's degree in Computer Science, Information Technology, Media Technology, or equivalent practical experience.
Required Skills, Experience, & Abilities:
  • 0-2 years of experience in systems administration, IT operations, video engineering, or SRE-related work.
  • Exposure to video delivery workflows (streaming, VOD, ABR packaging) is a plus but not required.
  • Basic understanding of Linux administration and troubleshooting.
  • Familiarity with scripting (Python, Bash, or PowerShell).
  • Exposure to monitoring/observability tools (Grafana, Prometheus, ELK, Splunk, DataDog, or similar).
  • General knowledge of networking fundamentals (TCP/IP, DNS, load balancing, HTTP).
  • Awareness of streaming protocols (HLS, MPEG-DASH, RTMP) is a plus.
  • Experience with cloud environments (AWS, Azure, GCP) is desirable but not required.
Job Overview:
We are seeking a Site Reliability Engineer (SRE I) to join our Video Platform Engineering team. In this role, you will support the operation, monitoring, and reliability of large-scale video delivery systems, including live streaming, video-on-demand (VOD), encoding, packaging, and content delivery networks (CDNs).As a Level 1 SRE, you will work closely with senior engineers to respond to incidents, perform routine operational tasks, and build foundational automation skills. This position is ideal for candidates with a background in systems or network administration who are eager to grow into video reliability engineering.

Responsibilities
  • Monitor the health and performance of video services (live, linear, and on-demand) using observability tools.
  • Support day-to-day operations of video platforms including encoding, transcoding, packaging, origin servers, DRM, and CDN delivery.
  • Assist in troubleshooting service-impacting issues and escalating to senior engineers when needed.
  • Participate in on-call rotations under supervision, responding to alerts and contributing to incident resolution.
  • Execute runbooks for routine operational tasks, deployments, and platform maintenance.
  • Write and maintain basic automation scripts (Python, Bash, PowerShell) to reduce manual work.
  • Document incident response steps, troubleshooting guides, and standard operating procedures.
  • Collaborate with cross-functional teams (network, storage, CDN, and applications) to support service delivery.
  • Only those lawfully authorized to work in the designated country associated with the position will be considered.
  • Please note that all Position start dates and duration are estimates and may be reduced or lengthened based upon a client's business needs and requirements.