$60 - 63/hr on W2 Job Description We are seeking a highly skilled Site Reliability Engineer (SRE) III to work hands-on across the technology stack, improving platform and application reliability, observability, and operational efficiency across Global Markets. The SRE will work closely with engineering and core technology teams to support platform modernization, cloud migration initiatives, production system hardening, monitoring improvements, and operational tooling. This position is critical to maintaining execution velocity, reducing operational risk, and ensuring platforms meet reliability and performance objectives. Responsibilities Design, implement, and maintain reliable, scalable, and highly available production systems. Define, implement, and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Perform performance engineering and analysis using Dynatrace and modern observability tools. Develop automation and operational tooling using Python or Perl. Develop and maintain applications and REST APIs using Python and Django. Troubleshoot complex issues across Linux, applications, databases, infrastructure, and cloud environments. Develop and maintain infrastructure automation using Ansible and Infrastructure-as-Code frameworks. Build, maintain, and improve CI/CD pipelines and DevOps automation. Support cloud migration and modernization initiatives across AWS and/or Azure environments. Implement and maintain observability solutions using Dynatrace, OpenTelemetry, and related monitoring technologies. Analyze production performance, identify reliability risks, and implement corrective actions. Improve system availability, scalability, resilience, and operational efficiency. Collaborate with development, infrastructure, cloud, and platform teams to resolve production issues. Support large-scale production environments with a strong focus on reliability, automation, and operational excellence. Work with financial services and trading platform environments when applicable. Required Skills Site Reliability Engineering (SRE) SLOs and SLIs Dynatrace Python or Perl AWS or Azure DevOps Strong Python development experience Hands-on Django and
REST API
development Strong MySQL and database skills Deep Linux administration and troubleshooting experience Infrastructure engineering and automation Ansible CI/CD pipelines Observability and monitoring OpenTelemetry Infrastructure-as-Code Automation frameworks Experience working in large-scale production environments Preferred Qualifications Financial Services, Capital Markets, or Trading Platform experience Experience with cloud migration and modernization Experience with performance engineering and application monitoring Experience with distributed systems and cloud-native technologies Strong production support, incident management, and troubleshooting experience