Find Jobs
Find Jobs Near You – Available Work in Your Location
Site Reliability Engineer
Job Description
Title:
Site Reliability Engineer Location:
Riverwoods, IL (hybrid - 2 to 3 days/week onsite)
Duration:
12+ Months. C2H
•Look for someone with monitoring engineering exp (Dynatrace, Datadog, etc.) Must have: Professional experience as a Site Reliability Engineer (SRE) Experience in performance testing, Ability to translate functional and non-functional requirements into appropriate NFT Automation tests. Experience of AWS Cloud Application (Must for sure) Experience Linux, AWS Cloud(Must) and Prem deployments . Good experience in Systems Observability and APM tools, preferably Datadog Experience in dashboarding tools such as Grafana and Kibana Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability. Expertise in one or more programming languages: Python, shell scripting (Unix/Linux), Java Nice to have: Hands-on experience on SNOW. Experience in container technology (OpenShift, Kubernetes) Strong JIRA knowledge Basic understating of Release Management. Experience in CI/CD pipelines preferably Jenkins expertise.
JD:
Experience 12to 15 years Professional experience as a Site Reliability Engineer (SRE) Software development hands on engineer with excellent understanding of SDLC Application delivery. Ability to translate functional and non-functional requirements into appropriate NFT Automation tests. Experience with DevOps, CI/CD tools. Good experience of Linux, AWS Cloud and on Prem deployments Good experience in Systems Observability and APM tools, preferably Datadog Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability. Strong JIRA knowledge Responsibilities Partner with Application Development teams to build resiliency for Payment application. Partner with our Application Develop teams to implement service level objectives. Partner with our Application Development teams and other SREs to build out end to end observability. Implement monitoring, alerting and dashboards needed for our apps. Automated operational processes. Help to develop our capacity management and performance management tools. Help to define the DR plan needed for our critical apps. Help to develop a chaos testing process. Participate in an on-call rotation and support production Incidents SRE Skillsets Expectations from Discover for
Pricing & Settlements:
Good understanding of hybrid infrastructure Expertise with AWS Expertise in one or more general purpose programming languages: Python, Go, shell scripting (Unix/Linux), Java Experience in CI/CD pipelines preferably Jenkins expertise. Experience in container technology (OpenShift, Kubernetes) Expertise in automation tools experience (preferably Ansible). Expertise in observability tools including APM (Datadog), synthetic monitoring and log aggregation (Elk) Experience in dashboarding tools such as Grafana and Kibana Understating of Agile concepts and experience in JIRA Basic understating of Release Management Hands-on experience on
SNOW SRE
Skillsets Expectations from Discover for
Data Platform:
Expertise in Message Broker (preferably Rabbit MQ, Kafka) Expertise on Hadoop, spark commands JSON formatting Skill Expertise in AWS-Lambda Services 4/5 Strong programming skills (Java, Python, Shell Optional Java Script) - 4/5 Proficiency in Database concepts, Strong knowledge of SQL and experience with MySQL 4/5 APM tools - Datadog 4/5 Hands-on experience on SNOW