Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details

Back to Results

Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

CATHEXIS

Site Reliability Engineer (req-307)

Review key factors to help you decide if the role fits your goals.
Pay Growth
?
out of 5
Not enough data
Not enough info to score pay or growth
Job Security
?
out of 5
Not enough data
Calculating job security score...
Total Score
99
out of 100
Average of individual scores

Were these scores useful?

Job Description

Site Reliability Engineer (req-307)

CATHEXIS

medical insurance, dental insurance, life insurance, vision insurance, parental leave, paid time off, long term disability, 401(k) United States, Virginia, Tysons Sep 22, 2026 Team

CATHEXIS

elevates the government contracting experience through rapid response, deep skill, and thoughtful problem-solving and communication. Our core capabilities are our top-tier program and project management, data analytics, and audit services, the backbone of which is our integrated approach to operational excellence. You worked hard to get to where you are. You strive to make every day better than the day before. So do we. Team

CATHEXIS

operates with an all-in mindset. We are working together to create a company that supports our shared values and individual goals. Our values are centered around leading with integrity, owning the outcome, growing together, and moving with purpose in everything we do for our employees, customers, partners, and communities. We believe success is best when we listen and lead with empathy; model high standards of ethics to provide a rewarding candidate experience; work hard, have fun, and appreciate the strengths we all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and optimize clusters on Kubernetes. Together, we're accelerating our clients' digital transformation through the building and deployment of data-driven, scalable AI solutions. The ideal candidate will have a deep understanding of Kubernetes, Cloud Infrastructure, and Infrastructure as Code (IaC) practices. You will be responsible for ensuring the reliability and scalability of our clients' Kubernetes clusters and Cloud Infrastructure. Responsibilities The responsibilities include, but are not limited to:

Monitor and Manage Kubernetes Clusters:

Ensure the stability, health, and scalability of Kubernetes Clusters, deploying applications and services on Kubernetes

Kubernetes Management:

Deploy, monitor, and scale applications on Kubernetes clusters. Maintain Helm charts, manage services, and ensure resource allocation for optimal cluster performance

Containerization & Deployment:

Design and maintain Docker-based microservices architecture, ensuring consistent and reproducible deployments across staging, QA, and production environments

Cloud Infrastructure Management:

Work with leading Cloud Platforms (AWS, Azure and/or GCP) to set up, configure, and manage infrastructure resources using Infrastructure as Code (Terraform, CloudFormation, etc.)

Monitoring & Incident Response:

Set up monitoring solutions, define alerts, an manage the incident response process for any issues related to Jenkins or Kubernetes clusters

Automate Infrastructure Processes:

Build automation tools for scaling, monitoring, and maintaining infrastructure using modern tools like Terraform, Ansible, Linux, or equivalent

Collaborate Across Teams:

Work closely with development, services, and operations teams to ensure a seamless integration between application development, deployment, and infrastructure

Security & Compliance:

Ensure all systems follow best practices in terms of security and compliance with relevant regulations. This includes role-based access, encryption, and automated vulnerability scanning Team

CATHEXIS

elevates the government contracting experience through rapid response, deep skill, and thoughtful problem-solving and communication. Our core capabilities are our top-tier program and project management, data analytics, and audit services, the backbone of which is our integrated approach to operational excellence. You worked hard to get to where you are. You strive to make every day better than the day before. So do we. Team

CATHEXIS

operates with an all-in mindset. We are working together to create a company that supports our shared values and individual goals. Our values are centered around leading with integrity, owning the outcome, growing together, and moving with purpose in everything we do for our employees, customers, partners, and communities. We believe success is best when we listen and lead with empathy; model high standards of ethics to provide a rewarding candidate experience; work hard, have fun, and appreciate the strengths we all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and optimize clusters on Kubernetes. Together, we're accelerating our clients' digital transformation through the building and deployment of data-driven, scalable AI solutions. The ideal candidate will have a deep understanding of Kubernetes, Cloud Infrastructure, and Infrastructure as Code (IaC) practices. You will be responsible for ensuring the reliability and scalability of our clients' Kubernetes clusters and Cloud Infrastructure. Responsibilities The responsibilities include, but are not limited to:

Monitor and Manage Kubernetes Clusters:

Ensure the stability, health, and scalability of Kubernetes Clusters, deploying applications and services on Kubernetes

Kubernetes Management:

Deploy, monitor, and scale applications on Kubernetes clusters. Maintain Helm charts, manage services, and ensure resource allocation for optimal cluster performance

Containerization & Deployment:

Design and maintain Docker-based microservices architecture, ensuring consistent and reproducible deployments across staging, QA, and production environments

Cloud Infrastructure Management:

Work with leading Cloud Platforms (AWS, Azure and/or GCP) to set up, configure, and manage infrastructure resources using Infrastructure as Code (Terraform, CloudFormation, etc.)

Monitoring & Incident Response:

Set up monitoring solutions, define alerts, an manage the incident response process for any issues related to Jenkins or Kubernetes clusters

Automate Infrastructure Processes:

Build automation tools for scaling, monitoring, and maintaining infrastructure using modern tools like Terraform, Ansible, Linux, or equivalent

Collaborate Across Teams:

Work closely with development, services, and operations teams to ensure a seamless integration between application development, deployment, and infrastructure

Security & Compliance:

Ensure all systems follow best practices in terms of security and compliance with relevant regulations. This includes role-based access, encryption, and automated vulnerability scanning

Active

TOP SECRET

clearance or higher is required

Bachelor's degree in Computer Science or related field

A minimum of two (2) years of experience working with on-premise and off-premise cloud environments

Experience with AWS and/or Azure

Hands-on experience with a range of open-source technologies, such as Linux, Docker, Kubernetes, K8s, Terraform, Helm, PostgreSQL, or similar technologies

Ability to program (structured and OOP) using one or more high-level languages, such as Python, Java, C/C++, Ruby, and JavaScript

Experience with distributed storage technologies such as NFS, HDFS, Ceph, and Amazon S3, as well as dynamic resource management frameworks (Apache Mesos, Kubernetes, Yarn)

Proactive approach to identifying problems, performance bottlenecks, and areas for improvement

Ability to lead and work independently in an Agile/Scrum environment

Real passion for developing team-oriented solutions to complex engineering problems

Thrive in an autonomous, empowering and exciting environment

Great verbal and written communication skills to collaborate multi-functionally and improve scalability

Interest in committing to a fun, friendly, expansive, and intellectually stimulating environment

Desired Skills

Hands-on experience deploying and operating applications using IaaS and PaaS on major cloud providers, such as Amazon AWS, Microsoft Azure, or Google Cloud Services

Experience with deep learning, natural language processing, computer vision, or reinforcement learning

Conveys highly technical concepts and information in written form to technical and non-technical audiences

The ability to work on multiple concurrent projects is essential. Strong self-motivation and the ability to work with minimal supervision

Must be a team-oriented individual, energetic, result & delivery oriented, with a keen interest on quality and the ability to meet deadlines

CATHEXIS

offers competitive compensation packages to all eligible employees. Our goal is to provide a compensation package that reflects the value you bring to our team, is competitive with national average market rates, and promotes your financial security and personal well-being. The annual salary range for this role is $100,000 - $160,000. Please note that the salary information provided is a general guideline.

CATHEXIS

considers various factors in its final offer, including location, qualifications, experience, and skills. Performance Bonuses

Medical Insurance

Dental Insurance

Vision Insurance

401(k) Plan (Traditional and ROTH)

Life Insurance (Basic, Voluntary & AD&D)

Paid Time Off

11 Federal Holidays

Parental Leave

Commute Benefits

Short Term & Long Term Disability

Training & Development

Wellness Program

Community Outreach Initiatives

CATHEXIS

is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, or protected veteran status and will not be discriminated against on the basis of disability

EEO IS THE LAW.

If you are an individual with a disability and would like to request a reasonable accommodation as part of the employment selection process, please contact the Recruiting DepartmentRecruitingTeam@cathexiscorp.com

Benefits

  • Paid Time Off (PTO)
  • 401(k) Plans
  • Health and Wellness Programs
  • Bonuses/Stipends