Cloverhealth
Cloverhealth

Senior Site Reliability Engineer

engineeringfull-timeRemote - USA
SALARY
Not listed
WORK TYPE
remote
JOB TYPE
full-time
INDUSTRY
healthcare
Apply for this position
✦ AutoApply Let us apply to roles like this on your behalf.
Learn more

About the role

About Counterpart Health

At Counterpart Health, we are transforming healthcare and improving patient care with our innovative primary care tool, Counterpart Assistant. By supporting Primary Care Physicians (PCPs), we are able to deliver improved outcomes to our patients at a lower cost through early diagnosis and longitudinal care management of chronic conditions.

We are looking for an experienced Site Reliability and Infrastructure Engineer to join our engineering team. You will support Counterpart Health’s existing technology infrastructure by reviewing and improving processes, developing automation tools to eliminate toil, and troubleshooting issues as they arise. You will collaborate with technical leads across engineering disciplines, as well as data scientists and technology professionals, to develop and maintain a modern, scalable infrastructure platform that supports domestic and international workloads across various compute, storage, and networking needs. We are seeking someone with prior experience deploying and maintaining containerized infrastructure and workloads. Kubernetes competency is highly valued.

Responsibilities

  • Build systems for declarative application and infrastructure lifecycle management, including continuous deployment, continuous integration, Kubernetes cluster management, and service/workload inventory.
  • Prioritize and troubleshoot infrastructure issues, minimizing downtime and responding to alerts efficiently.
  • Contribute to setting the direction of the Site Reliability Engineering (SRE) team, ensuring goals align with Counterpart Health’s company-wide objectives.
  • Foster a collaborative, high-performance culture that promotes motivation, innovation, and cross-disciplinary teamwork.
  • Streamline and automate infrastructure processes, including delivery pipelines and database changes.

Requirements

  • You have 5+ years of programming experience and are proficient in at least one of the following languages: Python, Go, or Shell Scripting.
  • You have in-depth knowledge of containerization technologies and orchestration, such as Docker, Containerd, and Kubernetes, along with experience with CNCF-based technologies like Helm, gRPC, and Prometheus.
  • You have experience with public cloud platforms such as GCP, Azure, or AWS.
  • You are knowledgeable in networking fundamentals, including TCP/IP, UDP, firewalls, routing, DNS, and load balancing.
  • You have experience with Linux system administration and a solid understanding of Linux design principles.
  • You understand key SRE concepts, such as monitoring, performance tuning, and automation.
  • You can work autonomously with limited guidance, proactively identifying and solving problems.
  • You have excellent communication and collaboration skills, with the ability to work effectively with cross-functional teams and adapt to new challenges and evolving technologies.

Benefits Overview

  • Financial Well-Being: Our commitment to attracting and retaining top talent begins with a competitive base salary and equity opportunities. Additionally, we offer a performance-based bonus program, 401k matching, and regular compensation reviews to recognize and reward exceptional contributions.
  • Physical Well-Being
✦ Let us apply for you
We find roles like this and apply on your behalf. Cover letter written for each one. Plans from $15/mo. Cancel anytime.
Get AutoApply
Apply now
Senior Site Reliability Engineer at Cloverhealth — Remote