Jobgether
Jobgether

SRE and Devops Team Manager

engineeringfull-timeIndia
SALARY
Not listed
WORK TYPE
remote
JOB TYPE
full-time
INDUSTRY
general
Apply for this position
✦ AutoApply Let us apply to roles like this on your behalf.
Learn more

About the role

Accountabilities:

    • Lead, mentor, and develop a team of SRE and DevOps engineers, working with engineering leadership to assess capabilities, define team needs, and support individual growth.
    • Own the end-to-end incident management process, including on-call rotations, escalation procedures, incident command, severity frameworks, blameless postmortems, root cause analysis, and preventative actions.
    • Establish and promote SRE principles across services, including SLIs, SLOs, error budgets, capacity planning, reliability standards, and observability practices.
    • Drive operational excellence by improving processes, automation, reliability, security, and cost efficiency while balancing operational priorities with feature delivery.
    • Coordinate a distributed support and engineering model across IST and US time zones, ensuring effective handoffs, coverage, communication, and escalation.
    • Partner with engineering leadership to prioritize infrastructure investments and align platform initiatives with business, security, compliance, and operational objectives.
    • Monitor and report on reliability metrics, team health, operational performance, incident trends, and progress against reliability goals.
    • Guide day-to-day technical priorities and remain hands-on enough to support the team through complex infrastructure and operational challenges.
    • Contribute to roadmap planning, department-wide initiatives, and the continuous improvement of the SRE/DevOps function.
    • Promote a culture of accountability, ownership, continuous learning, and proactive reliability engineering across the team.
    • Requirements:

      • 2–3 years of experience as a Team Lead or Manager within an SRE, DevOps, Platform Engineering, or similar infrastructure environment.
      • Proven experience managing and coordinating distributed teams across IST and US time zones, including on-call coverage, operational handoffs, and cross-regional collaboration.
      • Strong experience managing incident response processes, including on-call operations, escalation paths, incident command, severity frameworks, blameless postmortems, and RCA follow-through.
      • Strong hands-on expertise with cloud infrastructure and multi-cloud environments, including AWS, Kubernetes, GKE, EKS, managed databases, object storage, networking, and IAM.
      • Solid experience with Infrastructure-as-Code technologies such as Terraform, Pulumi, or equivalent tools.
      • Experience with modern CI/CD platforms and practices, including GitHub Actions, Jenkins, or similar technologies.
      • Strong understanding of observability across metrics, logging, tracing, monitoring, and alerting.
      • Experience establishing or scaling SRE/DevOps functions, on-call programs, reliability practices, or operational processes is highly desirable.
      • Experience in FinTech or another regulated environment, particularly where security, risk, and compliance requirements are important, is a plus.
      • Experience improving database and data infrastructure operations, as well as observability across multiple cloud technologies, is advantageous.
      • Strong understanding of designing cloud environments for cost efficiency, resilience, scalability, security, and high availability.
      • Excellent leadership, communication, prioritization, and problem-solving skills, with the ability to operate effectively in a distributed and fast-moving environment.
      • A hands-on mindset with the ability to balance technical depth, people leadership, strategic planning, and day-to-day operational demands.
      • Benefits:

        • 100% remote position based in India.
        • Opportunity to lead and develop an SRE/DevOps team within a high-growth SaaS environment.
        • Hands-on exposure to multi-cloud infrastructure, Kubernetes, automation, observability, and modern reliability engineering practices.
        • Opportunity to shape SRE processes, incident management, on-call practices, and operational standards.
        • Collaboration with engineering leadership and globally distributed technical teams.
        • Exposure to security, compliance, resilience, and cost optimization within a technology-driven financial services environment.
        • Career development opportunities through technical leadership, team management, and cross-functional initiatives.
        • Flexible collaboration across Indian and US time zones within a distributed working environment.
✦ Let us apply for you
We find roles like this and apply on your behalf. Cover letter written for each one. Plans from $15/mo. Cancel anytime.
Get AutoApply
Apply now
SRE and Devops Team Manager at Jobgether — Remote