Senior DevOps Engineer – Infra Services
About the role
We are seeking a highly experienced Senior DevOps Engineer – Infra Services with strong operational expertise in enterprise DNS platforms to implement and support our current and future DNS & Kubernetes architecture in private cloud. You will be directly responsible for ensuring the scalability, resilience, automation, and performance optimization of our DNS & Kubernetes environment by leveraging Infrastructure-as-Code (IaC) tools like Ansible, Python and Terraform
This is a deeply technical role requiring expert-level understanding of DNS, Kubernetes and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications.
Key Responsibilities
- Implement and support large scale DNS Architecture that stretches private & public cloud across multiple regions.
- Implement and support Infoblox DDI solutions (DNS, IPAM, DHCP), Cloudflare DNS management, Google & AWS DNS services.
- Implement and support multi-tenant Kubernetes clusters leveraging BareMetal servers and Software defined storage protocols.
- Implement Infrastructure as code leveraging CI/CD automation, Ansible, Terraform/Pulumi, Python and Bash for automated provisioning of DNS services.
- Implement continuous delivery pipelines for Infrastructure updates, including patch management, service upgrade testing, and rollback procedures.
- Develop automated monitoring, alerting, and healing mechanisms using GitOps principles and observability stacks (e.g., Prometheus, Loki, Grafana).
- Harden services for high availability, disaster recovery, and scale-out operations.
- Perform deep-dive troubleshooting and performance analysis of Infrastructure services across hypervisors, backend storage protocols and networking layers.
- Be involved in Change management and Global team collaboration
- Participate in on-call rotation, incident response, and root cause analysis for platform reliability issues.
Minimum Qualifications
- 5+ in managing global DNS & Kubernetes on Linux Operating Systems.
- Expertise with Infoblox DDI systems (or similar enterprise DNS systems), Bind9, CloudFlare, GCP and AWS DNS services.
- Expert knowledge in managing Kubernetes and Ubuntu Linux systems.
- Proven experience automating infrastructure processes and deployments using Ansible, Terraform, Python, and CI/CD pipelines, with demonstrated ability to drive and scale a culture of automation.
- Strong Linux (RHEL/CentOS/Ubuntu) systems engineering background with advanced scripting in Python, Bash, or
- Familiar with various Infrastructure technology stacks like Hypervisor Technologies (KVM, Vsphere,), storage protocols (SDS) (iSCSI, NFS, CEPH), L2/L3 Networking
- Strong experience supporting mission critical 24x7 customer facing production environments.
- Ability to write technical documentation and contribute to community wikis or knowledge bas