Jobgether
Senior DevOps & Infrastructure Engineer
engineeringfull-timeBrazil
SALARY
Not listed
WORK TYPE
remote
JOB TYPE
full-time
INDUSTRY
general
✦ AutoApply Sick of applying? We apply to roles like this for you, up to 20 a month.
Learn more
About the role
Accountabilities
- Own and continuously improve production cloud infrastructure across GCP, AWS, and Azure, focusing on reliability, scalability, performance, capacity, observability, and cost efficiency.
- Participate in a global on-call rotation, lead incident response, troubleshoot complex production issues, conduct root-cause analysis, and drive lasting remediation.
- Build and maintain developer platforms, release pipelines, deployment automation, and engineering tooling using technologies such as GitHub Actions, Google Cloud Build, Jenkins, and related CI/CD solutions.
- Design, maintain, and improve secure Terraform-based infrastructure as code, including reusable modules and automated infrastructure workflows.
- Develop automation and operational tooling using Python, Go, Shell, or comparable programming and scripting languages.
- Operate production Kubernetes and GKE environments, covering deployment, scaling, networking, security, observability, and troubleshooting.
- Design and troubleshoot cloud networking components including VPCs, firewalls, load balancers, routing, DNS, and WAFs.
- Partner with security and engineering teams to implement secure-by-default infrastructure, covering identity and access management, network security, secrets management, container security, and infrastructure protection.
- Use AI-assisted engineering tools to support log analysis, debugging, troubleshooting, optimization, and operational workflows.
- Collaborate with application teams on scalable architectures, deployments, production challenges, and operational best practices.
- Mentor fellow engineers and lead initiatives that raise the standards for infrastructure reliability, security, automation, and developer productivity.
- 5+ years of professional experience in cloud infrastructure, DevOps, SRE, platform engineering, or a related discipline, including hands-on experience operating GCP production environments.
- Strong practical experience with Terraform/IaC, CI/CD, and production Kubernetes, using tools such as GitHub Actions, Jenkins, Google Cloud Build, Cloud Deploy, or equivalent technologies.
- Solid understanding of cloud architecture, distributed systems, and networking, including VPCs, firewalls, load balancers, routing, DNS, and WAFs.
- Experience securing cloud environments, preferably within GCP, and applying infrastructure security and operational best practices.
- Strong programming or scripting capabilities in Python, Go, Shell, or equivalent languages.
- Demonstrated experience responding to production incidents, troubleshooting complex infrastructure problems, and performing root-cause analysis.
- Strong ownership, communication, analytical thinking, and problem-solving skills, with the ability to work independently in a distributed environment.
- Willingness and ability to participate in a shared global on-call rotation supporting infrastructure availability during scheduled periods.
- Experience independently operating AWS or Azure environments is a strong advantage.
- Familiarity with Ansible, Chef, or Puppet is a plus.
- Experience securing containers across Docker, Kubernetes, or GKE is desirable.
- Knowledge of advanced GCP security controls such as Organization Policies, VPC Service Controls, Access Context Manager, and Binary Authorization is beneficial.
- Relevant cloud certifications, particularly Google Cloud Professional certifications, are an advantage.
- A proactive, client-focused, agile, and AI-forward mindset, with an interest in using emerging technologies to improve engineering productivity.
- Opportunity to work at the forefront of AI and contribute to infrastructure supporting advanced AI development.
- Exposure to cutting-edge AI research, datasets, reinforcement learning environments, and evaluation benchmarks.
- Opportunity to contribute to work showcased at leading AI and technology conferences.
- Experience applying frontier AI innovation to real-world enterprise challenges.
- Collaboration with highly experienced colleagues from leading global technology companies.
- Startup-style pace, ownership, and the opportunity to make a meaningful impact.
- Remote working opportunity for candidates based in Brazil.
- Global, distributed working environment with opportunities to collaborate internationally.
- Strong focus on learning, innovation, and continuous professional development.
- Inclusive workplace that values diverse perspectives, backgrounds, and authentic contributions.
Requirements
Benefits
✦ Sick of applying to 40 jobs a month?
I rewrite your resume for ATS by hand first. Once you sign off on it, AutoApply applies to up to 20 roles like this a month, cover letter in your own voice each time. From $14.99/mo, cancel anytime.
Get AutoApply