We are looking for a Mid-Level or Senior SRE Analyst to work in a highly complex environment, contributing directly to the availability, reliability, performance, and evolution of technology environments.
If you have experience with AWS, Cloud, CI/CD, Observability, and Troubleshooting, this opportunity may be a great fit for you.
What will you do?
- Support and evolve AWS environments, ensuring high availability and reliability;
- Implement and maintain CI/CD pipelines, integrated with Git and automation tools;
- Develop and operate solutions using AWS Lambda, EKS, ECS, and EC2;
- Perform in-depth analysis (drill down) using observability and monitoring tools, identifying root causes and working on incident resolution;
- Administer Linux and Windows environments, applying security and automation best practices;
- Participate in incident response, root cause analysis (RCA), and postmortem processes;
- Work closely with Development and Operations teams, strengthening the DevOps/SRE culture.
Technical Requirements:
- Hands-on experience with AWS, especially Lambda, EKS, ECS, and EC2;
- Knowledge of CI/CD pipelines and version control using Git;
- Experience with Observability and Monitoring, using tools such as CloudWatch, Prometheus, Grafana, ELK, or similar solutions;
- Knowledge of Linux and Windows administration;
- Familiarity with automation tools and Infrastructure as Code (IaC), such as Terraform, Ansible, or equivalent solutions;
- Experience with advanced troubleshooting and the ability to perform in-depth analysis in complex environments;
- Analytical, collaborative, and problem-solving-oriented mindset.
What we’re looking for:
We are looking for professionals who take ownership of technology environments, take initiative in incident resolution, and are interested in contributing to the continuous improvement of reliability, automation, observability, and performance.