Job Closed
This listing is no longer active.
O Sucesso de sua empresa ao seu alcance!
SRE Analyst, Junior
Location
Brazil
Posted
128 days ago
Salary
0
Seniority
Junior
Job Description
SRE Analyst, Junior
Addvisor Group
• The analyst will be focused on Linux server administration with an emphasis on security.
Job Requirements
- Experience administering Linux servers
- Knowledge of WebLogic and Tomcat
- At least 1 year of experience with DevOps practices
- Familiarity with compute resources using VMware
- Monitoring and logging: Experience with Prometheus & Grafana for monitoring and visualizing metrics
- Knowledge of Ansible
- LPI/RHCSA certification: desirable
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Lead the end-to-end migration from Azure DevOps to GitHub Enterprise. • Design and implement reusable GitHub Actions workflows. • Drive the implementation and adoption of GitHub Advanced Security. • Define and enforce CI/CD standards. • Act as the technical lead for AWS modernization initiatives. • Design and maintain Terraform code for AWS infrastructure. • Support container engineering practices including Helm. • Contribute to serverless architectures using AWS Lambda. • Embed security and compliance controls into delivery pipelines. • Work closely with platform, security, cloud, and application teams. • Mentor junior engineers.
• Support migration activities from Azure DevOps to GitHub Enterprise, including repository organization, workflow conversion, validation, and documentation. • Contribute to GitHub repository administration, branching strategies, pull requests, and code review processes. • Create, update, and maintain GitHub Actions workflows for CI/CD tasks such as testing, building, packaging, and deployment. • Assist with Docker image creation, pushing images to Amazon ECR, and deployments to Amazon EKS under senior guidance. • Contribute to AWS Lambda development in Python or Node.js, including integration with API Gateway and event-driven services such as SQS and EventBridge. • Contribute to Terraform modules for AWS infrastructure and automation, including services such as S3, IAM, Lambda, and container-related resources. • Review and triage GitHub Advanced Security findings, including CodeQL, Dependabot alerts, and secret scanning, with support from senior engineers and security teams. • Participate in automated testing, quality gates, and CI/CD best practices. • Create and maintain technical documentation related to pipelines, standards, and migration activities. • Actively use GenAI tools such as GitHub Copilot, Claude, and ChatGPT to support engineering work. • Participate in Agile ceremonies and contribute to the broader cloud modernization effort.
• Own and evolve our infrastructure, reliability, and deployment practices • Build the foundational platform that enables our engineering teams to ship quickly and reliably • Design, implement, and maintain our AWS cloud infrastructure using infrastructure-as-code principles with Terraform • Build and optimize CI/CD pipelines to enable rapid, safe deployments across multiple environments • Own observability strategy—implement comprehensive monitoring, logging, and alerting systems using Datadog and other tooling • Architect and manage containerized workloads on ECS Fargate and evaluate migration paths to Kubernetes • Establish and enforce security best practices, working closely with compliance teams on financial services requirements • Design and implement disaster recovery, backup, and business continuity strategies • Optimize system performance, cost efficiency, and resource utilization across AWS services • Collaborate with engineering teams to improve service reliability, reduce toil, and establish SLOs/SLIs • Participate in incident response and conduct thorough post-mortems to drive continuous improvement • Mentor engineers on DevOps practices, cloud architecture patterns, and operational excellence
• Diagnose complex, intermittent, and high-impact issues to maintain system stability • Research and utilize advanced diagnostic tools to troubleshoot ongoing customer issues within live production environments • Identify single points of failure in the architecture to re-design systems for maximum redundancy and auto-recovery • Analyze application source code in Java and Angular to identify memory leaks, race conditions, or inefficient logic • Propose and implement code fixes directly to improve long-term system reliability rather than simply filing bug tickets • Adjust kernel parameters and network stack configurations to optimize low-level system performance • Build internal tooling to empower other engineering teams to self-serve their infrastructure needs • Develop high-quality automation to ensure that manually solved problems are never repeated • Tweak database queries and application thread pools to tune the performance of the entire software stack • Serve as a critical member of the on-call rotation to respond to and mitigate major system outages • Lead incident command efforts during high-pressure situations to restore service and protect critical data flows • Conduct post-incident reviews to convert outages into actionable architectural improvements



