Job Closed
This listing is no longer active.
#sejaSysMap #SysMap #soulSysMap
Site Reliability Engineer – Senior
Location
Brazil
Posted
80 days ago
Salary
0
Seniority
Senior
Job Description
Site Reliability Engineer – Senior
SysMap Solutions
- Garantir a disponibilidade, escalabilidade e desempenho das aplicações e serviços; - Implementar e evoluir práticas de **observabilidade**, incluindo métricas, logs e traces; - Criar e manter **dashboards (dashs)** para acompanhamento de indicadores de saúde dos sistemas; - Definir e gerenciar **alarmísticas**, com foco em alertas eficientes e redução de ruído; - Atuar na identificação e resolução de incidentes, realizando análise de causa raiz (RCA); - Trabalhar em conjunto com times de desenvolvimento para melhoria contínua (DevOps); - Automatizar rotinas operacionais e processos de monitoração; - Apoiar a definição e acompanhamento de SLIs, SLOs e SLAs; - Contribuir para a cultura de confiabilidade e engenharia de resiliência.
Job Requirements
- Experiência com práticas de **SRE/DevOps**;
- Conhecimento sólido em **observabilidade** (monitoramento, logging e tracing);
- Experiência na construção de **dashboards e visualização de dados operacionais**;
- Experiência com **gestão de alertas (alarmísticas)**;
- Vivência com ferramentas de monitoramento e observabilidade, como:
- Elastic Stack (Elasticsearch, Logstash, Kibana);
- Datadog;
- Splunk;
- Dynatrace;
- Conhecimento em ambientes cloud (AWS, Azure ou GCP);
- Experiência com automação (Python, Shell Script ou similares);
- Conhecimento em sistemas Linux e redes;
- Experiência com containers e orquestração (Docker/Kubernetes);
- Experiência com ferramentas de APM (Application Performance Monitoring);
- Conhecimento em infraestrutura como código (Terraform, CloudFormation);
- Experiência com pipelines CI/CD;
- Conhecimento em práticas de Chaos Engineering;
- Certificações em cloud ou SRE;
- Experiência com cultura de observabilidade orientada a negócio.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Analyze, design, develop and implement high-quality software solutions; • Work closely with the team ensuring continuous integration and continuous delivery (CI/CD); • Actively participate in agile ceremonies, ensuring alignment between technical objectives and business goals; • Develop and maintain applications using Microsoft technologies; • Build and evolve APIs, WebServices and system integrations; • Work on automation of pipelines and DevOps processes; • Identify, analyze and resolve technical and operational issues; • Ensure application stability, scalability and performance; • Contribute to the continuous improvement of development and deployment processes; • Work with cloud solutions, preferably Azure; • Use code assistants and AI to accelerate feature development; • Create effective prompts for code generation, refactoring and documentation; • Critically validate code and suggestions generated by AI before approval; • Ensure best practices for security and data privacy when using public and private AI
DevOps Engineer
NateraWe are a global leader in cell-free DNA (cfDNA) testing, dedicated to oncology, women’s health, and organ health.
• Designing, building and maintaining CI/CD pipelines using GitHub Actions to automate builds, testing, and deployment across development and production environments. • Implementing infrastructure-as-code solutions with Terraform to support repeatable, version-controlled, and auditable infrastructure deployments. • Automating deployment and configuration of applications, bioinformatics pipelines, and supporting services across both GCP and AWS. • Building and maintaining automation scripts in Python and Shell/Bash to reduce manual operations, improve deployment reliability, and increase developer velocity. • Managing containerized applications using Docker, Kubernetes, Google Cloud Run, and AWS container services. • Configuring and maintaining monitoring, logging, and altering solutions to provide full visibility into application and pipeline health. • Implementing security controls across CI/CD pipelines - including secrets management, access controls, and vulnerability scanning - with a regulated-environment mindset. • Managing configuration management for application environments and bioinformatics workstations. • Creating clear, comprehensive documentation for deployment processes, runbooks, and troubleshooting guides that support regulatory compliance requirements. • Driving continuous improvement initiatives to reduce deployment friction, improve build times, and make developers more productive. • Mentoring and providing technical guidance to development teams on DevOps best practices, cloud-native patterns, and tooling.
Cloud Engineer/DevOps Engineer
CapcoCapco, a Wipro company, is a management & technology consultancy dedicated to the financial services & energy industries
• Deliver cloud computing services in cloud environments such as Oracle Cloud. • Provide easy-to-consume services for software development teams. • Accelerate application development by creating infrastructure service patterns.
Senior DevOps, Site Reliability Engineer
AppspaceDiscover the easiest way to reach your workforce - at work, at home, or on the go.
• Be the technical anchor for a global platform footprint that includes a mix of Azure IaaS/PaaS, Google Cloud Platform (GCP), Kubernetes, and various data platforms. • Identify manual "toil" and replace it with automated workflows for monitoring, change management, and routine administration of large-scale VM environments to ensure a positive ROI. • Lead the integration of AI tools for automated code reviews, development frameworks, and predictive log analysis to drive departmental velocity and efficiency. • Design and maintain "self-service" deployment frameworks and CI/CD pipelines (GitHub Actions, Bamboo) using Infrastructure as Code (Bicep, Terraform). • Evaluate platform components to determine the most cost-effective path: automating the current state or migrating features to modern, shared architectures. • Design and maintain a comprehensive observability stack across Azure and GCP (metrics, logs, traces) to identify performance bottlenecks and proactively address system defects. • Partner with engineering, security and operations teams to ensure new features are "born" with reliability, security and automated delivery in mind; Ensure adherence to security best practices and compliance standards (SOC2, HIPAA, ISO 27001) and operational excellence with cost efficiency. • Investigate complex performance defects by following log trails across web, application, and database tiers (SQL Server, MongoDB, MySQL). • Ensure all platforms meet security standards (SOC2, HIPAA, ISO 27001) through automated policy enforcement across Azure and GCP.




