Verity Group logo
Verity Group

Somos Humanos. Somos Digitais. Somos Verity!

Senior SRE / DevOps Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteSeniorTeam 51-200Since 2010H1B No SponsorCompany SiteLinkedIn

Location

Brazil

Posted

1 day ago

Salary

0

Seniority

Senior

Job Description

Senior SRE / DevOps Engineer

Verity Group

• Design, implement and evolve CI/CD pipelines • Provision and maintain cloud infrastructure using GCP, AWS and Azure • Automate infrastructure using Terraform and Ansible • Operate, administer and scale Kubernetes (GKE) and Docker environments • Define, implement and track reliability metrics and practices such as SLIs, SLOs, SLAs, MTTR and MTTD • Build observability, monitoring, alerting and APM • Work with monitoring and observability tools such as Dynatrace, Datadog, Grafana, Prometheus and the ELK Stack (Elasticsearch and Kibana) • Monitor performance and availability indicators including Latency, Traffic, Errors and Saturation • Collaborate closely with squads, promoting platform engineering best practices, automation and reliability

Job Requirements

  • Hands-on experience with Cloud environments (GCP, AWS and/or Azure)
  • Strong knowledge of Kubernetes (GKE) and Docker
  • Solid experience with Terraform and Ansible
  • Experience with CI/CD pipelines (GitLab CI, GitHub Actions, Jenkins or similar)
  • Experience with observability and monitoring tools
  • Knowledge of SRE metrics and indicators (SLI, SLO, SLA, MTTR and MTTD)
  • Linux systems administration
  • Experience working in agile environments (Scrum/Kanban)

Benefits

  • Meal voucher
  • Food allowance (grocery voucher)
  • Home office allowance
  • Medical insurance
  • Dental insurance
  • Life insurance
  • Discount partnerships
  • Agreements with establishments and educational institutions
  • Recurring agility trainings
  • Alura licenses
  • Intervalo Verity
  • #VerityComVocê
  • Viva Engage

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Grafana Labs logo

Senior Software Engineer – Databases, SRE

Grafana Labs

Grafana Labs supports organizations’ monitoring, visualization and observability goals. 950,000+ active installations

Full TimeRemoteTeam 501-1,000Since 2014H1B Sponsor

• Partner closely with product engineering squads (embedded model) • Own production reliability for high-SLA and complex customer environments • Design and implement automation to scale our reliability practices • Ensuring our customers meet our SLO targets • Define and evolve per-tenant SLOs and reliability models • Proactively reduce SLO burn to prevent repeat incidents • Serving as a primary escalation point and on-call for relevant incidents • Lead customer-impacting incident response and post-incident reviews • Contribute to design docs and code reviews • Influence feature design to ensure production scalability and operability • Build automation to eliminate toil where needed • Improve alert quality and reduce noisy escalations

Canada
$164.5K - $197.4K / year

Role Description Seeking a senior-level DevSecOps Engineer with strong developer experience supporting containerized applications in an AWS cloud environment. Ideal candidate will: - Identify, prove out, test, and implement pipeline process improvements - including utilizing new or existing GitLab or third-party tools or features. - Exhibit excellent customer service skills to educate customers on pipeline functionality and troubleshoot pipeline issues. - Interact with all levels of organizational personnel, from training and mentoring teammates to developing and presenting briefs to leadership. Essential Functions: - Develop and implement cloud-native DevSecOps functionality for CI/CD pipeline solutions in AWS; improve and maintain GitLab pipeline configurations. - Understand/interpret Cyber Security guidelines to resolve code vulnerabilities. - Maintain, monitor, and proactively research/pilot solutions to optimize and improve the parent scan pipeline; provide recommendations for technology advancement to streamline CI/CD tools and processes. - Assist with GitLab upgrades as received from the vendor (i.e. bi-weekly, monthly, etc.; requires evening support). - Design and build secure, scalable, and automated container environments using Amazon ECS and EKS. - Configure customer projects/access to pipeline, including configuring git on customer assets and credentials in customer repositories. - Onboard new applications/customers to the CI/CD environment, working closely with application developers to provide training/technical guidance/troubleshooting assistance. - Craft/help customers craft gitlab-ci.yaml files to orchestrate their child pipelines or test projects; create, maintain, update, and monitor health checks. - Research, perform analysis of alternatives, recommend technical solutions, and architect new CI/CD environments and pipelines, such as Cloud migration or supporting classified systems. - Provide demos/overviews/briefs/newsletters regarding the pipeline and/or associated tools to customers and/or leadership of various levels. - Update and maintain documentation. - Provide mentorship to junior teammates. - Other duties as assigned or required. Qualifications - CI/CD implementation experience required. - Design/development of DevSecOps pipelines experience required. - Proficient communication and documentation skills with experience preparing technical guidance, how-to instructions, test plans, demos, and/or presenting training to internal and external stakeholders required. - Experience with Amazon ECS and EKS required. - CompTIA Sec+ certification is required and must show proof before interview. - BS/BA degree and 10 years related experience OR AA/AS degree and 14 years related experience OR HS and 16 years related experience. Requirements - Experience with specific CI/CD related tools such as GitLab Ultimate, Nexus, DORA metrics, and Prisma Cloud (formerly Twistlock) highly desired. - Experience working in a DoD environment highly desired. - Experience with OpenShift and Nexus is a plus. - Ability to work independently in a fast-paced technical environment. Benefits - Health Care Plan (Medical, Dental & Vision) - Retirement Plan (401k, IRA) - Life Insurance (Basic, Voluntary & AD&D) - Paid Time Off (Vacation, Sick & Public Holidays) - Short Term & Long Term Disability - Training & Development - Wellness Resources - Stock Option Benefit

United States
$140K - $170K / year
IRIUM logo

Ingeniero/a Cloud DevOps

IRIUM

Líderes en gestión de servicios integrados de infraestructuras y plataformas IT.

Full TimeRemoteTeam 501-1,000Since 2002H1B No Sponsor

• Colaborar en un proyecto en modalidad full-remote. • Diseñar y mantener pipelines CI/CD en Azure DevOps. • Implementar automatizaciones con scripting de PowerShell. • Administrar y operar en entornos Windows Server.

Spain
€33K - €40K / year
Full TimeRemoteTeam 5,001-10,000H1B Sponsor

Role Description Do you want to shape reliability practices for a new AI inference platform? Are you a senior technical leader who drives solutions across teams? Join the Akamai Inference Cloud Team! The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications. In this role, you'll lead reliability workstreams for Akamai's serverless inference platform, design SRE tooling and automation, and drive technical decisions. Opportunities exist to mentor other SREs, influence architecture decisions with product engineering teams, and shape SRE practices for AI inference workloads and GPU infrastructure at scale. As a Senior II Site Reliability Engineer, you will be responsible for: - Taking ownership of observability strategy for the serverless inference platform, designing telemetry, dashboards, and alerts, defining SLO/SLI frameworks, and driving improvements when targets are missed. - Building production-grade automation and tooling that reduces operational toil, improves incident response, and sets patterns that other SREs adopt. - Owning incident management integration for inference workloads, designing frameworks, leading incident response during on-call rotations, and driving systemic improvements from post-mortems. - Defining and implementing deployment safety practices including progressive rollouts, canary analysis, and rollback automation, establishing standards for the team. - Partnering with product engineering teams to influence architecture decisions, ensure operational readiness, and represent the SRE perspective in design reviews. - Mentoring Senior and mid-level SREs through code reviews, design discussions, and hands-on problem-solving. Qualifications - 8+ years of experience in SRE, infrastructure engineering, or platform engineering, working with large-scale distributed systems. - Possess a proven track record of defining SLO/SLI frameworks, building observability platforms, and running incident management processes at scale. - Have extensive Kubernetes and containerization experience at scale, including autoscaling, resource scheduling, and container orchestration for compute-intensive workloads. - Have experience building automation and tooling in Python or Go, with familiarity in CI/CD pipelines, deployment safety, and infrastructure-as-code. - Possess the ability to lead technical initiatives across teams, mentor other engineers, and drive complex reliability problems to resolution independently. - Have experience with or exposure to AI/ML infrastructure, model serving, or GPU workloads. Benefits - We support your health, well-being, finances, and life beyond work. - FlexBase adapts to your job's needs. - Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. - We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both. Compensation Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $146,400 - $263,600/year; a candidate’s salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.

United States
$146.4K - $263.6K / year