Job Closed
This listing is no longer active.
Software Development Partner. Result-driven. Quality-obsessed.
Senior DevOps Engineer
Location
Brazil
Posted
142 days ago
Salary
0
Seniority
Senior
Job Description
Senior DevOps Engineer
Dev.Pro
• Manage, scale, and optimize cloud environments used for data science workloads (primarily AWS, Databricks, dbt). • Provision, maintain, and optimize compute clusters for ML workloads (e.g., Kubernetes, ECS/EKS, Databricks, SageMaker). • Implement and maintain high-availability solutions for mission-critical analytics platforms. • Develop CI/CD pipelines for model deployment, infrastructure-as-code (IaC), and automated testing using industry standard toolchains. • Build monitoring, alerting, and logging systems for cloud and ML infrastructure (e.g., Datadog, CloudWatch, Prometheus, Grafana, ELK). • Automate provisioning, configuration, and deployments using tools such as Terraform and CloudFormation, GitHub actions, etc. • Collaborate with Data Engineering to maintain integrations between data pipelines and cloud systems. • Share responsibility for provisioning and operating application networking capabilities that support data platforms, including API gateways, CDNs, application load balancers, TLS, and WAFs. • Conduct periodic risk assessments, best practice reviews, and remediation efforts to strengthen security and resiliency.
Job Requirements
- A bachelor’s degree or higher in a STEM field, required
- 5+ years of experience in cloud operations, DevOps, platform engineering, SRE, sysadmin or related roles.
- Strong proficiency with at least one major cloud provider (AWS preferred).
- Hands-on experience with IaC tools (Terraform, CloudFormation, or similar).
- Strong scripting skills (Python, Bash, or PowerShell).
- Strong understanding of modern authentication and authorization technologies and secrets management (IAM, OIDC, OAuth2, RBAC, ABAC, privileged access management, JIT authorization, PKI).
- Experience with common CI/CD systems (GitHub Actions, Jenkins, GitLab CI, ArgoCD,, or similar).
- Familiarity with container orchestration (Docker Compose, EKS/Kubernetes, ECS).
- Experience supporting data-intensive or ML workloads.
Benefits
- 30 paid days off each year — use them for vacation, holidays, or personal time
- 5 paid sick days, up to 60 days of medical leave, and 6 paid days off for family events like weddings, funerals, or having a baby
- Partially covered health insurance - after probation
- Wellness bonus for gym memberships, sports nutrition, and similar needs
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Senior DevOps – Site Reliability Engineer
nDeavour ConsultingWe are a staffing and IT recruitment company based in Sofia, Bulgaria.
• Build, operate, and evolve all AWS environments (production and non-production), ensuring they meet availability, performance, and recovery requirements. • Own the security posture of the AWS environment. Design and enforce secure patterns for IAM, network segmentation (VPC), and ingress/egress controls. • Participate in an on-call as a primary AWS responder. You will drive technical triage and hands-on remediation for P1/P2 incidents, ensuring clear communication with stakeholders throughout. • Define SLIs/SLOs to measure service health. Proactively identify reliability risks and reduce operational toil through automation and self-healing infrastructure. • Provide hands-on support for ISO 27001 audits. You will manage technical documentation, maintain audit evidence, and ensure operational practices align with security control areas. • Maintain and secure our CI/CD pipelines and own the standards for Infrastructure as Code (IaC) to ensure all changes are controlled and auditable.
• Developing, testing, and distributing changes to automation, software, services, and tools the VHP team is responsible for. • Designing and implementing enhancements to VHP observability infrastructure in order to identify and correct problems before they impact our customers. • Comfortable working in new tooling, code and environments and automating what’s possible. • Create supporting tooling using Ansible or Profiling to aide in performance investigations. • Developing subject matter expertise in various components across our Compute environment. • Collaborating with our support, operations and engineering teams to investigate and troubleshoot complex problems • Participating in on-call rotations, guiding restoration and repair of service-impacting issues
• Collaborate closely with fellow devops engineers and the development team to deploy and maintain application infrastructure. • Assist in the development and support of tooling to streamline the deployment and maintenance of our products. • Work with Github, Jenkins, and Chef to deploy applications from development through to production environments. • Support both in-house and third-party applications, including handling deployments, upgrades, and troubleshooting. • Build and manage automation pipelines for application deployment and maintenance. • Engage in the day-to-day management of Linux servers via the command line • Create monitoring dashboards and alerts in Grafana leveraging Prometheus and Alertmanager. • Document processes and best practices clearly and concisely. • Participate in incident solving on-call rotation
• Design, implementation, and automation of large-scale distributed systems • Build tools and automation that help Five9 achieve higher availability, scalability, latency, and efficiency • Work with Engineering teams to deliver high quality software in a fast-paced environment • Monitor production and development environments to build preventive measures and provide a seamless customer experience • Work with delivery teams on software improvements to achieve higher availability and lower MTTD • Participate in on-call rotation




