This is an exciting opportunity to work on modern cloud security initiatives, protect enterprise-level infrastructure, and collaborate with global teams in a fast-paced and security-focused environment.
Senior DevOps Engineer – Platform Engineering
Location
Latin America (LATAM)
Posted
10 days ago
Salary
0
Seniority
Senior
No structured requirement data.
Job Description
Senior DevOps Engineer – Platform Engineering
Allied Technology Services
Role Description Are you passionate about building scalable platforms, automating software delivery, and improving developer experience? Join us as a Senior DevOps Engineer – Platform Engineering and play a key role in modernizing the software delivery lifecycle for our mission-critical Election Management platforms. This is a highly technical, hands-on engineering role where you'll design and build modern DevOps solutions—not just maintain existing infrastructure. If you enjoy creating CI/CD pipelines, developing automation, implementing Infrastructure as Code, and building engineering platforms that improve reliability and productivity, we'd love to hear from you. - Design and modernize enterprise CI/CD pipelines using Azure DevOps. - Build deployment automation and reusable engineering frameworks. - Implement Infrastructure as Code (Terraform, CloudFormation, or similar). - Improve cloud architecture, scalability, resiliency, and operational efficiency on AWS. - Develop automation using Python, PowerShell, Bash, Go, or similar languages. - Implement observability solutions (monitoring, logging, tracing, dashboards, and alerting). - Support Kubernetes and containerized workloads. - Drive DevSecOps, security automation, and disaster recovery initiatives. - Improve developer experience through self-service platform capabilities. - Collaborate with Engineering, QA, Security, Product, and Architecture teams to establish modern DevOps and Platform Engineering best practices. Qualifications - 8+ years of experience in DevOps, Platform Engineering, Site Reliability Engineering (SRE), or Software Engineering. - Strong hands-on experience with Azure DevOps and AWS. - Expertise in CI/CD, Infrastructure as Code, and cloud-native architectures. - Experience with Kubernetes and containerized applications. - Strong scripting/programming skills (Python, PowerShell, Bash, Go, or similar). - Experience implementing monitoring, observability, and deployment automation. - Excellent problem-solving, communication, and stakeholder management skills. Requirements - Bonus Points - Experience with GitOps (ArgoCD). - Knowledge of Datadog, Grafana, CloudWatch, ELK, or OpenTelemetry. - Experience with DevSecOps and security automation. - Background in AI-assisted operations (AIOps). - Experience with FedRAMP, GovCloud, or regulated cloud environments. - Previous work supporting government, election, or mission-critical platforms. Benefits At CIVIX, you'll help build secure, scalable technology that supports critical public services. You'll work alongside talented engineers, influence platform strategy, and have the opportunity to shape the future of our engineering practices while solving complex technical challenges.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Senior Lead Database Reliability Engineer
DraftKings Inc.Defining what it means to build and deliver the most extraordinary sports & entertainment experiences.The Crown is Yours
At DraftKings, AI is becoming an integral part of both our present and future, powering how work gets done today, guiding smarter decisions, and sparking bold ideas. It's transforming how we enhance customer experiences, streamline operations, and unlock new possibilities. Our teams are energized by innovation and readily embrace emerging technology. We're not waiting for the future to arrive. We're shaping it, one bold step at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Lead Database Reliability Engineer, you'll own the reliability, scalability, and operational excellence of the database infrastructure powering one of the most demanding real time platforms in sports betting and gaming. In this role, you'll combine deep database expertise with software engineering and infrastructure automation to build resilient, self healing systems, improve platform performance, and drive reliability across cloud and on premises database environments. What You'll Do - Drive the technical roadmap for database reliability across PostgreSQL, MySQL, MongoDB, Redis, ScyllaDB, Aerospike, and managed cloud services, while shaping architecture for high availability, replication, partitioning, storage, and connection management. - Design and build automation first database platforms by developing Kubernetes operators, infrastructure as code, GitOps workflows, and production quality tooling in Go or Python to automate provisioning, failover, backups, schema migrations, and lifecycle management. - Lead operational excellence by defining service level objectives, monitoring database health, capacity, and performance, eliminating recurring reliability issues, validating backup and recovery processes, and leading critical production incidents through resolution and continuous improvement. - Optimize database performance and cost across cloud and on premises environments by driving capacity planning, resource efficiency, storage optimization, workload consolidation, and performance tuning for large scale systems. - Partner closely with application engineering teams to establish safe database practices, including schema reviews, migration strategies, query optimization, connection management, and zero downtime deployment processes. - Leverage AI to improve engineering productivity and database operations through intelligent observability, anomaly detection, root cause analysis, documentation, predictive insights, and evaluation of AI generated code to ensure reliability and security. - Mentor engineers across the organization by sharing best practices, leading design and code reviews, influencing technical direction, supporting hiring efforts, and raising the overall maturity of database reliability engineering. What You'll Bring - At least 6 years of experience in Database Reliability Engineering, Database Platform Engineering, or Site Reliability Engineering with a strong database focus, including technical leadership experience delivering complex database infrastructure projects at scale. - Deep expertise in at least one major relational database, preferably PostgreSQL, along with operational experience supporting technologies such as MySQL, MongoDB, Redis, ScyllaDB, Aerospike, Aurora, Cloud SQL, and other managed cloud database services. - Strong experience building and operating stateful workloads on Kubernetes using technologies such as StatefulSets, Persistent Volumes, database operators, Terraform, Pulumi, FluxCD, ArgoCD, GKE, and EKS. - Hands on software development experience using Go or Python to create automation, platform tooling, Kubernetes controllers, APIs, and infrastructure that reduces manual effort and improves engineering efficiency. - A data driven, automation first mindset with proven experience improving reliability through observability, monitoring, service level objectives, capacity planning, performance optimization, and self service engineering solutions. - Practical experience using AI tools such as Claude, GitHub Copilot, Cursor, MCP, or similar technologies to improve design, coding, documentation, troubleshooting, and operational workflows while applying sound engineering judgment to validate AI generated outputs. - Excellent leadership and communication skills with a track record of mentoring engineers, influencing architectural decisions, collaborating across engineering teams, producing clear technical documentation, and driving continuous improvement in highly available production environments. Join Our Team We're a publicly traded (NASDAQ: DKNG) technology company headquartered in Boston. As a regulated gaming company, you may be required to obtain a gaming license issued by the appropriate state agency as a condition of employment. Don't worry, we'll guide you through the process if this is relevant to your role. The US base salary range for this full-time position is 168,000.00 USD - 210,000.00 USD, plus bonus, equity, and benefits as applicable. Our ranges are determined by role, level, and location. The compensation information displayed on each job posting reflects the range for new hire pay rates for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific pay range and how that was determined during the hiring process. It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.
Senior DevOps Engineer
Really Great ReadingOur PK-12 solutions ignite the power of literacy through easy-to-implement programs rooted in the Science of Reading. ✳️
• Ensure the stability, performance, and security of our AWS production environment through end of life • Monitor critical services, respond to incidents, and maintain high availability • Optimize AWS costs thoughtfully during the transition period • Handle day-to-day operations across GCP services — monitoring, security hardening, and observability • Maintain and improve CI/CD pipelines to support reliable, efficient deployments • Manage Infrastructure as Code (Terraform) with a focus on best practices, reviews, and continuous improvement • Support deployments and data migrations as needed • Promote and model best practices in IaC, GitOps, observability, and security • Document processes clearly and support teammates in understanding and using shared tools • Contribute to a culture of operational rigor, async collaboration, and continuous learning
DevOps Engineer
epensionZusammenhalt & Support: Wir sind ein Team, das sich trotz der Distanz gegenseitig unterstützt, Wissen aktiv teilt und sich bei der Lösungsfindung unter die Arme greift. Gelebte Wertschätzung: Ein hilfsbereites Umfeld und echte Wertschätzung füreinander sorgen dafür, dass du dich vom ersten Tag an einbringen kannst. Gemeinsam wachsen: Durch ein abgestimmtes Onboarding mit Mentoren sowie eine offene Feedbackkultur stellen wir sicher, dass wir als Team – trotz Remote-Fokus – eng zusammenwachsen.
Role Description Deine Stimme zählt! Verstärke unser DevOps-Team. Gestalte die zukünftige Infrastruktur aktiv mit und sei prägend im Aufbau der DevOps-Kultur. - Aufbau und Weiterentwicklung unserer Infrastruktur - Etablierung einer DevOps-Kultur - Crossfunktionale Zusammenarbeit und entwicklungsnahe Mitarbeit - Automatisierung und Standardisierung - Containerisierung und Migration - KI-Innovation und Prozessautomatisierung - Monitoring, Logging und Systemtransparenz - Mitgestaltung und Dokumentation - Strukturiertes Onboarding & individuelle Einarbeitung - Stabile Prozesse & moderne Technologien - Abwechslungsreiche Projekte mit technischem Tiefgang - Flexibles Arbeiten mit 100 % Homeoffice-Möglichkeit Qualifications - Mehrjährige Erfahrung im Bereich DevOps, Cloud-Integration oder Infrastruktur-Administration - Sichere Bewegung in Linux-Umgebungen (Ubuntu) mit fundierten Kenntnissen in Server-, Netzwerk- und Systemkonfiguration - Erfahrung mit Ansible, Docker und Proxmox; idealerweise Terraform oder andere IaC-Werkzeuge - Fähigkeit, Public- und Private-Cloud-Architekturen aufzubauen und weiterzuentwickeln - Nutzung von Monitoring- und Logging-Tools wie Grafana, Loki oder LibreNMS - Wohlfühlen in entwicklungsnahen Aufgaben und Verständnis von Kotlin-basiertem Code (inkl. Gradle-Builds) - Bereitschaft, DevOps-Prinzipien im Unternehmen zu verankern - Analytische, strukturierte und lösungsorientierte Arbeitsweise - Schätzung eines Arbeitsumfelds, in dem Ideen eingebracht und technische Entscheidungen mitgestaltet werden können Requirements - Du bringst Freude daran mit, Strukturen aufzubauen, nicht nur zu warten. - Du hast Lust, DevOps-Prinzipien im Unternehmen zu verankern. Benefits - Unbefristete Festanstellung - 30 Tage Urlaub - Flexible Arbeitszeiten - Zuschuss zur Altersvorsorge - Betriebliche Krankenzusatzversicherung - Dienstrad - Corporate Benefits - Essenszuschuss - Bezuschusstes Deutschlandticket - Workation-Möglichkeiten - Unabhängige Beratung in allen Lebenslagen - Weiterentwicklung nach deinen Vorstellungen Company Description Wir entwickeln digitale Lösungen rund um betriebliche Vorsorge. Komplexe Prozesse sollen so gestaltet sein, dass sie für alle Beteiligten – Arbeitgebende, Versicherungen und Beschäftigte – einfach, sicher und effizient ablaufen. Damit das gelingt, bauen wir unsere Infrastruktur und unsere Arbeitsweise konsequent weiter aus – hin zu einer modernen, Cloud-basierten und teamübergreifend automatisierten Plattform. Dafür suchen wir Menschen, die Lust haben, sich aktiv einzubringen, Verantwortung zu übernehmen und die DevOps-Kultur bei epension gemeinsam mit uns zu etablieren. Wir freuen uns darauf, dich kennenzulernen!
Systems Reliability Engineer
Bright Vision TechnologiesBright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications.
Role Description We are seeking an experienced Site Reliability Engineer to ensure the availability, performance, and operational excellence of large-scale distributed systems in production. As an SRE you will live at the boundary between development and operations, applying strong software engineering principles to infrastructure and operations problems, and continually pushing the platform toward higher reliability with lower operational toil. The ideal candidate will combine deep systems knowledge with strong programming skills, a measurement-driven mindset, and the discipline to design, automate, and operate complex services so that reliability becomes a first-class engineering deliverable rather than a reactive concern. Key Responsibilities - Define, instrument, and continually refine service-level objectives (SLOs), service-level indicators (SLIs), and error budgets for critical services. - Lead incident response and resolution for production issues, acting as a calm and effective incident commander when needed. - Ensure high-quality post-incident reviews that drive lasting improvements. - Design and implement comprehensive monitoring, logging, and tracing strategies using tools like Prometheus, Grafana, OpenTelemetry, ELK/EFK, Datadog, or similar. - Build and maintain robust on-call processes, runbooks, and escalation paths. - Automate operational toil aggressively by writing production-grade tooling in Python, Go, Bash, or similar languages. - Architect and operate large-scale Kubernetes clusters and container-based workloads. - Design CI/CD pipelines that promote safe, frequent, and observable releases. - Lead capacity planning and performance engineering activities. - Partner closely with application development teams to embed reliability practices early in design. - Strengthen the platform’s resiliency through chaos engineering and fault injection. - Drive continuous improvement of security posture in collaboration with security teams. - Contribute to the technical roadmap for reliability tooling and observability platforms. - Mentor engineers across the organization on SRE practices. Qualifications - Bachelor’s degree in Computer Science, Engineering, or a related technical discipline. - Five or more years of SRE, DevOps, or production engineering experience supporting large-scale distributed systems. - Strong programming skills in at least one of Python, Go, or Java. - Deep, hands-on experience operating Linux at scale. - Production experience operating Kubernetes and container-based workloads. - Strong working knowledge of observability tooling such as Prometheus, Grafana, OpenTelemetry, ELK/EFK, or commercial equivalents. - Hands-on experience designing and operating CI/CD pipelines. - Solid understanding of distributed system design. - Demonstrated experience leading incident response and conducting effective post-incident reviews. - Excellent communication and documentation skills. Preferred Qualifications - Experience defining and operationalizing SLOs and error budgets in real production environments. - Exposure to chaos engineering practices and tools such as Chaos Monkey, Gremlin, or Litmus. - Hands-on experience with at least one major cloud platform (AWS, Azure, or GCP). - Background in capacity planning, performance engineering, or large-scale load testing. - Familiarity with service mesh technologies such as Istio, Linkerd, or Consul. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 650-6699. Learn more about Bright Vision Technologies at www.bvteck.com . Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

