Wolters Kluwer logo
Wolters Kluwer

When you have to be right

Lead DevOps Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteLeadTeam 10,001+Since 1836H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

10 days ago

Salary

$116.4K - $204.1K / year

Seniority

Lead

No structured requirement data.

Job Description

Lead DevOps Engineer

Wolters Kluwer

Role Description As a Lead DevOps Engineer, you will be responsible for leading the design and implementation of the infrastructure and CI/CD pipelines for our applications, as well as maintaining and improving the reliability of our existing systems. You will work alongside the Wolters Kluwer Product Teams to successfully implement and bring any learnings through continuous improvements. You will work closely with the application development teams to ensure that our services and solutions are delivered efficiently and with the highest level of quality. In addition, you will be responsible for mentoring a team of DevOps engineers, fostering a culture of collaboration with continuous learning. - Design, implement and automate the creation of Cloud Infrastructure in Azure & AWS using Infrastructure as Code tools such as Terraform, Ansible and Jenkins. - Provide leadership, experience and mentoring on the adoption and usage of Infrastructure as Code automation across Service Operations and Delivery and as well as the broader organization. - Build and help transform towards container and serverless compute architectures, such as AKS, EKS, Docker and Azure Functions. - Build and help maintain CI/CD pipelines and orchestrations across multiple environments. - Advise, implement and support the adoption of Cloud Native solutions as appropriate to the situation and needs. - Build effective Observability of managed solutions, including troubleshooting issues as they arise. - Ensure the infrastructure, automation, monitoring, and CI/CD solutions for FCC/DXG are compatible with the broader organizational DevOps and Security patterns and processes. - Mentor a team of DevOps engineers, fostering a culture of collaboration and continuous learning. - Explore new technologies, development patterns, and partake in pilots/POC/technology evaluations. - Conduct work activities using SRE principles, such as resilience, metrics, capacity planning, toil, Incident management and security. - Provide input into architecture and engineering standards. - Deploy and maintain critical applications. - Provide input into designing new systems and modernize existing applications. - Identify areas in need of engineering improvements and propose replacement strategies. - Resolve Production Incidents and contribute to root cause analysis. - Improve knowledge sharing across the organization on automation best practices. - Have on call responsibilities in rotation within the Service Delivery and Operations team. Qualifications - Bachelor's Engineering and a Master's degree in Computer Science or a related field. - Minimum 8 years of software related experience required, with a mixture of Site Reliability, DevOps, or Release Engineering experience with Software Engineering experience in backend services, middleware or data systems. - Strong background in software development, with experience in languages such as Python, .NET or Java. - Possess recent experience with Cloud infrastructure (Azure and AWS). - Proficient in Shell Scripting Languages (PowerShell & Bash). - Understanding of the core Azure/AWS services (PaaS, IaaS, SaaS). - Proficiency with automation and source configuration management tools such as Git. - Experience with infrastructure as code tools such as Terraform or CloudFormation. - Experience with continuous integration and deployment tools such as Azure DevOps, Jenkins or Travis CI. - Experience in leading a team of DevOps engineers. - Able to clearly explain technical projects to diverse staff to gain their support. - Strong attention to detail with excellent problem-solving skills. - Experience and deep commitment to the transformation to a DevOps culture focusing on continuous integration – full lifecycle of building, automated and performance testing, and automating deployment. - Experience building platforms for Observability and defining standard operating procedures. - Experience with diagnostic tools and script automation of Windows and Linux. Benefits - A comprehensive benefits package that begins your first day of employment. - Medical, Dental, & Vision Plans. - 401(k), FSA/HSA, Commuter Benefits. - Tuition Assistance Plan. - Vacation and Sick Time. - Paid Parental Leave. - Full details of our benefits are available - Wolters Kluwer Benefits .

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Fonction publique Territoriale logo

DevOps Development Engineer – Service Manager

Fonction publique Territoriale

Vision stratégique et capacité d’analyse; Rigueur et sens de l’organisation; Pédagogie et capacité d’accompagnement des services; Capacité à travailler en transversalité; Force de proposition.

DevOps Engineer11 days ago

Role Description Située au cœur de la Région Grand Est, la Métropole du Grand Nancy regroupe 20 communes et offre un environnement de travail stimulant et varié. Avec ses 1680 agents, elle joue un rôle clé dans le développement urbain, l’aménagement du territoire et l’amélioration du quotidien des habitants. En rejoignant cette collectivité dynamique, vous aurez l’opportunité de participer à des projets innovants au service des citoyens, tout en évoluant dans un cadre professionnel collaboratif et enrichissant. La diversité des missions et l'engagement pour un avenir durable font de la Métropole du Grand Nancy un lieu où chaque talent peut s’épanouir. Le périmètre de gestion de la DSIT comprend 660 instances applicatives, au bénéfice de 5 000 utilisateurs. L’enjeu porté par ce recrutement est de poursuivre et d’accroître la stratégie d’amélioration continue, d’efficience et de souveraineté numérique de la direction : - Développements sur mesure d’applications et d’interfaces, internalisés ou sous-traités. - Définition et mise en place de développements s’appuyant sur des moteurs d’intelligence artificielle. - Mise en place et administration d’infrastructures dédiées. - Tutorat technique des processus d’automatisation. - Vous définissez et pilotez la feuille de route des développements sur mesure, des solutions d'IA et NoCode, en cohérence avec les orientations de la DSIT et les besoins des métiers. - Vous concevez et supervisez le développement de solutions innovantes (applications, API, automatisations, outils d'IA) en privilégiant des technologies souveraines et open source. - Vous encadrez et accompagnez une équipe de développeurs, favorisez sa montée en compétences et coordonnez les interventions des prestataires. - Vous garantissez la qualité, la sécurité et la pérennité des développements, en lien étroit avec les équipes Infrastructures et le RSSI. - Vous apportez votre expertise sur les sujets DevOps, IA et NoCode, assurez une veille technologique active et contribuez à l'évolution des pratiques de la collectivité. - Vous participez aux instances de pilotage de la DSIT afin d'accompagner la transformation numérique et de répondre aux enjeux des collectivités adhérentes. Qualifications - Vous justifiez d'une expérience en développement logiciel et/ou DevOps. Une expérience en administration systèmes constitue un atout. - Vous maîtrisez le développement en Python et êtes à l'aise avec les environnements DevOps (CI/CD, conteneurs, automatisation des déploiements). - Vous avez une appétence pour les solutions d'IA (n8n), le NoCode et les technologies open source, avec une sensibilité aux enjeux de souveraineté des données. - Vous connaissez les bonnes pratiques en matière de sécurité des applications et de protection des données. - Une expérience dans le secteur public ou dans un environnement fortement réglementé serait appréciée. Requirements - Vous savez expliquer des sujets techniques à des interlocuteurs non spécialistes et accompagner le changement. - Vous appréciez le travail en équipe et savez collaborer avec des partenaires internes comme externes. - Vous faites preuve d'écoute, d'esprit d'analyse et savez proposer des solutions adaptées aux besoins des utilisateurs. - Vous êtes autonome, force de proposition et savez vous adapter à un environnement en constante évolution. Benefits - La métropole du Grand Nancy mène une politique active en faveur du recrutement et du maintien dans l’emploi pour les personnes en situation de handicap. - A ce titre, les candidats ayant la reconnaissance de Bénéficiaire de l’Obligation d’Emploi des Travailleurs Handicapés (RQTH, AAH…) peuvent candidater sous réserve de la compatibilité à la tenue du poste.

France
Pear Tree. logo

Senior DevOps

Pear Tree.

Hire smarter, hire globally — scale your business while saving up to 80% on local costs. www.pear-tree.com

DevOps Engineer11 days ago
Full TimeRemoteTeam 1-10H1B No Sponsor

**Responsibilities:** - Maintain and support a Kubernetes-based hosting platform and AWS infrastructure for Drupal and Laravel applications - Migrate, configure, and optimize client applications into Kubernetes environments - Monitor platform health and respond to outages, incidents, and security events with a calm and methodical approach - Execute deployments and go-live activities while maintaining and improving CI/CD pipelines - Manage identity and access across Google Workspace, Atlassian, hosting platforms, and version control systems - Support employee onboarding and offboarding, including device management and access provisioning - Contribute to ISO 27001 and SOC 2 compliance by maintaining security controls, documentation, and audit readiness - Build, maintain, and improve monitoring, logging, and alerting systems to proactively identify issues - Manage DNS, domain records, and SSL/TLS certificate lifecycles for client and internal environments - Perform backup validation, disaster recovery testing, and recovery drills to ensure business continuity - Assist in managing cloud infrastructure costs, SaaS subscriptions, vendor relationships, and renewals - Drive automation initiatives by replacing manual operational tasks with scripted workflows and leveraging AI where appropriate - Collaborate with developers and platform engineers to continuously improve infrastructure reliability, security, and deployment processes - Stay up to date with DevOps best practices, cloud technologies, automation tools, and emerging industry trends

Philippines
$2.7K - $3.5K / month
Onebrief logo

Senior Site Reliability Engineer (Arlington, VA) - Relocation Provided

Onebrief

Software for rapid military planning: make planning fast enough for today's environment

DevOps Engineer11 days ago
Full TimeRemoteTeam 1-10H1B No Sponsor

Consequential Work. Dedicated People. About Onebrief Onebrief builds collaboration and AI-powered workflow software for military planning and operational coordination. Today, many critical planning workflows still rely on fragmented systems, static documents, and disconnected tools that make collaboration and decision-making unnecessarily difficult. Onebrief brings modern software, AI, and real-time collaboration into those environments, helping teams operate with greater clarity, coordination, and adaptability in situations where decisions carry real-world consequences. We are a distributed team of builders from military, operational, and technology backgrounds who care deeply about improving how important work gets done. Some team members work remotely, while others work directly alongside customers in operational environments around the world. Founded in 2019, Onebrief is backed by leading investors including General Catalyst, Battery Ventures, Insight Partners, Sapphire Ventures, and Human Capital. Valued at more than $2 billion, we continue to invest in product innovation, AI capabilities, and team growth. Security Clearance, Location, and Onsite Notice:This role requires regularly working on-site at customer locations in Arlington, VA. If you are not currently within commuting distance, you must be willing to relocate (note that Onebrief will provide relocation assistance). Active Top Secret Clearance required with the ability to obtain SCI eligibility. About The RoleWe are hiring a Site Reliability Engineer to join our Infrastructure & Security team. You’ll work closely with fellow SREs, security, and customer success. You will be the first line of support for our mission critical deployments, and responsible for ensuring best-in-class service quality and issue resolution. You will work in both on-premise DoD environments and AWS cloud environments. Your lessons from the field will shape how our team works, from policy to implementation. In addition to working at the customer, you will contribute directly to solutions that increase stability, performance, and security of our deployments, and improve the overall experience of deploying and managing Onebrief on premise. About YouYou care deeply about reliability and treat it as a core feature of any application or platform, with a bias toward “reliability over novelty.” You think about infrastructure and operability as products to be automated, well-documented, and continuously improved, and you aim to leave systems easier to operate than you found them. You are equally comfortable leading a post-incident review, or diving into a kubectl shell to triage a complex production issue. You don't just fix problems; you translate constraints and failure modes into clear, automated guardrails and scalable, resilient architecture. For you, robust monitoring, actionable alerting, and insightful runbooks are core parts of the engineering process, not afterthoughts. You mentor others, fostering a culture of blameless postmortems and proactive reliability. You collaborate naturally with application and platform teams, helping them move quickly but safely by building the tools, processes, and observability that make "fast recovery" a reality. What You'll DoYou'll own the reliability, scalability, and security of the production application and/or platform. You will do this by: - Implementing a World-Class Observability Platform: Design, implement, and manage our monitoring, logging, and alerting stack (e.g., Prometheus, Loki, Alloy, and Grafana). You won't just track metrics; you'll create the actionable insights and automated alerting that allow teams to identify and resolve issues before they impact users. - Defining and Upholding Reliability: Define, measure, and own alerting that feeds into our Service Level Indicators (SLIs) and Service Level Objectives (SLOs), increasing trust internally and externally. You will be the organization's expert on what it means for our systems to be reliable and how to measure it. - Leading Incident Response: Act as the incident responder and potentially incident commander during critical incidents who will lead blameless post-mortems / After Action Reviews (AARs) that identify true root causes and drive automated, long-term solutions to prevent recurrence. - Automating for Scale and Security: Partner with platform engineers to design, build, and manage secure, resilient Kubernetes clusters and cloud/on-prem environments using Infrastructure-as-Code (Terraform, Ansible). You will embed security and compliance controls (RMF, STIGs) directly into this automation. - Eliminating Toil and Scaling the Team: Proactively identify and eliminate operational toil by building automation. You will partner with other teams to share best practices for air-gapped environments and support their readiness for production. What We Look For - An active Top Secret clearance - 5+ years in Platform, DevOps, or Site Reliability Engineering with an infrastructure and operations focus. - Proven partner to DevOps/Platform and application teams; collaborates well across functions and shares context openly. - A deep understanding of incident response processes, with experience conducting thorough root cause analyses and driving continuous improvement. Technical expertise - Infrastructure as Code: Terraform (or CloudFormation), Ansible. - Containers and orchestration: Kubernetes design, deployment, and operations. - CI/CD: experience building and maintaining pipelines (GitLab CI/CD, Jenkins, GitHub Actions). - Scripting: proficiency with at least one of Python, Go, or Bash. - Cloud: Familiarity with AWS or AWS GovCloud. - Observability: Grafana stack, ELK stack, or Datadog. - Networking fundamentals: core protocols and secure configurations. Bonus points (nice to have) - Experience in DoD environments and compliance frameworks (RMF, STIGs, ICD 503). - GitOps practices and toolchains. - Security‑minded design for sensitive environments. - Experience designing and implementing meaningful SLIs/SLOs (including error budgets) for complex, distributed systems. - Familiarity with on‑prem virtualization(VMware, Proxmox, Nutanix, Hyper-V, etc). - Service mesh exposure (Istio, Linkerd). - Relevant certifications (e.g., AWS DevOps Engineer, CKA/CKAD). - Active Security+ or another DoD 8570.01-approved security credential, or the ability to obtain the valid credentials within 3 months of employment. Notice to Third Party Recruitment Agencies Please note that Onebrief does not accept unsolicited resumes from recruiters or employment agencies. In the absence of an executed Recruitment Services Agreement, there will be no obligation to any referral compensation or recruiter fee. In the event a recruiter or agency submits a resume or candidate without an agreement Onebrief explicitly reserves the right to pursue and hire those candidate(s) without any financial obligation to the recruiter or agency. Any unsolicited resumes, including those submitted to hiring managers, shall be deemed the property of Onebrief.

United States
$180K - $220K / year

GCP DevOps Engineer

Ruri Software Technologies LLC

MAGIC serves as the centralized ERP system for over 120 State agencies, supporting: Financial Management Procurement Grants Management Reporting These modules have been in production since 2015. The State is also currently implementing SAP SuccessFactors for: Human Capital Management (HCM) Payroll Processing

DevOps Engineer11 days ago

Role Description We are seeking a GCP DevOps Engineer for a 12+ month project with VZ. The ideal candidate will have extensive experience in building pipelines and I/CD (Continuous Integration and Continuous Delivery). - Experience in build automation tools like Jenkins, Docker, Kubernetes. - Expert in using different source code version control tools like Git. - Installation and configuration of DNS, Load Balancer, SSL, HTTP, FTP, and TCP/IP. - Experience with containers and orchestration services like Kubernetes and Docker. - Experience in public cloud architecture and engineering in Google Cloud Platform (GCP). - Experience in configuration, deployment, management, and maintenance of large cloud-hosted systems, including scaling, monitoring, performance tuning, troubleshooting, and disaster recovery. - Experience with Linux/UNIX environments and scripting for build & release automation. - Application deployments and environment configuration using Terraform and Ansible. - Involved in installing, configuring, and administering Hadoop clusters. - Hadoop administration, Spark tuning, data rebalancing, and tuning MapReduce and Spark jobs. - Extensively worked with Dataproc, GCS, Dataflow, GKE, BigQuery, JupyterLab, Jupyter Notebook, Vertex AI, and Domino. Qualifications - Proven experience in DevOps practices and methodologies. - Strong understanding of cloud architecture and services. - Familiarity with containerization and orchestration tools. Requirements - Minimum of 3 years of experience in a DevOps role. - Strong knowledge of Google Cloud Platform. - Experience with automation and configuration management tools. Benefits - Competitive salary. - Flexible working hours. - Opportunity for professional development.

United States