Technical DevOps / Infrastructure Engineer
Location
Mexico
Posted
97 days ago
Salary
0
Seniority
Mid Level
No structured requirement data.
Job Description
Technical DevOps / Infrastructure Engineer
Deep Origin
Role Description We are looking for a Senior DevOps / Infrastructure Engineer to join our existing DevOps team. This is a senior IC role with a broad technical scope: - Own complex initiatives end-to-end. - Drive collaboration across engineering and science teams. - Set a high bar for how we build and operate infrastructure. - Support R&D teams by running and evolving compute clusters that power bioinformatics pipelines, ML training, and other HPC workloads. Highly autonomous: able to operate with minimal guidance, prioritize work independently, and take full ownership of infrastructure decisions and outcomes. Qualifications - 10+ years of infrastructure and DevOps engineering experience, with a proven track record in senior or lead IC roles. - Ability to take end-to-end ownership of complex, multi-team initiatives and drive them from design through to production. - Hands-on experience running HPC or research compute clusters: bare-metal provisioning, Slurm (or equivalent), GPU infrastructure, and shared storage (NFS, Lustre, or similar). - Comfortable operating in environments with a mix of cloud, VPS, and bare-metal systems, including legacy or non-standard setups. - Experience supporting scientific or R&D teams with mixed workloads: long-running CPU batch jobs, GPU training jobs, and interactive compute. - Deep, hands-on AWS expertise: EKS/Kubernetes, IAM, VPC networking, S3, RDS, and cost management. - Solid Terraform skills and a principled approach to infrastructure-as-code. - Strong Linux fundamentals and experience managing multi-node environments at scale. - Experience owning and improving production observability systems (Prometheus/Grafana, OpenTelemetry, ELK, or similar). - Strong security fundamentals: threat modeling, least-privilege access design, vulnerability management, and compliance frameworks. - Experience owning incident management end-to-end, including process design and continuous improvement. - Excellent communication skills; able to work directly with researchers and scientists as well as with engineering and leadership. - Fluent English. Requirements - Background in biotech, bioinformatics, or scientific computing environments (Nice-to-Have). - SOC 2 Type II audit experience (Nice-to-Have). - Monorepo tooling and developer platform engineering (Nice-to-Have). Key Responsibilities - Own our cloud infrastructure across AWS and third-party hosting and compute providers; ensure it is reliable, scalable, and cost-efficient. - Own and operate bare-metal compute clusters: node provisioning, configuration management, networking, secure access, and ongoing reliability. - Build and maintain configuration management using Ansible (or similar), ensuring reproducible and scalable server provisioning. - Set up and maintain Slurm for job scheduling across CPU and GPU node pools; ensure researchers can submit, monitor, and manage jobs without DevOps involvement. - Design and manage cluster networking: management and storage networks, inter-node communication, DNS, and secure perimeter access, including bastion/jump host setup. - Deep hands-on experience managing Linux-based infrastructure, including networking, firewalls, VPNs, and performance tuning in distributed environments. - Own disaster recovery and business continuity: define RTO/RPO targets, maintain runbooks, and run regular tests. - Manage and optimize infrastructure spend through capacity planning, right-sizing, and intelligent use of reserved and spot capacity. - Manage Kubernetes clusters, networking, and workload scheduling across cloud and on-premise environments. - Enable infrastructure-as-code practices in Terraform; drive consistency, modularity, and auditability across the codebase. - Evolve our observability platform: improve coverage, reduce alert noise, and ensure engineering teams have the visibility they need to detect and resolve issues quickly. - Own security posture across the platform: IAM policies, secrets management, network segmentation, vulnerability management, and SOC 2 compliance. - Lead incident management: on-call processes, escalation policies, runbooks, and blameless post-mortems. - Drive CI/CD improvements and developer workflow initiatives that meaningfully increase engineering throughput. - Evolve internal tooling and CLI infrastructure that engineering teams depend on daily. Values & Working Style - Ownership mindset — you take responsibility from A to Z. - Comfortable navigating ambiguity in a fast-moving startup environment. - Clear communicator who can collaborate across technical and non-technical teams. - Pragmatic problem solver focused on impact. Why This Role Matters Now As we scale our AI platform and expand into new initiatives, engineering velocity and platform reliability directly impact research outcomes and product milestones. This role plays a key part in strengthening our technical foundation during the growth phase.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
DevOps Engineer
AXAAXA is a leading provider of wealth management, financial protection, and global insurance services for millions of clients nationwide. A French, multinational
Role Description Buscamos un Ingeniero DevOps que nos ayude a acompañar a equipos internacionales en la adopción de las mejores prácticas de DevOps. Dentro del equipo DevSecOps de AXA Partner, reportarás directamente al Líder de DevOps. Tus principales responsabilidades serán: - Mantener las plataformas DevOps para garantizar la continuidad del negocio. - Asesorar y capacitar a los equipos de desarrollo en la implementación de prácticas DevSecOps. - Gestionar la configuración de la plataforma Azure DevOps. - Ayudar a definir, promover y aplicar los estándares DevSecOps en todos los socios de AXA. - Colaborar con otros equipos transversales (Servicios en la Nube, Arquitectura, etc.) para alcanzar acuerdos y establecer puntos en común. - Interactuar con ingenieros DevOps y líderes técnicos de diferentes regiones de socios de AXA. - Trabajar en diversas iniciativas y proyectos (Simplificación e Industrialización). Qualifications - Al menos 5 años de experiencia en ingeniería de software. - Experiencia en Azure DevOps. - Experiencia en OpenShift. - Experiencia en Scripting (PowerShell, Bash o TypeScript). - Prácticas DevSecOps (Estrategias de CI/CD y estrategias de ramificación). - Experiencia en SAST. - Experiencia en IaC. - Mentalidad práctica y disposición para aprender. - Enfoque proactivo y colaborativo. - Capacidad para trabajar con equipos remotos. - Fuertes habilidades de atención al cliente. - Capacidad para adaptarse rápidamente a los cambios. - Dominio del español y del inglés. Application Process Para aplicar, haz clic en el botón ‘aplicar ahora’, luego necesitarás iniciar sesión o crear un perfil para enviar tu CV. Estamos orgullosos de ser un empleador que brinda igualdad de oportunidades y de realizar un proceso de reclutamiento justo, transparente e inclusivo. Si tienes una condición de salud o discapacidad que requiera un ajuste durante el proceso de aplicación o de reclutamiento, por favor, envía un correo electrónico a AXA Partners Global HR Response - globalhr@partners.axa. Company Description Somos AXA Partners: expertos en el diseño y la prestación de soluciones de asistencia y de seguros de vida, de crédito y especializados - en conjunto con y para nuestros socios en todo el mundo. - La experiencia y la pasión de nuestros más de 8500 empleados. - Una sólida red de más de 55.000 profesionales en todo el mundo. - Innovación tecnológica en el sector. - Compromiso con la protección del medio ambiente: plantamos un árbol por cada nuevo recluta (con un contrato permanente).
Principal DevOps Engineer
CES Family of CompaniesThe CES Family of Companies is a collection of strong brands and businesses providing food equipment, supplies, service.
• Design, develop, and maintain advanced Azure DevOps YAML pipelines for CI/CD automation. • Write and maintain robust PowerShell scripts for automation, monitoring, and deployment tasks. • Develop and manage infrastructure as code (IaC) using Bicep for Azure resource provisioning. • Operationalize GitHub Copilot to ground code generation, improving speed and accuracy of feature delivery. • Take complete ownership of assigned tasks, including requirement analysis, implementation, documentation, and support. • Collaborate with cross-functional teams to understand deployment needs and deliver scalable DevOps solutions. • Proactively monitor and support development, testing, and production environments, ensuring high availability and minimal downtime. • Troubleshoot and resolve issues across the DevOps lifecycle, including build failures, deployment errors, and infrastructure problems. • Continuously improve DevOps practices, tools, and processes to enhance team productivity and system reliability. • Monitor and optimize infrastructure performance, cost, and security. • Mentor junior engineers and contribute to knowledge sharing within the team.
• Use your experience to develop, maintain and implement Kubernetes clusters at scale that are ready for heterogeneous and elastic workloads • Work closely with Data Science and Data Engineering teams to implement, optimize and scale systems on Kubernetes using CI/CD, automation tools and scripting languages • Help Data Science and Data Engineering develop and implement specialized infrastructure to deliver tools, software, and platforms that improve the reliability and scalability of capabilities • Pioneer, implement, and encourage best practices for software deployment and code management through automation and education • Monitor the system and respond to incidents to maintain system SLO/SLA, review and follow up production incidents • Provide on call work as needed
Test-Focused DevOps Engineer
Booz Allen HamiltonBooz Allen Hamilton is an award-winning provider of strategic innovation, management consulting, technology, and engineering services. Founded in 1914, the comp
Title: Test-Focused DevOps Engineer Location: Chantilly, VA Work Type: Hybrid, Full Time Job ID: R0238293 Job Description: The Opportunity: Join our team as a DevOps Engineer with a focus on test and test systems, responsible for designing, implementing, and maintaining the testing infrastructure and processes that ensure the quality and reliability of software systems. You'll work closely with internal and external development and operations teams to create automated testing frameworks, manage test environments, and optimize testing workflows. How You'll Contribute: - Test Automation: Design and implement automated testing frameworks using tools such as Selenium, Appium, or Cypress, and integrate them into CI/CD pipelines - Test Environment Management: Manage and maintain test environments, ensuring they are up-to-date, scalable, and aligned with mission relevant environments - Test Data Management: Develop strategies for test data management, including data generation, masking and synchronization - CI/CD Pipelines: Collaborate with development teams in integrate testing into CI/CD pipelines, ensuring that tests are executed automatically and feedback in provided quickly - Test Infrastructure: Design and maintain the infrastructure required for testing, such as test servers, databases, and network configurations - Monitoring and Reporting: Implement monitoring and reporting tools to track test results, identify trends, and optimize testing processes - Collaboration and Communication: Work closely with development and operations teams to ensure that testing is aligned with mission objectives and that feedback is provided in a timely manner Join us. The world can't wait. You Have: - Experience with programming languages, such as Python, Java, or C#, and with testing frameworks and tools - Experience with test automation tools, such as Selenium, Appium, or Cypress, and test automation frameworks - Knowledge of testing methodologies, including black-box, white-box, and gray-box testing, and testing types, such as unit testing, integration testing, and end-to-end testing - Knowledge of the software development lifecycle, including Agile methodologies - Ability to analyze and troubleshoot complex technical issues related to testing - Ability to learn new testing technologies and adapt to changing project requirements - TS/SCI clearance - HS diploma or GED Nice If You Have: - Possession of excellent interpersonal skills to facilitate collaboration between teams and stakeholders - Possession of excellent problem-solving skills Clearance: Applicants selected will be subject to a security investigation and may need to meet eligibility requirements for access to classified information; TS/SCI clearance is required. Compensation At Booz Allen, we celebrate your contributions, provide you with opportunities and choices, and support your total well-being. Our offerings include health, life, disability, financial, and retirement benefits, as well as paid leave, professional development, tuition assistance, work-life programs, and dependent care. Our recognition awards program acknowledges employees for exceptional performance and superior demonstration of our values. Full-time and part-time employees working at least 20 hours a week on a regular basis are eligible to participate in Booz Allen's benefit programs. Individuals that do not meet the threshold are only eligible for select offerings, not inclusive of health benefits. We encourage you to learn more about our total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page. Salary at Booz Allen is determined by various factors, including but not limited to location, the individual's particular combination of education, knowledge, skills, competencies, and experience, as well as contract-specific affordability and organizational requirements. The projected compensation range for this position is $77,600.00 to $176,000.00 (annualized USD). The estimate displayed represents the typical salary range for this position and is just one component of Booz Allen's total compensation package for employees. Identity Statement As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud. Candidate AI Usage Policy AI is a part of our daily work at Booz Allen, and we are committed to the responsible and ethical use of AI tools. However, we want to ensure a fair candidate process based on your own skills and knowledge. As part of this commitment, the use of artificial intelligence (AI) or other tools to assist with responses during interviews (whether in-person or virtual) is prohibited unless permission is explicitly provided. Work Model Our people-first culture prioritizes the benefits of collaboration, whether it occurs in person or virtually. To support engagement and effective communication, employees working virtually are generally expected to have their cameras on during meetings. - Remote: If this position is listed as remote, there may still be occasions when you are required to work in person at a Booz Allen or customer facility. - Hybrid: If this position is listed as hybrid, you will be expected to work from a Booz Allen facility frequently, in alignment with leadership expectations and the needs of the role. You may also be required to work from or visit a customer facility. - Onsite: If this position is listed as onsite, work will primarily be performed at a Booz Allen office or customer facility, where employees will collaborate directly with colleagues and customers as required by the role. Commitment to Non-Discrimination All qualified applicants will receive consideration for employment without regard to disability, status as a protected veteran or any other status protected by applicable federal, state, local, or international law.


