Job Closed
This listing is no longer active.
Empowering people to understand and improve their heart health using technology and behavioral science.
Cloud Infrastructure Engineer
Location
United States
Posted
68 days ago
Salary
$150K - $170K / year
Seniority
Senior
Job Description
Cloud Infrastructure Engineer
Hello Heart
• Build, maintain, and scale production-ready cloud infrastructure across AWS and Kubernetes. • Support the development of machine learning pipelines and a full data lake architecture. • Improve build automation processes and help move the team from continuous integration to continuous delivery. • Secure, scale, and operate Kafka clusters on Kubernetes. • Partner with Engineering and Security teams to improve infrastructure reliability, security, and compliance. • Develop dashboards, alerts, internal tools, and response processes to identify and address security and reliability risks. • Improve logging, monitoring, and observability across production systems. • Support containerized application deployments using Docker and Kubernetes. • Help evaluate and adopt new tools that improve developer productivity, system reliability, and infrastructure scalability.
Job Requirements
- 5+ years of experience as a DevOps Engineer, Cloud Infrastructure Engineer, or similar infrastructure-focused engineering role.
- Ideally located in a Central or Eastern Time Zone US state, as well as previous experience working with globally-distributed teams across multiple time zones.
- Production experience with Kubernetes and AWS, or a similar major cloud provider, is required.
- Production experience building and running Docker containers is required.
- Programming experience with at least one of the following languages: Scala, Java, or Python.
- Experience with logging and monitoring stacks such as ELK, Prometheus, or similar tools.
- Experience with continuous integration and continuous delivery processes and tools.
- Deep understanding of production systems, scalable architecture, reliability, and security best practices.
- Experience deploying and managing Kafka clusters and Kafka streams is a strong advantage.
- Experience using AI coding tools, such as Claude Code, is a strong advantage.
- Independent, curious, and excited to learn and apply new technologies in a fast-paced environment.
Benefits
- Health insurance
- 401(k) matching
- Flexible work hours
- Paid time off
- Remote work options
Related Guides
Related Categories
Related Job Pages
More Infrastructure Engineer Jobs
Infrastructure Engineer – AI Platform
OpenVPN Inc.OpenVPN® helps businesses of all sizes create secure, virtualized, reliable networks that scale with your team.
• Own the rollout and operational management of AI-assisted development tools across engineering (e.g., Cursor, Copilot, Claude Code) • Define and implement access controls, license management, and usage policies that satisfy SOC2/ISO 27001 requirements • Build cost tracking and reporting so leadership has visibility into AI tool spend and usage patterns across the org • Reduce friction for engineers adopting these tools while maintaining security and auditability • Partner with teams across the org to identify, build, and support internal AI applications such as RAG pipelines, agents, and automation workflows • Evaluate and recommend tooling, frameworks, and patterns based on what teams actually need • Define where IaaS’s responsibility ends and consuming teams’ begins • Advise on data governance policies for LLM usage, including what data can go into which models, where outputs are stored, and how audit trails are maintained • Ensure AI infrastructure and tooling meets existing SOC2 and ISO 27001 controls and can be evidenced in audits • Provide leadership with clear, regular reporting on AI adoption, cost, risk, and usage across the org • Stand up and manage AI/ML infrastructure, primarily on GCP (Vertex AI) within OpenVPN’s existing environment • Design the Terraform modules and IaC patterns for AI infrastructure that follow the team’s existing conventions (e.g., Atlantis-driven GitOps workflows) • Build visibility into AI/ML infrastructure costs and implement controls consistent with how compute costs are managed elsewhere • Evaluate build-vs-buy decisions for AI/ML infrastructure components and managed services with an eye toward operational fit within existing patterns
Infrastructure Engineer – AI Platform
OpenVPN Inc.OpenVPN® helps businesses of all sizes create secure, virtualized, reliable networks that scale with your team.
• Own the rollout and operational management of AI-assisted development tools across engineering (e.g., Cursor, Copilot, Claude Code) • Define and implement access controls, license management, and usage policies that satisfy SOC2/ISO 27001 requirements • Build cost tracking and reporting so leadership has visibility into AI tool spend and usage patterns across the org • Reduce friction for engineers adopting these tools while maintaining security and auditability • Partner with teams across the org to identify, build, and support internal AI applications such as RAG pipelines, agents, and automation workflows • Evaluate and recommend tooling, frameworks, and patterns based on what teams actually need • Define where IaaS’s responsibility ends and consuming teams’ begins – this boundary doesn’t exist yet; you’ll help draw it • Advise on data governance policies for LLM usage, including what data can go into which models, where outputs are stored, and how audit trails are maintained • Ensure AI infrastructure and tooling meets existing SOC2 and ISO 27001 controls and can be evidenced in audits • Provide leadership with clear, regular reporting on AI adoption, cost, risk, and usage across the org • Stand up and manage AI/ML infrastructure, primarily on GCP (Vertex AI) within OpenVPN’s existing environment • Design the Terraform modules and IaC patterns for AI infrastructure that follow the team’s existing conventions (e.g., Atlantis-driven GitOps workflows) • Build visibility into AI/ML infrastructure costs and implement controls (spot instances, auto-scaling policies, idle resource cleanup) consistent with how compute costs are managed elsewhere • Evaluate build-vs-buy decisions for AI/ML infrastructure components and managed services with an eye toward operational fit within existing patterns
• Diseñar, implementar, administrar y soportar ambientes Oracle Cloud Infrastructure (OCI) • Gestionar servicios de networking, cómputo, almacenamiento y seguridad en Oracle Cloud • Monitorear y optimizar ambientes OCI en entornos productivos y no productivos • Atender incidentes, troubleshooting y actividades de soporte de infraestructura • Colaborar con equipos de arquitectura y desarrollo en soluciones cloud escalables • Participar en proyectos de modernización tecnológica y migración a la nube • Documentar procedimientos técnicos y estándares operativos • Garantizar disponibilidad, rendimiento y seguridad de las plataformas Oracle.
Senior Cloud Infrastructure Engineer
DragosDragos is a computer and network security company specializing in industrial cybersecurity, incident response, threat intelligence, and security software. Past
Role Description Dragos is seeking a Senior Cloud Infrastructure Engineer to join our Core Infrastructure Engineering team. This fully remote role is ideal for an experienced engineer passionate about Infrastructure as Code (IaC), cloud automation, and production reliability. - Design, build, and maintain scalable cloud infrastructure services in AWS and GCP - Contribute production-quality Go and Python code to existing cloud services - Develop and own automation and software deployment pipelines with maximum efficiency - Implement Infrastructure as Code (IaC) practices using Terraform, Ansible, Packer, or similar tools - Maintain and improve our custom RHEL-based OS platform - Design and enhance secure, reliable field software update capabilities - Manage production software release pipelines, ensuring stability and efficiency - Maintain system security plans, ensuring compliance with best practices for security, observability, reliability, and performance - Configure and optimize cloud networking components (firewalls, VPNs, load balancers, routing tables, etc.) - Support and improve internal DevOps tooling for engineering teams - Monitor and enhance service availability, ensuring high uptime and resilience - Collaborate with engineering teams to implement observability, logging, and security best practices Qualifications - 5-10 years of experience working with containerized applications (Docker, Kubernetes, etc.) - Strong proficiency in Go and/or Python - Experience provisioning virtual machines and infrastructure automation with Terraform, Packer, Ansible, or similar tools - Hands-on experience deploying and managing cloud-based infrastructure in AWS or GCP - Experience with maintaining OS package mirrors, package deployment infrastructure, and custom DEB or RPM packaging (preferred) - Strong knowledge of production software release pipeline and asset management concepts - Familiarity with cloud security best practices and the ability to author customer-facing security documentation - Experience maintaining system security plans for cloud services and their dependencies - Proven track record of building and maintaining cloud-based or virtualized appliance clusters for high-availability (HA) - Strong ownership mindset, accountability for production stability, and ability to thrive in a fast-paced environment - Excellent problem-solving, communication, and collaboration skills. Requirements - Salary: $165,000 - Competitive Equity Package - Comprehensive Benefits Plan Company Description Dragos is on a relentless mission to defend industrial organizations that provide us with the necessities of modern civilization; running water, functioning electricity, and safe industrial working environments. As the market leader in ICS/OT Cybersecurity, we are dedicated to arming our customers with best-in-class technology, threat intelligence, and services to protect their systems as effectively and efficiently as possible. We’re a remote-first culture with operations in North America, Europe, the Middle East, and APAC. We’re looking for mission-oriented teammates who embody our core values of authenticity, transparency, and trust. Are you ready to make a difference? Come join a mission that can save the world!



