DevOps Engineer Remote Jobs in Colorado (US)
This page tracks remote devops engineer openings that are location-eligible for Colorado.
This page tracks remote devops engineer openings that are location-eligible for Colorado.
Open jobs
2,185
Hiring companies this week
10
Salary sample
$112,700 - $150,000
Jobs added last hour
0
2185 Jobs
1367 Companies
Role Description As a DevOps Engineer at Fullthrottle.ai, you will work directly under the guidance of our DevSecOps Manager, contributing to the secure, reliable, and scalable delivery of our SaaS platform. You’ll collaborate closely with development, operations, and business teams, helping to embed security, compliance, reliability, and automation into every stage of our software delivery lifecycle. This is a hands-on role requiring ownership, technical curiosity, strong troubleshooting skills, and a proactive approach to problem-solving in a fast-moving startup environment. Key Responsibilities: - Platform Operations: Own the operational health, deployment, monitoring, maintenance, and continuous improvement of our AWS-based SaaS platform. Ensure systems remain secure, reliable, and scalable. - CI/CD Pipeline Development: Build, optimize, and maintain CI/CD pipelines to enable rapid, secure software delivery. Partner with development teams to improve deployment processes and developer productivity. - Security & Compliance: Implement and maintain security controls in line with SOC2 and other compliance requirements, working closely with the DevSecOps Manager. Support vulnerability remediation activities, access management, audit requests, and security monitoring initiatives. - Incident Response: Participate in incident response, business recovery planning, root cause analysis, and post-incident reviews. Help develop corrective and preventative actions to improve platform resilience. - Operational Excellence: Troubleshoot complex infrastructure, networking, application, and CI/CD issues. Demonstrate the ability to identify root causes, implement solutions, and prevent recurring problems. - Documentation & Knowledge Sharing: Create and maintain runbooks, recovery procedures, operational documentation, architectural diagrams, and technical knowledge articles to support team scalability and business continuity. - Collaboration: Work cross-functionally with engineering, operations, product, and business teams to deliver seamless, secure solutions. - Continuous Improvement: Identify opportunities to improve automation, reliability, security, observability, and operational efficiency across our infrastructure. Qualifications - Experience: 4+ years in a DevOps, Site Reliability, Platform Engineering, or Cloud Infrastructure role, preferably in a SaaS or cloud-native environment. Experience supporting production AWS workloads is strongly preferred. - AWS Skills: Hands-on experience with core AWS services (EC2, ECS, S3, VPC, IAM, Lambda, RDS) and familiarity with AWS security tools (GuardDuty, Inspector, Config, CloudTrail). - DevOps Methodologies: Solid understanding of DevOps principles, secure software development practices, release management, and CI/CD pipeline best practices. - Infrastructure as Code: Strong hands-on Terraform experience required. Experience with CloudFormation and infrastructure automation practices is beneficial. - Containers & Platform Technologies: Experience supporting containerized workloads using Docker, ECS, Kubernetes, or similar orchestration technologies. - Compliance Awareness: Exposure to compliance frameworks (SOC2, ISO, etc.) and security controls in cloud environments. - Technical Acumen: Proficiency in scripting (Python, Bash, PowerShell, etc.), infrastructure-as-code solutions, Jenkins, GitHub Actions, and monitoring or observability platforms. - Education: Bachelor’s degree in Computer Science, Information Security, or related field (or equivalent practical experience). - Soft Skills: Strong communication, collaboration, troubleshooting, and problem-solving abilities. Proactive, accountable, adaptable, and eager to learn. - Exposure/experience with AI architectures such as AWS Bedrock would be a plus. - Experience with observability platforms such as New Relic, Datadog, Grafana, or similar technologies is a plus. Requirements - You have limited experience supporting production cloud environments. - You prefer a slow-moving environment or are uncomfortable with the pace and ambiguity of a startup. - You are not interested in hands-on technical work, operational ownership, troubleshooting, or proactive problem-solving. - You require highly prescriptive direction and are uncomfortable working independently. - You prefer narrow responsibilities rather than broad ownership across infrastructure, security, automation, and operational support. Benefits - If you are an experienced DevOps professional who enjoys building reliable systems, solving challenging technical problems, taking ownership of outcomes, and driving continuous improvement, we want to hear from you. - Candidates who value clarity, consistency, and careful attention to detail tend to thrive here. - This is an opportunity to make a real impact at a dynamic and growing company.
Role Description As a Senior DevOps Engineer at BetterHelp, you’ll join a diverse team of licensed clinicians, engineers, product pros, creatives, marketers, and business leaders who share a passion for expanding access to therapy. We are looking for an individual to join our systems team and help us build our production and corporate systems and keep them up and running reliably. - Design, implement, deploy, and maintain reliable systems and processes. - Alongside software engineers, provide service outage escalation response. - Participate in on-call incident management rotation. - Manage configuration and security of instances across multiple environments. - Manage and build core tools, services, and infrastructure across the entire organization. - Monitoring of the production systems and networks. - Develop and deploy automation to make your job easier. - Have access to whatever training resources can help you succeed. Qualifications - 5+ years of experience in DevOps or as an SRE (or equivalent production engineering role). - Strong experience with AWS, Docker, and Kubernetes. - Experience with CI/CD pipelines and methodologies. - Experience with IaC tooling such as Terraform. - Very knowledgeable in Linux systems. - Strong verbal and written communication skills. - Strong ability to think on your feet, come up with innovative solutions, and solve problems that may be completely unique. Requirements - Remote work with regular in-person bonding experiences sponsored by the company. - Competitive compensation. - Holistic perks program (including free therapy, employee wellness, and more). - Excellent health, dental, and vision coverage. - 401k benefits with employer matching contribution. - The chance to build something that changes lives – and that people love. - Any piece of hardware or software that will make you happy and productive. - An awesome community of co-workers. Benefits - The base salary range for this position is $150,000 - $200,000. - This position is eligible for a performance bonus and the extensive benefits listed here (subject to eligibility requirements). - Total compensation is based on several factors – including, but not limited to, type of position, location, education level, work experience, and certifications.
Build software faster. The One DevOps Platform enables your entire org to collaborate around your code. We're hiring.
Role Description We're looking for an Intermediate Site Reliability Engineer to join the Runners Platform team. In this role, you'll build and operate Hosted Runners for GitLab Dedicated, the managed CI/CD compute platform that runs our customers' pipelines inside single-tenant Amazon Web Services (AWS) environments. You'll own infrastructure automation across the full lifecycle: - Provisioning runner fleets with Terraform - Extending the Go tooling and autoscaling stack - Ensuring customer continuous integration and continuous delivery (CI/CD) jobs run reliably across the platform What you’ll do: - Design, build, and operate AWS infrastructure for Hosted Runners across many single-tenant environments, including: - Elastic Compute Cloud (EC2) - Auto Scaling Groups - Virtual Private Cloud (VPC) networking - Subnets - Network Address Translation (NAT) - Network access control lists - PrivateLink - Identity and Access Management (IAM) - Elastic Container Registry (ECR) - Develop and maintain infrastructure as code using Terraform, contributing to common modules and the deployment tooling that provisions and upgrades runner stacks. - Write Go code for our runner tooling and autoscaling components, including: - Fleeting instance-lifecycle plugins - Zero-downtime deployment command-line interface - Reusable infrastructure toolkits - Build and improve the GitLab CI/CD pipelines that orchestrate: - Blue/green zero-downtime deployments - Automated upgrades - Quality assurance validation - Performance testing of runner stacks - Define and monitor service level objectives for CI job execution, including: - Queue times - Job success rates - Fleet saturation - Build the Grafana dashboards, alerts, and runbooks behind service level objectives. - Participate in an on-call rotation, handle incidents affecting customer CI/CD workloads, and automate away recurring toil. - Run performance and scale testing that reflects real customer workloads, and tune autoscaling parameters for cost and reliability. - Write documentation and runbooks for consistent operation of runner stacks. Qualifications - Professional experience operating production infrastructure on AWS at scale, including EC2, Auto Scaling Groups, IAM, and VPC networking. - Strong infrastructure-as-code experience with Terraform, including writing and refactoring modules used by other teams. - Proficiency in Go for building and debugging infrastructure tooling, or strong experience in another systems language and willingness to work in Go daily. - Practical knowledge of CI/CD systems and job execution. - Experience with observability practices such as metrics, dashboards, alerting, logging, and service-level-objective-based monitoring. - Experience with on-call rotations and incident management for customer-facing systems. - Strong problem-solving skills, excellent written communication, and comfort working asynchronously across multiple time zones. - Direct GitLab Runner experience, familiarity with configuration management such as Ansible, and container tooling such as Docker are a plus. Benefits - Flexible Paid Time Off - Team Member Resource Groups - Equity Compensation & Employee Stock Purchase Plan - Growth and Development Fund - Parental Leave
A leading provider of risk and compliance solutions, DFIN - Donnelley Financial Solutions offers data insights, industry expertise, and insightful technology to
Join a dynamic team at the pulse of global markets, where we deliver innovative software and service solutions for essential financial reporting and capital markets transactions. At DFIN, we are a values-driven organization that empowers you to build a fulfilling career while bringing your authentic self to work every day. Our "Win as One" mentality ensures that our team's success is directly linked to Client, Shareholder and Employee Satisfaction. In 2026, DFIN was named #1 on the 2026 Top 100 Global Most Loved Workplaces® by Best Practice Institute. We have also been recognized as one of America's Most Loved Workplaces® for five consecutive years and a Built In Best Place to Work for six years, reflecting our continued commitment to supporting employees' total well-being. Enjoy competitive compensation, a flexible workplace, comprehensive benefits, and opportunities for professional growth. Bring your passion and talents to DFIN - because being YOU thrives here. Summary: We are looking for technical team members at all levels who want to push themselves to deliver best in market SaaS solutions. We offer a challenging environment where you will have to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers. SRE's at DFIN take on availability, performance, managing change, monitoring, response and are guardians of non-functional requirements. You either have an SaaS infrastructure background with a programmatic, automated mindset or are someone that comes with a software engineering background with SaaS infrastructure experience. The SRE goal is to build automated systems that reduce or eliminate manual work to keep our products up and running and performing optimally. We are looking for someone who thrives on collaboration within the team and across other groups and can operate independently to deliver solutions. Responsibilities: • Champion and implement a culture of SRE to maintain a high-quality platform infrastructure in DFIN SaaS products • Leverage AI tools to enhance system reliability, including intelligent observability, incident prediction and automated remediation across cloud infrastructure • Evaluate and implement emerging AI powered operations and observability solutions to proactively improve system performance, reliability and scalability • Champion and implement application and infrastructure monitoring and alerting to prevent client impacting issues by ensuring system availability, performance and scalability to maintain SLOs and SLAs • Optimize application performance at scale • Automate everything including system operational runbooks • Define and support continuous integration and deployment pipelines (CI/CD) aligned to branching and quality assurance strategies • Dive deep into technology and stay on the forefront of the latest tools, technologies, and strategies; help evaluate, prototype, and integrate them into work processes • Perform with broad independence and deliver on project milestones and tasks on schedule while communicating progress regularly • Build strong relationships with SRE team members and software engineering teams to hold each other accountable for quality expectations • Learn continuously and apply lessons learned • Evangelize best practices, eliminate bottlenecks, and improve process • Participate in on-call duties 365/24/7 and lead the triage and RCA of production incidents Qualifications: • 5+ years experience designing, building, securing, monitoring and maintaining cloud infrastructure in Azure or AWS • Experience applying AI capabilities within CloudOps operations • Relevant certifications or training in AI, Cloud AI services or AIOps platforms are a plus • 5+ years experience writing software in any modern software language such as C# .NET, Java • 5+ years experience creating automated deployments with tools such as Harness, Azure DevOps, Ansible or Jenkins to manage Infrastructure as Code and software build and deployment in a continuous integration (CI) / continuous delivery (CD) environment • 5+ years experience implementing production performance, availability, and scalability monitoring and alerting using a tool such as New Relic, Dynatrace, DataDog or AppDynamics • 5+ years experience writing scripts in PowerShell or Python/Bash to automate system operations as runbooks for Windows or Linux environments. • 5+ years experience supporting public client facing revenue generating systems • Strong DevOps focus and experience building and deploying Infrastructure as Code with Terraform or similar technology • Experiencing monitoring and preventing issues with databases and database queries (SQL, Cosmos) using tools like Solarwinds Database Performance Analyzer, Idera SQL Diagnostic Manager, or Redgate SQL Monitor • Experience planning, coordinating, developing and executing all stages of post deployment verification test scripts • Experience securing Windows or Linux systems in 24x7 production environment • Experience with containerization and managing Kubernetes clusters (AKS or EKS) • Experience with common cloud networking, firewall and load balancing configuration • BS in Computer Science or equivalent work experience It is the policy of Donnelley Financial Solutions to select, place, and manage all its employees without discrimination based on race, color, national origin, gender, age, religion, actual or perceived disability, veteran status, actual or perceived sexual orientation, genetic information or any other protected status. If you are a qualified individual w ith a disability or a disabled veteran, you have the right to request a reasonable accommodation if you are unable or limited in your ability to use or access jobs.dfinsolutions.com as a result of your disability. You can request a reasonable accommodation by sending an email to talentacquisition@dfinsolutions.com . At DFIN, protecting your identity is a top priority. Please be aware of scammers impersonating DFIN recruiters. DFIN recruiters will never request personal information via email or text. You will only receive a text from us if you've already been in contact. All automated messages will come from talentacquisition@dfinsolutions.com . If you ever have doubts about the legitimacy of any communication from us, please do not hesitate to reach out for verification via talentacquisition@dfinsolutions.com (this email is for general TA questions and is not used for updates on your application status). #BI-Remote
Role Description We're seeking an experienced senior development operations (DevOps) engineer to own the build, test, and deployment pipeline for our distributed computer vision platform running at edge sites across customer manufacturing floors. This is a full-time remote position working directly with our engineering team to harden our release process, drive deployment automation, and pioneer LLM-driven test generation and validation at Rapta. - Own and evolve the end-to-end release pipeline — branching strategy, build orchestration, artifact promotion, and rollback — across our Bazel monorepo and Python deployable units. - Design and maintain Ansible-driven fleet automation for heterogeneous Linux edge nodes (Ubuntu LTS, NVIDIA driver stacks, Docker with NVIDIA runtime). - Manage all update tooling, currently written in Golang. - Build LLM-powered automated testing systems: test generation from specs, flake triage, log/failure analysis, regression diffing, and release-note synthesis from commit and ticket history. - Harden CI/CD for offline and bandwidth-constrained deployment targets (airgap wheel distribution, signed artifacts, deterministic builds). - Drive observability for releases — deployment telemetry, version drift detection, and post-deploy health validation across the fleet. - Mentor engineers on release hygiene, reproducible builds, and infrastructure-as-code practices. Qualifications - 10+ years of professional experience in release engineering, DevOps, or SRE roles shipping production Linux systems. - Deep curiosity for software, infrastructure, and applied AI — particularly using LLMs as production engineering tools, not just chat assistants. - Expert-level Python (3.8+) with a strong grasp of packaging, dependency resolution, and PEP 440 versioning discipline. - Demonstrated ownership of Linux fleets at scale — kernel, systemd, networking, package management. - Excellence in technical communication, runbook authorship, and post-incident documentation. - Strong systems thinking — comfortable reasoning about failure modes across hardware, OS, container, and application layers. Requirements - Expert proficiency with Ansible (roles, dynamic inventory, idempotent design); working knowledge of Terraform. - Expert proficiency with Docker, including creation and lifecycle management of containers, image hardening, registry management and installing & configuring the NVIDIA container runtime. - Production experience with Linux administration: systemd, networking (VLANs, DHCP, DNS), kernel/driver management (especially NVIDIA/DKMS), package and APT internals. - Strong Python skills focused on tooling, automation, packaging (wheels, pip, private indexes), and subprocess/CI integration. - Proficiency with Git workflows, branching strategies, and modern CI/CD systems (GitHub Actions, GitLab CI, or equivalent). - Experience designing and operating automated test infrastructure — unit, integration, hardware-in-the-loop, and end-to-end. - Practical experience using LLMs (Anthropic, OpenAI, or local) as part of engineering workflows — test generation, code review augmentation, log analysis, or agentic tooling. Nice to Have - Bazel or similar monorepo build systems. - Edge or embedded deployment experience. - Tailscale, WireGuard, or zero-trust networking in production. - gRPC/protobuf service ecosystems. - Vault, PKI, or secrets management at fleet scale. - Background in regulated or compliance-driven environments (CMMC, ISO 27001, SOC 2). Benefits - Work on cutting-edge AI infrastructure with real-world impact on American manufacturing. - Build the release engineering foundation for a late-seed company actively scaling. - Direct collaboration with the CTO and engineering leadership. - Remote-first culture with flexible hours. - Meaningful equity in a company solving hard problems. Location Requirements - Remote (US). Equal Opportunity Rapta is committed to hiring and retaining a diverse workforce. We are proud to be an Equal Opportunity/Affirmative Action Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. How to Apply No recruiters or agencies — we only accept applications directly from applicants.
Role Description The Site Enablement Lead (SEL) supports clinical study execution by identifying and enabling trained and qualified clinical research staff who are dedicated to a specific study or group of studies at a research site with the goal of accelerating study deliverables. The SEL also works in collaboration with the Patient Recruitment & Enablement (PRE) Team, in supporting clinical trial sites to optimize follow-through on protocol-specific patient referrals from Direct-to-Patient (DTP) campaigns by enabling sites to use IQVIA’s site-facing Referral Hub platform. The goal of the SEL role is to optimize overall performance of the site throughout the lifecycle of the study. Responsibilities - Introduce services to site, encouraging site adoption. - Train and educate site on site-facing tech tools and site worker services. - Engage with internal technology teams to ensure timely activation of sites. - Support clinical trial sites to optimize follow-through on protocol-specific patient referrals from Direct-to-Patient (DTP) campaigns. - Monitor referral flow at sites and follow up appropriately to ensure optimal referral funnel performance. - Collaborate with internal legal team to prepare, negotiate, and execute service agreement contracts with site. - Interview and hire temporary site staff. - Manage invoicing and effort forecasting process for the assigned sites. - Maintain collaborative site relationships to ensure continuous feedback loop to support delivery and assure quality. - Update department systems (e.g. WMT) with accuracy, quality and in a timely manner. - Work closely with functional lead to monitor impact of SES support. - Contribute to ad-hoc process development and process improvement projects. - Coach and mentor employees as they develop in their role. - Other duties may be assigned. Qualifications - In-depth knowledge of clinical trial conduct at clinical study sites for pharmaceutical research. - Experience engaging with sites and key stakeholders throughout the lifecycle of a clinical trial. - Advanced Level of Portuguese. - Advanced Level of Spanish. - Advanced Level of English. - Knowledge of patient recruitment practices and site-based processes and workflow. - Possess knowledge and ability to apply ICH/GCP and applicable regulatory guidelines in delivery of services. - Demonstrated ability to review and interpret data to identify trends and actions related to department delivery. - Possess strong computer skills including proficiency in aspects of using remote technology and presentation software such as TEAMs, Microsoft Word, Excel, and PowerPoint. - Excellent written and verbal communication, as well as presentation and training skills, including good command of English. - Effective time management and organizational skills and ability to manage competing priorities. - Strong attention to detail. - Ability to adapt and be flexible in a global and dynamic work environment with changing priorities. - Excellent interpersonal and problem-solving skills with ability to build and maintain strong relationships with IQVIA staff, site stakeholders, and other key stakeholders. Requirements - Bachelor’s degree or higher educational equivalent in health care or other scientific disciplines. - 3-5 years of relevant industry experience or equivalent combination of education, training, and experience. - Advanced level of English, Spanish, and Portuguese (Portuguese or Spanish could be native languages). - No certification required; Preferred: "CCRC" - Certified Clinical Research Coordinator (ACRP) or CCRP Certified Clinical Research Professional (SoCRA). - Rare Domestic Travel Possible. Benefits - Remote work opportunities.
GXO is a leading provider of cutting-edge supply chain solutions to the most successful companies in the world. We help our customers manage their goods most efficiently using our technology and services. Our greatest strength is our global team – energetic, innovative people of all experience levels and talents who make GXO a great place to work.
Role Description Are you ready to take your career to the next level with a rapidly growing global company? As a Senior DevOps Platform Engineer, you will establish and scale the enterprise DevOps operating model for GXO's Agentic AI Platform. This role is responsible for building secure, automated, and scalable cloud platform capabilities that enable AI application delivery across Google Cloud Platform, Kubernetes, Terraform, and CI/CD ecosystems. You'll partner closely with Cloud Engineering, Platform Architecture, Security, and Product teams to drive developer productivity, platform reliability, and operational excellence. If you're looking for an opportunity to make a significant impact on enterprise AI infrastructure, join us at GXO. Qualifications - Bachelor’s degree in computer science, Engineering, Information Technology, Cloud Computing, or a related technical field; equivalent hands-on experience may be considered. - Google Cloud Professional DevOps Engineer certification required. - Minimum of 8 years of platform engineering, DevOps, Site Reliability Engineering (SRE), infrastructure engineering, cloud engineering, or software delivery engineering experience. - Minimum of 5 years of hands-on Google Cloud Platform experience supporting production environments. - Deep expertise with Google Kubernetes Engine (GKE), including Kubernetes operations, workload identity, networking, autoscaling, ingress/egress, Helm or Kustomize, and production troubleshooting. - Expert-level experience developing and managing Terraform infrastructure, including reusable modules, state management, CI/CD integration, policy-as-code, infrastructure promotion, and drift management. - Strong experience designing and maintaining secure CI/CD pipelines using Cloud Build, GitHub Actions, GitLab CI, Azure DevOps, Jenkins, or similar platforms. - Experience implementing GitOps and DevSecOps practices, including code review automation, dependency scanning, container security, secrets management, signed artifacts, deployment approvals, and security guardrails. - Experience supporting cloud-native AI, machine learning, analytics, developer platform, or data platform workloads on Kubernetes and Google Cloud. - Ability to collaborate effectively with principal architects, cloud engineers, Information Security, product teams, and software developers to translate architectural vision into production-ready solutions. - Strong operational mindset with experience supporting incident response, root cause analysis, observability, production support, SLOs/SLIs, release readiness, and continuous operational improvement. - Excellent technical communication skills with the ability to develop engineering documentation, operating procedures, automation standards, and developer guidance. - Ability to influence engineering teams across a global matrix organization while driving adoption of modern DevOps and platform engineering practices. Requirements - HashiCorp Terraform Associate certification. - Certified Kubernetes Administrator (CKA) or Certified Kubernetes Application Developer (CKAD) certification. - Google Cloud Professional Cloud Architect, Cloud Security Engineer, or Machine Learning Engineer certifications. - Experience with AI platform technologies including LiteLLM, Agent Gateway, MCP servers, Vertex AI, Gemini, model routing, vLLM, or open-source model serving frameworks. - Experience enabling AI-assisted software development through secure coding assistants, automated testing, documentation generation, and developer productivity tooling. - Strong understanding of secure enterprise AI platform operations, cloud-native architecture, and scalable infrastructure automation. - Experience driving platform standardization, operational excellence, and developer enablement across large engineering organizations. - Self-starter with the ability to quickly establish credibility, operate independently, and make an immediate impact on the reliability, security, scalability, and velocity of enterprise AI platforms. Benefits - Competitive compensation. - Generous benefits package, including full health insurance (medical, dental, and vision). - 401(k) plan. - Life insurance. - Disability insurance. - Opportunity to participate in a company incentive plan. Company Description GXO is a leading provider of cutting-edge supply chain solutions to the most successful companies in the world. We help our customers manage their goods most efficiently using our technology and services. Our greatest strength is our global team – energetic, innovative people of all experience levels and talents who make GXO a great place to work. We are proud to be an Equal Opportunity employer including Disabled/Veterans. GXO adheres to CDC, OSHA and state and local requirements regarding COVID safety. All employees and visitors are expected to comply with GXO policies which are in place to safeguard our employees and customers. All applicants who receive a conditional offer of employment may be required to take and pass a pre-employment drug test.
Período: Indeterminado Idioma: Inglês Avançado Modelo de Atuação: Remoto
Role Description Engenheiro(a) de Plataforma Cloud (Azure) - Projetar, construir e manter infraestruturas baseadas em Azure utilizando módulos Terraform (recursos IaaS e PaaS); - Atuar em novos projetos e também em melhorias e sustentação de ambientes produtivos; - Administrar, monitorar e dar suporte a ambientes Azure em múltiplas subscriptions; - Troubleshooting de incidentes relacionados a IaaS e PaaS, garantindo cumprimento dos SLA’s; - Criar e manter documentações técnicas, diagramas de arquitetura, playbooks, runbooks e boas práticas; - Apoiar áreas internas com direcionamento técnico e recomendações sobre arquitetura e design de aplicações; - Trabalhar em conjunto com times multidisciplinares para garantir disponibilidade, performance e segurança dos ambientes; - Apoiar iniciativas de automação, CI/CD e infraestrutura como código. Qualifications - Experiência com serviços core do Azure: - Azure Resource Manager - Virtual Machines / Scale Sets - Storage (Blob e Files) - Networking (VNets, NSGs, Load Balancer, Application Gateway, DNS e Private Link) - Experiência em administração Linux e/ou Windows Server; - Experiência com Azure Kubernetes Service (AKS); - Experiência com Terraform; - Vivência com troubleshooting e sustentação de ambientes cloud; - Capacidade de atuar de forma autônoma e proativa; - Boa organização, comunicação e trabalho em equipe; - Excelente capacidade analítica e resolução de problemas. Requirements - Conhecimento em BICEP e/ou ARM Templates; - Experiência com CI/CD utilizando Azure DevOps Pipelines ou GitHub Actions (YAML); - Familiaridade com ferramentas de gerenciamento de configuração, preferencialmente Ansible; - Experiência com Kafka, MongoDB e/ou Elasticsearch; - Conhecimento em scripting utilizando Bash, Shell Script ou PowerShell; - Boa comunicação interpessoal e colaboração entre equipes. Benefits - Idioma: Inglês - Periodo: Indeterminado - Modelo de Atuação: Remoto
Role Description This role provisions and manages workload-specific infrastructure resources that power the Skylark insurance platform and the Legacy program. Sitting within the Infrastructure Build Pod, the Cloud Engineer translates architecture decisions and workstream team requirements into production-ready Terraform modules. - Covering compute, storage, PaaS services, and environment scaffolding — shipped through the shared DevSecOps CI/CD pipeline. - Partner closely with the Senior Cloud Network Engineer, application workstream leads, and the Security team to deliver governed, reproducible, and cost-aware infrastructure aligned to Calandra Blueprint and Azure Landing Zone standards. Qualifications - Experience with Azure resources — compute (VMs, AKS), storage, PaaS services (Container Apps, App Services, Azure Functions, Azure SQL, Cosmos DB). - Proficient in authoring and maintaining Terraform IaC modules. - Knowledge of governance guardrails: policy compliance, cost tagging, quota management, and secure defaults. Requirements - Provision workload-specific Azure resources per workstream team requirements and architecture-approved patterns. - Author and maintain Terraform IaC modules for workload resource provisioning, following Infrastructure Build Pod standards. - Scaffold and manage workload landing-zone scaffolding — child subscriptions, resource groups, RBAC assignments, tagging, and Azure Policy. - Integrate workload resources with the shared network layer — private endpoints, VNet integration, and DNS registration. - Maintain environment parity across dev / staging / production, enforcing governance guardrails. Company Description
UnitedHealth Group is a healthcare and well-being company that’s dedicated to improving the health outcomes of millions around the world. We are comprised of
Role Description As a senior member of our Kubernetes Platform Engineering team, you will be responsible for the reliability, scalability, security, and automation of our OpenShift ecosystem. Join our team to help steward our on-prem Cloud deployed on High-Performance Compute infrastructure. What You'll Do - Administer, maintain, and optimize enterprise OpenShift Kubernetes platforms deployed to our High-Performance Compute infrastructure - Manage large-scale OpenShift clusters using: - OpenShift Console - OpenShift CLI (oc) - Advanced Cluster Manager (ACM) - Infrastructure automation tools - Design and implement automated deployment, configuration, and lifecycle management solutions - Develop Infrastructure-as-Code and automation frameworks using Ansible and scripting - Troubleshoot complex Kubernetes, container, networking, and platform issues - Collaborate with developers to improve application reliability, observability, scalability, and performance - Implement enterprise security controls and platform hardening standards - Support high availability, disaster recovery, and multi-cluster architectures - Drive platform modernization initiatives and operational excellence through automation - Participate in architecture reviews and help define the future state of enterprise container platforms Qualifications - Bachelor's degree or equivalent experience - 5+ years administering Kubernetes platforms in production environments - Experience with Ansible and automation frameworks - Experience with container technologies and image creation - Experience troubleshooting enterprise-scale platform issues - Experience deploying and administering OpenShift Virtualization (KubeVirt) to support both containerized and virtual machine workloads - Experience implementing and supporting enterprise monitoring, logging, and alerting solutions using Grafana, Prometheus, Loki, AlertManager, and Thanos - Experience with Git Action pipeline deployments for OpenShift, bash scripting using oc for controls, deployments, and working closely with application teams on deployments, new namespaces, affinities, and sizing - Experience with cross-datacenter, high availability failover, and load balancing (F5 & haproxy) between multiple datacenters and K8s clusters - Solid knowledge of Kubernetes internals including: - Scheduling - Networking - Storage - Ingress and load balancing - Operators - Cluster lifecycle management - Proficiency in shell scripting and at least one programming language such as Python or Go - Proven deep expertise with Red Hat OpenShift - Proven solid Linux administration skills, preferably Red Hat Enterprise Linux - Proven excellent communication and technical documentation skills - Demonstrated security-first mindset with expertise in platform hardening, identity and access management, vulnerability remediation, encryption, and regulatory compliance Preferred Qualifications - Red Hat OpenShift certification - Experience with OpenShift deployed to IBM LinuxONE and IBM Z environments - s390x Linux architecture - Experience with GitOps, CI/CD pipelines, and Infrastructure-as-Code - Experience supporting high-availability Kubernetes environments - Experience supporting mission-critical applications with strict uptime requirements - Knowledge of enterprise networking concepts and Kubernetes networking architectures - Familiarity with GPFS (IBM Spectrum Scale) or other distributed storage technologies Benefits - Comprehensive benefits package - Incentive and recognition programs - Equity stock purchase - 401k contribution (all benefits are subject to eligibility requirements) Company Description At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone–of every race, gender, sexuality, age, location and income–deserves the opportunity to live their healthiest life. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes.
2,175more opportunities are still waiting for you.Log in now and take your next shot before someone else does.
Stack data is limited for this slice right now.