Cadence Solutions logo
Cadence Solutions

Helping Organizations with Digital Transformation

Staff DevOps Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteLeadTeam 11-50H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

56 days ago

Salary

$200K - $260K / year

Seniority

Lead

No structured requirement data.

Job Description

Staff DevOps Engineer

Cadence Solutions

Role Description We're hiring a Staff DevOps Engineer to design, operate, and scale the cloud infrastructure that powers care delivery for over 100,000 patients. You will own platform maturity across reliability, security, and observability - giving clinical and engineering teams the foundation they need to ship safely and move fast. This role sits at the intersection of infrastructure engineering and care delivery, where the systems you build directly affect how quickly Cadence can act on patient risk. What You'll Do - Own the design and continuous improvement of Cadence's cloud infrastructure, driving reliability, scalability, and secure software delivery across all environments. - Maintain and mature core Kubernetes services, including resilient networking, autoscaling, monitoring, and well-architected patterns for production workloads that support real-time patient monitoring. - Lead Terraform-based infrastructure-as-code practices, including authoring, reviewing, and enforcing standards for AI-generated IaC to ensure correctness and security before deployment. - Define and enforce DevSecOps controls across clusters, including least-privilege IAM, container image scanning, and runtime policy - ensuring infrastructure meets the compliance requirements of a regulated healthcare environment. - Sharpen observability practices using tools like Datadog, improving alerting, incident response, and the feedback loops that keep clinical systems available and performant. - Manage infrastructure spend with discipline, identifying and resolving cost inefficiencies without compromising system resilience. - Mentor engineers across the team on infrastructure best practices, raising the technical bar in pull requests, runbooks, and production operations. Qualifications - 8+ years of hands-on DevOps or platform engineering experience, with demonstrated ownership of production cloud infrastructure at scale. - Deep experience with AWS and Kubernetes, including designing, operating, and debugging production clusters under real load. - Proficiency with Terraform, Helm, and CI/CD pipelines using GitHub Actions or comparable tooling. - Strong command of observability tooling, including Datadog or equivalent platforms, with experience building alerting systems and leading incident response. - Experience in healthcare or another highly regulated industry, with working knowledge of relevant compliance and security requirements. - Track record of mentoring engineers and raising infrastructure standards across a team. - Fluency with LLM APIs, prompt engineering, and AI-assisted development tools; demonstrated experience building or evaluating AI-powered systems in production. Compensation Our job titles may span more than one career level. The base salary for this role typically ranges between $200,000 - $260,000, depending on experience, skills, seniority, and business needs. In addition to base salary, this role may be eligible for equity as part of the total compensation package. Actual compensation may vary by location. Benefits - Competitive pay & equity* - Fully remote - Comprehensive health coverage: Medical, dental & vision - Paid time off - 401k plan + matching - Paid parental leave - Home office stipend *Benefit offerings may vary depending on job profile, job level and worker type.

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Arbor Education logo

Site Reliability Lead

Arbor Education

Arbor MIS helps schools and MATs work more easily and collaboratively. Join a free webinar: http://bit.ly/Arbor-webinars

DevOps Engineer56 days ago
Full TimeRemoteTeam 51-200H1B No Sponsor

• Define and guide system architecture, balancing trade-offs between speed, scalability, maintainability, and security to meet business goals. • Champion accountability from design through to production by ensuring systems are observable and meet agreed Service Level Objectives (SLOs). Drive continuous improvement in platform reliability, performance, and efficiency. • Lead Root Cause Analysis (RCA) when issues occur and contribute to optimizing the incident response process and framework. • Drive automation initiatives across the team to reduce operational toil and improve system efficiency. • Uphold coding standards, promote automated testing, and work with the architecture community to drive technology adoption and share best practices across teams. Ensure production readiness standards for all services. • Lead technical estimation and feasibility assessments, ensuring plans are realistic and aligned with team capacity. Contribute to structured release planning and support post-release reviews. • Mentor and coach engineers through constructive feedback, knowledge sharing, and motivation. Foster alignment and help the team galvanise around technical solutions and goals. • Work closely with Product Managers, Engineering Managers, and other engineers to align technical direction with product strategy. Communicate complex technical concepts clearly to both technical and non-technical stakeholders.

United Kingdom
£80K - £90K / year
Openly logo

DevOps/SRE II

Openly

Premium, straightforward insurance

DevOps Engineer56 days ago
Full TimeRemoteTeam 201-500Since 2017H1B Sponsor

• Build internal tooling to help other engineers and the rest of the company understand and operate our system • Design and implement security best practices for our team and infrastructure • Reduce toil through automation, including building and maintaining CI/CD infrastructure • Build infrastructure as code using declarative provisioning tools • Develop high signal-to-noise ratio monitoring and alerting policies and technology to help us meet our SLOs • Lead incident response and postmortems • Contribute to important architectural and operational decisions like microservices vs. monoliths, deployment techniques, technologies, policies, etc.

United States
$115.2K - $129.6K / year
Job Closed
Arbor Education logo

Site Reliability Technical Lead

Arbor Education

Arbor MIS helps schools and MATs work more easily and collaboratively. Join a free webinar: http://bit.ly/Arbor-webinars

DevOps Engineer56 days ago
Full TimeRemoteTeam 51-200H1B No Sponsor

Role Description We are looking for an experienced and collaborative Site Reliability Technical Lead to join our Site Reliability team and take ownership of system and solution design to ensure our products are robust, scalable, and secure. The remit and focus of the role is to blend deep technical expertise with leadership, requiring you to mentor and coach engineers, embed a culture of quality and reliability, and guide the team in making sound technical decisions. It’s a broad and exciting role, so we’re looking for someone up for a challenge - if you’re highly technical and a good communicator, this is the role for you. Core Responsibilities - Architectural Leadership: Define and guide system architecture, balancing trade-offs between speed, scalability, maintainability, and security to meet business goals. - Reliability and Performance: Champion accountability from design through to production by ensuring systems are observable and meet agreed Service Level Objectives (SLOs). Drive continuous improvement in platform reliability, performance, and efficiency. - Incident Management: Lead Root Cause Analysis (RCA) when issues occur and contribute to optimizing the incident response process and framework. - Automation: Drive automation initiatives across the team to reduce operational toil and improve system efficiency. - Technical Standards: Uphold coding standards, promote automated testing, and work with the architecture community to drive technology adoption and share best practices across teams. Ensure production readiness standards for all services. - Planning and Delivery: Lead technical estimation and feasibility assessments, ensuring plans are realistic and aligned with team capacity. Contribute to structured release planning and support post-release reviews. - Mentorship and Coaching: Mentor and coach engineers through constructive feedback, knowledge sharing, and motivation. Foster alignment and help the team galvanise around technical solutions and goals. - Collaboration: Work closely with Product Managers, Engineering Managers, and other engineers to align technical direction with product strategy. Communicate complex technical concepts clearly to both technical and non-technical stakeholders. Qualifications - Extensive professional experience in SRE, DevOps, or Platform Engineering on complex, scalable systems. - Extensive expertise with AWS and distributed cloud architectures. - Proven experience operating platforms serving a high volume of requests (~1000 req/sec). - Advanced proficiency with Terraform and configuration management tools. - Strong skills in Python, Go, or a similar language for automation and tooling. - Deep experience with monitoring and observability platforms (e.g., DataDog, Prometheus, or equivalent), plus incident/problem management. - Expert understanding of distributed systems, microservices, and resilience patterns. - Hands-on experience with containerization and orchestration technologies (Docker, Kubernetes, ECS). - Practical experience with building and maintaining CI/CD pipelines for automated deployments. - Demonstrated ability in mentoring and supporting the growth of fellow engineers. Bonus Skills - Experience with chaos engineering and reliability testing. - Knowledge of security best practices and compliance frameworks. - Background in agile and lean methodologies (Scrum/Kanban). - Contributions to open-source projects or the SRE community. Benefits - The chance to work alongside a team of hard-working, passionate people in a role where you’ll see the impact of your work every day. - A dedicated wellbeing team who champion initiatives such as mindfulness, lunch n learns, manager training, mental health first aid training and much more! - 32 days holiday (plus Bank Holidays). This is made up of 25 days annual leave plus 7 extra company-wide days given over Easter, Summer & Christmas. - Life Assurance paid out at 3x annual salary. - Comprehensive wellness benefit provided by AIG Smart Health, which provides a 24/7 virtual GP service, Mental health support, Counselling, and personalised Health Checks. - Private Dental Insurance with Bupa. - Salary sacrifice Pension provided by Scottish Widows. - Enhanced maternity and adoption leave (20 weeks full pay) and paternity (6 weeks full pay) pay. - 5 free return to work maternity coaching sessions, helping you adapt to this new exciting time of life! - Access to services such as Calm and Bippit (financial wellbeing coaching). - All of our roles champion flexible working and we are happy to discuss what this means to you. - Social committees that plan team, office and company-wide events to bring people together and celebrate success. - Dedicated professional development training budget (CPD courses, upskilling resources, professional memberships etc). - Volunteer with a charity of your choice for a day each year. - Dog friendly offices! Interview Process - Phone screen - 1st stage - 2nd stage

United Kingdom
£80K - £90K / year
CodiLime logo

Senior DevOps/SRE Engineer, Kubernetes, Python

CodiLime

A strategic partner for technology-driven companies | Network engineering | Software engineering

DevOps Engineer56 days ago
ContractRemoteTeam 201-500Since 2011H1B No Sponsor

• Deploying and maintaining applications on-prem and in the cloud • Troubleshooting application and infrastructure issues • Creating documentation • Working in an agile methodology and collaborating with a team • Supporting teammates • Attending daily standups with the customer

Poland
zł22K - zł27K / month
Job Closed