Job Closed

This listing is no longer active.

Appspace logo
Appspace

Discover the easiest way to reach your workforce - at work, at home, or on the go.

Senior DevOps, Site Reliability Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteSeniorTeam 201-500H1B No SponsorCompany SiteLinkedIn

Location

Florida

Posted

67 days ago

Salary

0

Seniority

Senior

Job Description

Senior DevOps, Site Reliability Engineer

Appspace

• Be the technical anchor for a global platform footprint that includes a mix of Azure IaaS/PaaS, Google Cloud Platform (GCP), Kubernetes, and various data platforms. • Identify manual "toil" and replace it with automated workflows for monitoring, change management, and routine administration of large-scale VM environments to ensure a positive ROI. • Lead the integration of AI tools for automated code reviews, development frameworks, and predictive log analysis to drive departmental velocity and efficiency. • Design and maintain "self-service" deployment frameworks and CI/CD pipelines (GitHub Actions, Bamboo) using Infrastructure as Code (Bicep, Terraform). • Evaluate platform components to determine the most cost-effective path: automating the current state or migrating features to modern, shared architectures. • Design and maintain a comprehensive observability stack across Azure and GCP (metrics, logs, traces) to identify performance bottlenecks and proactively address system defects. • Partner with engineering, security and operations teams to ensure new features are "born" with reliability, security and automated delivery in mind; Ensure adherence to security best practices and compliance standards (SOC2, HIPAA, ISO 27001) and operational excellence with cost efficiency. • Investigate complex performance defects by following log trails across web, application, and database tiers (SQL Server, MongoDB, MySQL). • Ensure all platforms meet security standards (SOC2, HIPAA, ISO 27001) through automated policy enforcement across Azure and GCP.

Job Requirements

  • 6+ years in DevOps or SRE roles, with a proven track record of bridging development and operations in complex cloud environments
  • Extensive experience with Microsoft Azure (IaaS, PaaS, App Services, Networking) and/or Google Cloud Platform (GCP).
  • Expert-level PowerShell and Python skills. Hands-on experience with Bicep or Terraform is required
  • Strong background in Windows/Linux Server OS, Kubernetes (AKS/GKE), Helm, and container orchestration
  • Familiarity with various middleware and PaaS technologies (e.g. Event Hub, Service Bus, CosmosDB, RabbitMQ, MongoDB, etc.)
  • Expert-level troubleshooting and the ability to reason through complex process workflows to identify faults in large-scale platform environments.

Benefits

  • Health insurance
  • 401(k) plan
  • Paid parental leave
  • Generous PTO
  • Flexible work schedules
  • Remote work opportunities
  • Paid company holidays
  • Appspace Quiet Fridays (No non-essential internal meetings scheduled)
  • A casual dress work environment

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Full TimeRemoteTeam 201-500Since 2018H1B Sponsor

• Oversee a specialized SRE team focused on the design, deployment, and maintenance of automation toolsets as well as the systems they interact with. • Establish and enforce standards for IaC to ensure consistent, repeatable, and secure deployments across an entire infrastructure ecosystem. Strong proficiency in Terraform is required. • Lead the strategy for automated configuration and state management, ensuring Ansible playbooks and Packer image pipelines are optimized for both Windows, Linux, and ESXi Platforms. • Manage the monitoring and health of the automation platforms themselves. Implement SLIs/SLOs to ensure the 'tools that build the servers' are highly available and performant. • Drive the automated lifecycle of both physical and virtual assets, from initial template creation/deployment to automated patching, scaling, and decommissioning. • Lead the development of custom scripts and internal providers (Python, Go, PowerShell, Bash) to provide better insights and tooling for our systems. • Outside of the automation team you will need to be able to collaborate and foster workflows alongside the rest of the Datacenter team and be able to facilitate needs for the team as a whole. • Analyze system behavior and resource utilization in virtual environments to optimize the performance of automated deployments. • Provide technical guidance and career mentorship to SREs, fostering a culture of 'automate-first' and continuous improvement.

United States
Job Closed
viind GmbH logo

DevOps Engineer

viind GmbH

Chatbots für Verwaltungen und Privatunternehmen | Systeme mit KI-Anbindung | Recruiting via Messenger

DevOps Engineer67 days ago
Full TimeRemoteTeam 11-50H1B No Sponsor

• You take responsibility for the operation, maintenance and further development of our Kubernetes clusters (Hetzner Cloud & on‑premises) • You ensure the availability, scalability and security of our infrastructure and continuously optimize it • You operate and maintain our central platform components such as databases (PostgreSQL, Typesense, MongoDB), Keycloak and self‑hosted AI models • You develop and implement strategies for deployments, updates, backups and recovery of these systems • You implement and run monitoring and logging solutions for our entire infrastructure • You develop, operate and optimize our CI/CD pipelines based on GitLab and Docker • You contribute your own ideas to continuously improve and evolve our infrastructure, processes and tools • You manage our internal cloud and SaaS services (e.g. Atlassian, Microsoft 365, GitLab) • You ensure that all systems and processes meet the requirements of ISO 27001 and GDPR and actively support audits • You work closely with our development team and support infrastructure-related questions or backend development

Germany
Job Closed
Cribl logo

Senior Site Reliability Engineer

Cribl

Cribl, the Data Engine for IT and Security, empowers organizations to transform their data strategy.

DevOps Engineer67 days ago
Full TimeRemoteTeam 501-1,000Since 2017H1B Sponsor

• Engage with teams and improve service delivery and reliability across their entire lifecycle • Measure and monitor all production systems with an eye towards availability, latency and overall system health • Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence • Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability, resilience, and observability • Help identify and drive down toil with creative innovation and automation • This position will require stand-by, on-call, or off-hours duties

Poland
Resilient Co. logo

Senior DevOps

Resilient Co.

WE ARE RESILIENT CO. We adapt to your needs.

DevOps Engineer67 days ago
ContractRemoteTeam 11-50Since 2020H1B No Sponsor

• Design and implement infrastructure-as-code using Terraform for Azure services including AKS, Blob Storage and App Services. • Build, maintain and optimize CI/CD pipelines and mobile/web build pipelines. • Operate, troubleshoot and tune Kubernetes and Docker-based workloads running on AKS. • Implement and manage SSO and External ID flows using Microsoft Entra. • Create reusable templates, Terraform modules and pipeline templates to enable developer self-service. • Collaborate directly with technical leads to define platform direction and deployment patterns. • Mentor engineers on deployment best practices, observability and platform usage. • Own platform-level decisions and improvements, prioritizing strategic work over ticket-level execution. • Write clear, async-friendly documentation and communicate effectively in AI-augmented workflows. • Manage and support PostgreSQL-related deployment and operational concerns as they relate to platform infrastructure.

Argentina
Job Closed