DevOps Engineer Remote Jobs in Massachusetts (US)
This page tracks remote devops engineer openings that are location-eligible for Massachusetts.
This page tracks remote devops engineer openings that are location-eligible for Massachusetts.
Open jobs
2,171
Hiring companies this week
10
Salary sample
$65,000 - $1,124,480
Jobs added last hour
0
2171 Jobs
1352 Companies
Innovation / Venture Studio focusing on Software Development, Legacy and MVP Development: Supporting corporate clients in innovation / business modeling Helping technical and non-technical founders getting from 0-1 Technology focus: Cloud/DevOps, Embedded, C#, Java, SAP, AI, Quantum Computing, B2B SaaS Software Development: Supporting corporate customers in the areas of software development and DevOps Assisting customers in identifying automation potentials that can be exploited with software Developing customized AI agents for the continuous automation of increasingly challenging tasks in marketing, sales, customer success, HR and finance Helping customers in the application and technical integration of AI agents
Role Description Wir sind ein fast zwei Jahre altes IT Unternehmen, aktuell 50 MA, und suchen nun weiterhin Verstärkung, um Cloud-Projekte bei unseren Kunden durchzuführen. - Du willst Deine Arbeitszeiten selbst bestimmen? - von zuhause aus arbeiten (remote)? - den nächsten Schritt in Deiner Karriere gehen und lernen, wie man Teams leitet und ein Unternehmen führt? - einen Arbeitgeber, der Dich leistungs- und ergebnisorientiert bezahlt? Qualifications - bist seit 4 Jahren in Vollzeit berufstätig im Bereich Softwareentwicklung/DevOps und hast große Teile Deiner Erfahrung in Cloud-Infrastrukturen und Deployment-Prozessen gesammelt? - bringst Erfahrung in einem der großen Cloud-Dienstleister (z.B. AWS, Azure) mit? - bist firm in IaC (Terraform, Pulumi), Monitoring (z.B. ELK, Grafana) und CI/CD? - hast einen Fokus auf eine oder mehrere Programmiersprachen wie z.B. Java, Python, C++, C, Rust oder Go und bist dort ein Experte? - bist nicht kontaktscheu und kommunikationsfreudig? - bist gesegnet mit einem Growth Mindset und willst im Leben immer weiterkommen? - bist überzeugt davon, dass Dein Potenzial noch lange nicht ausgeschöpft ist? Benefits - Fixum: 65.000 € - 70.000 € - Zielgehalt: 75.000 € - 80.000 € (Umsatzbeteiligung an eigenem Umsatz, quartalsweise Ausschüttung) - 30 Tage Urlaub - IT Equipment deiner Wahl (Mac, Linux, Windows) - Gelebter interner Expertenaustausch und Support - Flexible Arbeitszeit und 95% - 100% Homeoffice (abhängig vom Kunden) - Remote-Arbeit im Ausland (GF war Co-Founder von rhome) Company Description Wir sind eine Gruppe von jungen und hungrigen Entwicklern, die gemeinsam die Firma betreiben, in der wir immer arbeiten wollten, die aber nicht existent war! - unterstützen unsere Kunden in Software Development Projekten - bauen Startups und gründen sie aus - arbeiten an und mit Top Notch Technologien (AI, Quantum Computing) - sind ein geiles Team
Clover is a healthcare technology company helping members live their healthiest lives with our Medicare Advantage plans.
• Build systems for declarative application and infrastructure lifecycle management, including continuous deployment, continuous integration, Kubernetes cluster management, and service/workload inventory. • Prioritize and troubleshoot infrastructure issues, minimizing downtime and responding to alerts efficiently. • Contribute to setting the direction of the Site Reliability Engineering (SRE) team, ensuring goals align with Counterpart Health’s company-wide objectives. • Foster a collaborative, high-performance culture that promotes motivation, innovation, and cross-disciplinary teamwork. • Streamline and automate infrastructure processes, including delivery pipelines and database changes.
Fulfilling the promise of precision medicine through quality and innovation.
• Design, deploy, and maintain Linux infrastructure in on-premises and cloud environments. • Automate infrastructure provisioning and configuration using tools such as Terraform, Ansible, or CloudFormation. • Manage and optimize AWS environments with a focus on performance, scalability, security, and cost efficiency. • Implement and maintain monitoring, logging, and alerting solutions (e.g., Datadog, Prometheus, Grafana, ELK, CloudWatch). • Architect, deploy, and operate production Kubernetes/AWS EKS clusters, including node group strategy, cluster upgrades, multi-tenant workload isolation, and cross-region disaster recovery (DR) architecture and build outs. • Define and lead cluster upgrade, security hardening, and disaster recovery strategies for production Kubernetes/AWS EKS environments at scale, while serving as a senior technical resource for complex production incidents. • Manage Kubernetes networking, including VPC CNI configuration and ingress controllers (ALB/NGINX/Traefik). • Implement IAM Roles for pod security standards, and network policies to secure EKS workloads. • Configure and tune cluster autoscaling (Cluster Autoscaler or Karpenter) and workload autoscaling (HPA/VPA) to optimize performance and cost. • Build and maintain Helm charts and GitOps-based deployment pipelines (e.g., ArgoCD, Flux) for Kubernetes workloads. • Manage Docker container builds and registries in support of EKS-based application deployment. • Deploy, scale, and maintain GitLab Runners (including Kubernetes executor runners on EKS) to support CI/CD pipeline throughput and reliability. • Support and help operate database platforms on AWS RDS (MySQL, PostgreSQL), collaborating with data owners on performance and reliability. • Ensure systems meet security and compliance requirements, including SOX and SOC 2 initiatives. • Execute and maintain Linux patching strategies, addressing security updates and CVEs in a timely manner. • Participate in incident response, root cause analysis, and recovery efforts. • Collaborate with development, QA, and cross-functional teams to improve reliability, release processes, and operational standards. • Participate in on-call rotations and provide after-hours support as required.
Axiad is a cybersecurity company that provides enterprise-grade identity and access management (IAM) solutions, helping organizations securely manage credentials and authentication
Role Description Axiad is seeking a skilled AI-First SRE/DevOps Engineer with 5–8 years of hands-on infrastructure and platform engineering experience to help build and run Mesh, our Identity Visibility and Intelligence Platform (IVIP) — a cloud-native microservices platform on Kubernetes spanning human identity, non-human identity (NHI), post-quantum cryptography, and agentic AI identity risk. The ideal candidate has a builder mentality and a strong AI-First mindset: automation and AI are the default, not the afterthought, and infrastructure is something you create, not just maintain. This is a startup environment. You will own real surface area end-to-end, move fast, and ship. The role requires deep operational expertise in Kubernetes, CI/CD, and infrastructure-as-code, along with practical experience running AI/LLM systems in production. If your instinct when facing a repetitive task is to script it, agent-ify it, or delete it entirely — you'll fit right in. Responsibilities - Own reliability, observability, and delivery for a multi-tenant, cloud-native Kubernetes platform — from design through production, yours to run and yours to improve. - Build (not just operate) CI/CD pipelines, infrastructure-as-code, and GitOps-driven progressive delivery that let a small team ship many times a day, safely. - Embrace and advocate AI-First operations: automate incident response, runbooks, and remediation, and put AI agents in the loop to triage, diagnose, and propose fixes where it makes sense. Treat toil as a bug. - Build the infrastructure that AI-native features run on: inference gateways, LLM cost/latency observability, prompt/version pipelines, eval harnesses, and guardrails for agentic workloads. - Instrument everything — SLOs, error budgets, and distributed tracing across services and data pipelines. - Harden the platform: secrets management, supply-chain security, and least-privilege everywhere. - Troubleshoot and resolve production issues, leveraging AI-powered debugging and observability tooling. - Collaborate directly with product and platform engineers to translate requirements into resilient infrastructure — no throwing tickets over a wall; if you see a problem, it's yours to solve. - Mentor engineers in adopting AI-first operational practices and automation-by-default culture. Qualifications - 5–8 years of professional experience in SRE, DevOps, or platform engineering roles. - Builder mentality: you'd rather create a tool, platform, or automation than run a manual process twice. You ship things and stand behind them. - Ownership: you take problems from ambiguity to resolution without waiting for a ticket, a spec, or permission. When something you own breaks, you're the first to know and the first to act. - Strong Kubernetes operational experience — running it in production, not just deploying to it. - Demonstrable adoption of an AI-First mindset and tools (Claude Code, Cursor, or Windsurf). Daily use of at least one AI development tool is a must. - Fluency with infrastructure-as-code, GitOps, and modern CI/CD; comfortable scripting and building tooling (Go or Python preferred). - Cloud-native depth on at least one major cloud provider. - Solid observability expertise and SLO-driven operations experience. - Experience with containerization (Docker) and service mesh concepts. - Strong problem-solving skills and a collaborative mindset; excellent communication within Agile teams. - A bias for shipping — startup pace energizes you rather than stresses you. Preferred Qualifications - Experience building or operating LLM infrastructure: inference gateways, eval/observability tooling, agentic orchestration. - Data-pipeline and streaming/CDC experience. - Security or identity background; familiarity with post-quantum cryptography or supply-chain security. - Prior experience at an early-stage startup. Benefits - 120,000 - 160,000 OTE + Equity + Benefits
Empowering providers to deliver high-quality care through real-time notifications and meaningful incentives.
• Collaborate daily with fellow DevOps engineers and cross-functional engineering teams to build a scalable, high-performing foundation. • Design and maintain HIPAA-compliant, multi-account AWS infrastructure using Terraform and Terragrunt with a strict GitOps mindset. • Build resilient GitHub Actions pipelines and implement safe release strategies (e.g., Blue/Green, automated rollbacks) for Python/Django applications. • Ensure high availability, automated backups, and version upgrades for core data stores, including Postgres, ElastiCache (Redis), and OpenSearch. • Mature full-stack Datadog/CloudWatch telemetry, refine on-call workflows, lead blameless postmortems, and drive key metrics (MTTR, deployment frequency). • Enforce least-privilege access and automated security guardrails across AWS, database layers, and GitHub. • Create paved-path tooling and self-serve infrastructure that reduce friction, shorten test cycles, and lower cognitive load for product teams. • Author clear design docs, conduct thorough code reviews, mentor team members, and proactively identify operational risks before they impact production.
• DevSecOps Engineer to strengthen our software development lifecycle by embedding security practices into every stage of delivery • Work across development, operations, and security teams to ensure applications and infrastructure are secure, compliant, and resilient, while maintaining speed and efficiency in deployment • Responsible for streamlining our development and operational processes, ensuring efficient deployment and management of applications in cloud environments • Implement CI/CD pipelines, manage cloud resources, and enhance system performance
First Citizens Bank offers a full line of financial services and focuses on individuals, as well as small to medium-sized businesses. As an employer, the compan
Role Description This is a remote role that may only be hired in the following locations: NC, TX, AZ, GA. This position is responsible for all phases of data processing system projects, from requirements definition to installation. Leads technical efforts in the development, implementation, and maintenance of complex systems. Develops test plans, software, and procedures that improve processing capabilities. Supports production systems by resolving complicated issues and ensuring ongoing functionality. Serves as a technical expert and may provide a leadership role for less experienced associates in the work group. Responsibilities - CI/CD implementation: Develop, implement, and maintain CICD pipelines to automate the build, test, and deployment process. Standardize and ensure adoption of the process. - Infrastructure management: Automate the process to provision, configure, and manage infrastructure on cloud platforms using "infrastructure as code" tools like Terraform and Ansible. - Automation: Identify and automate manual processes to improve efficiency and reduce errors, including writing scripts for various operational and development tasks. - System monitoring and reliability: Implement and manage monitoring, logging, and alerting tools to track system performance and availability. Perform root cause analysis and troubleshoot production issues. - Security as a code: Ensure the security of the software development pipeline and infrastructure by implementing security best practices and tools throughout the lifecycle. - Collaboration: Work closely with development, QA, Application Security, Governance, and IT teams to improve the overall delivery process. - Containerization and orchestration: Implement and manage containerization technologies like Docker and orchestration platforms such as Kubernetes. - System Enhancement: Evaluate and improve department systems, processes, and applications. Provide new feature time estimates for system changes and assist in implementing modifications. - Analysis: Collect data related to user requests and determine scope, time estimates, and system impacts. Inspect business specifications, programming specifications, coding, test plans, documentation, and implementation plans for accuracy. - Business Support: Provide technical support to production systems by addressing reported issues, anticipating maintenance requirements, and ensuring functionality for end user needs. - Technical Expertise: Responsible for complex involvement in the software development life cycle including the creation, enhancement, implementation, and evaluation of software. - On-call Support: Provide 24/7 on-call support via rotations. Qualifications - Bachelor's Degree and 4 years of experience in Software application development and maintenance OR High School Diploma or GED and 8 years of experience in Software application development and maintenance. - Preferred skills: - 8+ years of experience in DevOps, DevSecOps, Platform Engineering, Cloud Engineering, or related disciplines. - 6+ years of experience in architecting, developing, maintaining, and optimizing enterprise-scale CI/CD pipelines using GitLab. - 5+ years of experience in implementing DevSecOps practices, security controls, and automated compliance checks within CI/CD pipelines. - Experience in developing automation solutions using Python, Bash, and other scripting languages. - Knowledge in major tech stacks of modern application development. - Experience in developing build processes for Java, .NET, iOS, and Android, and automating deployment processes for OpenShift and VM. - Knowledge in AWS & Azure is a plus. - Hands-on experience in Linux environments (RHEL), Ansible, writing playbooks to automate tasks, deploy artifacts, configuring monitoring tools (AppDynamics, Dynatrace), tools upgrade, patching and migrations, security vulnerability remediations. - Experience in automating deployment and release processes across OpenShift, Kubernetes, Docker, virtual machines, and cloud environments. - Experience in implementing Infrastructure as Code (IaC) solutions using Terraform for cloud infrastructure provisioning and DevSecOps tool deployments. - Deep knowledge in security scans and experience in implementing them as part of the CI/CD pipeline. - Experience in working with cross-functional teams, including App security, Application engineering, Quality automation engineering, and Infrastructure teams to integrate security and automation throughout the software development lifecycle. - Experience in migrating applications from one SCM to another. - Working experience on Docker, OpenShift/Kubernetes/Docker Swarm. - Implementation, management, and administration of Enterprise systems tools and processes - AWS stack, JIRA, Confluence, GitLab, Jenkins, Azure DevOps, Nexus, Artifactory, Snyk, SonarQube. Benefits - Benefits are an integral part of total rewards and First Citizens Bank is committed to providing a competitive, thoughtfully designed and quality benefits program to meet the needs of our associates. - More information can be found at https://jobs.firstcitizens.com/benefits .
SimplePractice offers an all-in-one platform used by more than 160,000 health and wellness providers to manage their private practices. As an employer, the comp
Role Description We are hiring a DevOps Engineer to support and scale our Data and AI platform in production. This role focuses on building reliable infrastructure for data pipelines and ML systems, standardizing deployment patterns, and ensuring performance, observability, and cost efficiency across compute-intensive workloads. Responsibilities - Build and operate infrastructure for data pipelines and AI/ML workloads - Develop and maintain CI/CD for application and model lifecycle (build, train, deploy) - Manage Infrastructure as Code (Terraform) across environments - Support containerized workloads and orchestration (Docker, Kubernetes) - Partner with Machine Learning teams and engineering to productionize models - Implement monitoring, logging, and tracing for data flow and model performance - Improve reliability, scalability, and cost efficiency of data systems - Enforce security and access controls for data and infrastructure - Reduce operational overhead through automation and tooling Qualifications - 3+ years of experience in DevOps, SRE, or infrastructure engineering - End-to-End MLOps/LLMOps Expertise: Experience deploying and maintaining ML/AI workflows. Familiarity with the unique nature of promoting AI assets (models, datasets, and code) through the lifecycle. - Strong cloud experience (AWS preferred) - Proficiency with Terraform (or similar IaC tools) - Experience with Docker and Kubernetes - Familiarity with CI/CD and Git-based workflows - Experience supporting data platforms (e.g., Airflow, Kafka, Spark, or similar) - Programming/scripting (Python, Bash, or similar) - Experience with observability tools and practices Preferred Qualifications - Experience with MLOps tooling (e.g., MLflow, SageMaker, Kubeflow) - Familiarity with LLM-based systems and AI observability (token usage tracking, prompt versioning) and evaluation loops - Experience with real-time or high-throughput data systems - Exposure to security and compliance requirements (e.g., SOC 2, HIPAA) - Experience with specific MLOps tooling (Outerbounds, SageMaker, Metaflow) and vector database Benefits - Privatized Medical, Dental & Vision Coverage - Supplemental health and wellness benefit - Modern Health and Vivawell - Work From Home stipend - Flexible Time Off (FTO), wellbeing days, Summer Fridays, and Mid-Year and Year-End Reflect & Recharge (Company Holidays), Paid Holidays - Christmas Bonus (15-day aguinaldo) - Monthly meal/grocery voucher via Si Vale card - Catered Lunch - A relocation bonus for candidates joining us from a different city - Annual Bonus - Tuition Reimbursement - Saving Funds - Employee Resource Groups (ERGs)
• In this role you will be involved with the operations, security and application teams to maintain and operate the web infrastructure’s availability and the ability to scale and fail over. • Develop and support secure CI/CD pipelines with validation, scanning, and approval gates • Implement the web architecture including high availability, failover, and multi-region designs • Support platform standardization and governance, ensuring auditability and repeatable deployments • Monitor and optimize resources for performance, cost, and reliability • Implement and support security controls, including vulnerability remediation and compliance requirements • Troubleshoot platform and deployment issues across environments (Dev,Staging, Prod) • Maintain documentation, runbooks, and deployment standards to support consistent execution • Participate in release activities and ensure safe, controlled deployments • Strong troubleshooting skills across infrastructure and application deployment layers • Ability to work in Agile delivery models and manage concurrent priorities.
Companies trust VPS with their training needs. From instructor to simulation, to sales and technical, we deliver it all.
• Design, implement, and maintain automated CI/CD pipelines that carry code from development through security scanning, compliance validation, and deployment into Navy and DoD environments • Build and maintain hardened Kubernetes environments aligned to DISA STIG requirements across cloud and restricted network deployment contexts • Automate security artifact generation including SBOM production, CVE scanning, and continuous compliance validation • Drive adoption of Infrastructure as Code, GitOps practices, and controls-as-code across the team • Leverage AI tooling to accelerate pipeline development, vulnerability triage, compliance remediation, and operational documentation • Partner closely with software engineers, systems engineers, and ISSE's to embed security and compliance requirements from the start of development • Maintain and evolve deployment infrastructure across multiple secure environments, including cloud and air-gapped or intermittently connected contexts • Support ATO processes through automated evidence generation, documentation as code, and direct collaboration with the security team • Establish and promote standards for pipeline design, container security, secrets management, and deployment consistency • Contribute to feature development when team capacity requires, applying security-first development practices to application code • Maintain operational documentation including runbooks, deployment guides, and architecture diagrams as version-controlled artifacts
2,161more opportunities are still waiting for you.Log in now and take your next shot before someone else does.
Cloud, Python, AWS, Kubernetes, Terraform, Docker