DevOps Engineer Remote Jobs in District of Columbia (US)
This page tracks remote devops engineer openings that are location-eligible for District of Columbia.
This page tracks remote devops engineer openings that are location-eligible for District of Columbia.
Open jobs
2,094
Hiring companies this week
10
Salary sample
$95 - $154,445
Jobs added last hour
0
2094 Jobs
1315 Companies
• In this role you will be involved with the operations, security and application teams to maintain and operate the web infrastructure’s availability and the ability to scale and fail over. • Develop and support secure CI/CD pipelines with validation, scanning, and approval gates • Implement the web architecture including high availability, failover, and multi-region designs • Support platform standardization and governance, ensuring auditability and repeatable deployments • Monitor and optimize resources for performance, cost, and reliability • Implement and support security controls, including vulnerability remediation and compliance requirements • Troubleshoot platform and deployment issues across environments (Dev,Staging, Prod) • Maintain documentation, runbooks, and deployment standards to support consistent execution • Participate in release activities and ensure safe, controlled deployments • Strong troubleshooting skills across infrastructure and application deployment layers • Ability to work in Agile delivery models and manage concurrent priorities.
Companies trust VPS with their training needs. From instructor to simulation, to sales and technical, we deliver it all.
• Design, implement, and maintain automated CI/CD pipelines that carry code from development through security scanning, compliance validation, and deployment into Navy and DoD environments • Build and maintain hardened Kubernetes environments aligned to DISA STIG requirements across cloud and restricted network deployment contexts • Automate security artifact generation including SBOM production, CVE scanning, and continuous compliance validation • Drive adoption of Infrastructure as Code, GitOps practices, and controls-as-code across the team • Leverage AI tooling to accelerate pipeline development, vulnerability triage, compliance remediation, and operational documentation • Partner closely with software engineers, systems engineers, and ISSE's to embed security and compliance requirements from the start of development • Maintain and evolve deployment infrastructure across multiple secure environments, including cloud and air-gapped or intermittently connected contexts • Support ATO processes through automated evidence generation, documentation as code, and direct collaboration with the security team • Establish and promote standards for pipeline design, container security, secrets management, and deployment consistency • Contribute to feature development when team capacity requires, applying security-first development practices to application code • Maintain operational documentation including runbooks, deployment guides, and architecture diagrams as version-controlled artifacts
• Responsible for designing, implementing, and maintaining cloud infrastructure • Ensure the reliability and performance of applications • Work closely with development teams to establish CI/CD pipelines • Automate deployment processes, leveraging expertise in cloud platforms and DevOps practices
Granicus is driven by the excitement of building, implementing, and maintaining technology that is transforming the Govtech industry by bringing governments and its constituents together. We are on a mission to support our customers with meeting the needs of their communities and implementing our technology in ways that are equitable and inclusive. Consistently appeared on the GovTech 100 list over the past 5 years Recognized as one of the best companies to work for on BuiltIn Served 5,500 federal, state, and local government agencies More than 300 million citizen subscribers power an unmatched Subscriber Network Comprehensive cloud-based solutions for communications, government website design, meeting and agenda management software, records management, and digital services Empowers stronger relationships between government and residents across the U.S., U.K., Australia, New Zealand, and Canada
Role Description The Senior DevOps Engineer is an advanced engineering position responsible for designing, implementing, and supporting automation, deployment, monitoring, and operational reliability across cloud and hosted application environments, with expertise in applying AI to optimize platform operations. This role partners closely with software engineering, site reliability, security, infrastructure, and architecture teams to improve continuous integration and continuous delivery processes, infrastructure automation, platform operations, and service resilience across increasingly complex environments and delivery pipelines. The successful candidate will demonstrate deep hands-on experience with DevOps practices, cloud infrastructure, systems administration, automation, and troubleshooting, along with the ability to independently lead complex technical efforts from planning through implementation. This position is intended for senior engineers with demonstrated AI/ML experience who are expected to apply advanced technical judgment and drive significant improvements in platform reliability and efficiency by applying AI across CI/CD, observability, incident response, and infrastructure designing and operation. - Own and support software delivery activities, including build, deployment, release, and environment management processes across multiple complex applications and infrastructure environments. - Design, maintain, and optimize CI/CD pipelines used to build, test, secure, and deploy applications with greater efficiency, consistency, resiliency, and scalability. - Architect, provision, configure, and maintain infrastructure in cloud and hosted environments using automation and infrastructure as code practices. - Lead complex technical projects or cross-functional workstreams, including planning, dependency coordination, risk identification, execution oversight, and successful delivery of outcomes. - Provide technical leadership across DevOps and platform engineering initiatives by setting direction for implementation approaches, reviewing designs, promoting best practices, and influencing technical decisions. - Mentor and guide engineers through technical coaching, knowledge sharing, solution reviews, and hands-on support to improve team capability and delivery quality. - Troubleshoot and resolve complex application, deployment, infrastructure, and environment issues with minimal supervision, including active participation in incident response, root cause analysis, and follow-through on corrective actions. - Monitor system health, review logs and telemetry, and implement improvements in reliability, performance, scalability, recoverability, and operational efficiency. - Drive standardization of engineering practices, reusable automation, deployment patterns, architecture guardrails, and operational runbooks that improve consistency across environments. - Collaborate across engineering, support, operations, security, and architecture teams to resolve issues, support releases, manage dependencies, and improve platform and delivery processes. - Apply established security, compliance, and operational standards in all engineering activities and identify opportunities to strengthen controls, reduce risk, improve resilience, and increase audit readiness. - Participate in operational support and on-call activities, and take ownership of critical issues through resolution, communication, and follow-through. Qualifications - Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience. - Significant progressive professional experience in DevOps (8+ Years), platform engineering, systems engineering, cloud infrastructure, or a related engineering discipline. - Demonstrated advanced working knowledge of DevOps concepts, including CI/CD, automation, infrastructure as code, cloud operations, release management, system reliability, and operational resilience. - Strong hands-on experience with Linux and/or Windows administration, scripting, source control, cloud platforms, and troubleshooting in production or pre-production environments. - Practical experience with containers, infrastructure as code, monitoring, logging, operational support processes, and modern platform engineering practices. - Demonstrated ability to independently lead complex assignments and technical initiatives, manage competing priorities, and deliver reliable results in a fast-paced engineering environment. - Demonstrated experience leading complex projects or technical workstreams, including organizing tasks, coordinating stakeholders, managing dependencies, identifying risks, and driving work to completion. - Demonstrated ability to provide technical leadership, mentor engineers, and influence engineering practices, design approaches, and operational standards without formal people management responsibility. - Effective written and verbal communication skills and the ability to collaborate across technical teams while clearly communicating technical issues, risks, recommendations, project status, and implementation strategy. - Demonstrated accountability, sound judgment, attention to detail, and commitment to operational excellence, continuous improvement, and technical quality. - Ability to use approved AI-enabled tools responsibly to improve productivity, support troubleshooting, enhance documentation, and assist engineering workflows. Requirements - Lead the design and architecture of AI-driven DevOps platforms across CI/CD, observability, and operations. - Drive organization-wide adoption of AI, MCP frameworks, and AIOps practices. - Define standards for AI usage, governance, security, and compliance across engineering workflows. - Architect end-to-end AI-enabled automation ecosystems (agents, workflows, deployment intelligence). - Use AI to influence strategic decisions (capacity planning, cost optimization, and reliability engineering). - Mentor engineers and lead initiatives to scale AI capabilities, frameworks, and best practices. - Have experience using Python (or similar) to build operational tools with guardrails. Security and Privacy Requirements - Responsible for Granicus information security by appropriately preserving the Confidentiality, Integrity, and Availability (CIA) of Granicus information assets in accordance with the company's information security program. - Responsible for ensuring the data privacy of our employees and customers, their data, as well as taking all required privacy training in a timely manner, in accordance with company policies. The Team We are a remote-first company with a globally distributed workforce across the United States, Canada, United Kingdom, India, Armenia, Australia, and New Zealand. The Culture - At Granicus, we are building a transparent, inclusive, and safe space for everyone who wants to be a part of our journey. - Employee Resource Groups to encourage diverse voices. - Coffee with Mark sessions – Our employees get to interact with our CEO on very important and sometimes difficult issues ranging from mental health to work-life balance and current affairs. - Microsoft Teams communities focused on wellness, art, furbabies, family, parenting, and more. - We bring in special guests from time to time to discuss issues that impact our employee population. The Impact We are proud to serve dynamic organizations around the globe that use our digital solutions to make the world a better place — quite literally. We have so many powerful success stories that illustrate how our solutions are impacting the world.
Role Description We're looking for a DevOps Engineer to join our CMS BDAMAX team, supporting a federal program that directly impacts the Medicare experience for millions of Americans. - Build and maintain CI/CD pipelines including automated scanning, testing, and deployment workflows - Provision and manage infrastructure using Terraform with emphasis on reusable modules and secure configuration baselines - Manage and optimize AWS infrastructure including EKS, ECS, Fargate, EC2, S3, RDS Aurora PostgreSQL, and Secrets Manager - Work with Kubernetes and containerized environments including Argo Workflows - Support enterprise adoption of AI engineering platforms including Amazon Bedrock, GitHub Copilot, Gemini, and Cursor - Partner with DevOps, Architecture, and Development teams to implement reliable, scalable infrastructure patterns - Build internal tooling to support development and operational workflows Qualifications - Hands-on experience with AWS and core cloud infrastructure management - Experience building and maintaining CI/CD pipelines with Jenkins - Proficiency with Terraform for infrastructure provisioning and environment management - Experience with Kubernetes and containerized deployments - Familiarity with Argo Workflows is a plus - Thrives in a remote, collaborative Agile environment and genuinely enjoys working closely with a cross-functional team - Communicates clearly and openly, whether documenting infrastructure decisions or coordinating with engineering teams - Performs other related duties as assigned Requirements - Applicants must be authorized to work in the United States. - In alignment with federal contract requirements, certain roles may also require U.S. citizenship and the ability to obtain and maintain a federal background investigation and/or a security clearance. - Education: Bachelor’s degree Benefits - Fully remote - Annual stipend - Comprehensive Benefits Package - Company Match 401(k) plan - Flexible PTO, Paid Holidays Compensation At Oddball, it’s important each employee is compensated competitively and fairly. In alignment with state legal requirements, a range for the included position is listed below. Be advised, actual offer details are determined by job category, job location, and candidate skill level. United States Wage Range: $100,000 – $145,000
Grafana Labs supports organizations’ monitoring, visualization and observability goals. 950,000+ active installations
• Partner closely with product engineering squads (embedded model) • Own production reliability for high-SLA and complex customer environments • Design and implement automation to scale our reliability practices • Ensuring our customers meet our SLO targets • Define and evolve per-tenant SLOs and reliability models • Proactively reduce SLO burn to prevent repeat incidents • Serving as a primary escalation point and on-call for relevant incidents • Lead customer-impacting incident response and post-incident reviews • Contribute to design docs and code reviews • Influence feature design to ensure production scalability and operability • Build automation to eliminate toil where needed • Improve alert quality and reduce noisy escalations
Role Description Seeking a senior-level DevSecOps Engineer with strong developer experience supporting containerized applications in an AWS cloud environment. Ideal candidate will: - Identify, prove out, test, and implement pipeline process improvements - including utilizing new or existing GitLab or third-party tools or features. - Exhibit excellent customer service skills to educate customers on pipeline functionality and troubleshoot pipeline issues. - Interact with all levels of organizational personnel, from training and mentoring teammates to developing and presenting briefs to leadership. Essential Functions: - Develop and implement cloud-native DevSecOps functionality for CI/CD pipeline solutions in AWS; improve and maintain GitLab pipeline configurations. - Understand/interpret Cyber Security guidelines to resolve code vulnerabilities. - Maintain, monitor, and proactively research/pilot solutions to optimize and improve the parent scan pipeline; provide recommendations for technology advancement to streamline CI/CD tools and processes. - Assist with GitLab upgrades as received from the vendor (i.e. bi-weekly, monthly, etc.; requires evening support). - Design and build secure, scalable, and automated container environments using Amazon ECS and EKS. - Configure customer projects/access to pipeline, including configuring git on customer assets and credentials in customer repositories. - Onboard new applications/customers to the CI/CD environment, working closely with application developers to provide training/technical guidance/troubleshooting assistance. - Craft/help customers craft gitlab-ci.yaml files to orchestrate their child pipelines or test projects; create, maintain, update, and monitor health checks. - Research, perform analysis of alternatives, recommend technical solutions, and architect new CI/CD environments and pipelines, such as Cloud migration or supporting classified systems. - Provide demos/overviews/briefs/newsletters regarding the pipeline and/or associated tools to customers and/or leadership of various levels. - Update and maintain documentation. - Provide mentorship to junior teammates. - Other duties as assigned or required. Qualifications - CI/CD implementation experience required. - Design/development of DevSecOps pipelines experience required. - Proficient communication and documentation skills with experience preparing technical guidance, how-to instructions, test plans, demos, and/or presenting training to internal and external stakeholders required. - Experience with Amazon ECS and EKS required. - CompTIA Sec+ certification is required and must show proof before interview. - BS/BA degree and 10 years related experience OR AA/AS degree and 14 years related experience OR HS and 16 years related experience. Requirements - Experience with specific CI/CD related tools such as GitLab Ultimate, Nexus, DORA metrics, and Prisma Cloud (formerly Twistlock) highly desired. - Experience working in a DoD environment highly desired. - Experience with OpenShift and Nexus is a plus. - Ability to work independently in a fast-paced technical environment. Benefits - Health Care Plan (Medical, Dental & Vision) - Retirement Plan (401k, IRA) - Life Insurance (Basic, Voluntary & AD&D) - Paid Time Off (Vacation, Sick & Public Holidays) - Short Term & Long Term Disability - Training & Development - Wellness Resources - Stock Option Benefit
Role Description Do you want to shape reliability practices for a new AI inference platform? Are you a senior technical leader who drives solutions across teams? Join the Akamai Inference Cloud Team! The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications. In this role, you'll lead reliability workstreams for Akamai's serverless inference platform, design SRE tooling and automation, and drive technical decisions. Opportunities exist to mentor other SREs, influence architecture decisions with product engineering teams, and shape SRE practices for AI inference workloads and GPU infrastructure at scale. As a Senior II Site Reliability Engineer, you will be responsible for: - Taking ownership of observability strategy for the serverless inference platform, designing telemetry, dashboards, and alerts, defining SLO/SLI frameworks, and driving improvements when targets are missed. - Building production-grade automation and tooling that reduces operational toil, improves incident response, and sets patterns that other SREs adopt. - Owning incident management integration for inference workloads, designing frameworks, leading incident response during on-call rotations, and driving systemic improvements from post-mortems. - Defining and implementing deployment safety practices including progressive rollouts, canary analysis, and rollback automation, establishing standards for the team. - Partnering with product engineering teams to influence architecture decisions, ensure operational readiness, and represent the SRE perspective in design reviews. - Mentoring Senior and mid-level SREs through code reviews, design discussions, and hands-on problem-solving. Qualifications - 8+ years of experience in SRE, infrastructure engineering, or platform engineering, working with large-scale distributed systems. - Possess a proven track record of defining SLO/SLI frameworks, building observability platforms, and running incident management processes at scale. - Have extensive Kubernetes and containerization experience at scale, including autoscaling, resource scheduling, and container orchestration for compute-intensive workloads. - Have experience building automation and tooling in Python or Go, with familiarity in CI/CD pipelines, deployment safety, and infrastructure-as-code. - Possess the ability to lead technical initiatives across teams, mentor other engineers, and drive complex reliability problems to resolution independently. - Have experience with or exposure to AI/ML infrastructure, model serving, or GPU workloads. Benefits - We support your health, well-being, finances, and life beyond work. - FlexBase adapts to your job's needs. - Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. - We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both. Compensation Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $146,400 - $263,600/year; a candidate’s salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.
Top world’s largest social discovery company uniting 70+ brands with 500M+ users
• Design and develop internal services and tools in Go • Design, build, and maintain scalable CI/CD pipelines using GitLab CI/CD • Improve and support Kubernetes-based environments • Contribute to the evolution of the internal deployment platform (Packman) • Optimize stage and ephemeral environments for performance, reliability, and cost-efficiency • Implement and maintain infrastructure as code using Terraform and Ansible • Improve observability through monitoring, logging, and alerting • Collaborate with engineering teams to enhance developer experience and delivery speed
Procuramos uma pessoa que: Goste de trabalhar em equipe e seja colaborativa em suas atribuições; Tenha coragem para se desafiar e ir além, abraçando novas oportunidades de crescimento; Transforme ideias em soluções criativas e busque qualidade em toda sua rotina; Tenha habilidades de resolução de problemas; Possua habilidade e se sinta confortável para trabalhar de forma independente e gerenciar o próprio tempo; Tenha interesse em lidar com situações adversas e inovadoras no âmbito tecnológico. Big enough to deliver – small enough to care. #VempraGFT #VamosVoarJuntos #ProudToBeGFT
Role Description Profissional de nível Pleno/Sênior que atue com (Devops/AWS). - Liderar a instrumentação ponta a ponta de métricas (infraestrutura e aplicação), logs estruturados e tracing distribuído, garantindo visibilidade do ecossistema. - Implementar, evoluir e gerenciar ferramentas de Application Performance Monitoring (APM) para identificar gargalos e otimizar a experiência do usuário. - Definir, implementar e monitorar SLIs, SLOs e Error Budgets, apoiando o equilíbrio entre inovação e estabilidade. - Planejar e executar experimentos de Chaos Engineering para validar hipóteses de falha e fortalecer a arquitetura. - Definir políticas de alertas preditivos, reduzindo fadiga de alertas e acelerando a resposta a incidentes. - Atuar de forma transversal apoiando pipelines de Engenharia de Dados e arquiteturas de microsserviços (APIs REST) na AWS. Qualifications - Domínio avançado de Datadog para dashboards, monitores, APM e Log Management. - Experiência sólida com AWS, Docker e Kubernetes (EKS). - Vivência em automação de infraestrutura e monitoramento. - Conhecimento prático dos pilares de SRE, incluindo gerenciamento de incidentes, SLIs, SLOs e Error Budgets. - Conhecimento de arquitetura de sistemas distribuídos, Alta Disponibilidade (HA), tolerância a falhas e resiliência. Requirements - Infraestrutura como Código com Terraform, Pulumi ou CloudFormation. - Observabilidade em Engenharia de Dados, incluindo pipelines de Big Data (Apache Airflow, Spark ou similares). - Conhecimento em SecOps/DevSecOps e Observability-driven Security. - Certificações AWS (ex.: AWS Certified DevOps Engineer) ou Kubernetes (CKA/CKAD). Benefits - Goste de trabalhar em equipe e seja colaborativa em suas atribuições. - Tenha coragem para se desafiar e ir além, abraçando novas oportunidades de crescimento. - Transforme ideias em soluções criativas e busque qualidade em toda sua rotina. - Tenha habilidades de resolução de problemas. - Possua habilidade e se sinta confortável para trabalhar de forma independente e gerenciar o próprio tempo. - Tenha interesse em lidar com situações adversas e inovadoras no âmbito tecnológico.
2,084more opportunities are still waiting for you.Log in now and take your next shot before someone else does.
Cloud, Kubernetes, Linux, Terraform, Python, AWS