Dropzone AI logo
Dropzone AI

Reinforcements have arrived.

Senior DevOps Engineer

Location

United States

Posted

4 days ago

Salary

$170K - $185K / year

Seniority

Senior

Job Description

Senior DevOps Engineer

Dropzone AI

• Own and evolve the infrastructure that powers our AI platform • Work closely with teammates throughout engineering and security to design scalable, resilient, and secure systems • Enhance production monitoring and observability in containerized environments • Improve our infrastructure-as-code deployments (Pulumi) to prepare for scaling to the next customer growth inflection point • Develop and refine internal tooling to support efficiency and reliability • Strengthen the reliability, performance, and scalability of our core SOC analyst product • Participate actively in a 24x7 on-call rotation, maintaining high availability and rapid response • Contribute to future expansion efforts to support multi-cloud infrastructure (GCP, Azure) • Write new product features when interested and excited to do so

Job Requirements

  • 4+ years of experience as a DevOps, Infrastructure, or Site Reliability Engineer with practical, hands-on depth
  • Skills in observability and monitoring (e.g. Prometheus, Grafana), distributed tracing
  • Strong experience with containerization technologies, especially Docker and Kubernetes (k8s)
  • Software development skills with Python, capable of writing production-level code
  • Solid understanding of distributed systems, including troubleshooting, identifying bottlenecks, and performance optimization
  • Infrastructure-as-code experience (Pulumi preferred), maintaining fully automated CI/CD deployment pipelines
  • System administration and infrastructure management capabilities
  • Familiarity with modern security best practices, common vulnerabilities, and securing distributed environments
  • Ability to differentiate when a production-grade service is necessary versus implementing a one-time or temporary solution
  • Early-stage startup mindset; you conquer ambiguity and move with lightspeed execution
  • Data-driven decision making - it's part of your DNA
  • A strong opinion on emacs vs vim

Benefits

  • company paid health insurance
  • 401K Plan with employer match
  • Self-Managed PTO
  • parental leave
  • significant above market new hire equity grants
  • generous benefits package

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Saipos | Sistema para Restaurante logo

Security Engineer – DevSecOps, AppSec

Saipos | Sistema para Restaurante

Tornando o dia a dia do seu restaurante mais simples, ágil e inteligente. 🐿️

DevOps Engineer4 days ago
Full TimeRemoteTeam 51-200Since 2017H1B No Sponsor

• Implement and maintain SAST, DAST and SCA tools in the CI/CD pipeline (Bitbucket/GitHub/GitLab) • Implement, maintain, operate and evolve the Secrets Manager and Vault • Perform continuous secret scanning in repositories and drive remediation of findings with development teams • Conduct security-focused code reviews on critical features — authentication, authorization, external integrations, handling of sensitive data • Harden infrastructure: Terraform state, IAM policies, AWS configurations, Security Groups • Support remediation of technical findings collected from security analysis tools, external pentests and internal scans • Build and lead the Security Champions program — identify focal points within dev squads and structure ongoing training • Conduct threat modeling for new features and critical integrations with product and engineering teams • Define and document security requirements in the SDLC — from design to deployment • Assess the security of internal and external APIs, partner integrations and authentication flows

Brazil
Inmetrics logo

Analista de SRE Sr

Inmetrics

We make a difference, solve outstanding problems and make the digital transformation of our clients possible.

DevOps Engineer4 days ago
Full TimeRemoteTeam 501-1,000Since 2002H1B No Sponsor

• Atuar como referência técnica de SRE/Infraestrutura Cloud dentro dos times/clientes internos atendidos • Realizar diagnóstico de ambientes cloud (GCP/AWS) identificando oportunidades de otimização técnica e financeira • Propor e implementar ações de FinOps: rightsizing, otimização de recursos, controle e redução de custos • Manter e evoluir pipelines de CI/CD • Gerenciar infraestrutura como código (Terraform) • Garantir disponibilidade, escalabilidade e performance de ambientes em produção • Apoiar migrações de ambientes cloud (ex: AWS ↔ GCP) • Administrar e otimizar clusters Kubernetes • Implementar monitoramento e observabilidade (Prometheus, Grafana, DataDog, Cloud Logging etc.) • Atuar de forma consultiva, comunicando-se diretamente com stakeholders técnicos e de negócio

Brazil
Golden 1 Credit Union logo

DevOps Engineer II – IT Infrastructure Systems

Golden 1 Credit Union

Golden 1 is California's leading credit union. Insured by NCUA. Equal Housing Opportunity. NMLS #669333

DevOps Engineer4 days ago
Full TimeRemoteTeam 1,001-5,000Since 1933

• Independently lead infrastructure-as-code development using Terraform and scripting languages such as Python and PowerShell to support scalable and reliable deployments. • Manage Linux/Kubernetes cluster environments. • Deploy solutions in accordance with Change Management Processes. • Support development teams on API integration strategy and standards development. • Ensure systems are secure against cybersecurity threats. • Identify technical problems and develop software updates and fixes. • Strong Splunk skills for administration, query optimization, alerting, and dashboard development. • Build tools to reduce errors and improve customer experience. • Propose ideas and solutions within the Infrastructure Department to reduce workload through automation. • Design, implement, and optimize CI/CD pipelines for faster and more reliable software releases. • Independently conduct root cause analysis and implement corrective actions. • Design and write tests to investigate infrastructure failure and scaling. • Create and maintain response playbooks across incident management and monitoring tools. • Develop automation to ensure repeatability, eliminate toil, and reduce time to action and repair services. • Analyze key operational metrics to identify opportunities to improve availability. • Implement effective monitoring, alerting, and reduction of alert fatigue. • Manage container orchestration environments and optimize deployment workflows to enhance scalability, reliability, and operational efficiency. • Design, build, and manage containerized environments using Docker. • Create and maintain SLIs, SLOs, and error budgets. • Design and optimize monitoring dashboards and alerting systems to proactively detect and address application performance and uptime issues. • Implement code branching strategies using GitHub functions. • Advanced Terraform syntax and GitLab CI/CD configuration, pipelines, jobs. • Provisioning and setting up metrics in Prometheus, Thanos, and Grafana, creating and managing alerts. • Implement cloud engineering standards, reusable modules, and platform patterns in Microsoft Azure. • Operate shared cloud platform services according to Cloud Engineering defined architectures. • Ensure infrastructure changes comply with reliability, security, and cost controls established by Cloud Engineering. • Maintain operational documentation and runbooks for cloud platform services.

California
$123.6K - $135K / year
Full TimeRemoteTeam 11-50Since 2021

• Own and operate CI/CD pipelines and blue/green/canary deployment frameworks across Production, Staging, and Sandbox EKS clusters • Build and enforce policy-as-code guardrails: automated compliance checks against NIST 800-53, TRM, and VA-specific controls before production deployment • Manage Terraform IaC for all VAEC infrastructure provisioning; ensure all changes are code-reviewed and version-controlled • Lead SRE functions: define SLIs, track SLOs against PWS SLA tiers, manage error budgets, and own reliability reviews • Operate the Istio service mesh: traffic management, mTLS enforcement, and sidecar lifecycle across all EKS namespaces • Lead SAST, DAST, container scanning, and Fortify remediation — maintaining ATO with zero unmitigated critical CVEs • Maintain the onboarding pipeline: IaC-driven runbook bringing new engineers to readiness within contractual timeframes • Support Optional Task surge efforts for DevSecOps/SRE as defined in the Performance Work Statement

Florida