Reinforcements have arrived.
Senior DevOps Engineer
Location
United States
Posted
4 days ago
Salary
$170K - $185K / year
Seniority
Senior
Job Description
Senior DevOps Engineer
Dropzone AI
• Own and evolve the infrastructure that powers our AI platform • Work closely with teammates throughout engineering and security to design scalable, resilient, and secure systems • Enhance production monitoring and observability in containerized environments • Improve our infrastructure-as-code deployments (Pulumi) to prepare for scaling to the next customer growth inflection point • Develop and refine internal tooling to support efficiency and reliability • Strengthen the reliability, performance, and scalability of our core SOC analyst product • Participate actively in a 24x7 on-call rotation, maintaining high availability and rapid response • Contribute to future expansion efforts to support multi-cloud infrastructure (GCP, Azure) • Write new product features when interested and excited to do so
Job Requirements
- 4+ years of experience as a DevOps, Infrastructure, or Site Reliability Engineer with practical, hands-on depth
- Skills in observability and monitoring (e.g. Prometheus, Grafana), distributed tracing
- Strong experience with containerization technologies, especially Docker and Kubernetes (k8s)
- Software development skills with Python, capable of writing production-level code
- Solid understanding of distributed systems, including troubleshooting, identifying bottlenecks, and performance optimization
- Infrastructure-as-code experience (Pulumi preferred), maintaining fully automated CI/CD deployment pipelines
- System administration and infrastructure management capabilities
- Familiarity with modern security best practices, common vulnerabilities, and securing distributed environments
- Ability to differentiate when a production-grade service is necessary versus implementing a one-time or temporary solution
- Early-stage startup mindset; you conquer ambiguity and move with lightspeed execution
- Data-driven decision making - it's part of your DNA
- A strong opinion on emacs vs vim
Benefits
- company paid health insurance
- 401K Plan with employer match
- Self-Managed PTO
- parental leave
- significant above market new hire equity grants
- generous benefits package
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Security Engineer – DevSecOps, AppSec
Saipos | Sistema para RestauranteTornando o dia a dia do seu restaurante mais simples, ágil e inteligente. 🐿️
• Implement and maintain SAST, DAST and SCA tools in the CI/CD pipeline (Bitbucket/GitHub/GitLab) • Implement, maintain, operate and evolve the Secrets Manager and Vault • Perform continuous secret scanning in repositories and drive remediation of findings with development teams • Conduct security-focused code reviews on critical features — authentication, authorization, external integrations, handling of sensitive data • Harden infrastructure: Terraform state, IAM policies, AWS configurations, Security Groups • Support remediation of technical findings collected from security analysis tools, external pentests and internal scans • Build and lead the Security Champions program — identify focal points within dev squads and structure ongoing training • Conduct threat modeling for new features and critical integrations with product and engineering teams • Define and document security requirements in the SDLC — from design to deployment • Assess the security of internal and external APIs, partner integrations and authentication flows
Analista de SRE Sr
InmetricsWe make a difference, solve outstanding problems and make the digital transformation of our clients possible.
• Atuar como referência técnica de SRE/Infraestrutura Cloud dentro dos times/clientes internos atendidos • Realizar diagnóstico de ambientes cloud (GCP/AWS) identificando oportunidades de otimização técnica e financeira • Propor e implementar ações de FinOps: rightsizing, otimização de recursos, controle e redução de custos • Manter e evoluir pipelines de CI/CD • Gerenciar infraestrutura como código (Terraform) • Garantir disponibilidade, escalabilidade e performance de ambientes em produção • Apoiar migrações de ambientes cloud (ex: AWS ↔ GCP) • Administrar e otimizar clusters Kubernetes • Implementar monitoramento e observabilidade (Prometheus, Grafana, DataDog, Cloud Logging etc.) • Atuar de forma consultiva, comunicando-se diretamente com stakeholders técnicos e de negócio
DevOps Engineer II – IT Infrastructure Systems
Golden 1 Credit UnionGolden 1 is California's leading credit union. Insured by NCUA. Equal Housing Opportunity. NMLS #669333
• Independently lead infrastructure-as-code development using Terraform and scripting languages such as Python and PowerShell to support scalable and reliable deployments. • Manage Linux/Kubernetes cluster environments. • Deploy solutions in accordance with Change Management Processes. • Support development teams on API integration strategy and standards development. • Ensure systems are secure against cybersecurity threats. • Identify technical problems and develop software updates and fixes. • Strong Splunk skills for administration, query optimization, alerting, and dashboard development. • Build tools to reduce errors and improve customer experience. • Propose ideas and solutions within the Infrastructure Department to reduce workload through automation. • Design, implement, and optimize CI/CD pipelines for faster and more reliable software releases. • Independently conduct root cause analysis and implement corrective actions. • Design and write tests to investigate infrastructure failure and scaling. • Create and maintain response playbooks across incident management and monitoring tools. • Develop automation to ensure repeatability, eliminate toil, and reduce time to action and repair services. • Analyze key operational metrics to identify opportunities to improve availability. • Implement effective monitoring, alerting, and reduction of alert fatigue. • Manage container orchestration environments and optimize deployment workflows to enhance scalability, reliability, and operational efficiency. • Design, build, and manage containerized environments using Docker. • Create and maintain SLIs, SLOs, and error budgets. • Design and optimize monitoring dashboards and alerting systems to proactively detect and address application performance and uptime issues. • Implement code branching strategies using GitHub functions. • Advanced Terraform syntax and GitLab CI/CD configuration, pipelines, jobs. • Provisioning and setting up metrics in Prometheus, Thanos, and Grafana, creating and managing alerts. • Implement cloud engineering standards, reusable modules, and platform patterns in Microsoft Azure. • Operate shared cloud platform services according to Cloud Engineering defined architectures. • Ensure infrastructure changes comply with reliability, security, and cost controls established by Cloud Engineering. • Maintain operational documentation and runbooks for cloud platform services.
• Own and operate CI/CD pipelines and blue/green/canary deployment frameworks across Production, Staging, and Sandbox EKS clusters • Build and enforce policy-as-code guardrails: automated compliance checks against NIST 800-53, TRM, and VA-specific controls before production deployment • Manage Terraform IaC for all VAEC infrastructure provisioning; ensure all changes are code-reviewed and version-controlled • Lead SRE functions: define SLIs, track SLOs against PWS SLA tiers, manage error budgets, and own reliability reviews • Operate the Istio service mesh: traffic management, mTLS enforcement, and sidecar lifecycle across all EKS namespaces • Lead SAST, DAST, container scanning, and Fortify remediation — maintaining ATO with zero unmitigated critical CVEs • Maintain the onboarding pipeline: IaC-driven runbook bringing new engineers to readiness within contractual timeframes • Support Optional Task surge efforts for DevSecOps/SRE as defined in the Performance Work Statement




