Golden 1 is California's leading credit union. Insured by NCUA. Equal Housing Opportunity. NMLS #669333
DevOps Engineer II – IT Infrastructure Systems
Location
California
Posted
6 days ago
Salary
$123.6K - $135K / year
Seniority
Senior
Job Description
DevOps Engineer II – IT Infrastructure Systems
Golden 1 Credit Union
• Independently lead infrastructure-as-code development using Terraform and scripting languages such as Python and PowerShell to support scalable and reliable deployments. • Manage Linux/Kubernetes cluster environments. • Deploy solutions in accordance with Change Management Processes. • Support development teams on API integration strategy and standards development. • Ensure systems are secure against cybersecurity threats. • Identify technical problems and develop software updates and fixes. • Strong Splunk skills for administration, query optimization, alerting, and dashboard development. • Build tools to reduce errors and improve customer experience. • Propose ideas and solutions within the Infrastructure Department to reduce workload through automation. • Design, implement, and optimize CI/CD pipelines for faster and more reliable software releases. • Independently conduct root cause analysis and implement corrective actions. • Design and write tests to investigate infrastructure failure and scaling. • Create and maintain response playbooks across incident management and monitoring tools. • Develop automation to ensure repeatability, eliminate toil, and reduce time to action and repair services. • Analyze key operational metrics to identify opportunities to improve availability. • Implement effective monitoring, alerting, and reduction of alert fatigue. • Manage container orchestration environments and optimize deployment workflows to enhance scalability, reliability, and operational efficiency. • Design, build, and manage containerized environments using Docker. • Create and maintain SLIs, SLOs, and error budgets. • Design and optimize monitoring dashboards and alerting systems to proactively detect and address application performance and uptime issues. • Implement code branching strategies using GitHub functions. • Advanced Terraform syntax and GitLab CI/CD configuration, pipelines, jobs. • Provisioning and setting up metrics in Prometheus, Thanos, and Grafana, creating and managing alerts. • Implement cloud engineering standards, reusable modules, and platform patterns in Microsoft Azure. • Operate shared cloud platform services according to Cloud Engineering defined architectures. • Ensure infrastructure changes comply with reliability, security, and cost controls established by Cloud Engineering. • Maintain operational documentation and runbooks for cloud platform services.
Job Requirements
- Over 4 years as a DevOps Engineer in medium to large-scale environments.
- Proficient in Windows Server, Linux, and hybrid cloud deployments using Microsoft Azure and VMWare.
- Skilled in Git/GitHub workflows, Terraform, Python, PowerShell, and container orchestration (Tanzu, Docker, Kubernetes, OpenShift).
- Experienced with CI/CD tools (Jenkins, GitLab CI, Azure DevOps) and observability platforms (Datadog, Prometheus, Grafana, ThousandEyes).
- Knowledgeable in log management (ELK Stack) and database technologies (PostgreSQL, MySQL, NoSQL).
- Strong background in automating infrastructure provisioning and application deployment using Terraform, Ansible, and Kubernetes.
- Proficient in creating and maintaining monitoring dashboards, SLIs, SLOs, and error budgets to ensure application uptime and performance.
- Experienced in ensuring infrastructure security, driving automation initiatives, and collaborating across teams to improve reliability and scalability.
- Experienced in building observability pipelines and performing advanced queries in log management tools like Splunk for troubleshooting.
- Experience implementing and operating Azure-based shared services defined by platform or cloud engineering teams.
- Microsoft Azure DevOps Engineer Expert Certification (Required).
- Kubernetes Administration Certification (Required).
- Linux Certification (Desired).
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Own and operate CI/CD pipelines and blue/green/canary deployment frameworks across Production, Staging, and Sandbox EKS clusters • Build and enforce policy-as-code guardrails: automated compliance checks against NIST 800-53, TRM, and VA-specific controls before production deployment • Manage Terraform IaC for all VAEC infrastructure provisioning; ensure all changes are code-reviewed and version-controlled • Lead SRE functions: define SLIs, track SLOs against PWS SLA tiers, manage error budgets, and own reliability reviews • Operate the Istio service mesh: traffic management, mTLS enforcement, and sidecar lifecycle across all EKS namespaces • Lead SAST, DAST, container scanning, and Fortify remediation — maintaining ATO with zero unmitigated critical CVEs • Maintain the onboarding pipeline: IaC-driven runbook bringing new engineers to readiness within contractual timeframes • Support Optional Task surge efforts for DevSecOps/SRE as defined in the Performance Work Statement
• You define and enforce SLOs, SLIs, and SLAs across production • You monitor system health, plan capacity, and automate deployments, patching, and infrastructure provisioning • You diagnose and resolve production incidents fast – including 24/7 on-call participation • You lead post-mortems and turn findings into prevention; maintain runbooks and escalation procedures • You manage cloud infrastructure via IaC and own CI/CD pipeline design and maintenance • You drive scalability, fault tolerance, disaster recovery, and security compliance • You partner with Dev teams on production readiness, Shift Left practices, and error budget management
DevOps Engineer, Marketplace Team
eMAGWe create e-commerce solutions for customers & partners to make their everyday lives easier.
• Participate in migrating on-prem services to the Cloud; • Configure and tune a variety of services in order to increase system stability, performance and security; • Develop, maintain and improve automation for all our systems; • Capacity planning and performance tuning for high traffic events such as Black Friday; • Get involved in designing and implementation of new features; • Troubleshoot issues and provide support for internal teams; • Identify repetitive issues or tasks and focus on automation and implementation of permanent solutions; • Continuous improvement of monitoring and alerting systems.
Senior DevOps Engineer
Group-IBGlobal Threat Hunting and Adversary-Centric Cyber Intelligence Company
• Maintain and optimize core server infrastructure, including bare-metal servers, LXC containers, virtual machines, and cloud environments. • Operate and support core infrastructure services such as Nginx, Puppet, GitLab, Artifactory, Nexus, Harbor, Grafana, etc. • Manage and evolve infrastructure following the Infrastructure as Code (IaC) paradigm. • Handle and resolve incidents related to infrastructure operations. • Collaborate closely with cross-functional teams (network engineers, developers, and other technical stakeholders). • Design and implement high-availability, fault-tolerant, and scalable software solutions. • Monitor service performance and availability using modern observability tools, ensuring system reliability and optimal resource utilization.




