Where partnerships drive potential.
Senior Site Reliability Engineer
Location
California
Posted
2 days ago
Salary
$165K - $195K / year
Seniority
Senior
Job Description
Senior Site Reliability Engineer
Juniper Square
• Own reliability and observability across the organization — SLAs/SLOs, instrumentation, and on-call health • Design, deploy, and maintain Kubernetes infrastructure (Helm, EKS/ECS) and core AWS services (RDS/Aurora Postgres, networking, scaling) • Build and maintain CI/CD pipelines in GitHub Actions and Argo/Helm • Drive adoption of Infrastructure as Code standards (Terraform, and increasingly Crossplane) across multiple teams • Partner with 16+ engineering teams to roll out new standards, tools, and processes • Participate in on-call rotation and lead incident response/root-cause analysis for issues that cross team boundaries • Use AI tools daily to work faster and more effectively • Represent SRE's perspective in planning conversations with engineering leadership
Job Requirements
- 8-10 years of experience as a Senior SRE, DevOps, or Infrastructure Engineer
- Background at smaller-to-mid-size, high-growth SaaS companies (Series A/B through C/D)
- Deep hands-on experience with Kubernetes (EKS or ECS), GitHub Actions, Terraform, CI/CD pipelines, and Helm/Argo
- Real database experience with Postgres and DocumentDB (Mongo) across RDS/Aurora
- A track record of working cross-functionally — managing timelines, setting expectations, and collaborating with other teams' engineers
- Already using AI as part of your actual workflow (coding, debugging, design)
Benefits
- Health, dental, and vision care for you and your family
- Life insurance
- Mental wellness coverage
- Fertility and growing family support
- Flex Time Off in addition to company-paid holidays
- Paid family leave, medical leave, and bereavement leave policies
- Retirement saving plans
- Allowance to customize your work and technology setup at home
- Annual professional development stipend
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Site Reliability Developer
Vena SolutionsTake your entire business from reactive to proactive with the leading AI-Powered Complete FP&A Platform.
• Support key ITIL processes, including Incident management, request management, problem management and change management. • Define and document runbooks and standard operating procedures. • Field operational requests from our Application Support team and other internal stakeholders • Triage and solve issues within defined SLA’s to ensure an excellent customer experience and to unblock other development and support teams • Maintain services once they are live by measuring and monitoring availability, latency and overall system health. • Identify and troubleshoot problems, investigate root causes, and champion fixes across the organization. • Work with infrastructure-as-Code (IaC) with a focus on continuous improvement. • Collaborate with cross-functional team members on features and implementation within an agile environment. • Report on SLAs and performance metrics as part of the Operations function. • Participate in on-call rotation.
• proaktywne dostrzeganie wyzwań i ich adresowanie wspólnie z zespołami backendowymi i devops • tworzenie szablonów i standardów infrastrukturalnych dla najczęstszych potrzeb zespołów developerskich (np. nowe serwisy, projekty Cloud Run) • rozwój i utrzymanie stacku observability (Loki, Grafana, Prometheus) oraz optymalizacja wykorzystania zasobów i kosztów chmury (GCP) • utrzymanie i rozwój wewnętrznych narzędzi self-hosted (m.in. NocoDB, n8n, Outline) • wsparcie infrastrukturalne zespołów Data Engineering i Data Science (m.in. Airflow, pipeline'y forecastingowe czy serwery pod ML) • zarządzanie infrastrukturą sieciową w darkstore'ach (sieć, CCTV, drukarki paragonów, urządzenia handheld) we współpracy z IT Managerem • budowanie i utrzymanie infrastruktury on-prem (serwery pod ML i workloady wymagające dużych zasobów, runnery CI/CD dla projektów mobilnych) • utrzymanie i usprawnianie procesów CI/CD (GitLab) oraz developer experience • budowanie infrastruktury pod rozwiązania AI (autonomiczne agenty, boty Slack, agenty code review, remote coding agents) • rozwój praktyk SRE: alerting, procesy zgłaszania i obsługi incydentów (PagerDuty, Slack)
Senior DevOps Engineer
MKS2 TechnologiesAustin-based SDVOSB delivering application development, cybersecurity, instructional design and training to DOD and VA.
• Conduct DevOps and DevSecOps activities for an Azure integration platform • Create and maintain GitHub CI/CD pipelines • Conduct Azure DevOps activities • Conduct Azure system administration • Build platform automation • Create and maintain PowerShell scripting • Understanding and provisioning of Azure Infrastructure services and network • Create and maintain containers and infrastructure (IaC, Terraform) • Deploy solutions through CI/CD pipelines • Embed DevOps best practices and process/tech improvements across the team and platform • Leverage AI tools to accelerate activities • Collaborate across technical teams • Create and maintain detailed technical documentation • Support troubleshooting and resolution of Production issues • Participate in Agile ceremonies • Report on progress and status of development
Senior DevOps Engineer
FiservFounded in 1984, Fiserv is a global provider of ecommerce and information management systems for the financial services industry. In 2013, Fiserv acquired Open
• Design, implement, and support secure Azure IaaS and PaaS solutions across enterprise environments • Build and maintain CI/CD pipelines using Azure DevOps and GitLab to enable efficient, automated software delivery • Develop and manage Infrastructure as Code solutions using Terraform and Ansible for provisioning and configuration management • Implement and maintain Zero Trust security architectures, including identity and access management (IAM), data encryption, and security controls • Identify, assess, and remediate infrastructure and application security vulnerabilities • Create and maintain monitoring, alerting, and operational dashboards using Azure Monitor, Log Analytics, and Application Insights • Troubleshoot production issues, restore services, and maintain operational runbooks and knowledgebase documentation • Collaborate with cross-functional teams to support cloud modernization initiatives




