Job Closed
This listing is no longer active.
A leading provider of AI-driven cloud business communications, contact center, video, and hybrid event solutions.
Site Reliability Engineer
Location
Colorado
Posted
132 days ago
Salary
$94.9K - $135.5K / year
Seniority
Mid Level
Job Description
Site Reliability Engineer
RingCentral
• Operate and maintain Linux based telephony and platform services in production • Troubleshoot SIP signaling and RTP media flows, including call routing, provisioning, registration, and signaling behavior • Diagnose issues across the network stack affecting realtime voice and media traffic, including analysis of packet and signaling flows • Deploy and administer Kubernetes services using Helm and GitOps workflows (e.g. CI/CD, FluxCD), including tracing and debugging configuration through layered rendering pipelines • Manage stateful and pinned workloads, including understanding of Kubernetes scheduling primitives such as taints, tolerations, and node affinity • Monitor systems and participate in on-call incident response for production infrastructure • Implement production changes using testing, rollback planning, and risk mitigation practices • Contribute to automation, observability, and operational tooling improvements • Coordinate with infrastructure, network, storage, and platform teams to resolve cross-domain issues and maintain highly available global services
Job Requirements
- Strong background administering UNIX/Linux systems and troubleshooting via the command line
- Solid grasp of TCP/IP networking fundamentals including routing, NAT, load balancing, and container networking
- Experience supporting SIP based VoIP or realtime communications systems, with strong understanding of SIP proxies, applications, session border controllers, RTP media servers
- Familiarity with Git based version control systems (e.g. GitLab) and common repository workflows
- Experience deploying services using GitOps automation (e.g. FluxCD, CI/CD)
- Skilled in analyzing network traffic using Wireshark, tcpdump, or similar tools
- Comfortable working with observability platforms including Prometheus, Grafana, Loki, and the ELK stack
- Hands on experience operating containerized platforms using Kubernetes, including interaction with the Kubernetes API, container registries, and Helm-based deployments
- Working knowledge of persistent storage in Kubernetes (e.g. PVCs)
- Experience operating infrastructure in AWS or GCP, including services such as S3, EC2, CloudFront, and certificate/secret management
- Familiarity with change management processes and production change approval workflows
- Experience automating operational tasks using Python or shell scripting
Benefits
- Comprehensive medical, dental, vision, disability, life insurance
- Health Savings Account (HSA), Flexible Spending Account (FSAs) and Commuter benefits
- Voluntary supplemental health coverage and life insurance
- 401K match and ESPP
- Paid time off and paid sick leave
- Paid parental and pregnancy leave
- Family-forming benefits (IVF, Preservation, Adoption etc.)
- Emergency backup care (Child/Adult/Pets)
- Employee Assistance Program (EAP) with counseling sessions available 24/7
- Free legal services that provide legal advice, document creation and estate planning
- Employee bonus referral program
- Student loan refinancing assistance
- Employee 1:1 coaching, perks and discounts program
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Senior Site Reliability Engineer – Metrics and Observability
CVS HealthBringing our heart to every moment of your health.
• Define, implement, and maintain key performance metrics, SLOs, and SLIs to measure system reliability and performance • Manage error budgets effectively, collaborating with development teams to balance reliability and feature delivery • Design and implement comprehensive monitoring solutions to provide real-time visibility into system health • Develop and implement automated quality gates that ensure all releases meet defined reliability and performance standards • Assist in incident response efforts by providing insights from metrics and monitoring tools • Drive initiatives to enhance monitoring, observability, and reliability practices
Principal DevOps Engineer
Perforce SoftwareThe DevOps Edge for the Outperformers: Enable teams to build, manage & maintain apps — from code to business-ready.
• Responsible for building platforms and frameworks to create consistent, verifiable, and automatic management of applications and infrastructure between non-production and production environments • Mentored exceptional engineers and DevOps developers on Cloud technology and practice. • Implement application security best practices throughout the agile SDLC • Foster and advocate for a DevOps culture at Perforce to ensure efficient testing, delivery, and deployment of all software artifacts • Lead the development and enhancements of our CI/CD pipeline infrastructure/tools. • Establish technical design principles and practices and drive them across all product portfolios to make operation design a must-have phase of the development lifecycle • Your daily tasks will include developing a technical design for our cloud platform, developing a framework, and working with • Dev and DevOps teams to ensure that the product meets our quality standards.
Senior Autonomy Release Engineer
May MobilityTransforming cities through autonomous technology to create a safer, greener, more accessible world.
• Release ownership and release execution end-to-end across: • Major autonomy releases • Incremental/performance releases • Hotfix/safety patches • Manage branching strategy, versioning, and release cut processes • Drive release readiness and go/no-go decisions • Partner cross-functionally with Autonomy, Infra, Validation, and Fleet Ops • Act as a system owner for release readiness • Investigate and resolve complex issues arising from: • Software/hardware interactions • Distributed systems behavior • On-vehicle vs simulation discrepancies • Develop deep understanding of: • Sensor stack, middleware, autonomy stack • Compute platforms, networking, configurations • Be the go-to for: • “Why does this fail in the real world?” • “What changed between releases?” • Enforce stage-gated release framework: • Feature Complete → Code Freeze → Validation → Release Candidate • Integrate validation signals: • Simulation corpus results • Regression testing • Vehicle testing (HIL / on-road) • Ensure safety-critical issues are identified, tracked, and gated • Take initiative to find and permanently solve challenging system level issues caused by the interplay between different software and hardware components. • Collaborate and lead system-wide improvements when working with other teams without having direct ownership or management responsibility. • Assess and develop approaches that scale and improve performance in a variety of ways (e.g. CPU performance, memory usage, disk usage, network usage).
• Implements complex solutions • Mentors junior engineers • Manages technical risks • Demonstrates a clear understanding of business impact




