Job Closed

This listing is no longer active.

Senior Site Reliability Engineer

Location

California

Posted

97 days ago

Salary

$150K - $180K / year

Seniority

Senior

Job Description

Senior Site Reliability Engineer

Tango

• Own reliability outcomes for Tango’s cloud platform across production and non-production environments • Design, implement, and operate SLOs/SLIs, error budgets, and reliability reporting • Drive prioritization of reliability work with Engineering and Product • Build and maintain observability foundations: metrics, logging, tracing, dashboards, and alerting • Lead incident response and post-incident reviews • Engineer and evolve CI/CD and release safety practices • Improve infrastructure-as-code and environment consistency • Partner with Security and Compliance to support secure operations • Optimize cloud cost and capacity through right-sizing and performance tuning • Enable engineering teams with reliable internal tooling and automation • Mentor engineers on reliability best practices

Job Requirements

  • 8+ years of experience in Site Reliability Engineering, DevOps, or Production Engineering supporting distributed SaaS applications
  • Strong background in Linux systems engineering
  • Networking fundamentals (TCP/IP, DNS, load balancing)
  • Proficiency with at least one programming language used for automation (e.g., Python, Go, or Java)
  • Strong scripting skills
  • Hands-on experience with cloud infrastructure (AWS, Azure, or GCP)
  • Deep experience with infrastructure-as-code and configuration management (e.g., Terraform, CloudFormation, Ansible)
  • Expertise in containerization and orchestration (Docker, Kubernetes)
  • Strong observability practices with tools such as Prometheus/Grafana, Datadog, New Relic, ELK/Splunk
  • Incident management leadership with a focus on root cause analysis
  • Experience designing and operating CI/CD pipelines and release management practices
  • Ability to work cross-functionally with Engineering, Product, Support, and Security
  • Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience
  • Relevant certifications are a plus (e.g., AWS/Azure/GCP, Kubernetes CKA/CKAD, ITIL)

Benefits

  • Competitive Compensation
  • Comprehensive Benefits Including health, dental, and vision insurance
  • 401(k) plan with company match
  • Generous paid time off
  • Flexible Work Environment
  • Inclusive & Collaborative Culture

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Coralogix logo

Site Reliability Engineer Team Lead

Coralogix

Full-stack observability for logs, metrics, traces and security events with built-in cost optimization.

DevOps Engineer97 days ago
Full TimeRemoteTeam 201-500Since 2014H1B No Sponsor

Role Description Coralogix is a modern, full-stack observability platform transforming how businesses process and understand their data. Our unique architecture powers in-stream analytics without reliance on expensive indexing or hot storage. We specialize in comprehensive monitoring of logs, metrics, trace and security events with features such as APM, RUM, SIEM, Kubernetes monitoring and more, all enhancing operational efficiency and reducing observability spend by up to 70%. We are looking for a Site Reliability Engineer Team Lead to lead our Cloud Infrastructure Team, focusing on Enterprise FedRAMP Cloud Infrastructure. In this role, you will: - Lead and mentor a team of engineers, including hiring, onboarding, and performance management. - Work in high scale environments - Coralogix data pipeline processes 55Tb of data each day. - Adopt cutting edge technologies with end-to-end responsibility. - Build internal tools to expand our platform capabilities. - Collaborate with R&D to improve stability & reliability of the system. - Lead the product roadmap - our product is designed for engineers. Therefore, our engineers promote, enhance, and take a crucial part in influencing the product roadmap. - Perform operational duties for FedRAMP cloud products, including deployments, on-call support, and incident management. This role is remote; employees must be within EST / CT time zone. Our tech stack is unique and in constant growth: - Kubernetes - Kops - AWS - Kafka - Prometheus - Thanos - Coralogix - Git - Argo CD - Istio - and many more! Qualifications - 2+ years of experience as a Team Lead / Tech Lead. - At least 5 years of experience as a DevOps Engineer/ SRE in production environments. - At least 2 years of experience with FedRAMP compliance (High/Moderate levels), vulnerability management, and continuous monitoring, including scanning, patching, and reporting. - In-depth experience with Kubernetes - operating & monitoring are key parts. - High familiarity with monitoring tools such as Coralogix, Grafana, Prometheus. - Experience in AWS or other cloud providers. - Experience with infrastructure as code (Terraform, Crossplane, etc.). - Understanding of networking - from networking layers to different networking protocols (http, grpc, ssl). - Some software engineering experience, preferably in Golang. - An advantage - operating data pipelines. - An advantage - familiarity with Apache Kafka. Cultural Fit We’re seeking candidates who are hungry, humble, and smart. Coralogix fosters a culture of innovation and continuous learning, where team members are encouraged to challenge the status quo and contribute to our shared mission. If you thrive in dynamic environments and are eager to shape the future of observability solutions, we’d love to hear from you. Compensation and Rewards The earnings range for this role is $230,000 - $270,000. When determining your salary, we consider your experience, skills, education, and work location. Our total compensation package includes comprehensive and inclusive employee benefits for healthcare, dental, and mental health benefits, a 401(k) plan and match, paid sick time, and paid time off. Coralogix is an equal opportunity employer and encourages applicants from all backgrounds to apply.

EST (UTC-5) + 1 moreAll locations: EST (UTC-5) | CST (UTC-6)
$230K - $270K / year
Job Closed
Coralogix logo

Site Reliability Engineer

Coralogix

Full-stack observability for logs, metrics, traces and security events with built-in cost optimization.

DevOps Engineer97 days ago
Full TimeRemoteTeam 201-500Since 2014H1B No Sponsor

Role Description We are looking for a Site Reliability Engineer to work as part of our Cloud Infrastructure Team, focusing on Enterprise FedRal Cloud Infrastructure. - Work in high scale environments - Coralogix data pipeline processes 55Tb of data each day - Adopt cutting edge technologies with end-to-end responsibility - Building internal tools to expand our platform capabilities - Collaborate with R&D to improve stability & reliability of the system - Lead the product roadmap - our product is designed for engineers. Therefore, our engineers promote, enhance, and take a crucial part in influencing the product roadmap. - Perform operational duties for FedRAMP cloud products, including deployments, on-call support, and incident management. This role is remote, employees must be within EST / CT time zone. Our Tech Stack Is Unique And In Constant Growth: - Kubernetes - Kops - AWS - Kafka - Prometheus - Thanos - Coralogix - Git - Argo CD - Istio - and many more! Qualifications - At least 5 years of experience as a DevOps Engineer/ SRE in production environments - In-depth experience with Kubernetes - operating & monitoring are key parts - At least 2 years of experience with FedRAMP compliance (High/Moderate levels), vulnerability management, and continuous monitoring, including scanning, patching, and reporting - advantage - High familiarity with monitoring tools such as Coralogix, Grafana, Prometheus - Experience in AWS or other cloud providers - Experience with infrastructure as code (Terraform, Crossplane, etc.) - Understanding of networking - from networking layers to different networking protocols (http, grpc, ssl) - Some software engineering experience, preferably in Golang. - An advantage - operating data pipelines - An advantage - familiarity with Apache Kafka Cultural Fit We’re seeking candidates who are hungry, humble, and smart. Coralogix fosters a culture of innovation and continuous learning, where team members are encouraged to challenge the status quo and contribute to our shared mission. If you thrive in dynamic environments and are eager to shape the future of observability solutions, we’d love to hear from you. Compensation and Rewards The earnings range for this role is $170,000-$220,000. When determining your salary, we consider your experience, skills, education, and work location. - Our total compensation package includes comprehensive and inclusive employee benefits for healthcare, dental, and mental health benefits - A 401(k) plan and match - Paid sick time and paid time off Coralogix is an equal opportunity employer and encourages applicants from all backgrounds to apply.

Michigan + 39 moreAll locations: Michigan | Indiana | Kentucky | Tennessee | Georgia | Florida | Ohio | North Carolina | South Carolina | West Virginia | Virginia | Pennsylvania | District Of Columbia | Connecticut | New Jersey | New York | Rhode Island | New Hampshire | Maine | Maryland | Delaware | Vermont | Massachusetts | North Dakota | South Dakota | Nebraska | Kansas | Oklahoma | Texas | Minnesota | Iowa | Missouri | Arkansas | Louisiana | Wisconsin | Illinois | Mississippi | Alabama | EST (UTC-5) | UTC-5 to UTC-3
$170K - $220K / year
Job Closed
KBR, Inc. logo

Senior Software/DevSecOps Engineer

KBR, Inc.

We deliver science, technology and engineering solutions to governments and companies around the world.

DevOps Engineer97 days ago
Full TimeRemoteTeam 10,001+Since 1901H1B No Sponsor

• Work as part of the team supporting the Test Resource Management Center’s (TRMC) Test and Training Enabling Architecture (TENA), Cloud Hybrid Edge to Enterprise Test Analysis Suite (CHEETAS), and Joint Mission Environment Testing Capability (JMETC) User Support Teams • Develop software applications that interface with the TENA and CHEETAS products • Lead and mentor junior members of the team on development efforts • Documenting, managing configuration, testing, analysis and bug fixing involved in creating and maintaining applications and frameworks involved within an agile software release life cycle

Alabama + 5 moreAll locations: Alabama | Colorado | District Of Columbia | Florida | Virginia | Washington
$166.8K - $220K / year
Job Closed

Senior Java Engineer

Evolve Today

Evolve Today is a recruitment agency connecting top engineering talent with world‑class opportunities.

DevOps Engineer97 days ago

Role Description We are hiring a Senior Java Backend Engineer on behalf of our client — a global next‑generation technology partner delivering intelligent, scalable digital solutions across media, finance, healthcare, and banking. Your role will involve: - Designing and building distributed backend systems, high‑throughput microservices, and large‑scale data processing pipelines. - Working in complex architectures with high‑volume data flows. - Integrating ML models into production. This role suits a Senior Java Engineer with strong JVM expertise, deep backend fundamentals, and hands‑on experience — or strong motivation — in Big Data and ML‑driven systems. Qualifications - Strong proficiency in Java and a deep understanding of the JVM ecosystem (performance tuning, concurrency, memory model). - Solid backend engineering fundamentals: API design, microservices, observability, CI/CD. - Experience with distributed systems (horizontal scaling, resilience patterns, distributed coordination). - Enterprise‑level engineering experience in data‑intensive environments. - Strong command of clean code and design patterns. Requirements - Build and optimize Java microservices with strict SLAs (latency, throughput, reliability). - Design and evolve distributed systems handling large‑volume, high‑velocity data. - Develop and maintain Big Data pipelines for ML training, inference, and personalization. - Integrate ML models into production with robust deployment, monitoring, and rollback strategies. - Apply clean code, design patterns, and engineering best practices in high‑complexity environments. - Contribute to architectural decisions involving scalability, partitioning, caching, concurrency, and fault tolerance. - Collaborate with Data Science teams to operationalize ML workloads at scale. Company Description Evolve Today is a recruitment agency connecting top engineering talent with world‑class opportunities.

Worldwide
Job Closed