Job Closed

This listing is no longer active.

Fieldguide logo
Fieldguide

Powering the future of trust with modern software for assurance & advisory firms.

Staff Site Reliability Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteLeadTeam 11-50Since 2020H1B SponsorCompany SiteLinkedIn

Location

California

Posted

90 days ago

Salary

$210K - $247K / year

Seniority

Lead

Job Description

Staff Site Reliability Engineer

Fieldguide

• Lead the design and evolution of highly scalable, fault-tolerant distributed systems across our cloud infrastructure. • Define and drive adoption of SLOs, SLIs, and error budgets across engineering teams. • Architect and continuously improve observability platforms (metrics, logging, tracing). • Own reliability strategy and roadmap, proactively identifying risks and driving long-term improvements. • Lead cross-team initiatives to improve system performance, scalability, and resilience. • Establish and enforce best practices for incident response, on-call, and operational excellence. • Drive root cause analysis and systemic improvements through blameless postmortems. • Champion automation and reduction of operational toil. • Guide capacity planning, load testing, and performance optimization efforts. • Design and validate disaster recovery, failover strategies, and resilience testing. • Mentor and coach engineers to elevate reliability engineering maturity. • Partner with Staff engineers across the organization to drive meaningful change • Partner with leadership to align business goals with reliability investments.

Job Requirements

  • 10+ years of experience in software engineering, with a focus on distributed systems and production infrastructure.
  • Extensive experience operating and scaling distributed systems in cloud environments, with a strong preference for AWS.
  • Deep expertise in system reliability, scalability, and performance engineering at scale.
  • Demonstrated experience implementing SLO-driven engineering practices and reliability frameworks.
  • Strong background building and owning observability ecosystems (e.g., Datadog, Prometheus, Grafana).
  • Proficiency with Infrastructure as Code tooling, particularly Terraform or equivalent.
  • Proven experience leading incident management, post-mortems, and production operations.
  • Strong software engineering fundamentals with the ability to contribute to and review complex codebases.
  • Track record of technical leadership and cross-functional influence across engineering and product teams.
  • Ability to balance tactical short-term needs with strategic long-term architectural improvements.
  • Excellent written and verbal communication skills, with the ability to translate complex technical concepts for diverse audiences.

Benefits

  • Competitive compensation packages with meaningful ownership
  • Flexible PTO
  • 401k
  • Wellness benefits, including a bundle of free therapy sessions
  • Technology & Work from Home reimbursement
  • Flexible work schedules

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Chooch AI logo

Software Engineer – Web/ML Dev Ops

Chooch AI

Stop stockouts. Cut inventory waste. Gain real-time visibility with automated supply chain management.

DevOps Engineer90 days ago
Full TimeRemoteTeam 51-200Since 2015H1B No Sponsor

- Web Development and Implementation: Design, develop, and optimize our web analytics services learning models for computer vision applications and key customer stakeholders. - Data Management: Collaborate with ML engineers to ensure the successful deployment and implementation of computer vision and LLM models. - Algorithm Implementation: Implement state-of-the-art computer vision algorithms to improve system accuracy and performance in real-world scenarios. - Cross-Functional Collaboration: Work closely within the solutions engineering team to integrate machine learning models into the broader product and systems. - On-Site Travel and Field Testing: Many of our projects are live deployments with major fortune 500 customers – testing new, innovative ML technologies. Traveling on site is sometimes needed to work with customers, gather requirements, and ensure the technology is operating as expected. - Performance Evaluation: Regularly evaluate the performance of computer vision models, using both qualitative and quantitative methods, and iterate to enhance accuracy and efficiency. - Research and Innovation: Understanding of the latest developments in web, backend, and working with our ML team to understand the necessary components to implement machine learning and computer vision, applying innovative approaches and technologies to solve complex challenges. - Stakeholder Engagement: Collaborate with internal and external stakeholders to understand their needs and translate them into effective technical solutions. - Problem-Solving: Address and troubleshoot complex problems that arise during the development and deployment of computer vision systems.

California
Job Closed

Role Description We are looking for a DevOps Engineer to help us build functional systems that improve customer experience. DevOps Engineer responsibilities include: - Implement integrations requested by customers - Deploy updates and fixes - Provide Level 2 technical support - Build tools to reduce occurrences of errors and improve the customer experience - Develop software to integrate with internal back-end systems - Perform root cause analysis for production errors - Investigate and resolve technical issues - Develop scripts to automate visualization - Design procedures for system troubleshooting and maintenance Qualifications - Work experience as a DevOps Engineer or similar software engineering role - Good knowledge of Ruby or Python - Working knowledge of databases and SQL - Problem-solving attitude - Team spirit - BSc in Computer Science, Engineering or relevant field

United States
Job Closed
CI&T logo

Senior DevOps, AWS

CI&T

Navigate Change

DevOps Engineer90 days ago
Full TimeRemoteTeam 5,001-10,000Since 1995H1B No Sponsor

• Design and implement infrastructure using Terraform, manage environments, and ensure security. • Develop and automate CI/CD pipelines for efficient deployment and environment promotion. • Establish and standardize observability practices, optimizing monitoring and alerting across platforms. • Integrate Data Quality processes with observability; document the framework and provide knowledge transfer to teams.

Colombia
Job Closed
CI&T logo

Senior DevOps, AWS

CI&T

Navigate Change

DevOps Engineer90 days ago
Full TimeRemoteTeam 5,001-10,000Since 1995H1B No Sponsor

• Design and implement infrastructure using Terraform, manage environments, and ensure security. • Develop and automate CI/CD pipelines for efficient deployment and environment promotion. • Establish and standardize observability practices, optimizing monitoring and alerting across platforms. • Integrate Data Quality processes with observability; document the framework and provide knowledge transfer to teams.

Brazil
Job Closed