Job Closed
This listing is no longer active.
At OpenTable every employee has an impact on how we help restaurants around the world succeed.
Site Reliability Engineer II, Data Platforms
Location
India
Posted
146 days ago
Salary
0
Seniority
Senior
Job Description
Site Reliability Engineer II, Data Platforms
OpenTable
• Act as the primary SRE partner for the DBA team, bringing deep systems and reliability expertise to all managed database platforms. • Own the end-to-end observability stack for databases – define and implement metrics, logs, traces, dashboards, and actionable alerts for PostgreSQL, MongoDB, Redis, OpenSearch, YugabyteDB and related services. • Design, implement, and continuously improve monitoring and alerting specifically for database reliability (replication lag, connection saturation, query latency, storage and WAL usage, vacuum/bloat, backup health, etc.). • Lead and participate in on-call and incident response for database or system incidents • Build and maintain automation and self-service workflows for the DBA team using any Infrastructure and configuration management (Puppet/Ansible) – e.g. cluster provisioning, configuration rollout, user/role management, backup/restore orchestration, and failover procedures. • Develop and maintain runbooks, playbooks, and standard operating procedures for SRE aspects of database operations • Champion an automation-first, “no manual changes” culture for database and platform changes, and actively work to reduce toil through tooling and platform improvements. • Be a member of a weekly 24/7 on-call rotation • Strong collaboration skills as well as the ability to work independently
Job Requirements
- Bachelor's degree in Computer Science, Information Technology, or a related field (or equivalent practical experience).
- 4–7 years of total experience, with strong focus on SRE / Production Engineering / Platform Engineering
- Solid experience running Linux-based production systems at scale
- Proficiency in Infrastructure and configuration management (Puppet)
- Familiarity with containerization and orchestration technologies (Docker, Kubernetes) is a plus.
- Strong scripting / programming skills in one or more:
- Python, Go, Shell (bash), or similar
- Solid experience with Git and GitHub – branching and PR workflows, code reviews, and automating operations via GitHub Actions or similar CI systems.
- Experience building/maintaining CI/CD pipelines and integrating operational checks and tests
- Excellent problem-solving skills and a proactive, "can-do" attitude.
- Strong communication and collaboration skills to work effectively with cross-functional teams.
- Practical experience with:
- Metrics (Prometheus, CloudWatch, Grafana, etc.)
- Alerting (PagerDuty.)
Benefits
- Work from (almost) anywhere for up to 20 days per year
- Focus on mental health and well-being
- Company-paid therapy sessions through SpringHealth
- Company-paid subscription to Headspace
- Annual company-wide week off a year - the whole team fully recharges (and returns without a pile-up of work!)
- Paid parental leave
- Generous paid vacation + time off for your birthday
- Paid volunteer time
- Focus on your career growth
- Development Dollars
- Leadership development
- Access to thousands of on-demand e-learnings
- Travel Discounts
- Employee Resource Groups
- Quarterly team offsite
- Tax optimisation options
- Generous health insurance
- Pension fund
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Expanding the functionality of internal automation tools written in Bash and Python • Handling system configuration automation using SaltStack • Handling CI/CD mechanisms • Reacting to and investigating incidents and external support requests • Creating and expanding Terraform projects • Setting up monitoring and creating Grafana dashboards
• Develop and maintain cloud infrastructures and web applications focusing on FedRAMP standards • Develop, configure, and maintain CI/CD pipelines to automate build, test, and deployment processes • Monitor CI/CD pipeline performance, identify and resolve issues, and optimize pipelines • Work closely with engineering teams to design and shape the product deployment • Document CI/CD processes, pipeline configurations, and best practices
Senior DevOps Engineer (Remote - US)
NextechWe are a leader in specialty healthcare technology solutions. We’re committed to hiring and retaining talent, which is why we invest in our employees through competitive pay, a generous bonus structure, great healthcare, a comprehensive wellness program, and many other benefits. If you are a software engineer, finance or accounting professional, customer support specialist, or a business development expert with a passion for healthcare technology (just to name a few), we want to hear from you. We are an equal opportunity employer with a commitment to diversity. All individuals, regardless of personal characteristics are encouraged to apply.
Why join Nextech? We are a leader in specialty healthcare technology solutions. We’re committed to hiring and retaining talent, which is why we invest in our employees through competitive pay, a generous bonus structure, great healthcare, a comprehensive wellness program, and many other benefits. If you are a software engineer, finance or accounting professional, customer support specialist, or a business development expert with a passion for healthcare technology (just to name a few), we want to hear from you. We are an equal opportunity employer with a commitment to diversity. All individuals, regardless of personal characteristics are encouraged to apply. If you are a candidate in need of assistance or an accommodation in the application process, please contact talent@nextech.com. Job Summary The Senior DevOps Engineer role is focused on continually improving Nextech’s infrastructure, automation and operations. In collaboration with other DevOps Engineers, Software Development Engineers and Cloud Engineers, you will support our efforts to deliver high-quality, high-value, scalable solutions to our customers. All activities must be in compliance with Equal Employment Opportunity laws, HIPAA, ERISA and other regulations, as appropriate. Essential Functions - Infrastructure Automation & CI/CD: Design, implement, and maintain scalable infrastructure and automation processes, including Infrastructure as Code (IaC), CI/CD pipelines, configuration management, and monitoring systems. - Cross-Functional Collaboration: Partner with engineering teams to understand their technical needs and develop automated solutions that improve efficiency, scalability, and system reliability. - Monitoring & Optimization: Continuously enhance monitoring, alerting, and observability frameworks to proactively identify system performance and reliability improvements. - Continuous Improvement: Drive ongoing improvements in automation, infrastructure management, and process optimization to ensure high system availability, stability, and performance. - Mentorship & Leadership: Provide technical leadership and mentorship, conduct training sessions, and lead by example in best practices for DevOps, CI/CD, and infrastructure management. - Operational Excellence: Ensure the stability and reliability of production systems through day-to-day operations, incident response, and root cause analysis, including after-hours on-call support as needed. - Strategic Initiative Ownership: Take ownership of key projects, from concept to execution, ensuring alignment with business objectives, timelines, and technical requirements. - Effective Communication: Communicate technical concepts clearly and effectively to stakeholders at all levels, fostering collaboration and transparency across teams. - Carry out additional responsibilities as assigned based on business need. Minimum Requirements - 5+ years of working experience in the tech industry - 4+ years of DevOps experience - Strong understanding of Azure - Proven record of implementing DevOps in a Windows-based environment - Experience using Infrastructure as Code (IaC) tools such as Terraform to automate the creation of infrastructure - Experience implementing configuration management tools such as Ansible, Puppet, or Chef - Experience building secure CI/CD pipelines leveraging tools like Azure DevOps, GitHub Actions, or Jenkins - Experience with at least 1 scripting or programming language such as PowerShell, Python, .NET - Hands-on experience working with APM solutions such as Dynatrace, Datadog or New Relic - Experience working in an Agile method environment Preferred Qualifications - At least one Cloud related certification Working Environment/Physical Demands - 100% Remote - Activities require a significant amount of work in front of a computer monitor Total Rewards Generous annual bonus opportunity 401(k) with Employer Match Flexible Time Off: take time off when you need it without worrying about available hours 11 paid holidays Volunteer Time Off Insurance: Choice of Medical, Dental, and Vision plans Health Savings Account with employer match Flexible Spending Account 100% Company-Paid Parental leave (After 6 months with the company) 100% Company-Paid Life Insurance and Short/Long Term Disability Insurance Nextech Luminary Peer Recognition Program Wellness Program including discounts on medical premiums Employee Assistance Program with free counseling sessions available Corporate Discounts on Retail, Travel, and Entertainment Pet Insurance options
Senior Site Reliability Engineer
ArctiqArchitecting intelligent IT solutions in Enterprise Security, Modern Infrastructure & Platform Engineering.
• Define the strategy for Service Level Objectives (SLOs) and Error Budgets. • Design complex telemetry pipelines for full-stack observability. • Design and govern the enterprise Infrastructure as Code (IaC) standards. • Develop custom tooling to automate complex recovery procedures and system scaling. • Act as the Incident Commander for major system outages, leading the technical response and directing the Root Cause Analysis (RCA) process. • Lead the integration of security-as-code within DevSecOps pipelines, ensuring full compliance with RMF and NIST 800-53 standards. • Provide technical guidance and mentorship to Mid-Level SREs and developers, fostering a culture of reliability across the organization.



