Job Closed

This listing is no longer active.

Arbor Education logo
Arbor Education

Arbor MIS helps schools and MATs work more easily and collaboratively. Join a free webinar: http://bit.ly/Arbor-webinars

Site Reliability Engineer

Location

United Kingdom

Posted

74 days ago

Salary

£60K - £70K / year

Seniority

Mid Level

Job Description

Site Reliability Engineer

Arbor Education

Role Description We are looking for an enthusiastic and proactive Site Reliability Engineer to join our SRE team and help us ensure we provide world-class resilience and performance across the platform. The remit and focus of the role is to advise on all aspects of site reliability including availability, scalability, observability and capacity planning. It’s a broad and exciting role, so we’re looking for someone up for a challenge - if you’re an energetic and a collaborative Site Reliability Engineer, this is the role for you. Core Responsibilities - Proactively monitor and analyse platform performance. - Collaborate with engineering teams to address performance bottlenecks and ensure scalability. - Assist engineering teams with implementing and reviewing SLOs. - Continually improve observability through monitoring and alerting, and dashboards, using tools such as DataDog or Prometheus. - Work with other teams to ensure it is effective and provides full coverage. - Ensure the service is highly available and resilient. - Champion best practices in design for high availability. - Devise runbooks and run game sessions to test our DR plan, H/A and backups. - Conduct assessments of capacity and plan for scaling to meet current and future business needs. - Work closely with the Head of Platform Engineering and Head of SRE to strategize and implement scalable solutions. - Work closely with the Platform team, feature teams and 2nd line support and other stakeholders to ensure a good level of service is provided for our customers and embed SRE practices. - Key player in the response and troubleshooting of incidents, ensuring rapid resolution and minimising downtime. - Participate in blameless postmortems to identify root cause and corrective actions. - Develop and maintain playbooks and documentation. Qualifications - Experience in performance monitoring and analysis. - Capacity planning experience. - Scripting and automation skills, with experience in relevant technologies. - Experience with Infrastructure as Code, in particular, Terraform. - Understanding of relational database technologies and their cloud versions (e.g. AWS Aurora). - Experience with messaging and distributed asynchronous workloads. - Experience with nginx or similar technologies. - Familiarity with SRE processes. - Aware of DevOps principles like the 3 ways and 5 ideals. Bonus Skills - Experience with other database technologies and cloud platforms. - Past experience with Enterprise solutions running at scale. - Familiarity with Kanban and Agile development processes. - Experience with containerisation, for example Docker. - Familiarity with software best practices such as Refactoring, Clean Code, Domain-Driven Design and Test-Driven Development. Benefits - The chance to work alongside a team of hard-working, passionate people in a role where you’ll see the impact of your work every day. - A dedicated wellbeing team who champion initiatives such as mindfulness, lunch n learns, manager training, mental health first aid training and much more! - 32 days holiday (plus Bank Holidays). This is made up of 25 days annual leave plus 7 extra company wide days given over Easter, Summer & Christmas. - Life Assurance paid out at 3x annual salary. - Comprehensive wellness benefit provided by AIG Smart Health, which provides a 24/7 virtual GP service, Mental health support, Counselling, and personalised Health Checks. - Private Dental Insurance with Bupa. - Salary sacrifice Pension provided by Scottish Widows. - Enhanced maternity and adoption leave (20 weeks full pay) and paternity (6 weeks full pay) pay. - 5 free return to work maternity coaching sessions, helping you adapt to this new exciting time of life! - Access to services such as Calm and Bippit (financial wellbeing coaching). - All of our roles champion flexible working and we are happy to discuss what this means to you. - Social committees that plan team, office and company wide events to bring people together and celebrate success. - Dedicated professional development training budget (CPD courses, upskilling resources, professional memberships etc). - Volunteer with a charity of your choice for a day each year. - Dog friendly offices! Interview Process - Phone screen - 1st stage - 2nd stage

Related Categories

Related Job Pages

More Engineer Jobs

Nokia logo

Senior Software Engineer

Nokia

At Nokia, we create technology that helps the world act together.

Engineer74 days ago
Full TimeRemoteTeam 10,001+Since 1865H1B Sponsor

• Responsible for the development and/or maintenance of assigned deliverables in your area. • Ensure timely fault resolutions for any issues that arise. • Prepare and document designs for review. • Participate in specification reviews to ensure quality deliverables.

Poland
Job Closed
Switzerland Global Enterprise logo

Turbine Resident Engineer

Switzerland Global Enterprise

We support Swiss SMEs in their international business and help innovative foreign companies to establish in Switzerland.

Engineer74 days ago
Full TimeRemoteTeam 51-200Since 1927H1B No Sponsor

• Assume overall technical responsibility to support the customer in identifying plant risks and problems • Recommend suitable solutions to improve plant performance and availability • Respond to customer requests for technical issue resolution relating to steam turbines • Perform plant walk downs, including a comprehensive visual inspection of the plant • Submit a monthly report to the client • Attend a 3 monthly technical discussion meeting between similar designed Eskom Power Stations • Provide detailed reports to relevant client representatives when requested • Be available for emergency consultation during unplanned turbine failure incidents • Inform the client of the latest engineering designs and developments that may improve the plant design • Regularly inform the client about the fleet experience on similar turbine islands supplied by the Engineer • Assist with compiling technical specifications for the refurbishment of turbine components • Provide turbine related training as per the needs of the client’s turbine system engineers • Communication channel to the OEM on technical matters • Provision of information and documentation as required • Revise, monitor and update relevant GE Vernova documentation • Compile technical specifications for plant operation and maintenance • Identify and optimize spares to be held to support planned overhauls • Determine maintenance requirements and the scope of work prior to overhauls • Recommend scope of work for all major planned outages • Investigate and report on turbine occurrences • Record monthly all significant abnormal events on the turbine • Advise on the required turbine history recording

United States
Job Closed

Role Description Echo360 is seeking a Senior / Staff Full Stack Engineer to design, build, and operate high-impact features across our learning platform used by global education and enterprise customers. You will work across the full stack—owning features from frontend to backend to production systems—while collaborating closely with infrastructure teams to deliver reliable, scalable applications in AWS. This role is scoped for senior engineers with strong ownership as well as staff-level engineers who influence architecture, systems, and engineering practices across teams. What You’ll Do - Build and maintain frontend features using React and TypeScript - Design and implement backend services and APIs using Node.js - Own features end-to-end, including deployment, monitoring, and performance - Collaborate with infrastructure teams to manage AWS environments using Terraform - Improve service reliability, observability, and performance (monitoring/APM) - Contribute to code reviews, testing practices, and engineering standards - Mentor and collaborate with engineers across teams For more experienced candidates (Staff scope): - Lead design of complex systems and influence architectural decisions - Drive improvements to engineering practices, CI/CD, and developer workflows - Partner across teams to align on technical direction and system design Qualifications - 6+ years of professional software development experience - Strong experience with JavaScript/TypeScript, Node.js, and React - Experience with databases such as Postgres, MySQL, and Redis - Experience building and operating production systems in AWS Requirements - Experience with CI/CD pipelines and production deployments - Strong testing, documentation, and code review practices - Understanding of system performance, monitoring, and reliability - Experience working with SRE/Infrastructure teams - Ability to take ownership of systems in production - Experience mentoring or supporting other engineers Additional Experience (Nice to Have) - Familiarity with Python, Scala, or Backbone - Experience using AI-assisted development tools to improve productivity and code quality What Success Looks Like - You deliver production-ready features and systems across the stack - You improve system reliability, performance, or developer workflows - You contribute to strong engineering practices and team effectiveness At Staff level: - You influence architecture and technical decisions across teams - You elevate engineering standards and mentor other engineers Additional Details - Fully remote role (U.S.-based candidates only) - Preference for candidates in Pacific, Mountain, or Central time zones - Must be eligible to work in the United States without sponsorship - Salary Range: $130,000 – $185,000 annually (Compensation may vary based on experience and level)

United States
$130K - $185K / year
Job Closed
Five9 logo

Telecom Engineer

Five9

Helping Companies Bring Joy to CX.

Engineer74 days ago
Full TimeRemoteTeam 1,001-5,000Since 2001H1B Sponsor

• Interface directly with customer’s telecom engineers and IT teams to deploy customized solutions and trouble shoot issues; • Assist in the day-to-day operational support of the Telecommunications network, analyzing problems affecting network availability and customer quality reports; • Escalation point of contact for the Network Operations Center to resolve critical alerts generated by the SBC’s Analyze history of telecommunication related incidents and perform preventive measures; • Assist in the day-to-day operation of the telecommunications network, where necessary analyzing problems affecting network availability and customer/vendor service quality; • Provide root-cause analyses on service outages; • Respond to Telecom alerts and alarms and provide corrective actions in coordination with the Network Operations Center; • Create technical documentation and call flow drawings as needed Manage Telecommunications Service Provider & Vendors Implement hardware and software deployments on the telephony network as required; • Deploy new services including interop testing with telecom carriers and customers; • Provide emergency assistance and technical recovery in the event of an emergency situation affecting the availability and/or service quality of the network; • Participate in on-call emergency rotation.

United Kingdom