Job Closed
This listing is no longer active.
Empowering you to share your message and create a lasting impression since 1995
Site Reliability Engineer
Location
Philippines
Posted
123 days ago
Salary
0
Seniority
Senior
Job Description
Site Reliability Engineer
DiscountMugs
• Focus on improving service reliability through automation • Define, measure, and manage SLIs, SLOs, and error budgets • Analyze system performance and identify opportunities for improvement • Build and optimize monitoring, logging, and alerting systems • Reduce toil through automation and build reliability-focused tools • Enhance CI/CD pipelines for safe deployments • Participate in on-call rotations and lead incident response
Job Requirements
- Bachelor's degree in Computer Science, Software Engineering, or related field
- 4+ years in SRE, DevOps, or Platform Engineering
- Strong programming or scripting experience
- Hands-on cloud experience
- Understanding of distributed systems
- Experience with observability tools
- Familiarity with chaos engineering and resilience patterns
Benefits
- Flexible working arrangements
- Professional development opportunities
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Participate in the monitoring, maintenance, and evolution of the system infrastructure supporting the TuoTempo web application. • Design, create, and support system distribution; test and monitor application code. • Support our internal development teams in the effective use of our organizational systems. • Propose and implement new solutions based on innovative technologies. • Planning and assignment of team activities. • Providing technical support to resolve complex issues. • Mentoring and professional development of team members. • Management of performance evaluations.
Site Reliability Engineer 5 – CDN, Live Streaming, Open Connect
NetflixPlay, pause, and resume watching anytime and anywhere.
• Support the CDN delivery and day-to-day live-streaming operations for Netflix • Participate in the preparation, validation, and execution of live streaming focused initiatives in collaboration with related production and engineering teams • Impact multiple areas of the live event lifecycle, from the planning phase through testing and event launch days • Lead innovation initiatives, implementing new features, and driving enhancements in the streaming services delivery • Drive continual improvement in resilience, observability, monitoring, instrumentation, and automation to maintain highly scalable and reliable CDN services with excellent quality of experience (QoE) • Implement, automate, execute, and analyze the results from a broad range of streaming CDN delivery focused functional, performance, resilience, and fault injection testing • Coordinate, collaborate, and partner across multiple stakeholders for the smooth execution of live-streaming events • Aggregate, analyze, and correlate large amounts of server and application performance data • Use the innovative Netflix Big Data platform as a toolset for service delivery optimization and system reliability improvements • Participate in an on-call rotation and work flexible hours based on live events schedule, including weekends and holidays
• Work on the automation and standardization of environments using Ansible, Git, and CI/CD pipelines. • Support and evolve workloads running on Kubernetes/OpenShift clusters. • Participate in the full deployment and release cycle, always prioritizing stability, risk mitigation, and continuous improvement. • Collaborate daily with international teams (in English). • Monitor, analyze, and respond to incidents in production environments. • Contribute to the maturity of the DevOps team in Brazil by sharing knowledge and best practices. • Occasionally participate in weekend deployment windows (rotational/on-call schedule).
Deployment Engineer
Collaborative RoboticsCollaborative Robotics' mission is to create a world where humans and robots collaborate in a trusted partnership.
• Deliver exceptional deployment support and elevate the customer experience. • Lead onsite hardware and software testing to ensure optimal cobot performance. • Troubleshoot, document, and resolve issues to enhance system reliability, and user experience. • Perform parts replacements and robot retrofits to maintain high operational standards. • Build and maintain strong customer relationships grounded in trust, reliability and mutual success. • Prepare sites for deployment through readiness assessments, checklists, and testing. • Create clear operational and technical documentation to build a foundation for excellence. • Partner with Program Manager to drive consistent communication and project progress.




