Managed SASE solutions that securely connect hybrid IT environments.
Senior Site Reliability Engineer
Location
Switzerland
Posted
145 days ago
Salary
0
Seniority
Senior
Job Description
Senior Site Reliability Engineer
Open Systems
• Building Operational Automation: Design, build, and evolve the automation framework and tooling that powers the MC platform — primarily in Golang — with a strong focus on maintainability, scalability, and reliability. • Developing Self-Service APIs: Build and maintain the service APIs and self-service operational tooling of the MC platform that enable customers and teams to safely and efficiently operate services in production without manual intervention. • Applying Site Reliability Engineering (SRE) Principles: Define, implement, and continuously improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, and SLA measurements so that reliability is measurable and actionable across the MC platform and the services it operates. • Owning Reliability and Operations Initiatives: Take ownership of reliability, automation, and Mission Control projects, driving them independently from problem identification through implementation and long-term operation. • Collaborating with AI Engineering: Work closely with our AI team and tooling, integrating AI-assisted capabilities into our automation and operational workflows. • Incident Response and Learning: Participate in incident response and the on-call rotation, leading root cause analysis and driving sustainable corrective and preventive actions. • Lead midsize to large automation and reliability initiatives, from early concept and design through production deployment and ongoing operations.
Job Requirements
- Strong software engineering background, ideally with production-grade Go (Golang) experience — proficiency in another modern language is fine if you are a quick learner and willing to adapt, as Go is our primary language.
- Solid understanding of distributed systems and scalable architecture.
- Proven experience designing, building, and operating services and their APIs (e.g. REST, gRPC) in production.
- Experience operating production systems, including incident response, on-call, and root cause analysis.
- Experience with SRE, DevOps, platform engineering, or reliability-focused roles.
- Hands-on experience with infrastructure and operations tooling, such as: Kubernetes Terraform / Infrastructure as Code GitOps principles and CI/CD tooling Prometheus, Loki, Tempo, and modern observability stacks.
- Major cloud platforms (AWS, Azure, GCP).
- Knowledge of Linux system administration, networking concepts, and major Internet protocols (TCP/IP, IPsec, SSL , SSH, SMTP, HTTPS, DNS).
- Ability to think in terms of systems, failure modes, and trade-offs.
- Strong communication skills and the ability to build trust across engineering and operations teams.
- A proactive mindset: you see problems, propose solutions, and take responsibility for delivering them.
- University degree in Computer Science or similar educational level.
Benefits
- Caring PASSIONATELY about keeping our customers safe – We’re dedicated to solving problems. Whatever it takes.
- Thinking UNCONVENTIONALLY to stay ahead – The world never fails to surprise us. So let’s surprise it first.
- Doing the hard work to make things SIMPLE – Craft and hone something that delights in its simplicity.
- Working COLLABORATIVELY to build success – The power of the team will always make us faster and better.
- Open Systems has been recognized as an outstanding place to work.
- Surrounded by smart teams who enrich your experience and provide opportunities you will need to develop your skills and advance your career.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Release Management Engineer
Arizona Department of AdministrationThe Attorney General's Office offers a comprehensive benefits package. For a complete list of benefits provided by The State of Arizona, please visit our benefits page.
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description Would you like to be part of an amazing team that helps Arizonans thrive? At the Department of Economic Security (DES) we strengthen individuals, families, and communities for a better quality of life. DES is looking for individuals who are committed to service, community, and teamwork. The Information Technology Service Management (ITSM) Engineer will be responsible for managing the full life cycle of major enhancements, as well as the creation of new solutions to support the ITSM team’s process improvement, user experience enhancement, and automation efforts. The position will work with internal DES customers to identify opportunities for process improvement based on established business needs to then design and implement relevant solutions that meet requirements supporting those business needs. This position is available for remote work on a full-time basis within Arizona, based upon the department's business needs and continual meeting of expected performance measures. The State of Arizona strives for a work culture that affords employees flexibility, autonomy, and trust. Across our many agencies, boards, and commissions, many State employees participate in the State’s Remote Work Program and are able to work remotely in their homes, in offices, and in hoteling spaces. Qualifications - Bachelor’s degree OR 8 or more years of extensive experience in programming, analysis, change management, and release management (or equivalent experience). - Candidates have or meet the requirements to obtain prior to their first day of employment, a valid Level One Arizona fingerprint clearance card that meets DES requirements for a Level One card pursuant to Arizona Revised Statute 41-1969. Requirements - Successfully pass background and reference checks; employment is contingent upon completion of the above-mentioned process and the agency’s ability to reasonably accommodate any restrictions. - All newly hired State employees are subject to and must successfully complete the Electronic Employment Eligibility Verification Program (E-Verify). - If this position requires driving or the use of a vehicle as an essential function of the job to conduct State business, then the following requirements apply: Driver’s License Requirements. Benefits - Career Advancement & Employee Development Opportunities - Flexible schedules to create a work/life balance - Tuition Reimbursement - Participation in the Arizona State Retirement System (ASRS) - 10 paid holidays per year - Stipend Opportunities - Infant at Work Program - RideShare and Public Transit Subsidy - Affordable Medical, Dental, and Vision - Life and Disability Insurance By providing the option of a full-time or part-time remote work schedule, employees enjoy improved work/life balance, report higher job satisfaction, and are more productive. Remote work is a management option and not an employee entitlement or right. An agency may terminate a remote work agreement at its discretion. Contact Us For questions about this career opportunity, please contact us at (602) 542-3809 or email OODHRstaffing@azdes.gov. The State of Arizona is an Equal Opportunity/Reasonable Accommodation Employer. Persons with a disability may request a reasonable accommodation such as a sign language interpreter or an alternative format by contacting Shannon Anguiano by calling or texting at 480-876-0851 or ShannonAnguiano@azdes.gov. Requests should be made as early as possible to allow sufficient time to arrange the accommodation.
Principal Site Reliability Engineer
DeimosThe Cloud Native Developer and Security Operations Company.
• Design and build advanced cloud-native infrastructure • Guide technical discussions with clients and build technical roadmaps • Collaborate with the Engineering Director(s) to (re)design architecture • Assist the Site Reliability Manager with resource planning • Assist engineering managers with building career paths for individuals wishing to be promoted to Principal Engineers • Teach, mentor, grow, and provide advice to other domain experts, individual contributors, and across several teams. • Document processes and monitor performance metrics • Guide conversations to remove blockers and encourage collaboration across teams. • Constantly improve the stability, scalability, security, cost-effectiveness, and operational excellence of our clients' systems. • Continuously discover, evaluate, and implement new technologies to maximize development efficiency and security. • Conduct infrastructure planning, testing, and development • Provide technical leadership on multiple projects.
• Lead the design, implementation, and evolution of cloud infrastructure using Infrastructure as Code. • Drive automation strategies for deployment and infrastructure to ensure high availability, reliability, and efficiency. • Own the architecture and management of AWS networking environments. • Lead the design and operation of containerized platforms and orchestration. • Architect, maintain, and continuously improve CI/CD pipelines at scale. • Define and implement monitoring, observability, and alerting strategies. • Ensure end-to-end security, performance, scalability, and cost optimization of the infrastructure. • Act as a technical partner for engineering and operations teams to enable continuous delivery. • Evaluate, propose, and lead the adoption of new tools, technologies, and best practices. • Provide technical leadership, mentorship, and guidance to engineers.
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description Our mission at HubSpot is to help millions of organizations grow better. On the Marketing Web Team, you’ll play a critical role in ensuring the systems behind HubSpot’s digital experiences are reliable, scalable, and built to grow with our customers. This is a highly visible team responsible for delivering innovative web platforms and applications that power HubSpot’s global brand and are experienced by millions every day. As a Senior DevOps Engineer, you’ll design and operate the infrastructure that supports high-traffic, data-intensive marketing platforms and systems. You’ll partner closely with developers and marketers to improve delivery velocity while upholding strong standards for security, performance, and operational excellence. Together, we’re on a mission to build the best web experience in the world, and the reliability of our platforms is foundational to that goal. What You’ll Do - Design, build, and maintain scalable, secure, and reliable cloud infrastructure to support high-traffic, business-critical digital applications. - Proactively improve system reliability, availability, and performance through automation, observability, and continuous optimization. - Collaborate cross-functionally with Web Development, QA, and Product teams to align infrastructure decisions with business and delivery goals. - Lead incident response, root cause analysis, and post-incident reviews, ensuring learnings are translated into systemic improvements. - Build and maintain CI/CD pipelines that reduce friction and increase deployment confidence. - Establish and evolve infrastructure-as-code standards to ensure consistency, scalability, and long-term maintainability. - Identify opportunities to incorporate AI-assisted tooling into DevOps workflows and partner with leadership to measure its impact on reliability, delivery velocity, and cost efficiency. - Drive security, resiliency, and cost-awareness across environments. - Own medium-to-large infrastructure initiatives from design through production, including documentation and long-term support. - Contribute to sprint planning and backlog refinement within an Agile environment, ensuring work is prioritized and delivered effectively. - Contribute to roadmap planning by identifying infrastructure investments that improve reliability, scalability, and developer velocity. - Evaluate technical trade-offs and clearly communicate risks, impact, and recommendations to stakeholders. - Develop and maintain runbooks, architectural diagrams, and system documentation to support operational excellence. - Mentor engineers by sharing best practices in DevOps and cloud architecture. Qualifications - 5+ years of experience in DevOps, platform engineering, or infrastructure-focused software engineering roles. - Significant experience operating production systems in at least one major public cloud environment (AWS, Azure, or GCP). - Practical experience with infrastructure as code and configuration management (e.g., Terraform, CloudFormation, Ansible, Chef). - Hands-on experience building and maintaining CI/CD pipelines using tools such as Jenkins, GitHub Actions, or similar platforms. - Experience working with CDN, DNS, and edge security platforms (e.g., Cloudflare). - Solid understanding of Linux systems, networking, and distributed systems. - Hands-on experience with containers and orchestration tools (e.g., Docker, Kubernetes). - Ability to troubleshoot complex system issues and drive them to resolution. - Ability to prioritize competing initiatives and manage work with minimal oversight. - Strong written and verbal communication skills, with the ability to explain complex technical concepts clearly. Requirements - Experience supporting large-scale, customer-facing SaaS platforms. - Experience operating highly distributed, scalable systems in production environments. - Exposure to observability tooling such as Prometheus, Grafana, or OpenTelemetry. - Familiarity with web technologies such as TypeScript, JavaScript, and Node.js. - Experience leveraging AI-assisted engineering or operations tools to improve productivity or system reliability. - Experience with security best practices in cloud environments. - Exposure to cost optimization strategies in high-growth systems. - Prior experience mentoring or leading technical initiatives. Benefits - Annual Cash Compensation Range: $108,000 — $162,000 USD. - Base salary, on-target commission for eligible roles, and annual bonus targets under HubSpot’s bonus plan. - Participation in HubSpot’s equity plan to receive restricted stock units (RSUs) for eligible roles. - Flexible work arrangements, including remote options and in-person onboarding. - Support for candidates needing accommodations during the hiring process. Company Description HubSpot (NYSE: HUBS) is an AI-powered customer platform with all the software, integrations, and resources customers need to connect marketing, sales, and service. HubSpot's connected platform enables businesses to grow faster by focusing on what matters most: customers. At HubSpot, bold is our baseline. Our employees around the globe move fast, stay customer-obsessed, and win together. Our culture is grounded in four commitments: Solve for the Customer, Be Bold, Learn Fast, Align, Adapt & Go!, and Deliver with HEART. These commitments shape how we work, lead, and grow. We’re building a company where people can do their best work. We focus on brilliant work, not badge swipes. By combining clarity, ownership, and trust, we create space for big thinking and meaningful progress. And we know that when our employees grow, our customers do too. Recognized globally for our award-winning culture by Comparably, Glassdoor, Fortune, and more, HubSpot is headquartered in Cambridge, MA, with employees and offices around the world.



