Senior Site Reliability Engineer – Traffic Engineering
Location
Poland
Posted
4 days ago
Salary
zł23.3K - zł34K / month
Seniority
Senior
Job Description
Senior Site Reliability Engineer – Traffic Engineering
Nord Security
• Ensure content accessibility across a globally distributed edge infrastructure • Design, operate, and improve load balancing and traffic shaping systems at scale • Troubleshoot connection, latency, and accessibility issues end to end • Own HTTP traffic flow - from client to edge to origin and back • Investigate and resolve website and service reachability problems • Work across protocol layers to diagnose issues at the right depth
Job Requirements
- Load balancing and traffic shaping - architecture, algorithms, health checking, failover
- NGINX and/or HAProxy - configuration, tuning, debugging in high-traffic environments
- Networking fundamentals - DNS, routing, firewall, packet capture and analysis
- OSI model fluency - comfortable working across L3–L7, TCP internals, SSL/TLS handshake and termination
- SaltStack
- Linux administration
- Python programming - automation, tooling, scripting
- Troubleshooting network and application-layer issues in production at scale
- Observability knowledge - metrics, logging, and monitoring for traffic systems
- Docker knowledge - full lifecycle and dependency management with advanced networking
Benefits
- Health insurance
- Physical well-being programs
- Mental & emotional health support
- Company events & team-building
- Workation opportunities
- Flexible work arrangements
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Drive the evolution of the company’s DevSecOps function by defining security initiatives, long-term priorities, and promoting a security-first engineering mindset. • Partner with the DevOps team to improve the security and resilience of Kubernetes (EKS), Istio, Service Mesh, and related cloud-native infrastructure. • Support engineering teams during service transitions by embedding security best practices into operational processes and ensuring effective knowledge sharing. • Develop and maintain policy-as-code frameworks and infrastructure validation mechanisms using technologies such as OPA and other compliance automation tools. • Improve the security of infrastructure provisioning and deployment workflows across Terraform, Ansible, and other Infrastructure-as-Code solutions. • Integrate, maintain, and optimise security testing and scanning capabilities throughout the software development and release lifecycle. • Build automation around security monitoring, compliance checks, and operational controls to increase efficiency and reduce manual effort.
Customer Deployment Engineer
harrison.aiOn a mission to raise the standard of healthcare for millions of patients every day. Through our clinical Al solutions.
• Deploy Harrison.ai solutions, including software installation, configuration, and system testing. • Remotely build, configure, and manage Ubuntu Linux enterprise systems. • Work with our healthcare industry vendor partners to create and maintain integrations for our customer base. • Implement and design integrations with our partners using messaging standards like RESTful APIs and HL7. • Maintain and manage test environments for problem replication. • Provide systems and demos to internal cross-functional teams and for client demonstrations. • Collaborate with customers to validate the proposed solution architecture and ensure alignment with their workflows and objectives. • Maintain close communication with customers regarding ongoing tasks and expectations. • Take ownership of and manage customer cases and customer expectations professionally. • Provide effective and high-quality technical support (Level 2 and 3) to our growing customer base • Troubleshoot and diagnose customer issues using documentation, knowledge base articles, and the wider technical team. • Document troubleshooting and problem resolution steps for future use. • Correctly escalate technical issues to other internal teams, and work with our development group to resolve advanced issues. • Document all work in line with organisational policies and regulatory guidelines in a timely manner.
• Architect, deploy, and operate scalable, secure production environments (AWS preferred). • Lead reliability improvements across multiple engineering streams. • Design and evolve Kubernetes-based infrastructure, including migration and optimisation initiatives. • Build and enforce strong Infrastructure-as-Code standards. • Define and operationalise SLIs, SLOs, and error budgets. • Strengthen observability across applications, infrastructure, data pipelines, and ML systems. • Work closely with product and data teams to integrate model analytics and product telemetry into reliability insights. • Work across and optimise the entire CI/CD pipeline, from build to deploy to rollback. • Improve release safety, deployment frequency, and predictability of SLAs. • Lead incident response for complex cross-system failures and drive postmortems. • Reduce operational toil through automation and platform engineering improvements. • Design processes and tooling to absorb, standardise, and troubleshoot customer environments. • Support and productionise ML workloads (MLOps practices including model deployment, monitoring, retraining workflows). • Ensure infrastructure aligns with enterprise-grade security and regulatory requirements. • Mentor engineers and raise the overall reliability bar across teams.
• Own and lead all SRE-related strategy, standards, and execution. • Embed SRE culture and operational excellence across engineering teams. • Review the current infrastructure and operational model; redesign and rebuild where needed. • Architect, deploy, and maintain scalable, secure production environments. • Define and implement SLIs, SLOs, and uptime targets. • Establish robust monitoring, alerting, and observability practices. • Design and implement incident management, RCA and postmortem processes. • Build and manage sustainable on-call frameworks and escalation models. • Automate the software delivery lifecycle to improve release predictability and safety. • Create reproducible environments and IaaC provisioning templates. • Improve system performance, availability, and reliability. • Support and productionise data platforms and ML workloads. • Partner closely with QA and Engineering leadership to improve release quality and stability. • Ensure infrastructure meets enterprise-grade security and regulatory requirements. • Hire, manage, and mentor a team of SRE engineers.



