Five9 logo
Five9

Helping Companies Bring Joy to CX.

Senior Network DevOps Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteSeniorTeam 1,001-5,000Since 2001H1B SponsorCompany SiteLinkedIn

Location

India

Posted

7 days ago

Salary

0

Seniority

Senior

Job Description

Senior Network DevOps Engineer

Five9

• Design, engineer, and oversee highly scalable CI/CD pipelines for automated testing, digital-twin dry-running, and zero-downtime deployment of configurations across massive physical and virtual network fabrics. • Champion the "Network-as-Code" paradigm, establishing enterprise-wide GitOps workflows and enforcing source control (Git) as the absolute single source of truth for all network state, policy, and infrastructure design. • Drive the strategic vision for network automation, designing complex, reusable code frameworks and Playbooks to permanently eliminate manual configuration and operational "toil" at a global scale. • Design and optimize massive-scale software-defined networks (SDN) and container networking interfaces (CNIs) supporting ultra-high-throughput, multi-region cloud and bare-metal environments. • Serve as the ultimate technical authority (SME / Tier 4) for complex network, routing, and overlay issues, rapidly deciphering and resolving the most difficult challenges across physical underlays and software-defined layers. • Design and implement next-generation telemetry, monitoring, and alerting-as-code platforms, leveraging advanced heuristics for proactive auto-remediation of performance, security, and capacity issues. • Partner at a strategic level with Enterprise Designers, SRE leaders, and Platform teams to build and implement highly resilient, secure, and multi-tenant-isolated network infrastructure. • Establish coding standards and write/review sophisticated Infrastructure-as-Code (IaC) using Ansible, Terraform, Helm, Python, and Bash. • Lead the engineering, deployment, and optimization of modern overlay and container networking solutions, particularly VXLAN, OVN (Open Virtual Network), and Project Calico. • Design the automation of complex security and traffic management policies across enterprise-grade firewalls (Palo Alto/Fortinet) and load balancers (F5 BIG-IP/Cloudflare/Octavia). • Embed "Security-as-Code" directly into CI/CD pipelines, engineering automated validation gates for security groups, ACLs, and firewall rules prior to production pushes. • Engineer robust, automated network validation suites to continuously verify network health, state, and complex routing tables (BGP, OSPF) pre- and post-deployment. • Drive the adoption of code-based infrastructure definitions and dynamic, living technical documentation (e.g., NetBox, Git Markdown), setting the standard for documentation across the engineering organization.

Job Requirements

  • 8+ years of progressive experience in Network Engineering, DevOps, or SRE roles, with a deeply established focus on large-scale network design, automation, and Infrastructure-as-Code (IaC).
  • Expert-level proficiency in engineering complex network automation tooling utilizing Ansible, Terraform, and Python (including deep experience with network-specific libraries like Netmiko, Nornir, or NAPALM).
  • Deep design and operational expertise with massive-scale software-defined networks (SDN) and overlays, specifically VXLAN, OVN, and Project Calico.
  • Advanced mastery of Kubernetes networking designs (CNIs, Ingress controllers, Service Mesh, Kube-Proxy behaviors) at an enterprise scale.
  • Extensive, expert-level experience engineering and automating enterprise edge, routing, and security fabrics, including Cisco Nexus (9k/7k/5k/2k), Arista, Versa SD-WAN, and firewalls (Palo Alto and Fortinet).
  • Authoritative knowledge of API-driven application delivery and load balancing across monolithic and cloud-native stacks (e.g., F5 BIG-IP AS3, Cloudflare APIs, Octavia LBaaS).
  • Deep Linux internals and networking expertise, including advanced knowledge of Linux network namespaces, virtual interfaces, eBPF/iptables, and kernel-level routing.
  • CCIE-level (or equivalent) mastery of core routing protocols (BGP, OSPF, MPLS, EVPN) and a comprehensive understanding of how virtual/container overlay networks map to physical underlays.
  • Extensive experience designing and scaling CI/CD platforms (e.g., GitLab CI, Jenkins, GitHub Actions) and implementing test-driven development (TDD) for infrastructure changes.
  • Unparalleled troubleshooting capabilities utilizing advanced packet analysis (tcpdump/Wireshark), NetFlow, and log aggregation platforms to trace asymmetric routing and microburst traffic across hybrid fabrics.
  • Proven track record of technical leadership, including mentoring mid-level engineers, leading design review boards, and driving cross-functional engineering initiatives.
  • Deep telecommunications background—specifically VoIP, RTP, SIP peering/trunking, and Session Border Controllers (SBC)—is highly advantageous.
  • Exceptional verbal and written communication skills, with the ability to translate complex technical concepts to executive stakeholders while fostering a collaborative, DevOps-oriented engineering culture.

Benefits

  • Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills
  • The more inclusive we are, the better we are. Five9 is an equal opportunity employer.

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Weekday (YC W21) logo

DevOps Engineer

Weekday (YC W21)

We are a Y-Combinator-backed startup building your AI-powered Recruiter Agent

DevOps Engineer7 days ago
Full TimeRemoteTeam 11-50Since 2021H1B No Sponsor

Role Description This is a full-time Remote role for an experienced Senior Platform Architect / DevOps Engineer — Microsoft Fabric who enjoys solving complex data problems and takes ownership from design through delivery. You can work from home (anywhere in INDIA, job location is in India). This role is ideal for someone who can build scalable data pipelines on Microsoft Azure Fabric, leverage AI to improve engineering productivity, and work closely with business and technical stakeholders. We're looking for a Senior Platform Architect with strong DevOps expertise to design, implement, and operationalize our Microsoft Fabric data platform. You'll own the architecture of our unified data estate — spanning OneLake, Data Factory pipelines, Synapse Data Engineering, Power BI, and Real-Time Analytics — while building the CI/CD, governance, and automation backbone that lets our data teams ship reliably and securely. This is a hands-on architecture role: you'll set technical direction and also build it yourself. What You'll Do - Design and own the end-to-end architecture for Microsoft Fabric, including OneLake data lake structure, workspace and capacity strategy, domain/workspace segregation, and medallion (bronze/silver/gold) layering. - Define and implement DevOps practices for Fabric: Git integration, deployment pipelines, environment promotion (dev/test/prod), and infrastructure-as-code for Fabric items (lakehouses, warehouses, notebooks, pipelines, semantic models). - Build and maintain CI/CD pipelines (Azure DevOps or GitHub Actions) for Fabric artifacts, including automated testing, validation, and rollback strategies. - Architect and enforce governance, security, and access controls across Fabric — workspace roles, OneLake data access policies, sensitivity labels, row-level security, and integration with Microsoft Purview. - Design capacity planning and cost optimization strategies for Fabric capacity units (CUs), monitoring consumption and tuning workloads to control spend. - Partner with data engineers, analysts, and platform teams to standardize development patterns, reusable templates, and best practices across Fabric workloads (Data Factory, Data Engineering/Spark, Data Warehouse, Real-Time Intelligence, Power BI). - Own monitoring, alerting, and observability for the platform using Fabric monitoring hub, Azure Monitor, and Log Analytics. - Lead migration efforts from legacy platforms (Synapse Analytics, Azure Data Factory, on-prem SSIS/SQL Server, or other cloud data warehouses) onto Microsoft Fabric. - Mentor engineering teams on Fabric architecture patterns, DevOps practices, and operational excellence. - Act as the primary technical escalation point for platform reliability, performance, and incident response. Qualifications - 8+ years of experience in data platform architecture, cloud infrastructure, or DevOps engineering roles. - 2+ years of hands-on experience with Microsoft Fabric, or deep equivalent experience with Azure Synapse Analytics, Power BI, and Azure Data Factory with a clear path to Fabric. - Strong command of OneLake, Lakehouse and Warehouse architecture, Delta Lake/Parquet formats, and medallion architecture design. - Proven DevOps expertise: Git-based version control, CI/CD pipeline design (Azure DevOps or GitHub Actions), and infrastructure-as-code (Bicep, Terraform, or ARM templates) for Azure/Fabric resources. - Solid understanding of Fabric security model — workspace roles, item-level permissions, OneLake data access roles, sensitivity labels, and integration with Microsoft Entra ID (Azure AD) and Purview. - Experience with Fabric capacity management, performance tuning, and cost governance (CU monitoring, throttling, smoothing). - Hands-on skills in at least one of: PySpark/Spark SQL, T-SQL, Python, or KQL for pipeline and query development. - Experience integrating Fabric with external tools — APIs, Power Automate, Azure Functions, or third-party data sources. - Strong communication skills with the ability to translate architecture decisions for both technical teams and business stakeholders. - Prior experience leading platform migrations to modern data platforms is a strong plus. Requirements - Must have at least 5+ years of experience. Nice to Have - Microsoft certifications: DP-600 (Fabric Analytics Engineer), DP-203 (Azure Data Engineer), or Azure Solutions Architect Expert. - Experience with Power BI semantic model optimization and DirectLake mode. - Familiarity with data mesh or domain-driven data architecture principles. - Background in regulated industries requiring strict data governance (finance, healthcare, etc.). Must-have skills - Microsoft Fabric - Azure DevOps - Data Platform Good-to-have skills - Infrastructure as code - Spark SQL

India
₹1,100K - ₹5,000K / year
Social Discovery Group logo

DevOps Engineer

Social Discovery Group

Top world’s largest social discovery company uniting 70+ brands with 500M+ users

DevOps Engineer7 days ago
Full TimeRemoteTeam 1,001-5,000Since 20 yearsH1B No Sponsor

• Design and develop internal services and tools in Go • Design, build, and maintain scalable CI/CD pipelines using GitLab CI/CD • Improve and support Kubernetes-based environments • Contribute to the evolution of the internal deployment platform (Packman) • Optimize stage and ephemeral environments for performance, reliability, and cost-efficiency • Implement and maintain infrastructure as code using Terraform and Ansible • Improve observability through monitoring, logging, and alerting • Collaborate with engineering teams to enhance developer experience and delivery speed

Worldwide
Xsolla logo

Site Reliability Engineer - Monetization

Xsolla

Xsolla's video game business engine helps game developers and publishers operate more efficiently and sell more games.

DevOps Engineer7 days ago
Full TimeRemoteTeam 201-500Since 2005H1B Sponsor

Title: Site Reliability Engineer (Monetization) Location: Montreal Type: Full time Workplace: remote Category: Infrastructure Team Job Description: ABOUT YOU We are looking for a Site Reliability Engineer (Monetization) who is pragmatic, product-minded, and equally comfortable writing code and running production systems to join our Infrastructure department's SRE team. The best candidate will be someone who thrives in a fast-paced, highly collaborative, and exceptionally dynamic setting and is excited to own the application-level infrastructure and reliability of a high-traffic commerce domain end to end - from deploy pipelines and Kubernetes manifests to SLOs, capacity planning, and production readiness. Strong Kubernetes, observability, and software engineering skills are essential, along with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product development teams. The ability to hold a dual perspective - understanding both how developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions early will be key to your success in this role. This is a hybrid embedded role: you remain part of the SRE organization (practices, standards, duty rotation) while being functionally embedded into the Monetization product domain. You'll build long-term working relationships with the domain's engineering teams, own a meaningful share of their application infrastructure execution, and co-author the reliability practices used company-wide. If you're passionate about making complex distributed systems boringly reliable and love building the commerce and monetization backbone that lets game developers around the world get paid, we would love to hear from you! ABOUT US Xsolla is a global commerce company with robust tools and services to help developers solve the inherent challenges of the video game industry. From indie to AAA, companies partner with Xsolla to help them fund, distribute, market, and monetize their games. Grounded in the belief in the future of video games, Xsolla is resolute in the mission to bring opportunities together, and continually make new resources available to creators. Headquartered and incorporated in Los Angeles, California, Xsolla operates as the merchant of record and has helped over 1,500+ game developers to reach more players and grow their businesses around the world. With more paths to profits and ways to win, developers have all the things needed to enjoy the game. Responsibilities - Own the application-level infrastructure of the Monetization domain: Helm charts, Terraform configurations, Kubernetes deployments, runtime configuration, and service-level networking and integrations - Own the domain's observability: design and implement SLOs/SLIs, monitors, alerts, and dashboards for critical services on Datadog and OpenTelemetry-based tooling - Help to set up and evolve CI/CD pipelines for domain services (GitLab CI, GitHub Actions), including deploy and rollback automation - Perform capacity planning and performance tuning ahead of expected load - product launches, sales events, and regional rollouts - including load testing and performance regression investigation - Run Production Readiness Reviews for new services and major changes; define and enforce what "production-ready" means for the domain - Support domain incident response: assist with deep investigation of complex incidents, contribute to post-mortems, drive follow-up reliability improvements, and maintain runbooks - Build domain-specific automation that reduces operational toil: runbook automation, deploy helpers, recurring operational scripts - Maintain and drive a forward-looking reliability roadmap for the domain together with product engineering leads - Participate in product team planning, refinements, and architecture reviews, bringing the reliability perspective before design decisions become expensive to change - Co-author company-wide SLO/SLI, capacity, and operational standards together with the broader SRE team; contribute improvements directly to shared SRE-operated subsystems - Participate in the SRE duty rotation, supporting developers across the company Qualifications & Skills - 3+ years of proven SRE, DevOps, or platform engineering experience: on-call or incident response duty, SLO/monitoring ownership, deploy pipeline and infrastructure work for production services - Software development background: you have built and shipped backend services, not only operated them - comfortable reading application code during an investigation and writing production-quality automation in at least one language (e.g., Go, PHP) - Hands-on Kubernetes experience: Helm, manifests, deploy strategies, debugging application-level performance and networking issues (GKE or another managed Kubernetes) - Solid observability practice: building monitors, dashboards, and SLOs/SLIs on a modern platform (Datadog preferred; Prometheus/Grafana experience also relevant), familiarity with OpenTelemetry - Infrastructure as Code exposure (Terraform/Terragrunt) for collaboration with platform teams - GCP experience (IAM, networking, managed services) - Experience building and maintaining CI/CD pipelines (GitLab CI and/or GitHub Actions) - Programming/scripting proficiency sufficient to build automation and tooling (e.g., Python, Go, or Bash) - Practical experience with incident response, post-mortems, and driving reliability improvements from incidents - Strong collaboration and communication skills — this role works embedded with product development teams daily - Experience in payments, fintech, e-commerce, or gaming — high-traffic transactional systems Nice to Have: - Kubernetes certifications - Google Cloud Platform certifications - HashiCorp certifications Salary varies depending on experience level and location. Benefits We are passionate about fostering a supportive environment for our team, so we prioritize the physical, mental, and emotional well-being of our employees and their families through a comprehensive Benefits Program. This includes medical, dental, and vision, PTO, and a personalized career roadmap for each employee. By investing in professional development through training and educational opportunities, we ensure that our team thrives both personally and professionally. Together, we're not just building a business; we're cultivating a community that values creativity, collaboration, and the transformative power of play. Equal Employment Opportunity Statement Xsolla is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate based on race, color, religion, sex, national origin, age, disability, sexual orientation, gender identity, or any other characteristic protected by law. We consider qualified applicants with criminal histories in accordance with the Fair Chance Act. Criminal History Consideration For the Site Reliability Engineer (Monetization) position, we will conduct a background check that may include the following: - Criminal history check - Employment verification - Education verification Relevance to Job Responsibilities The background check is relevant to this position because of the following role responsibilities: - Accessing confidential company data - Handling infrastructure that processes sensitive financial transactions - Ensuring compliance with regulatory requirements Rights Under the Fair Chance Act Applicants are encouraged to inquire about their rights under the Fair Chance Act. If you have questions regarding our hiring practices. By submitting the following job application form, you consent to Xsolla processing your data for career-related inquiries and potential employment opportunities. We process your data in accordance with this Xsolla Privacy Notice for Job Applicants.

Canada
Altimea logo

Ingeniero DevOps

Altimea

Impulsando cambios

DevOps Engineer7 days ago
Full TimeRemoteTeam 11-50Since 2005H1B No Sponsor

• Diseño, implementación y mantenimiento de pipelines de CI/CD para automatizar los despliegues de aplicaciones. • Aprovisionamiento y administración de la infraestructura cloud principal (Azure) mediante Terraform. • Soporte, optimización y resolución de incidencias en entornos de Google Cloud Platform (GCP) cuando el proyecto lo requiera. • Monitoreo, escalabilidad y aseguramiento de la alta disponibilidad de clústeres de Kubernetes (AKS). • Coordinación directa con los equipos de desarrollo y arquitectura sobre aspectos técnicos y de despliegue. • Investigación y propuesta de mejoras en seguridad, rendimiento y optimización de costos en la nube.

Peru