Job Closed
This listing is no longer active.
Here you can create the extraordinary. Join us.
Site Reliability Engineer
Location
New York
Posted
96 days ago
Salary
$96K - $137K / year
Seniority
Lead
Job Description
Site Reliability Engineer
NBCUniversal
• Design, implement, and maintain Nexthink dashboards, investigations, campaigns, and services focused on device performance and user experience • Leverage Nexthink telemetry and analytics to develop a device health scoring framework and identify candidates for PC refresh based on performance, lifecycle stage, failure trends, and user sentiment • Develop proactive maintenance initiatives (e.g., patching, storage optimization, software updates) based on predictive insights and device analytics • Lead the configuration, administration, and continuous optimization of the Nexthink platform in alignment with NBCU’s endpoint engineering roadmap • Develop and implement automation strategies, including Nexthink remote actions and scripts, to proactively remediate endpoint issues and streamline operational workflows • Partner with cross-functional teams including Local IT/Operations, Cyber Security, and Software Governance to improve endpoint visibility and reliability • Provide technical guidance and best practices for monitoring and improving enterprise-wide digital employee experience • Translate platform insights and analytics into actionable recommendations that drive operational improvements and innovation • Identify new opportunities to leverage Nexthink data for proactive engineering solutions and improved decision-making • Provide training and guidance to operations and engineering teams on Nexthink capabilities and best practices • Stay current on emerging technologies, digital experience monitoring tools, and IT service management best practices to continuously improve endpoint engineering processes
Job Requirements
- Bachelor’s degree in Computer Science, Information Technology, or equivalent experience
- 8+ years of experience in desktop or end-user systems engineering, preferably in a large enterprise environment (30,000+ endpoints)
- Experience working with enterprise security and compliance standards
- Nexthink Platform Experience – Strong experience configuring, monitoring, and automating Nexthink to improve Digital Employee Experience (DEX)
- Experience with similar platforms (Microsoft Endpoint Analytics, Riverbed Aternity, ControlUp) is also valued
- Scripting & Automation – Experience with automation and scripting tools such as PowerShell, Python, Bash, NQL, or Ansible to improve operational efficiency
- Endpoint & Infrastructure Knowledge – Deep understanding of endpoint performance monitoring, endpoint lifecycle management, and cloud platforms (Azure, AWS)
- Incident Prevention & Troubleshooting – Ability to analyze large data sets, identify trends, and implement preventative solutions to reduce incidents
- Security & Network Awareness – Familiarity with enterprise security policies, endpoint protection tools, and IT compliance standards
- Strong experience with endpoint management platforms including Intune, Autopilot, SCCM, Jamf, Active Directory, and Microsoft 365
- Excellent communication skills and ability to collaborate effectively with operations, infrastructure, cybersecurity, application teams, and leadership
- Ability to work both independently and collaboratively in a fast-paced enterprise environment.
Benefits
- medical, dental and vision insurance
- 401(k)
- paid leave
- tuition reimbursement
- a variety of other discounts and perks
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Design, implement, and maintain CI/CD pipelines to enable reliable and automated build, test, and deployment processes • Automate infrastructure provisioning and environment configuration using Infrastructure as Code (IaC) principles to ensure consistency, scalability, and maintainability • Manage and optimise containerised environments • Support and manage cloud infrastructure, ensuring scalability, reliability, and cost efficiency • Monitor application performance and infrastructure health, proactively identifying and resolving performance, availability, and security issues • Collaborate closely with development teams to improve deployment frequency, reduce lead time, and enhance system stability • Integrate security best practices into development and deployment processes (DevSecOps approach) • Contribute to defining standards, documentation, and best practices across development and operations • Perform troubleshooting and root cause analysis across applications, infrastructure, and environments
DevOps Engineer Internship
RAINNNation's largest anti-sexual violence organization. Created & operates the National Sexual Assault Hotline: 800.656.4673
• Learn to assist with maintaining and supporting cloud-based infrastructure (primarily AWS) within approved environments. • Learn to support CI/CD pipeline processes, including building, testing, and deployment workflows. • Learn to assist with system monitoring, logging, and alerting to ensure performance and reliability. • Learn how to help automate routine operational tasks using scripts and infrastructure-as-code tools (as appropriate). • Participate in troubleshooting system issues and documenting resolutions. • Collaborate with team members on system improvements, performance optimization, and reliability efforts. • Learn and apply best practices related to security, scalability, and system maintenance. • Learn to contribute to technical documentation, including system configurations and standard operating procedures. • Participate in team meetings, stand-ups, and knowledge-sharing sessions.
• Rightsize workloads for efficient resource utilization in Kubernetes and cloud service • Contribute to the development, deployment and operations for new microservices within the platform • Review technical specifications to provide guidance and help development teams drive operational excellence • Work hand-in-hand with software developers to facilitate the development and adoption of "Paved Road" solutions and DevSecOps processes • Support large-scale services across multiple environments • Assist in resolution efforts for problems ranging from infrastructure network layers to application scaling • This role includes participation in an on-call rotation - we believe in shared ownership of our platform and aim to build systems that are resilient, observable, and require minimal intervention.
• Develop and architect complex integrations using MuleSoft • Implement and maintain automated tests (MUnit), ensuring quality and preventing failures • Create and maintain CI/CD pipelines to automate deployments • Design resilient messaging-based solutions (Anypoint MQ) • Configure security policies and integrate with Azure AD • Administer APIs and environments on the Anypoint Platform (API Manager, Runtime Manager, Exchange) • Ensure governance, security, and high availability standards across integrations




