Deutsche Telekom IT Solutions logo
Deutsche Telekom IT Solutions

As Hungary’s most attractive employer in 2025 (according to Randstad’s representative survey), Deutsche Telekom IT Solutions is a subsidiary of the Deutsche Telekom Group. The company provides a wide portfolio of IT and telecommunications services with more than 5300 employees. We have hundreds of large customers, corporations in Germany and in other European countries. DT-ITS received the Best in Educational Cooperation award from HIPA in 2019, acknowledged as the Most Ethical Multinational Company in 2019. The company continuously develops its four sites in Budapest, Debrecen, Pécs and Szeged and is looking for skilled IT professionals to join its team.

DevOps Engineer GroudOS (REF5181Z)

DevOps EngineerDevOps EngineerFull TimeRemoteMid LevelTeam 5,001-10,000

Location

Hungary

Posted

111 days ago

Salary

0

Seniority

Mid Level

No structured requirement data.

Job Description

DevOps Engineer GroudOS (REF5181Z)

Deutsche Telekom IT Solutions

Company Description As Hungary’s most attractive employer in 2025 (according to Randstad’s representative survey), Deutsche Telekom IT Solutions is a subsidiary of the Deutsche Telekom Group. The company provides a wide portfolio of IT and telecommunications services with more than 5300 employees. We have hundreds of large customers, corporations in Germany and in other European countries. DT-ITS recieved the Best in Educational Cooperation award from HIPA in 2019, acknowledged as the the Most Ethical Multinational Company in 2019. The company continuously develops its four sites in Budapest, Debrecen, Pécs and Szeged and is looking for skilled IT professionals to join its team. Job Description Are you an expert in deploying, observing, and maintaining distributed fleets of devices? Do you build infrastructure that scales effortlessly and recovers automatically from mass reconnections? Join our team to oversee the operational backbone of our edge-to-cloud ecosystem. If you love automating complex deployments and diving deep into observability metrics, you are the right fit for us! Project Description: Our project, GroundOS, is not just another screen manager. It is a next-generation Universal Display System (UDS) built to power the future of global mobility. We are building an "Operating System for Reality" that orchestrates massive, data-driven signage networks across critical infrastructure, from major international airports to sprawling public transport systems. GroundOS moves beyond static displays; it uses a state-of-the-art digital twin to process and react to real-time operational data. To guarantee continuous operation, the platform features a resilient, offline-first edge architecture that ensures screens keep running smoothly even if the network fails. Join us to blend high-performance Rust edge computing with modern TypeScript cloud services and help us set a new global standard for how hundreds of millions of passengers experience their journey. Tasks - Manage the deployment, observability, and lifecycle of thousands of remote mini-PCs alongside Cloud components. - Execute Over-The-Air (OTA) updates reliably across a massive edge fleet. - Configure and manage NATS JetStream, including Leaf Nodes for edge-cloud bridging, stream retention, and cluster HA. - Setup and maintain tracing and metrics using OpenTelemetry to monitor cross-system health. - Architect resilient systems capable of withstanding mass fleet reconnection events (thundering herd) without performance loss. - Manage secrets, certificates, and secure mTLS communication between edge devices and the central control plane. - Lead incident management and root-cause analysis for fleet-wide issues. - Design scalable operations workflows to keep maintenance effort constant as the fleet grows. Qualifications - Extensive experience with infrastructure automation and remote fleet management. - High proficiency in containerization (Docker), specifically optimized for edge devices (multi-arch builds, ARM/x64). - Deep operational knowledge of NATS JetStream or similar high-throughput event brokers. - Strong background in observability, tracing, and metric collection. - Solid understanding of Zero-Trust security architectures and certificate management. - Ability to remain calm and analytical during high-pressure incident response situations. - Expert knowledge of agile development - Solid knowledge of Scrum - Experience working in agile projects and teams - Excellent English skills, both written and spoken (B2–C1) - Excellent technical and analytical skills, as well as problem-solving abilities - Ability to handle stressful situations and work independently Advantages: - Experience with Google Clouds GKE for the central cloud control plane. - Prior experience with specific edge orchestration tools Additional Information * Please be informed that our remote working possibility is only available within Hungary due to European taxation regulation. - Company: Deutsche Telekom TSI Hungary Kft.

Related Categories

Related Job Pages

More DevOps Engineer Jobs

QuantumLoopAI logo

Azure DevOps / Infrastructure Engineer - India Remote

QuantumLoopAI

QuantumLoopAI is one of the UK's fastest-growing healthtech companies. Our AI-powered platform, EMMA, is used by hundreds of organisations across the UK to manage communications, automate workflows and deliver better outcomes through intelligent automation. Backed by leading investors and scaling rapidly ahead of our Series A, we are building the infrastructure that will define the next generation of AI-driven service delivery.

DevOps Engineer111 days ago

ABOUT QUANTUMLOOPAI QuantumLoopAI is one of the UK's fastest-growing healthtech companies. Our AI-powered platform, EMMA, is used by hundreds of organisations across the UK to manage communications, automate workflows and deliver better outcomes through intelligent automation. Backed by leading investors and scaling rapidly ahead of our Series A, we are building the infrastructure that will define the next generation of AI-driven service delivery. ABOUT THE ROLE We are looking for a senior Azure DevOps / Infrastructure Engineer to take ownership of our Azure cloud infrastructure. This is a hands-on role focused on building, securing and optimising the platform that underpins everything we do — from access control and monitoring to scalability, cost governance and deployment automation. You will establish best practices, mentor the engineering team on infrastructure and DevOps topics, and ensure our cloud environment is robust, observable and cost-efficient as we scale. This role has an immediate focus on infrastructure setup, monitoring and optimisation, with longer-term ownership of our cloud strategy as the platform grows. WHAT YOU WILL DO RBAC and Access Management - Design and implement robust Role-Based Access Control (RBAC) policies across our Azure environment. - Set up and manage users, groups and service principals to ensure secure, least-privilege access to all Azure resources. Monitoring and Observability - Build dashboards to monitor performance, utilisation and health of Azure resources, providing actionable insights to the engineering team. - Implement centralised logging and telemetry solutions using Azure Monitor, Application Insights and related tooling. - Set up proactive alerting to detect and resolve potential issues before they impact users. Scalability and Performance - Architect and implement scalable infrastructure to handle increasing workloads as the platform grows. - Optimise infrastructure and application performance for maximum reliability and responsiveness. Cost Governance - Analyse resource utilisation, identify waste and recommend cost-saving measures. - Ensure cost-efficient deployment, scaling and lifecycle management of Azure resources. DevOps and Automation - Own and improve CI/CD pipelines, containerisation (Docker, Kubernetes) and deployment automation. - Establish infrastructure-as-code practices using Terraform, Bicep or equivalent tooling. - Enforce security best practices across all environments. Technical Leadership - Establish and promote best practices for Azure infrastructure and DevOps processes across the engineering team. - Provide guidance and mentorship to team members on cloud architecture, security and operational excellence.

India

Role Description Cloud Engineering/Platform Engineering. Cloud/DevOps Engineer responsible for designing and deploying secure Azure infrastructure using Terraform and Bicep. This project is on modernizing CIAM infrastructure in Azure by implementing secure, scalable, and compliant cloud-native solutions. The initiative includes provisioning and deploying infrastructure using Terraform and Bicep. The objective is to enhance encryption, strengthen security controls, and ensure high availability and compliance with enterprise security standards. Qualifications - 13+ Years of experience - Must be a hands-on Microsoft Azure System Engineer - Must have Azure App Service - Must have Azure native deployment - Must have Landing Zones (Very important) - Must have Azure Key Vault (Premium) - Must have Terraform Landing Zones - Must have Bicep - Must have worked on Azure DevOps (ADO) - Must have Azure infra using IaC - Must have Implementing CI/CD pipelines in Azure DevOps for infrastructure deployments - Must have Azure Private Endpoints - Must have Azure Networking Git - Must have Service Principals - Must have RBAC - Experience with set/troubleshoot security mechanisms like mTLS - Must have AppInsights experience - Experience with certificate management, manual and implementing automation - Experience with PowerShell - Familiarity with IAM and user authentication (such as OAuth and SAML) is preferred - Experience with ServiceNow preferred - Strong troubleshooting skills required - Proactive ownership of the system engineer space is necessary, including identifying and prioritizing work as a recommended roadmap - Managing role-based access control (RBAC) and service principal configurations - Collaborating with architects and security teams to ensure compliance with encryption and networking standards - Troubleshooting pipeline failures and network access issues (e.g., firewall and private endpoint configurations) - Supporting sandbox, dev, and environment-specific infrastructure deployments - Willing to work on On Call rotation Requirements - Must be willing to use your own laptop - Must be willing to work in PST Hours - 3 levels of interviews - 1st interview - Introduction on Azure skills - 2nd interview - 1 hour coding interview - 3rd interview with manager Company Description

United States
$60 - $70 / hour
Job Closed
Full TimeRemoteTeam 201-500H1B No Sponsor

• Contribute to our global Technology Operations organization as a DevOps Engineer. • Provide agile, secure, and automated delivery pipelines and platform resilience for cloud-native workloads running on AWS and Azure. • Lead end-to-end CI/CD infrastructure design, container orchestration, observability integration and system reliability in development, staging, and production environments. • Collaborate with Site Reliability, Application, Network and Security Engineering teams to align infrastructure delivery with evolving business and technical requirements. • Design, maintain, and document CI/CD infrastructure, supporting canary deployments, blue/green strategies, and infrastructure-as-code best practices. • Automate and manage Kubernetes (AKS) workloads using Helm, Kustomize, and GitOps patterns (ArgoCD).

Malaysia
Job Closed
Natuvion logo

Senior Site Reliability Engineer, SRE

Natuvion

Natuvion supports its customers in moving business-critical data and processes from one technology platform to another.

DevOps Engineer111 days ago
Full TimeRemoteTeam 201-500H1B Sponsor

• Operating and optimizing Kubernetes clusters (EKS) and AWS infrastructure • Debugging complex issues (performance, scheduling, OOM, CrashLoopBackOff) • Deploying and operating self-hosted services (e.g., Istio, OpenSearch, RabbitMQ) • Implementing GitOps (Argo CD / Flux) and observability (logging, metrics, tracing) • Defining SLIs/SLOs and alerting strategies • Developing backup and disaster recovery concepts (including RTO/RPO) • Analyzing and improving system architectures (scalability, security, single points of failure)

Germany
Job Closed