Tech Lead DevOps & Infrastructure
Location
Germany
Posted
4 days ago
Salary
0
Seniority
Lead
No structured requirement data.
Job Description
Tech Lead DevOps & Infrastructure
anwalt.de services AG
Role Description We're looking for an experienced Tech Lead DevOps & Infrastructure with solid infrastructure and cloud experience, who can break down complex business and technical topics and communicate them clearly — fluent in English, German a strong plus. You'll lead a team of 3–4 colleagues and be responsible for: - Analyzing and troubleshooting business-critical issues (DDoS attacks, performance degradation, delivery chain failures, partner outages) - Designing, operating, and refactoring cloud-based environments - particularly AWS infrastructure - Building and operating monitoring and observability stacks at production level - Making strategic decisions on infrastructure architecture, automation, and DevOps concepts Qualifications - Solid experience analyzing and troubleshooting business-critical issues - Traffic and CDN management (Cloudflare, AWS CloudFront) - Extensive practical experience with Infrastructure-as-Code (CloudFormation, Terraform) - Deep knowledge of Linux-based server operations (Debian, Amazon Linux, CentOS) - Strong knowledge of Identity, Access and Permissions Management (IAM) - Solid experience designing, operating and refactoring cloud-based environments (AWS-focused) - Very good understanding of container principles (building, architectures, optimization, registry management) - Solid understanding of cloud-based networking, including cross-account and VPC integration - Experience designing and managing monitoring and observability stacks (push and pull models) - Good grasp of automation, scheduling and CI/CD concepts Requirements - Practical experience with AWS Elastic Container Service (ECS) - Database administration experience (PostgreSQL, MySQL) - Experience working in regulated environments (e.g., ISO 27001) - Good understanding of AWS Organizations - Experience designing, testing and managing disaster recovery solutions (backup strategies, standby systems, playbooks) - Experience with Active Directory - Working exposure to programming languages (PHP, Python, Go, JavaScript) - Understanding of managed Kubernetes deployments Benefits - 100% home office and/or an attractive office in Nuremberg or Hanover - High level of personal responsibility and further development - Successful legal tech company with the leading lawyer platform in Germany - Family-friendly and agile team spirit - Monthly home office subsidy - 30 vacation days and 24.12. and 31.12. off - Company and team events - Fitness/health: Subsidy for USC, FitX or EGYM Wellpass - Subsidy for lunch - Mobility: subsidy for „Deutschlandticket“ (70%)
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Own and improve the reliability, availability and performance of critical production services across the Runware platform • Define and evolve our reliability practices, including SLIs, SLOs, alerting, observability and production-readiness standards • Investigate complex production issues across distributed systems, APIs, networking, queues, databases and GPU-backed workloads, participating in our engineering on-call rotation • Lead and contribute to incident reviews and RCAs, turning recurring failure modes into lasting engineering improvements • Reduce operational toil through automation, automated remediation and improvements to deployment safety, recovery and system resilience • Work closely with Engineering and DevOps teams on capacity planning, performance, scaling and architectural improvements as the platform grows
Senior Engineer, SRE
ZencoderZencoder is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.
• We’re looking for an Engineer to help build and operate the infrastructure behind Zencoder’s AI-powered products. • You’ll work across our production platform, improving its reliability, security, scalability and cost efficiency. • This includes our Kubernetes foundations, cloud infrastructure, networking, data systems and the internal tooling that enables engineers to deploy and operate services confidently. • You’ll write code, automate infrastructure, investigate production issues and design systems that reduce operational complexity.
• Design, engineer, and oversee highly scalable CI/CD pipelines for automated testing, digital-twin dry-running, and zero-downtime deployment of configurations across massive physical and virtual network fabrics. • Champion the "Network-as-Code" paradigm, establishing enterprise-wide GitOps workflows and enforcing source control (Git) as the absolute single source of truth for all network state, policy, and infrastructure design. • Drive the strategic vision for network automation, designing complex, reusable code frameworks and Playbooks to permanently eliminate manual configuration and operational "toil" at a global scale. • Design and optimize massive-scale software-defined networks (SDN) and container networking interfaces (CNIs) supporting ultra-high-throughput, multi-region cloud and bare-metal environments. • Serve as the ultimate technical authority (SME / Tier 4) for complex network, routing, and overlay issues, rapidly deciphering and resolving the most difficult challenges across physical underlays and software-defined layers. • Design and implement next-generation telemetry, monitoring, and alerting-as-code platforms, leveraging advanced heuristics for proactive auto-remediation of performance, security, and capacity issues. • Partner at a strategic level with Enterprise Designers, SRE leaders, and Platform teams to build and implement highly resilient, secure, and multi-tenant-isolated network infrastructure. • Establish coding standards and write/review sophisticated Infrastructure-as-Code (IaC) using Ansible, Terraform, Helm, Python, and Bash. • Lead the engineering, deployment, and optimization of modern overlay and container networking solutions, particularly VXLAN, OVN (Open Virtual Network), and Project Calico. • Design the automation of complex security and traffic management policies across enterprise-grade firewalls (Palo Alto/Fortinet) and load balancers (F5 BIG-IP/Cloudflare/Octavia). • Embed "Security-as-Code" directly into CI/CD pipelines, engineering automated validation gates for security groups, ACLs, and firewall rules prior to production pushes. • Engineer robust, automated network validation suites to continuously verify network health, state, and complex routing tables (BGP, OSPF) pre- and post-deployment. • Drive the adoption of code-based infrastructure definitions and dynamic, living technical documentation (e.g., NetBox, Git Markdown), setting the standard for documentation across the engineering organization.
DevOps Engineer
Weekday (YC W21)We are a Y-Combinator-backed startup building your AI-powered Recruiter Agent
Role Description This is a full-time Remote role for an experienced Senior Platform Architect / DevOps Engineer — Microsoft Fabric who enjoys solving complex data problems and takes ownership from design through delivery. You can work from home (anywhere in INDIA, job location is in India). This role is ideal for someone who can build scalable data pipelines on Microsoft Azure Fabric, leverage AI to improve engineering productivity, and work closely with business and technical stakeholders. We're looking for a Senior Platform Architect with strong DevOps expertise to design, implement, and operationalize our Microsoft Fabric data platform. You'll own the architecture of our unified data estate — spanning OneLake, Data Factory pipelines, Synapse Data Engineering, Power BI, and Real-Time Analytics — while building the CI/CD, governance, and automation backbone that lets our data teams ship reliably and securely. This is a hands-on architecture role: you'll set technical direction and also build it yourself. What You'll Do - Design and own the end-to-end architecture for Microsoft Fabric, including OneLake data lake structure, workspace and capacity strategy, domain/workspace segregation, and medallion (bronze/silver/gold) layering. - Define and implement DevOps practices for Fabric: Git integration, deployment pipelines, environment promotion (dev/test/prod), and infrastructure-as-code for Fabric items (lakehouses, warehouses, notebooks, pipelines, semantic models). - Build and maintain CI/CD pipelines (Azure DevOps or GitHub Actions) for Fabric artifacts, including automated testing, validation, and rollback strategies. - Architect and enforce governance, security, and access controls across Fabric — workspace roles, OneLake data access policies, sensitivity labels, row-level security, and integration with Microsoft Purview. - Design capacity planning and cost optimization strategies for Fabric capacity units (CUs), monitoring consumption and tuning workloads to control spend. - Partner with data engineers, analysts, and platform teams to standardize development patterns, reusable templates, and best practices across Fabric workloads (Data Factory, Data Engineering/Spark, Data Warehouse, Real-Time Intelligence, Power BI). - Own monitoring, alerting, and observability for the platform using Fabric monitoring hub, Azure Monitor, and Log Analytics. - Lead migration efforts from legacy platforms (Synapse Analytics, Azure Data Factory, on-prem SSIS/SQL Server, or other cloud data warehouses) onto Microsoft Fabric. - Mentor engineering teams on Fabric architecture patterns, DevOps practices, and operational excellence. - Act as the primary technical escalation point for platform reliability, performance, and incident response. Qualifications - 8+ years of experience in data platform architecture, cloud infrastructure, or DevOps engineering roles. - 2+ years of hands-on experience with Microsoft Fabric, or deep equivalent experience with Azure Synapse Analytics, Power BI, and Azure Data Factory with a clear path to Fabric. - Strong command of OneLake, Lakehouse and Warehouse architecture, Delta Lake/Parquet formats, and medallion architecture design. - Proven DevOps expertise: Git-based version control, CI/CD pipeline design (Azure DevOps or GitHub Actions), and infrastructure-as-code (Bicep, Terraform, or ARM templates) for Azure/Fabric resources. - Solid understanding of Fabric security model — workspace roles, item-level permissions, OneLake data access roles, sensitivity labels, and integration with Microsoft Entra ID (Azure AD) and Purview. - Experience with Fabric capacity management, performance tuning, and cost governance (CU monitoring, throttling, smoothing). - Hands-on skills in at least one of: PySpark/Spark SQL, T-SQL, Python, or KQL for pipeline and query development. - Experience integrating Fabric with external tools — APIs, Power Automate, Azure Functions, or third-party data sources. - Strong communication skills with the ability to translate architecture decisions for both technical teams and business stakeholders. - Prior experience leading platform migrations to modern data platforms is a strong plus. Requirements - Must have at least 5+ years of experience. Nice to Have - Microsoft certifications: DP-600 (Fabric Analytics Engineer), DP-203 (Azure Data Engineer), or Azure Solutions Architect Expert. - Experience with Power BI semantic model optimization and DirectLake mode. - Familiarity with data mesh or domain-driven data architecture principles. - Background in regulated industries requiring strict data governance (finance, healthcare, etc.). Must-have skills - Microsoft Fabric - Azure DevOps - Data Platform Good-to-have skills - Infrastructure as code - Spark SQL



