Workato is a computer software company that has developed an enterprise automation platform with easy-to-use automation and integrations. The company fosters a
Senior Infrastructure Engineer
Location
Netherlands
Posted
2 days ago
Salary
0
Seniority
Senior
Job Description
Senior Infrastructure Engineer
Workato
Role Description We are looking for a Senior+ Platform Infrastructure Engineer who thrives in a highly autonomous environment and enjoys taking ownership of complex technical challenges. Our Platform Infrastructure team is responsible for the foundation that powers all production services. While the team's structure continues to evolve as we grow, its core responsibilities are well established. You'll have the opportunity to influence technical direction, improve platform reliability, and help shape future infrastructure initiatives. This role is best suited for engineers who enjoy solving problems across multiple infrastructure domains rather than focusing on a single specialization. You will do: - Design, build, and operate scalable infrastructure and platform services. - Improve platform reliability through monitoring, alerting, incident management, and operational excellence. - Develop and maintain our Kubernetes-based compute platform. - Work with AWS production infrastructure, including compute, networking, and load balancing. - Build and enhance observability solutions (monitoring, logging, tracing, alerting). - Participate in security and compliance initiatives, including vulnerability remediation and infrastructure hardening. - Support onboarding of new services and facilitate their transition between engineering teams. - Own projects end-to-end from identifying dependencies and coordinating with stakeholders to successful delivery. - Drive improvements that reduce operational overhead through automation and better platform design. - Participate in production releases and on-call rotations. - Engineers may have opportunities to move between infrastructure teams based on business needs and personal interests, allowing exposure to different technical domains over time. Qualifications - Extensive experience as a Senior Platform Engineer , Infrastructure Engineer , or Site Reliability Engineer . - Strong sense of ownership and ability to work independently with minimal supervision. - Ability to identify problems, define priorities, and drive initiatives without waiting for detailed instructions. - Broad infrastructure background with experience across multiple domains rather than deep specialization in only one area. - Strong understanding of Linux internals, including: - Virtual memory - Process resource management - cgroups - Hands-on experience with: - Kubernetes - Distributed systems - AWS production infrastructure (EC2, VPC, Load Balancers, etc.) - Networking fundamentals - Proven ability to deliver complex technical initiatives from concept to production. Requirements - 10+ years of experience in infrastructure or systems engineering roles. - Proven expertise in distributed systems and cloud-native architectures. - Strong understanding of Linux internals, including networking and system-level performance tuning. - Experience operating infrastructure on AWS at scale. - Proficiency with Terraform and Kubernetes. - Solid understanding of network security and cloud security best practices. - Excellent analytical, troubleshooting, and communication skills. Nice to Have - Experience with one or more of the following: - Infrastructure security and compliance - Vulnerability management and security hardening - Identity and access management - Firewall configuration and privilege management - Compliance frameworks and interpreting security requirements - Observability platforms, monitoring, logging, and alerting - Database infrastructure and performance troubleshooting - CI/CD pipelines and deployment automation - Release engineering What Makes You Successful - Works effectively without micromanagement. - Takes responsibility for technical decisions and can clearly justify design choices. - Pays close attention to detail, especially when working on security-sensitive systems. - Thinks beyond immediate fixes and strives to reduce long-term operational burden. - Collaborates effectively across engineering teams to unblock dependencies and deliver results. - Is comfortable working across multiple areas of infrastructure and continuously learning new technologies. Why Join Us? - Work on infrastructure that powers production at scale. - Influence architectural decisions and long-term platform strategy. - Solve complex engineering challenges across infrastructure, Kubernetes, reliability, security, and cloud platforms. - Enjoy a high level of autonomy, trust, and technical ownership. - Grow across different infrastructure domains as the platform and team continue to evolve. Tech Stack - AWS - Terraform - Kubernetes - ArgoCD - Kafka - PostgreSQL - Redis - ClickHouse - Grafana - VictoriaMetrics - Vault
Related Guides
Related Categories
Related Job Pages
More Infrastructure Engineer Jobs
Senior Infrastructure Engineer
UnitedHealth GroupUnitedHealth Group is a healthcare and well-being company that’s dedicated to improving the health outcomes of millions around the world. We are comprised of
Role Description The Epic Systems Configuration Architect (ECSA) / Sr I O Engineer serves as the enterprise technical authority for Epic infrastructure, client deployment strategy, interoperability services, API management, and technical architecture governance. This role is dedicated to ensuring secure, highly available, and performant access to enterprise Epic applications while eliminating single points of failure. - Lead technical strategy, drive complex troubleshooting, and optimize system infrastructure utilizing advanced tools such as Epic System Pulse analytics. - Work with cutting-edge healthcare interoperability technologies to directly improve clinical application performance across our global network. Qualifications - Bachelor’s degree in Computer Science, Information Technology, or a related field (or equivalent practical experience) - Active Epic certification in Hyperspace Deployment Administration (Client Access to Hyperspace) - Active Epic certification in Web and Service Servers Administration (covering Hyperspace Web, MyChart, Interconnect, Web BLOB, BCA, and Epic Printing Service) - 5+ years of experience with Epic client/server architecture and technical infrastructure deployment - 5+ years of enterprise systems administration experience with Windows Server - 3+ years of experience managing SSL/TLS certificates, security tokens, and enterprise authentication protocols (such as OAuth, SAML, or OpenID Connect) - 3+ years of experience with network traffic routing, load balancers, and Citrix NetScaler configuration - 3+ years of experience using monitoring and analytics tools (such as Epic System Pulse) to troubleshoot complex system performance issues Requirements - Deploy, configure, maintain, and troubleshoot Epic Hyperspace, Hyperdrive, Kuiper, and Client Installer environments. - Administer Web and Service servers, including Hyperspace Web, MyChart, Interconnect, and Web BLOB infrastructure, ensuring optimal high availability and load balancing. - Configure and manage Epic APIs and interoperability services, supporting integrations such as Care Everywhere, FHIR, and SMART on FHIR. - Implement, maintain, and test Business Continuity Access (BCA) workflows to guarantee continuous disaster recovery readiness. - Proactively monitor infrastructure health, utilization, and capacity using Epic System Pulse and other analytics tools. - Manage security compliance across Epic environments, including SSL/TLS certificate installation and OAuth/SAML token management. - Configure and administer Citrix NetScaler environments, load balancing, and enterprise print/fax services for reliable routing. - Leverage enterprise-approved AI and automation tools to streamline infrastructure configuration workflows, automate repetitive systems management tasks, and drive continuous improvement. - Evaluate emerging technological trends and automated solutions to inform Epic infrastructure design and strategic innovation. - Design, develop, and deploy AI-powered solutions to address complex business challenges with emphasis on responsible use of AI. Benefits - Comprehensive benefits package - Incentive and recognition programs - Equity stock purchase - 401k contribution (subject to eligibility requirements)
• Provide hands-on IT support for plant floor network infrastructure, including: • Network cabling, switches, and patching • LAN/WAN connectivity • VLAN configuration and segmentation • IP addressing and network device configuration • Wired and wireless (Wi-Fi) network support • Troubleshoot and resolve network connectivity issues for plant floor devices, including: • CNC machines, PLCs, HMIs, and industrial PCs • IoT/edge devices and data collection systems • Manufacturing execution systems (MES) interfaces • Support infrastructure readiness assessments for integrating machines into Smart Connected Factory (SCF) systems by: • Validating network availability and performance and ensuring proper IP schema and connectivity standards • Identifying gaps and recommending solutions • Ensure site IT hardware and software readiness to support: • Support Manufacturing Digital Transformation initiatives and Digital Thread enablement across systems and equipment • Collaborate with cross-functional teams (Manufacturing, Engineering, Automation, and Corporate IT) to: • Enable secure and reliable machine connectivity. • Support deployment of smart factory technologies • Maintain backup and develop strategies to ensure Continuity of manufacturing equipment connected to the IT system • Maintain compliance with IT and cybersecurity standards • Assist with installation, configuration, and maintenance of plant floor IT equipment and systems • Document network configurations, troubleshooting steps, and infrastructure standards
• Own the services driving order lifecycles, internal ledgers, multi-currency funding flows, and automated reconciliation from day one. • Build high-throughput, fault-tolerant infrastructure capable of expanding from early users to hundreds of thousands of active investors without compromising auditability or data integrity. • Partner directly with the CTO to define company-wide standards for correctness, distributed system reliability, and production operational excellence. • Design, implement, and maintain end-to-end order routing, execution lifecycles, internal double-entry ledgers, and automated funding/reconciliation systems. • Build production systems optimized for concurrency and partial network failures using idempotent handlers, durable event history, replayable state, and self-healing mechanisms. • Integrate directly with partner clearing broker-dealers, international brokers, banking rails (ACH, wires), FX mechanisms, and automated KYC/AML verification providers. • Take end-to-end operational ownership—including sandbox testing environments, monitoring/alerting, incident response, on-call rotations, and postmortems. • Build robust internal tools and dashboards for operations, compliance, and support teams to audit, investigate, and resolve issues proactively. • Support underlying data pipelines powering our real-time AI research and content-translation engines (no dedicated ML background required).
Role Description At Sunrise Robotics, this role ensures that our robots can be deployed, updated and supported reliably at scale. As deployments grow, we need consistent environments, secure access control and automated workflows that remove friction from engineering and reduce operational risk. Through deep troubleshooting, incident investigation, clean documentation, and clear communication, you will keep our internal systems and engineering infrastructure reliable. You’ll work with Linear as the ticketing/triage system and partner closely with Engineering to eliminate recurring infrastructure problems. You will own deployment environments, cloud provisioning, endpoint configuration, networking and security practices end to end. Your accountability is straightforward: ensure our infrastructure is consistent, secure and reliable so the team can ship and deploy robots without unnecessary blockers. What You’ll Do - Automate and maintain deployment workflows to streamline robot cell rollouts, reduce manual setup, and eliminate recurring operational issues. - Containerise and stabilise our robotics software stack to ensure consistent, reproducible environments across development, cloud and deployed production systems. - Standardise and manage Linux and Windows environments, establishing repeatable configuration and provisioning practices. - Provision and maintain cloud infrastructure, including secure access, permissions and identity management (SSO, SSH), with clear ownership of reliability and recovery practices. - Configure and manage networking infrastructure, including VPNs and secure remote access. - Own endpoint security and system hardening across internal machines and deployed systems. - Design and operate monitoring and alerting systems, leading incident response, root cause analysis and preventative improvements. - Own operational intake and prioritisation of infrastructure issues, ensuring clear communication and structured resolution. - Maintain clear documentation of environments, incident learnings, processes and recovery procedures to ensure continuity. Qualifications - Proven experience automating infrastructure and deployment workflows using scripting and programming (e.g. Ansible, Python, Bash). - Strong hands-on Linux systems administration experience. - Experience with scalable containerised environments in production. - Practical experience with cloud infrastructure provisioning, reliability and access management (AWS, GCP or Azure). - Solid understanding of networking fundamentals, VPN configuration and secure remote access. - Experience implementing and managing credentials, SSO and identity/access control systems. - Demonstrated experience configuring, standardising and hardening Linux and Windows endpoints. - Experience building monitoring and alerting systems, and using logs and metrics for structured diagnosis. - Experience leading incident response, including clear communication, root cause analysis and preventative follow-up actions. - Experience establishing repeatable configuration or environment management practices in production systems. - Strong troubleshooting capability, with evidence-based investigation and durable problem resolution. - Ability to write clear, structured technical documentation for repeatable processes, incident learnings and handovers. - Experience working in GPU-enabled or hardware-adjacent environments (e.g. NVIDIA stack, CUDA, driver-level debugging). What Makes You Stand Out - Experience operating infrastructure in robotics, autonomy, AI-heavy or other hardware-adjacent environments. - Experience with SaltStack or similar infrastructure management frameworks and their use from early-stage foundations to multi-site production-grade deployments. - Experience with Infrastructure as Code tools (e.g. Terraform, Pulumi or similar). - Experience designing backup, disaster recovery and resilience testing processes. - Experience building internal tooling to improve developer productivity. - Prior experience serving as a senior escalation point or mentoring engineers in operational best practices. Benefits - High exposure: Work directly with founders and leadership on decisions that shape the company’s trajectory. - Career acceleration: This role sits at the core of how we scale. You’ll gain deep exposure to robotics infrastructure, cloud systems, security and deployment at a formative stage of the company. - Real impact: Your work will directly reduce deployment time, improve system stability, strengthen security and enable faster customer rollouts. When a robot launches smoothly, it will be because the foundation you built made it possible.


