Zencoder logo
Zencoder

Zencoder is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.

Senior Engineer, Infrastructure

Infrastructure EngineerInfrastructure EngineerFull TimeRemoteSeniorTeam 11-50Since 2010H1B No SponsorCompany SiteLinkedIn

Location

Europe

Posted

5 days ago

Salary

0

Seniority

Senior

Job Description

Senior Engineer, Infrastructure

Zencoder

• We’re looking for an Engineer to help build and operate the infrastructure behind Zencoder’s AI-powered products. • You’ll work across our production platform, improving its reliability, security, scalability and cost efficiency. • This includes our Kubernetes foundations, cloud infrastructure, networking, data systems and the internal tooling that enables engineers to deploy and operate services confidently. • This is a hands-on engineering role rather than a traditional operations position. • You’ll write code, automate infrastructure, investigate production issues and design systems that reduce operational complexity as the company grows. • The exact problems will evolve quickly. You should be comfortable taking ownership of unfamiliar systems, identifying the highest-leverage improvements and moving between immediate production needs and longer-term platform investments.

Job Requirements

  • You have strong software-engineering skills and regularly write production code.
  • You have experience building and operating infrastructure on GCP, AWS or another major cloud platform.
  • You have hands-on experience with Kubernetes in production.
  • You understand cloud networking and security concepts such as VPCs, IAM, load balancing, firewalls, WAFs, CDNs, DNS and service identity.
  • You have experience with infrastructure-as-code and automated deployment systems.
  • You are comfortable debugging problems across application, infrastructure, networking and data-system boundaries.
  • You have operated distributed systems such as OpenSearch, Elasticsearch, PostgreSQL or similar technologies at scale.
  • Experience deploying or operating large language models with serving frameworks such as vLLM or SGLang is a plus, but not required.
  • You care about reliability, security, developer experience and cost - not just whether infrastructure is technically running.
  • You look for ways to remove operational toil rather than accepting repetitive manual work.
  • You take ownership of important problems and are comfortable working across traditional team boundaries.

Benefits

  • Shape the Future of Software Creation: We’re not just improving how developers write code — we’re redefining how ideas turn into reality. By closing the gap between concept and execution, we’re creating tools that will influence every industry that relies on software.
  • Massive Impact, Real Ownership: At Zencoder, you’ll have full visibility into how your work moves the product and the company forward. You’ll ship features that matter, see the immediate impact of your decisions, and get feedback directly from users — fast.
  • ICs Are the Core: Individual Contributors are the highest-status role at Zencoder. Our culture celebrates those who lead by doing — who create momentum, inspire others, and turn ideas into shipped products.
  • High-Caliber Team & Founder: Work alongside exceptional AI and software engineers, and learn directly from Andrew Filev, founder of a unicorn startup, who brings deep expertise in scaling world-class technology companies.
  • Global & Flexible: We hire talent, not coordinates. Work from wherever you’re happiest and most productive — as long as you bring the energy, focus, and results.
  • Aligned Incentives: Our equity plan ensures that when we succeed, you succeed. Your impact compounds as the company grows.

Related Categories

Related Job Pages

More Infrastructure Engineer Jobs

Graduate - Infrastructure

Ampol

Ampol is a leading Australian energy company committed to delivering high-quality, reliable, and sustainable energy solutions designed to meet the evolving need

Title: Graduate - Infrastructure Location: Australia Job Description: - Permanent full time opportunity post 2 year graduate program - Our program is designed to support your learning and development - Location: Alexandria Head Office / Kurnell Terminal (Sydney, NSW) About Ampol Ampol is a leader in the trans-Tasman transport fuels industry, supplying Australia's largest branded petrol and convenience network, and the owner of Z Energy Limited - a trusted and iconic New Zealand brand. We have also expanded our supply chain and operations into international markets, which operate out of Singapore and Houston, USA. Our Graduate Program Our program is built around structured rotations that immerse you in different areas of the business over two years. You'll gain hands-on experience, develop a well-rounded skill set, and build a strong understanding of how the organisation operates-setting you up for long-term career success. Along the way, you'll benefit from tailored learning and development designed specifically for early-career professionals. From self-paced growth through LinkedIn Learning to interactive workshops, you'll build key skills in communication, leadership, and time management to support your ongoing development. Your rotations: During your time at Ampol, you will join the Infrastructure team and gain experience across Distribution and Infrastructure Strategy & Projects in your first two rotations. We are committed to supporting your development and value your feedback throughout the program, ensuring your rotations provide meaningful learning opportunities and help prepare you for your permanent placement. About you Our ideal candidates are driven to develop commercial, operational, and stakeholder engagement skills, with a passion for continuous improvement and delivering value across our Infrastructure and Commercial Fuels operations. Eligibility/Qualification criteria - Must be a Citizen / Permanent Resident to be considered for this program - Tertiary qualification in Engineering (any discipline) or a Science related field and recently graduated Our benefits - Our total remuneration is competitive. This is across base salary, a performance incentive, employee share offers and a 25% discount on Fuel for two privately used cars! - Access to Ampol''s Benefits & Recognition platform: providing you access to retail discounts and cashbacks at over 500+ retailers in Australia that assist with everyday living expenses. - We are flexible. Many of our teams have embraced hybrid work, balancing time spent remote working, with time spent at an office to connect and work together where it adds value. - Remote Working: Support for up to 3 months remote international working (conditional to 5 days paid leave for every 30 days of remote work). - We value recognition. We have an internal recognition platform amplifying the achievements of those who do great work and demonstrate our capabilities and values. - Career development and learning opportunities including LinkedIn Learning and other tailored training solutions. - Paid Parental Leave - up to 12 weeks paid Parental leave, and up to a year off (unpaid). In addition to the 12 months of unpaid parental leave, employees may apply for a further 12 months of unpaid parental leave (a total of 24 months for each birth) - BabyCare Package - financial and flexible support for parents transitioning back to work. - Need some wheels? Novated Lease options are available. - Invest in your future with the Employee Share Scheme - Leave Options - We offer wellbeing leave and leave purchasing - Care for your Community. Spend one paid day a year volunteering with one of our Ampol Foundation partners. We're an equal opportunity workplace. We not only embrace diversity and inclusion; we celebrate what makes you unique. We welcome applications from people of all ages, cultural backgrounds, and diverse sexualities and genders (including if you identify as transgender). We also highly encourage Aboriginal and Torres Strait Islander peoples to apply for roles with Ampol.

Australia

Senior/ Lead Cloud Infrastructure Engineer

GlobalLogic

GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world's largest and most forward-thinking companies. Since 2000, we've been at the forefront of the digital revolution - helping create some of the most innovative and widely used digital products and experiences.

Role Description The Client is the UK's leading fee-free mortgage broker. The role focuses on supporting the Client's ongoing IT transformation and modernizing its technology and infrastructure stack. The primary initiative involves planning, designing, documenting and executing the migration of systems from traditional datacenter hosted infrastructure into a secure, automated and cost-effective Microsoft Azure and MS365 cloud environment. Qualifications - Solid experience with the following technologies and tools (must-have): - MS Azure Expertise: Tenants and Subscriptions, Cloud Adoption Framework/ Azure Landing Zones, Azure architecture, Azure Entra ID (Enterprise applications, conditional access etc), Azure storage accounts, Azure Virtual Machines, Azure networking, and Azure performance, security and cost management. - MS365 Applications: SharePoint, Teams, OneDrive and Intune. - MS Exchange Online - On-prem Infrastructure: Building and configuration of MS Windows Servers, Active Directory domain knowledge, Hypervisor virtualization management (ideally VMWare), Backup solutions (ideally VEEAM), and Infrastructure performance monitoring. - Additional Technologies: Vulnerability Management, SIEM Technologies, PowerShell and Infrastructure As Code (ideally Terraform and Ansible). - Excellent written and verbal communication (English language). - Desirable technologies and tools: - SQL. - Networking experience (Cisco switch, firewalls). - Cisco Umbrella. - Nutanix Hyperconverged Infrastructure support. - Netapp Experience. - Knowledge of and/or certification in ITIL 4. Requirements - Enable Cloud technology adoption through strategy planning, design, and implementation. - Lead Infrastructure strategy planning alongside the IT Operations Manager to ensure systems are fit-for-purpose and meet business needs. - Act as a Subject Matter Expert on Azure and MS365 applications, and support mentoring and development of other IT colleagues. - Serve as a technical escalation point, collaborating with team members to resolve complex problems and provide root cause analysis. - Enhance automation of cloud enablement tools, such as Infrastructure As Code, and participate in their design and evolution. - Plan and implement system migrations from datacenter hosted infrastructure to the cloud. - Create and review comprehensive technical documentation, including high and low-level designs, architectural statements, and standard operating manuals. - Lead technical projects alongside project managers to define and meet critical timescales. - Design and implement cloud storage and backup solutions. - Provide third-line technical support on L&C IT infrastructure across on-prem, datacenter and cloud environments. - Monitor system performance, costs, availability and functionality and take proactive or remedial action to prevent outages. - Track asset lifecycle and participate in the replacement, retirement and decommissioning of assets. - Assist in planning Disaster Recovery and Business Continuity playbooks, carry out testing and perform remedial work. - Plan, monitor, and test system and data backups in line with IT policies. - Monitor software versions, and plan and implement system upgrades or patches. - Maintain a security-first mindset, follow best practices, and manage technologies to protect systems from internal and external malicious activities. Benefits - Empowering Projects: With 500+ clients spanning diverse industries and domains, we provide an exciting opportunity to contribute to groundbreaking projects that leverage cutting-edge technologies. - Empowering Growth: We foster a culture of continuous learning and professional development, providing timely and comprehensive assistance for every consultant through our dedicated Learning & Development team. - DE&I Matters: We deeply value and embrace diversity, providing equal opportunities for all individuals and fostering an inclusive work environment. - Career Development: Our corporate culture emphasizes career development, offering abundant opportunities for growth and regular interactions with our teams. - Comprehensive Benefits: We provide a comprehensive benefits package that prioritizes the overall well-being of our consultants. - Flexible Opportunities: We prioritize work-life balance by offering flexible opportunities tailored to your lifestyle. Company Description GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world's largest and most forward-thinking companies. Since 2000, we've been at the forefront of the digital revolution - helping create some of the most innovative and widely used digital products and experiences.

Worldwide
24-MAG logo

ML Infrastructure & Kernel Optimization Engineer

24-MAG

This opportunity is available through a leading AI-driven work platform.

Role Description We are sharing a specialised full-time consulting opportunity for US-based MLOps and ML systems engineers with production experience in JAX, PyTorch, distributed training infrastructure, and custom GPU kernel development using Pallas or Triton. This role supports a high-impact generative AI initiative focused on developing and evaluating advanced ML infrastructure tasks for frontier model training. Selected engineers will: - Design technically challenging problems - Produce rigorous solutions - Assess model-generated outputs - Help establish evaluation standards across training pipelines, distributed systems, framework-level optimisation, and GPU kernel performance Qualifications - At least 2 years of dedicated professional experience in MLOps, ML infrastructure, or ML systems engineering - Production experience with JAX, PyTorch, or both at meaningful scale - Hands-on experience writing or optimising custom GPU kernels using Pallas or Triton - Strong knowledge of model-training pipelines, distributed systems, accelerators, and performance optimisation - Experience working within a recognised technology, AI research, or high-performance engineering organisation - Demonstrable professional growth and increasing technical responsibility - Strong written communication and the ability to explain complex engineering decisions clearly - Reliable availability for a full-time, 40-hour weekday schedule Requirements - A degree in computer science, machine learning, electrical engineering, applied mathematics, or a related technical field is highly relevant - Graduate-level education in machine learning systems, distributed computing, or high-performance computing may be helpful - Equivalent professional experience in production ML infrastructure may also be considered - Advanced technical work involving GPU programming, compiler systems, or large-scale model training is especially valuable Benefits - Contribute to advanced generative AI and large-scale model-training initiatives - Apply deep expertise in JAX, PyTorch, Pallas, Triton, and ML infrastructure - Work on challenging problems spanning training systems, distributed computing, and GPU optimisation - Influence the quality of technical training data used in frontier AI development - Join a full-time remote engagement with competitive hourly compensation Contract Details - Full-time W-2 contingent employment arrangement - Fully remote role available to candidates based in the United States - Expected commitment of 40 hours per week during weekdays - This engagement requires full professional availability without conflicting employment or external commitments - Competitive rates between $65–$105 per hour depending on expertise and project scope - Immediate availability is preferred - Work may include onboarding, technical calibration, and ongoing quality-review activities - Project scope and duration may be adjusted according to programme requirements and performance About the Platform This opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams. By submitting this application, you acknowledge that your information may be processed by 24-MAG LLC for recruitment and opportunity matching in accordance with our Privacy Policy: https://www.24-mag.com/privacy-policy

United States
$65 - $105 / hour
UserTesting logo

IT Infrastructure Engineer – AI

UserTesting

Experience what your customers experience with real-time video feedback.

Full TimeRemoteTeam 501-1,000Since 2007H1B No Sponsor

• Set up, configure, and maintain MCP (Model Context Protocol) servers that connect AI systems to internal tools and data sources, including scoping access, managing credentials, and troubleshooting integrations • Design, provision, and support hosting environments for internally built AI applications ranging from lightweight PaaS deployments to full cloud infrastructure on AWS/GCP (VPCs, compute, containers, load balancing) depending on the app’s complexity and scale requirements • Familiarity with AI features within existing SaaS platforms (e.g., Atlassian Rovo, Slack AI, Okta AI, ChatGPT, Google Gemini) • Provision and support AI agent / application infrastructure: service accounts, API key management, permissions scoping, rate limit monitoring, and access auditing for agents operating across internal systems and cloud environments • Build and maintain CI/CD pipelines and infrastructure-as-code (e.g., Terraform, CloudFormation) to support repeatable, secure deployment of AI-hosted applications and supporting services • Maintain documentation for all AI integrations and hosted environments, including architecture diagrams, access maps, runbooks, cost/usage tracking, and troubleshooting guides • Stay current on the AI tooling and cloud ecosystem and proactively identify tools, platforms, or integrations that could benefit the IT team or the broader organization • Support the administration of UserTesting’s IT cloud infrastructure (AWS and/or GCP) and DevOps toolset across the organization as it relates to IT’s AI Infrastructure • Become a subject matter expert for AI-hosted platforms, cloud integrations, and supporting infrastructure • Monitor uptime, performance, and cost of hosted AI applications; implement alerting and observability (e.g., Datadog, CloudWatch, GCO) to catch issues before they affect users • Identify, prioritize, and address technical debt across cloud infrastructure, deployment pipelines, and integrations • Partner with other IT System Admins and Engineering/ Security teams to improve reliability, automation, documentation, and supportability of hosted environments • Act as an escalation point for complex infrastructure and hosting issues for the IT support team • Collaborate with cross-functional teams (Security, Engineering, People Ops, etc.) to improve cloud governance, security posture, and AI tool/app hosting adoption • Stakeholder management, this role will act (at times) as a project lead in some capacity. The ability to interact, demo and collaborate with non technical counterparts is critical. • Document procedures, system changes, and troubleshooting workflows.

United Kingdom