Flatgigs logo
Flatgigs

Scaling Investor-Backed Startups & Growth Companies

Senior DevOps Engineer (MLOps Focus)

DevOps EngineerDevOps EngineerFull TimeRemoteSeniorTeam 1-10Since 2023H1B No SponsorCompany SiteLinkedIn

Location

Pakistan

Posted

96 days ago

Salary

0

Seniority

Senior

English

Job Description

Senior DevOps Engineer (MLOps Focus)

Flatgigs

At AHOY, we reimagine and redefine movement. We’re building a next-generation technology ecosystem across mobility, logistics, and infrastructure, enabling innovators to create, scale, and operate smarter systems. We’re looking for a hands-on Senior DevOps Engineer to help us build, scale, and secure our multi-cloud infrastructure. This role sits at the intersection of DevOps, AI infrastructure, and cloud security, where you’ll take ownership of systems powering advanced machine learning workloads. As part of a fast-moving engineering team, you’ll also play a key role in supporting internal IT operations, ensuring reliability across both platform and people. What You’ll Work on - Design and maintain secure, scalable cloud infrastructure across Azure and GCP - Build and optimise CI/CD pipelines and deployment workflows - Implement and manage Infrastructure as Code (IaC) (Terraform or similar) - Deploy and operate machine learning pipelines in production environments - Manage and scale multi-cluster GPU environments for AI workloads - Set up monitoring, logging, and alerting systems to ensure reliability - Apply cloud security best practices across infrastructure and access layers - Troubleshoot and resolve system, deployment, and performance issues - Provide hands-on IT support (user access, devices, internal systems) to ensure smooth day-to-day operations - Collaborate with engineering teams to improve scalability, performance, and automation

Job Requirements

  • 4+ years of experience in DevOps / SRE / Infrastructure Engineering
  • Strong experience with Azure and Google Cloud Platform (GCP)
  • Hands-on experience building and managing CI/CD pipelines
  • Solid expertise in Infrastructure as Code (Terraform, Pulumi, etc.)
  • Experience managing Kubernetes environments and GPU workloads
  • Proven experience deploying machine learning models or MLOps pipelines
  • Strong understanding of cloud security, networking, and system reliability
  • Comfortable handling IT support responsibilities alongside core DevOps work
  • Strong problem-solving mindset with the ability to work independently

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Okta logo

Site Reliability Engineer

Okta

The World's Identity Company

DevOps Engineer96 days ago
Full TimeRemoteTeam 5,001-10,000Since 2010H1B Sponsor

Role Description As a mid-level Site Reliability Engineer, you'll join our SRE team based in Europe to ensure our production systems are not only operational but also resilient, scalable, and ready for exponential growth. This isn't just about keeping the lights on; it's about directly contributing to the platform's core resiliency and robustness. You'll be a hands-on builder, crafting solutions that make our system more reliable by design. - Design and build custom software in Go to enhance the platform's reliability, resiliency, and redundancy. - Partner with engineering teams to embed reliability principles, improving the availability, performance, and observability of our services. - Use your deep understanding of infrastructure and observability principles to identify opportunities for improvement within the product and implement solutions. - Contribute to our on-call rotation, providing rapid, effective response to critical incidents and using your expertise to troubleshoot, mitigate or accurately escalate production issues. - Develop and refine our SRE tooling and processes, focusing on automation and operational efficiency. - Define, document, and champion reliability best practices across the organisation. Qualifications - A proactive and systematic approach to problem-solving, with a high degree of ownership. - Proven experience in a production environment supporting large-scale, mission-critical applications with a high degree of autonomy. - Proficiency in at least one programming language, with a strong preference for Go. You should be comfortable writing custom applications, not just scripts. - Experience with infrastructure as code (Terraform), container orchestration (Kubernetes, Docker) and GitOps (ArgoCD). - Demonstrable expertise in a major cloud provider (Azure, AWS, or GCP). - A strong grasp of microservices architecture, databases (SQL, NoSQL), and networking fundamentals, so you can understand how custom code can solve platform-level issues. - An understanding of core SRE principles, including SLIs, SLOs, and error budgets. - Experience in an on-call rotation for a 24/7 cloud-based environment. - Exceptional communication and collaboration skills, with a proven ability to work effectively in a remote, distributed team, where tasks may be self-driven. Benefits - Supporting Your Well-Being - Driving Social Impact - Developing Talent and Fostering Connection + Community Company Description Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one. Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran.

Worldwide
Job Closed
NBCUniversal logo

DevOps Intern, Direct-to-Consumer Engineering

NBCUniversal

Here you can create the extraordinary. Join us.

DevOps Engineer96 days ago
InternshipRemoteTeam 10,001+Since 2004H1B Sponsor

• As an NBCUniversal Academic Year intern, you’ll work on real projects and be part of our collaborative culture. • Contribute to meaningful work while building skills that matter. • Work closely with development and operations partners to ensure reliable, scalable, and efficient systems that power TVE platforms. • Responsibilities include building, testing, and maintaining infrastructure and technology stacks, implementing and optimizing CI/CD pipelines, monitoring system health and automating maintenance tasks, and contributing to process improvement initiatives aimed at enhancing quality while reducing time and costs.

New York
$30 / hour
Job Closed
Nagarro logo

Principal Engineer, Python DevOps

Nagarro

Nagarro (Frankfurt: NA9) is a leader in digital product engineering and drives technology-led business breakthroughs.

DevOps Engineer96 days ago
Full TimeRemoteTeam 10,001+Since 1996H1B Sponsor

• Design and develop scalable web applications using Python and modern frontend frameworks. • Build and maintain backend services and APIs for integrations. • Develop responsive frontend applications using JavaScript frameworks. • Implement microservices and integrate with databases and cloud platforms. • Ensure application security, performance, and scalability. • Contribute to CI/CD pipelines and DevOps processes. • Participate in code reviews, technical discussions, and mentoring. • Implement monitoring, logging, and system reliability improvements. • Collaborate with cross-functional teams to deliver end-to-end solutions.

India
Job Closed
Corning logo

DevSecOps Lead

Corning

Headquartered in Corning, New York, Corning is a leading global manufacturer of specialty glass and ceramics. This company has a long history of innovation and

DevOps Engineer96 days ago

Lead the security and compliance program while managing security tools and cloud infrastructure. Collaborate across teams to implement automated solutions and enhance security processes, ensuring readiness for audits and compliance standards.

Canada