Job Closed

This listing is no longer active.

Michael Baker International logo
Michael Baker International

We Make a Difference

Senior Cloud DevOps Engineer – Azure, AWS

DevOps EngineerDevOps EngineerOtherRemoteSeniorTeam 1,001-5,000H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

149 days ago

Salary

$130K - $170K / year

Seniority

Senior

Job Description

Senior Cloud DevOps Engineer – Azure, AWS

Michael Baker International

• As a Senior Cloud DevOps Engineer at MBI, you will take a hands-on role in designing, building, and maintaining our cloud infrastructure across Microsoft Azure and Amazon Web Services. • You will work in close partnership with the CTO, VP of Product, VP of Infrastructure, and CISO to ensure cloud platforms align with product delivery timelines, maintain operational excellence, and enforce security best practices. • Ensure fast, secure, and reliable software delivery through automation and DevOps best practices. • Manage multi-cloud environments on Azure and AWS, covering compute, storage, networking, and platform services. • Develop and maintain robust CI/CD pipelines to automate software build, test, and deployment processes. • Employ Infrastructure as Code (IaC) tools to automate provisioning and configuration of cloud resources. • Monitor and analyze system performance metrics, ensuring scalability, reliability, and cost optimization. • Work directly with the CISO and security team to ensure cloud environments meet security requirements and compliance standards.

Job Requirements

  • Extensive Cloud Expertise: 5+ years of hands-on experience designing and supporting Azure and AWS cloud infrastructure in production environments.
  • Automation & DevOps Tools: Strong proficiency with CI/CD pipelines (Jenkins, Azure DevOps, GitLab CI, GitHub Actions), Infrastructure as Code (Terraform, CloudFormation, ARM/Bicep), and configuration management tools (Ansible, Chef, PowerShell DSC).
  • Infrastructure as Code Best Practices: Deep understanding and demonstrated experience implementing DevOps infrastructure as code best practices, including modular template design, state management, drift detection, policy-as-code, and environment promotion strategies.
  • Programming/Scripting: Solid scripting and coding abilities in Python, Bash, and PowerShell for automating tasks, integrating systems, and building DevOps workflows.
  • Cloud Resource Management: Proven experience managing cloud resources at enterprise scale, including resource tagging strategies, governance frameworks, budget controls, subscription/account structures, and FinOps practices.
  • Monitoring & Troubleshooting: Experience with monitoring/alerting frameworks (Azure Monitor, CloudWatch, Prometheus, Elastic stack, Grafana) and incident management processes.
  • Security & Compliance Knowledge: Strong understanding of cloud security practices—IAM, network segmentation, encryption, zero-trust architecture, and compliance requirements.
  • Collaboration & Communication: Excellent communication skills and a collaborative mindset.
  • Education: Bachelor’s degree in Computer Science, Engineering, or related field preferred. Equivalent practical experience is also highly valued.

Benefits

  • Medical, dental, vision insurance ​
  • 401 (k) Retirement Plan ​
  • Health Savings Account (HSA) ​
  • Flexible Spending Account (FSA) ​
  • Life, AD&D, short-term, and long-term disability ​
  • Professional and personal development ​
  • Generous paid time off​
  • Commuter and wellness benefits

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Initiate Government Solutions, LLC. logo

DevOps Engineer

Initiate Government Solutions, LLC.

We provide the framework to build solid foundations that allow you to leverage and grow your revenue and capabilities.

DevOps Engineer149 days ago
OtherRemoteTeam 51-200H1B No Sponsor

• Create and foster Self-Service Principles - create automated process for creating on-demand platform for deploying containerized applications and exposing API endpoints to federal data consumers • Architectural Vision & Strategy: Shape and communicate the architectural blueprint for Lighthouse with a focus on scalability, security, automation, and developer experience in a cloud-native environment • Stakeholder Engagement: Collaborate with product owners, security teams, and stakeholders to facilitate and prioritize features that Program Manager(s) can use to meet business needs and create technical roadmaps, ensuring compliance with VA, federal, and industry requirements • Establish and Enforce Standards: Define, document, and advocate for engineering best practices across the VA Lighthouse program, leveraging your expertise in DevSecOps and Kubernetes • Platform as a Service (PaaS): Provide clients with PaaS solutions—delivering low-code/no-code development platforms and enterprise-grade API engineering to accelerate application delivery, streamline integration, and enable scalable digital solutions • End-to-End Software Development: Execute efforts across the full software development lifecycle, automating processes, boosting operational efficiency, and strengthening system reliability • DevSecOps Ownership: Design, implement, and maintain secure, automated CI/CD pipelines and increased monitoring and platform related business intelligence • Automate build, test, deployment, and infrastructure processes; champion shift-left security practices • Continuous Innovation: Continuously refine the platform and deliver innovative, reliable solutions to meet the mission • Governance & Compliance: Ensure solutions meet all compliance requirements, including audit readiness and documentation

Washington
Job Closed
Asymmetric logo

Site Reliability Engineer

Asymmetric

Early stage capital for disruptive technology companies.

DevOps Engineer149 days ago
OtherRemoteTeam 1-10H1B No Sponsor

• Manage and maintain a globally distributed blockchain infrastructure fleet • Design, architect, deploy, and operate production-grade infrastructure services • Implement and maintain infrastructure-as-code across development, staging, and production environments • Ensure high availability and performance of mission-critical systems • Contribute to automation, CI/CD pipelines, and operational tooling • Monitor system health and respond to incidents with strong troubleshooting fundamentals • Uphold the highest standards of integrity, professionalism, and operational discipline

United States
Job Closed

This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description To build and maintain the automated production line for PHIN's Physical Superintelligence. You will own the plumbing that allows our simulation engine to seamlessly scale, ensuring that our team can deploy updates multiple times a day and ingest massive amounts of simulation data without friction. - Greenfield Observability: Architect and implement a comprehensive logging, monitoring, and alerting stack across our platform from the ground up. - Compute Architecture, Scaling & FinOps: Provision, manage, and optimize highly concurrent scaling clusters. Act as a cloud-agnostic thinker to direct future architecture and implement rigorous FinOps practices to minimize the cost of running thousands of simultaneous jobs. - Infrastructure as Code (IaC): Own, maintain, and expand our Terraform footprint. - Continuous Deployment (CD): Design and maintain high-velocity CI/CD pipelines supporting multiple deployments per day. Ensure "code to production" is a seamless, automated journey. - Backend Robustness: Manage the API layer that sits between the infrastructure and the application layer. Read and refactor services to optimize data movement, squash bottlenecks, and maintain security. - Data Pipeline Architecture: Build the underlying pipelines to move, store, and process the massive datasets generated by atomic-scale simulations. - Platform DevEx & MLOps: Build self-serve tooling and event-driven pipelines that empower the entire organization. Create seamless abstractions so our developers can focus on what they do best. - DevOps & Intelligence Automation: Ruthlessly automate manual toil. Use and build AI-driven tools to manage logs, infrastructure provisioning, and business workflows. - Standard Enterprise Security: Implement and maintain security best practices (SOC2/ISO focus) required for enterprise-grade contracts. Qualifications - 5–8 years as a high-output Individual Contributor in Infrastructure or Backend roles. - Comfortable touching any part of the system—from networking and security to API design and data engineering. - Familiarity with Python and TypeScript/Node.js. - Deep experience with major cloud providers. - Familiarity with high-performance computing (HPC) schedulers like Slurm is a major plus. - Not married to one framework; you choose the best tool for the job (K8s, Serverless, HPC Schedulers, etc.). - Expert user of intelligence tools (Claude, Cursor, Codex, Copilot, Agents, etc.) to 10x your own productivity and automate business tasks. - Previous experience working closely with machine learning teams, supporting ML workflows, or building MLOps pipelines is highly desirable.

United States
Job Closed
Weekday (YC W21) logo

Staff Engineer – DevOps

Weekday (YC W21)

We are a Y-Combinator-backed startup building your AI-powered Recruiter Agent

DevOps Engineer149 days ago
Full TimeRemoteTeam 11-50Since 2021H1B No Sponsor

• Architect and evolve our DevOps ecosystem, champion cloud cost governance, and implement best-in-class container orchestration practices. • Work cross-functionally with engineering, security, and finance teams to ensure operational excellence while proactively managing infrastructure spend. • Lead end-to-end DevOps strategy, including CI/CD pipelines, automation, infrastructure-as-code, and release engineering. • Design scalable, resilient cloud-native architectures aligned with business growth. • Establish DevOps best practices, reliability standards, and operational governance. • Architect and manage large-scale Kubernetes environments for production workloads. • Optimize workloads across clusters for performance, reliability, and cost efficiency. • Build and maintain containerized applications using Docker and Kubernetes, ensuring portability and scalability. • Drive multi-cluster, multi-region deployments where necessary. • Own infrastructure cost visibility and optimization initiatives. • Implement cloud cost-saving strategies including rightsizing, reserved capacity planning, auto-scaling optimization, and workload scheduling. • Create dashboards and reporting mechanisms to track infrastructure ROI and spend trends. • Continuously identify inefficiencies and implement measurable cost-reduction initiatives without compromising performance. • Design and implement comprehensive monitoring systems using Grafana and related observability tools. • Build real-time dashboards for system health, performance metrics, and cost insights. • Establish alerting frameworks to minimize downtime and improve incident response. • Drive improvements in system reliability through data-driven monitoring and post-incident analysis. • Automate provisioning, deployments, scaling, and recovery processes. • Improve system resilience, availability, and disaster recovery strategies.

India
Job Closed