DevOps Engineer Remote Jobs in Montana (US)
This page tracks remote devops engineer openings that are location-eligible for Montana.
This page tracks remote devops engineer openings that are location-eligible for Montana.
Open jobs
2,194
Hiring companies this week
9
Salary sample
$60 - $175,000
Jobs added last hour
0
2194 Jobs
1367 Companies
• Own reliability targets across our backend/API, worker services, applications and CV pipeline; MTD, MTM, MTR, and follow-through on root causes. • Level up our monitoring and alerting, and build out auto-remediation, so on-call load scales with automation, not headcount. • Partner with our agentic engineering work to build agents that triage alerts and handle routine remediation. • Harden and optimize our GCP infrastructure (Cloud Run, Cloud SQL, GCS) for cost and performance as load scales. • Own database scale and performance; connection pooling, query optimization and indexing, read replicas, and capacity planning, so Postgres doesn't become the bottleneck as data volume grows. • Improve the reliability of our ML training and monitoring infrastructure, in partnership with the CV/ML team. • Run blameless postmortems and drive fixes for root causes, not just symptoms. • Participate in on-call rotation.
We provide cross-platform linking and attribution solutions to the world's leading digital brands.
• Partner with Developers to produce high-performing and robust services through rigorous testing and release procedures • Design infrastructure, monitoring, processes, and standards for systems and applications • Support services through design, development, load testing, and launch phases • Develop, measure, and monitor key performance and service level indicators including availability, latency, and overall system health • Define and establish SLIs, SLOs, and error budgets with service owners, and drive adoption across platform teams • Profile and optimize platform performance, resilience, and efficiency, including latency, throughput, and capacity planning under load • Participate in incident response and root cause analysis • Remediate tasks and develop preventative and automated measures to meet SLAs/SLOs/SLIs • Manage monitoring services utilized by applications
We provide cross-platform linking and attribution solutions to the world's leading digital brands.
• Lead, and take ownership of our core database systems to ensure high levels of performance, availability, and security • Communicate critical data changes to the data warehousing team • Work alongside both engineering teams and data teams to shepherd schema changes from design to deployment to ensure architectural and downstream compatibility • Build and maintain automated test and deployment data pipelines, including production level data testing at scale, with SRE and DevOps teams • Proactively work on detecting weak areas and incorporate into automated monitoring and production alerting with our Cloud Ops teams • Document and automate processes to eliminate time spent on recurring tasks • Evaluate new database features (and products if needed) to meet scalability and global uptime requirements • Own database backup and recovery, and help establish disaster recovery targets and practices as the systems scale • Support, debug, and perform RCA on production database issues, and participate in incident response as needed
• Maintain our CI/CD pipelines and containerized infrastructure (Docker, ECS/ECR Fargate, Terraform), keeping deployments fast and reliable day to day • Monitor cloud infrastructure on AWS, catching and resolving issues before they affect customers • Support incident response, helping triage and resolve production issues as they come up • Execute steady-state maintenance and operational work, freeing up senior team members to focus on architecture and security priorities • Ramp into our existing systems and tooling, building depth over time as you take on more ownership
Accelerating speed to care by optimizing provider schedules, streamlining clinical communication, and engaging patients.
• Help configure and rollout customers of our existing AI voice agent • Build and run evaluation suites to validate behavior, catch regressions, and tune prompts as new customers are deployed • Develop new features and fixes to the core product • Optimize latency, inference cost, and hallucination guardrails • Run and improve model evaluation, benchmarking, and observability across the product • Monitor cost controls, usage monitoring, and security/PHI handling constraints • Partner with Product and Customer Success to turn customer requirements into working, validated deployments, and own customer go-live, post-launch monitoring, and production issue debugging • Help streamline and automate the onboarding process to reduce the engineering effort required to bring each new customer live
Abercrombie & Fitch Co. is a leading, global specialty retailer of apparel and accessories for men, women and kids.
Role Description We are seeking a Framework Engineer who can build and support Azure-based data and analytics frameworks with Databricks as the primary compute platform. This role is focused on implementing reusable ingestion and transformation patterns, strengthening platform standards, and enabling engineering teams to deliver reliable data solutions at scale. In this role, you will implement and maintain modern data platform frameworks across Azure Data Factory, Databricks, ADLS, and related services. This includes: - Contributing to metadata-driven ingestion patterns (batch, streaming, and near real-time) - Supporting CI/CD and automated testing practices - Helping teams adopt shared accelerators, playbooks, and engineering standards through the Azure Data and Analytics Center for Enablement (C4E) You will partner with engineering, enterprise architecture, and DevOps teams to deliver scalable, secure data integrations and support modernization from on-prem platforms to Azure. Success in this role requires: - Strong hands-on engineering skills - The ability to translate platform standards into practical delivery - Clear collaboration with application and business partners This role will report directly to the Manager, Data & Analytics DevOps. What Will You Be Doing? - Implement and maintain Azure and Databricks frameworks, standards, accelerators, and reusable Data Factory/Databricks ingestion and transformation patterns - Build and support migration of data workloads from on-prem technologies onto Azure using Databricks, ADLS, ADF, and related services - Contribute to metadata-driven frameworks for large-scale ingestion across batch, streaming, and near real-time patterns - Partner with engineering, enterprise architecture, and technology teams to apply integration standards and deliver scalable, flexible, and secure solutions - Work with DevOps to implement CI/CD pipelines, automated testing, and release processes for ADF and Databricks solutions - Contribute to playbooks and enablement materials that improve knowledge sharing, onboarding, and adoption of platform standards through C4E - Participate in code reviews and provide design feedback in Azure, Databricks, and modern data engineering practices - Support modernization from incumbent technologies (e.g., Hadoop, Netezza, legacy BI) to Azure, Databricks, Snowflake, Power BI, and related cloud services - Deliver framework and integration solutions through design, development, testing, deployment, and production support, with guidance from senior engineers as needed Qualifications - Hands-on experience building data integrations and platform frameworks in complex enterprise environments - Solid expertise in PySpark and SQL/PLSQL, including performance tuning and cost awareness for Azure and Databricks workloads - Working knowledge of Databricks (Spark clusters, Unity Catalog, Delta Lake, notebooks, workflows, Delta Live Tables, jobs, and pipelines) - Working knowledge of Azure Data Factory across common ingestion patterns and activities - Experience with Azure DevOps and Git, including contributing to deployment pipelines for ADF and Databricks - Familiarity with common data exchange formats such as XML, XSD, YAML, and JSON - Understanding of modern data engineering practices, SDLC, CI/CD, automation, and observability - Strong collaboration skills and a self-starter mindset, with the ability to solve problems and adapt in a fast-moving environment Requirements - Bachelor's degree in computer science or a related field, or equivalent work experience - 3–7 years of experience in data engineering, integration, or related platform engineering roles - Experience implementing Azure data platform solutions using Databricks and/or Azure Data Factory - Experience supporting metadata-driven or reusable framework patterns for data ingestion and transformation - Experience with Snowflake, Kafka, API management, and/or REST interfaces preferred - Retail, ecommerce, merchandising, payment, or store systems integration experience is a plus Company Description Abercrombie & Fitch Co. (A&F Co.) is a global, digitally led omnichannel specialty retailer of apparel and accessories catering to kids through millennials with assortments curated for their specific lifestyle needs. The company operates a family of brands, including Abercrombie brands and Hollister brands, each sharing a commitment to offer products of enduring quality and exceptional comfort that support global customers on their journey to being and becoming who they are. Abercrombie & Fitch Co. operates stores under these brands across North America, Europe, Asia and the Middle East, as well as the e-commerce sites abercrombie.com, abercrombiekids.com and hollisterco.com. Learn more about A&F Co. by visiting our corporate website. ABERCROMBIE & FITCH CO. IS AN EQUAL OPPORTUNITY EMPLOYER.
Advancing the power of technology and innovation to serve and protect our world.
Role Description We are seeking a highly skilled and motivated Senior DevSecOps Engineer with proven experience leading Agile teams. The selected candidate will play a critical role in designing, developing, and implementing DevSecOps practices and tools to enable efficient and secure software delivery pipelines. This position requires strong leadership skills, an in-depth understanding of cloud technologies, security, and Agile frameworks, as well as hands-on expertise in automated software delivery and deployment. The ideal candidate will be passionate about fostering collaboration, driving innovation through automation, ensuring security is embedded throughout the SDLC (Secure Software Development Lifecycle), and mentoring cross-functional teams in an Agile environment. - Establish, maintain, and enforce DevSecOps best practices throughout the software development lifecycle (SDLC). - Evaluate, integrate, and maintain DevSecOps tools and technologies to build and improve automated CI/CD pipelines. - Lead Agile teams by facilitating daily stand-ups, sprint planning, reviews, retrospectives, and other Agile ceremonies. - Ensure that security is a top priority in the design, implementation, and delivery of all development projects. - Collaborate across teams, including software developers, security engineers, and IT operations, to ensure seamless integration of security practices within workflows. - Drive team collaboration and communication, fostering a culture of innovation, accountability, and continuous improvement. - Provide mentorship to development team members on DevSecOps practices, tooling, and Agile methodologies. - Generate reports and provide updates to leadership on project progress, risks, and implemented security measures. Company Description SAIC® is a premier mission integrator focused on advancing the power of technology and innovation to serve and protect our world. Our robust portfolio of offerings across the defense, space, intelligence, and civilian markets includes secure high-end solutions in mission IT, enterprise IT, engineering services, and professional services. We integrate emerging technology, rapidly and securely, into mission critical operations that modernize and enable critical national imperatives. - We are approximately 23,000 strong; driven by mission, united by purpose, and inspired by opportunities. - SAIC is an Equal Opportunity Employer. - Headquartered in Reston, Virginia, SAIC has annual revenues of approximately $7.3 billion. - For more information, visit saic.com . - For ongoing news, please visit our newsroom .
Role Description We're looking for an experienced Senior DevOps Engineer to help drive the reliability, scalability, and efficiency of our infrastructure and deployment processes. In this role, you'll tackle complex technical challenges, designing and implementing solutions that improve our system architecture and developer workflows. You'll play a pivotal role in guiding projects from design through production support, and you'll mentor junior engineers in DevOps best practices. - Oversee the reliability, performance, and security of critical production services from design to deployment, ensuring they meet our uptime and performance targets. - Collaborate with development, QA, and product teams to build and maintain resilient infrastructure and efficient deployment pipelines. - Automate infrastructure provisioning and software deployments using Infrastructure as Code and CI/CD tools, reducing manual work and errors. - Participate in and improve our 24×7 on-call process, swiftly troubleshooting incidents and performing root cause analysis to prevent recurrence. - Document and standardize processes and configurations, sharing knowledge to uplift the entire engineering team’s capabilities. Qualifications - 5+ years of experience in DevOps, SRE, or Software Engineering roles, with increasing responsibility in system design and operations. - Extensive experience with containerization (Docker) and orchestration (Kubernetes) in production environments, including managing and scaling clusters. - Proficiency in Infrastructure as Code (Terraform, CloudFormation, etc.) and configuration management tools (Ansible, Puppet) to automate infrastructure provisioning. - Strong coding and scripting skills in languages like Python, Go, or Ruby, with the ability to build automation tools for system management. - Deep knowledge of cloud platforms (AWS and/or GCP) and their services, with experience designing and operating cloud-based infrastructure at scale. - Solid understanding of networking and security fundamentals in cloud and on-prem environments. - Experience setting up and tuning monitoring/alerting systems (Prometheus, Grafana, etc.), and a thorough understanding of SRE best practices (SLIs, SLOs, incident management). - Strong problem-solving and communication skills, with a track record of working effectively in collaborative team environments. Requirements - Preferred (Not Essential): Kafka - MySQL/Postgres - Redis - Elasticsearch Benefits - Empower your impact at Cision. Be seen, be understood, be you.
• Design, deploy, and operate enterprise observability platforms. • Build and maintain Splunk Enterprise/Splunk Cloud infrastructure including Indexers, Search Head Clusters, Heavy Forwarders, and Deployment Servers. • Deploy and operate large-scale Elasticsearch clusters for log analytics and search. • Design, deploy, and support distributed tracing platforms using Grafana Tempo and OpenTelemetry. • Build and maintain end-to-end tracing pipelines, instrumentation standards, and trace retention strategies. • Scale Prometheus, Grafana, Kafka, Tempo, and OpenTelemetry-based monitoring solutions. • Develop dashboards, alerts, analytics, and trace visualizations using Splunk SPL, Grafana, Kibana, and Tempo. • Automate infrastructure using Terraform and configuration management tools.
• Design and implement enterprise-scale multi-cloud infrastructure across Microsoft Azure and AWS, with a focus on Azure as the primary platform, ensuring high availability, fault tolerance, and performance for all environments and services. • Develop and deploy Infrastructure as Code (IaC) using Terraform as the primary tool, with proficiency in ARM/Bicep and AWS CloudFormation to standardize cloud provisioning, enforce configuration governance, and enable repeatable, version-controlled deployments. • Build and maintain CI/CD pipelines using GitHub Actions and GitHub as the primary version control system to automate build, test, and deployment processes aligned with product delivery timelines. • Establish comprehensive monitoring, observability, and incident response capabilities using Azure Monitor, CloudWatch, Datadog, and Prometheus to uphold enterprise SLAs, detect anomalies, and drive continuous reliability improvement. • Manage Azure infrastructure fundamentals including subscriptions, tenants, VNets, VWAN, Network Security Groups, and firewalls to establish secure, scalable, and compliant cloud foundations. • Drive cloud cost optimization through FinOps practices—leveraging Azure Cost Management and AWS Cost Explorer to right-size resources, implement governance frameworks, and eliminate waste across the enterprise. • Enforce security and compliance standards across cloud environments in partnership with the CISO, implementing IAM controls, network segmentation, encryption, and zero-trust architecture aligned with SOC 2, FedRAMP, and CMMC requirements. • Serve as MBI's primary technical liaison with Microsoft Azure and AWS—managing vendor partnerships, coordinating architecture reviews, and aligning the cloud roadmap with provider capabilities and strategic initiatives. • Write automation scripts and tools in Python, Bash, or PowerShell to streamline infrastructure deployment, system integration, and DevOps workflows.
2,184more opportunities are still waiting for you.Log in now and take your next shot before someone else does.
Cloud, Python, Docker, Google Cloud Platform, Kubernetes, Terraform