DevOps Engineer Remote Jobs in Georgia (US)
This page tracks remote devops engineer openings that are location-eligible for Georgia.
This page tracks remote devops engineer openings that are location-eligible for Georgia.
Open jobs
2,195
Hiring companies this week
8
Salary sample
$105,700 - $175,000
Jobs added last hour
0
2195 Jobs
1369 Companies
On a mission to create a world of software delivered without friction from developer to device.
• Work with JFrog customers to design, setup and build CI/CD pipelines and DevOps platform using JFrog products and other cutting edge technologies and practices like Docker, Kubernetes, IaC and Cloud Native • Team up with Sales, Customer Success, Support and Development while working directly with Devs and DevOps Pros to ensure success with the customer’s DevOps journey using the JFrog platform • Setup, design and build CI pipelines with Docker, NPM, Java, Pypi etc.. (Binary Repositories, Distribution to Devices and Continuous Build Tooling) • Train the open source community and JFrog customers • Influence the features and roadmap of JFrog tools based on customer needs • Keep current with the latest technology trends related to DevOps and the landscape of CI/CD Technology • Work closely with our customers and community to build solid relationships
On a mission to create a world of software delivered without friction from developer to device.
Role Description At JFrog, we’re reinventing DevOps to help the world’s greatest companies innovate. If you love working with brilliant people and being part of an energetic team, you might be the perfect Frog to join our Swamp! Come and join the Professional Services team and help us to continue to lead the rapidly evolving space of DevOps and DevSecOps! As a Professional Services DevOps Senior Engineer in JFrog you will: - Work with JFrog customers to design, setup and build CI/CD pipelines and DevOps platform using JFrog products and other cutting edge technologies and practices like Docker, Kubernetes, IaC and Cloud Native. - Team up with Sales, Customer Success, Support and Development while working directly with Devs and DevOps Pros to ensure success with the customer’s DevOps journey using the JFrog platform. - Setup, design and build CI pipelines with Docker, NPM, Java, Pypi etc. (Binary Repositories, Distribution to Devices and Continuous Build Tooling). - Train the open source community and JFrog customers. - Influence the features and roadmap of JFrog tools based on customer needs. - Keep current with the latest technology trends related to DevOps and the landscape of CI/CD Technology. - Work closely with our customers and community to build solid relationships. Qualifications - 5-7+ years experience with Continuous Integration tools: CI Server, Git, Artifactory, Jenkins, Maven, Docker, NPM. - Ability to build software delivery pipelines with Docker, npm, Java, Pypi etc. with various DevOps tools such as Git, Binary repositories management, Binary scanning, and Continuous integration. - Good understanding of infrastructure & operations - storage, network, computer, security, cloud (public, on-prem). - Experience with Continuous Deployment and Delivery tools: Chef, Puppet, Ansible, Kubernetes. - Hands on experience in Linux - Mandatory. - Hands on experience with cloud infrastructure - AWS / Azure / GCP - Mandatory. - Experience with customer facing with great interpersonal and customer service oriented skills. - Experience with server side software on-premise and in the cloud. - Experience with Software Architecture design and product development a plus. - Uncompromising will to learn. - Open Source state-of-mind. Benefits - Open to remote work for candidates outside a reasonable commuting distance to the Atlanta office. - Base salary range between $135,000 to $145,000, based on skills, qualifications, experience and location. - Eligible for discretionary bonuses or commission payments. - Equity package of restricted stock units (RSU). - Participation in Employee Stock Purchase Plan. - Comprehensive benefits including medical, dental, vision, retirement, wellness and much more! Company Description JFrog is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, creed, religion, sex, sexual orientation, national origin or nationality, ancestry, age, disability, gender identity or expression, marital status or any other category protected by law.
• Own the reliability and stability of production data pipelines and data platform services • Diagnose and resolve data pipeline failures, delays, and data quality issues in production environments • Investigate issues across distributed data systems (e.g., Spark/EMR workloads, ingestion pipelines, warehouse performance) • Lead or support incident response, including triage, mitigation, and long-term resolution • Perform root cause analysis (RCA) and implement durable fixes to prevent recurrence • Define and improve data SLAs (freshness, latency, completeness) and ensure adherence • Design and enhance monitoring, alerting, and observability for data systems • Develop automation and tooling to reduce operational toil and improve system resilience • Contribute to disaster recovery (DR) and resiliency planning, including backup validation and recovery workflows • Partner with engineering teams to improve pipeline design, reliability, and operational readiness • Create and maintain runbooks, SOPs, and operational documentation • Participate in occasional off-hours support for production data systems when required
About itD: We are part of a new generation of consulting and software development company that blends diversity, innovation, and integrity with real business results. Our structure rejects any strong hierarchy, empowering us to deliver excellent results. We are a woman- and minority-led firm. Every day, we challenge ourselves to be considerate, fair and to re-think what great outcomes mean for our customers. This permeates down to how we approach every interaction, on every project, for every client. You’ll thrive here if you are a dynamic self-starter, a difference-maker or someone who wants to deliver great results, without constraints. The itD Digital Experience: Joining us means you’ll be part of our global community, you have a say about your own career journey, and you’ll get a chance to give back to causes that matter. You will experience working with Fortune 500 companies and high-performance teams across numerous industries. itD offers our employees excellent benefits such as medical, dental, vision, life insurance, paid holidays, 401K + matching.
Role Description itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering, infrastructure automation, and cloud operations, with a proven track record of delivering reliable automation, scalable deployment pipelines, and resilient production environments. Location: 100% Remote within the United States. Duration: 24 months Responsibilities - Develop and maintain infrastructure automation solutions using Ansible to improve the reliability, scalability, and operational efficiency of cloud environments. - Design, implement, and enhance CI/CD pipelines, testing frameworks, and operational tooling to support infrastructure growth and software delivery. - Troubleshoot complex Linux-based infrastructure and distributed systems issues to maintain high availability and platform performance. - Build automation that enables the rapid, repeatable deployment of regional, sovereign, and purpose-built cloud environments. - Collaborate with engineering teams, product management, and cross-functional stakeholders to identify opportunities for operational improvements and increased platform reliability. - Monitor infrastructure performance and implement enhancements that reduce operational overhead and improve system scalability. - Contribute to automation and engineering best practices that support reliable, efficient cloud platform operations. Internal Responsibilities - Attend regular internal practice community meetings. - Collaborate with your itD practice team on industry thought leadership. - Complete client case studies and learning material (blogs, media material). - Build out material to contribute to the Digital Transformation practice. - Attend internal itD networking events (in person and virtual). - Work with leadership on career fast-track opportunities. Qualifications - 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments. - Experience developing and maintaining infrastructure automation using Ansible. - Experience programming in Ruby and developing automated tests using RSpec or comparable testing frameworks. - Experience administering and troubleshooting Linux-based systems and distributed infrastructure environments. - Experience designing, implementing, and maintaining CI/CD pipelines, including GitLab CI. - Experience supporting large-scale infrastructure environments consisting of hundreds or thousands of systems. - Must be eligible to work on FedRAMP projects. - Must be a U.S. citizen working from U.S. soil. Preferred Qualifications and Skills - Experience with AWS or other public cloud platforms and hybrid infrastructure environments. - Knowledge of monitoring, observability, and site reliability engineering practices and tools. - Familiarity with Kubernetes concepts and containerized application platforms. - Experience leveraging AI-assisted development tools to improve software development, infrastructure automation, operational analysis, and engineering productivity. Education - Bachelor's degree in a relevant field or equivalent work experience required. Benefits - Comprehensive medical benefits. - 401(k) plan. - Paid holidays. - Networking & career learning and development programs. Company Description About itD: We are part of a new generation of consulting and software development company that blends diversity, innovation, and integrity with real business results. Our structure rejects any strong hierarchy, empowering us to deliver excellent results. We are a woman- and minority-led firm. Every day, we challenge ourselves to be considerate, fair and to re-think what great outcomes mean for our customers. The itD Digital Experience: Joining us means you’ll be part of our global community, you have a say about your own career journey, and you’ll get a chance to give back to causes that matter. You will experience working with Fortune 500 companies and high-performance teams across numerous industries.
• Own reliability targets across our backend/API, worker services, applications and CV pipeline; MTD, MTM, MTR, and follow-through on root causes. • Level up our monitoring and alerting, and build out auto-remediation, so on-call load scales with automation, not headcount. • Partner with our agentic engineering work to build agents that triage alerts and handle routine remediation. • Harden and optimize our GCP infrastructure (Cloud Run, Cloud SQL, GCS) for cost and performance as load scales. • Own database scale and performance; connection pooling, query optimization and indexing, read replicas, and capacity planning, so Postgres doesn't become the bottleneck as data volume grows. • Improve the reliability of our ML training and monitoring infrastructure, in partnership with the CV/ML team. • Run blameless postmortems and drive fixes for root causes, not just symptoms. • Participate in on-call rotation.
We provide cross-platform linking and attribution solutions to the world's leading digital brands.
• Partner with Developers to produce high-performing and robust services through rigorous testing and release procedures • Design infrastructure, monitoring, processes, and standards for systems and applications • Support services through design, development, load testing, and launch phases • Develop, measure, and monitor key performance and service level indicators including availability, latency, and overall system health • Define and establish SLIs, SLOs, and error budgets with service owners, and drive adoption across platform teams • Profile and optimize platform performance, resilience, and efficiency, including latency, throughput, and capacity planning under load • Participate in incident response and root cause analysis • Remediate tasks and develop preventative and automated measures to meet SLAs/SLOs/SLIs • Manage monitoring services utilized by applications
We provide cross-platform linking and attribution solutions to the world's leading digital brands.
• Lead, and take ownership of our core database systems to ensure high levels of performance, availability, and security • Communicate critical data changes to the data warehousing team • Work alongside both engineering teams and data teams to shepherd schema changes from design to deployment to ensure architectural and downstream compatibility • Build and maintain automated test and deployment data pipelines, including production level data testing at scale, with SRE and DevOps teams • Proactively work on detecting weak areas and incorporate into automated monitoring and production alerting with our Cloud Ops teams • Document and automate processes to eliminate time spent on recurring tasks • Evaluate new database features (and products if needed) to meet scalability and global uptime requirements • Own database backup and recovery, and help establish disaster recovery targets and practices as the systems scale • Support, debug, and perform RCA on production database issues, and participate in incident response as needed
• Maintain our CI/CD pipelines and containerized infrastructure (Docker, ECS/ECR Fargate, Terraform), keeping deployments fast and reliable day to day • Monitor cloud infrastructure on AWS, catching and resolving issues before they affect customers • Support incident response, helping triage and resolve production issues as they come up • Execute steady-state maintenance and operational work, freeing up senior team members to focus on architecture and security priorities • Ramp into our existing systems and tooling, building depth over time as you take on more ownership
Accelerating speed to care by optimizing provider schedules, streamlining clinical communication, and engaging patients.
• Help configure and rollout customers of our existing AI voice agent • Build and run evaluation suites to validate behavior, catch regressions, and tune prompts as new customers are deployed • Develop new features and fixes to the core product • Optimize latency, inference cost, and hallucination guardrails • Run and improve model evaluation, benchmarking, and observability across the product • Monitor cost controls, usage monitoring, and security/PHI handling constraints • Partner with Product and Customer Success to turn customer requirements into working, validated deployments, and own customer go-live, post-launch monitoring, and production issue debugging • Help streamline and automate the onboarding process to reduce the engineering effort required to bring each new customer live
• Design and operate scalable cloud infrastructure across AWS and GCP. • Build and improve Kubernetes, Linux, and cloud networking environments. • Standardize infrastructure using Terraform, Terragrunt, Helm, and GitOps practices. • Improve CI/CD pipelines with GitHub Actions and ArgoCD. • Build monitoring, logging, and alerting using Prometheus, Grafana, Loki, CloudWatch, and Better Stack. • Strengthen security, disaster recovery, and platform resilience. • Drive platform strategy while remaining hands-on with implementation. • Improve developer productivity through automation and self-service infrastructure.
2,185more opportunities are still waiting for you.Log in now and take your next shot before someone else does.
Cloud, AWS, Docker, Google Cloud Platform, Kubernetes, Python