Oracle, headquartered in Austin, Texas, is a global leader in computing solutions. The company specializes in database management systems, cloud-engineered systems, and enterprise
[Remote] Principal Site Reliability Developer- USC Required
Location
United States
Posted
130 days ago
Salary
$86.4K - $199K / year
Seniority
Lead
Job Description
[Remote] Principal Site Reliability Developer- USC Required
Oracle
Are you a creative person who loves a challenge? Solve the complex puzzles you’ve been dreaming of as our Engineer. If you have a passion for innovation in tech, we want you on our team! Thrive in this crucial automation role. Oracle is a technology leader that’s changing how the world does business. We’re looking for an experienced and self-motivated person. We appreciate you taking the time to review the list of qualifications and to apply for the position. Come and join us! Building off our Cloud momentum, Oracle has formed a new organization - Oracle Health. This team will focus on product deployment, sustainability, troubleshooting and product strategy for Oracle Health, while building out a complete platform supporting modernized, automated healthcare. This is a net new line of business, constructed with an entrepreneurial spirit that promotes an energetic and creative environment. We are unencumbered and will need your contribution to make it a world class engineering center with the focus on excellence. As a Site Reliability DevOps Engineer, you will be responsible for defining and deploying key services with deep focus on architecture, production operations, capacity planning, performance management, deployment, and release engineering. You will work with multiple cross-functional teams helping deliver new and outstanding experiences to our collaborators while ensuring reliability and performance. Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States. Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Full Stack DevOps Engineer
GuidehouseGuidehouse, a "next-generation consultancy" and a portfolio company of Veritas Capital, provides management, risk consulting, and technology services to help cl
• Architect, automate, and operate enterprise analytics platforms leveraging Terraform, GitHub, Kubernetes, and GitOps to deliver secure, repeatable, and scalable cloud environments • Design and build next‑generation analytics infrastructure, modernizing legacy deployments and implementing hardened CI/CD workflows aligned with Zero Trust and cloud governance best practices • Develop reusable modules, templates, and reference architectures to accelerate product teams and improve platform standardization • Integrate and operationalize third‑party libraries, APIs, and partner technologies into end‑to‑end cloud solutions • Contribute to engineering excellence through code commits, PR reviews, automation improvements, and platform observability enhancements • Improve containerized workloads on both Azure and AWS and enhance workload efficiency across analytic platforms such as Databricks, Posit Workbench, and Posit Connect • Collaborate across IT and security organizations to ensure compliance, resiliency, and policy alignment • Leverage tools such as GitHub, Jira, and ServiceNow as part of a mature DevSecOps operating model
Senior Site Reliability Engineer - Government Cloud
Ping IdentityIdentity Security for the Global Enterprise
About Ping Identity: At Ping Identity, we believe in making digital experiences both secure and seamless for all users, without compromise. We call this digital freedom. And it's not just something we provide our customers. It's something that inspires our company. People don't come here to join a culture that's built on digital freedom. They come to cultivate it. Our intelligent, cloud identity platform lets people shop, work, bank, and interact wherever and however they want. Without friction. Without fear. While protecting digital identities is at the core of our technology, protecting individual identities is at the core of our culture. We champion every identity. One of our core values, Respect Individuality, reminds us to celebrate differences so you are empowered to bring your authentic self to work. We're headquartered in Denver, Colorado and we have offices and employees around the globe. We serve the largest, most demanding enterprises worldwide, including more than half of the Fortune 100. At Ping Identity, we're changing the way people and businesses think about cybersecurity, digital experiences, and identity and access management. As a Site Reliability Engineer, you will be a key contributor to our mission-critical, cloud-based identity services. You will establish and optimize solutions for building, deploying, and maintaining one of the world’s largest identity platforms within highly regulated environments. Our engineering team follows a robust DevOps model where Development and Operations are fully integrated. You will collaborate daily to ensure our infrastructure is not only scalable and resilient but also meets the rigorous security standards required for FedRAMP and other significant audit cycles. You Will: - Design and Deliver: Work both collaboratively and independently to design production infrastructure that prioritizes resiliency, observability, and cost-efficiency. - Automate Compliance: Shape mission-critical solutions using optimized CI/CD pipelines that integrate automated security controls and compliance checks. - Support Regulated Environments: Help build and maintain infrastructure within FedRAMP authorized boundaries, ensuring continuous compliance with NIST 800-53 controls. - Manage Kubernetes at Scale: Orchestrate containerized workloads to ensure high availability, implementing self-healing patterns and automated scaling. - Collaborate and Grow: Proactively communicate technical concepts to diverse audiences and share your expertise to help the team develop. - Maintain Reliability: Participate in planning, identify areas for improvement, and join an on-call rotation to maintain the health of our cloud solutions. You Have: - Cloud Expertise: Significant experience with Google Cloud Platform (GCP), including familiarity with Assured Workloads for regulated data. - Programming Proficiency: Experience programming in Go (Golang) or similar high-performance development languages. - Orchestration Mastery: Deep experience with Kubernetes and containerization technologies like Docker. - Compliance Experience: Proven track record running cloud environments in compliance with FedRAMP (Moderate or High), including experience with Continuous Monitoring (ConMon) and audit readiness. - Infrastructure as Code: Experience defining automated service deployments with provisions for networking, security, and configuration management. - Distributed Systems: A solid understanding of scalable, distributed systems architecture and team-based version control (Git). - Education: A CS Degree or equivalent professional experience. You Will Have an Advantage If: - Network Deep Dive: You have an in-depth understanding of networking (routing, FIPS-validated encryption, and boundary protection). - Audit Navigation: You have participated in 3PAO audits or have experience resolving customer-facing deployment issues in regulated sectors. - IAM Knowledge: You possess an in-depth understanding of Identity and Access Management (IAM). - Global Collaboration: You have experience working with distributed, global teams. Ping Identity is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status. Salary: $105,000 - $130,000 In accordance with Colorado’s Equal Pay for Equal Work Act (SB 19-085) the approximate compensation range for this role in Colorado is listed above. Final compensation for this role will be determined by various factors, such as knowledge, skills, and abilities. Life at Ping: We believe in and facilitate a flexible, collaborative work environment. We’re growing quickly, but remain true to the innovative, can-do startup values that got us here. Most importantly, we keep hiring talented, smart, fun, and genuinely nice people because that’s who we want to succeed with every day. Here are just a few of the things that make Ping special: - A company culture that empowers you to do your best work. - Employee Resource Groups that create a sense of belonging for everyone. - Regular company and team bonding events. - Competitive benefits and perks. - Global volunteering and community initiatives Our Benefits: - Generous PTO & Holiday Schedule - Parental Leave - Progressive Healthcare Options - Retirement Programs - Opportunity for Education Reimbursement - Commuter Offset (Specific locations) Ping is the collective sum of all our individual experiences, backgrounds and influences and we pride ourselves in growing and learning together. We are committed to building an inclusive and diverse environment where everyone’s individuality is respected and everyone has an Identity. In recruiting for new colleagues, we welcome the unique contributions you can bring and encourage you to be your best self. We are an Equal Opportunity/Affirmative Action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex including sexual orientation and gender identity, national origin, disability, protected Veteran Status, or any other characteristic protected by applicable federal, state, or local law.
Senior Staff Site Reliability Engineer
JobgetherWe use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1 We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Role Description This role offers a unique opportunity to lead the reliability, scalability, and operational excellence of a large-scale, globally distributed platform. You will shape infrastructure strategy, guide deployment automation, and ensure systems perform seamlessly across multi-region environments. The position blends hands-on engineering with high-level leadership, providing influence over both core product architecture and operational best practices. You will mentor senior engineers, drive IaC and observability standards, and partner cross-functionally with Product, Security, and Customer Success teams. This is a highly dynamic, remote-friendly environment where technical vision and proactive problem-solving directly impact customer satisfaction and organizational resilience. - Lead the strategy and execution of infrastructure, deployment, and reliability initiatives across cloud, Kubernetes, and multi-region platforms - Design, build, and maintain scalable, high-availability systems, automation tools, and core frameworks in Go and Python - Serve as a technical leader and mentor for SREs and engineering teams, elevating practices across reliability, monitoring, and operational excellence - Oversee Infrastructure as Code (IaC) implementation and continuous deployment systems to ensure consistency and reproducibility - Act as senior escalation point for high-severity incidents and implement systemic solutions to prevent recurring failures - Define and enforce standards for observability, SLIs/SLOs, disaster recovery, chaos engineering, and production readiness - Collaborate with Engineering leadership, Product, Security, and Customer Success to align infrastructure capabilities with business and customer needs Qualifications - 10+ years of experience in software engineering or site reliability roles with distributed systems - Deep expertise with cloud platforms (AWS, GCP, Azure), Kubernetes, containerized environments, and CI/CD systems - Strong coding skills in Go (preferred) and Python, capable of contributing to core frameworks and platform tooling - Hands-on experience with observability platforms (Prometheus, Grafana, Loki) and multi-region architectures - Experience managing infrastructure, SRE, or platform engineering teams supporting production systems - Advanced understanding of microservices, event-driven architectures, networking, and security best practices - Familiarity with databases such as MongoDB or Redis at scale - Excellent leadership, mentorship, and communication skills with the ability to drive cross-functional initiatives - Bonus: open-source contributions, U.S. security clearance eligibility, or prior experience with large-scale SaaS platforms Benefits - Fully remote work with flexible hours - Flexible paid time off, holidays, and vacation - Company-provided laptop and remote work benefits - Professional development through courses, books, and learning platforms - Stock options and equity participation - Multicultural and inclusive work environment - Vibrant, collaborative company culture Company Description
DevOps & Site Reliability Engineer
Oowlish TechnologyOowlish, one of Latin America's rapidly expanding software development companies, is seeking experienced technology professionals to enhance our diverse and vibrant team. As a valued member of Oowlish, you will collaborate with premier clients from the United States and Europe, contributing to pioneering digital solutions.
Role Description We are seeking a DevOps & Site Reliability Engineer to join a growing AI-focused SaaS startup. In this role, you’ll be responsible for maintaining, optimizing, and scaling the infrastructure that supports our platform, ensuring high availability, performance, and reliability. You’ll work closely with development and product teams to improve deployment processes, monitor systems, and respond to incidents proactively. If you are passionate about DevOps culture, automation, and ensuring systems are always running smoothly, this is the perfect opportunity for you! Key Responsibilities - Deploy and manage web, mobile, and API applications across cloud environments - Implement and maintain monitoring and observability tools like NewRelic, Datadog, or Prometheus/Grafana - Design and optimize CI/CD pipelines using tools like Azure Pipelines, Jenkins, or CircleCI - Manage containerized environments with Docker, Kubernetes, and Helm - Build and manage cloud infrastructure on Azure, AWS, or GCP - Write automation scripts using Bash and other scripting languages - Develop and maintain incident response processes and disaster recovery strategies - Collaborate with development, product, and operations teams to improve system reliability and deployment efficiency Qualifications - 3+ years of experience in a DevOps, Site Reliability Engineering (SRE), or related role - Strong hands-on experience with the deployment of web, mobile, and API applications - Expertise in monitoring and observability tools (e.g., NewRelic, Datadog, Prometheus/Grafana) - Strong experience with CI/CD pipelines and associated tools (Azure Pipelines, Jenkins, CircleCI) - Proficiency with Docker, Kubernetes, and Helm - Experience working with cloud platforms like Azure, AWS, or GCP - Scripting proficiency in Bash - Familiarity with incident response and disaster recovery planning Benefits - Home office - Competitive compensation based on experience - Career plans to allow for extensive growth in the company - International Projects - Oowlish English Program (Technical and Conversational) - Oowlish Fitness with Total Pass - Games and Competitions



