Job Closed
This listing is no longer active.
Empowering IT at Every Endpoint. We're hiring!
Senior DevOps Engineer
Location
United States
Posted
115 days ago
Salary
$130K - $160K / year
Seniority
Senior
Job Description
Senior DevOps Engineer
Recast Software
• Maintain up-time requirements and SLA’s for our cloud infrastructure. • Plan and build reusable Infrastructure as Code to maintain standardization and scalability efforts. • Work within a global DevOps team to develop best practices in our cloud ecosystem. • Drive Change throughout our DevOps process and infrastructure management. • Work with developers to create/refine our CI/CD pipelines. • Participate in an on-call rotation acting as the infrastructure SME during incidents.
Job Requirements
- 3+ years of working in a Highly Available and Scalable cloud environment (Azure preferred, AWS, GCP).
- 2+ years of CI/CD pipeline experience (e.g. ADO, Jenkins, or similar technology).
- 2+ years of experience with cloud monitoring solutions and automation.
- 3+ years of previous experience in software or systems engineering.
- Preferred understanding of coding languages and/or scripting (e.g. Powershell, Python, C++, etc).
- IaC experience is a plus, Azure Bicep a double plus.
- Experience in containerization technology such as Kubernetes.
- 2 years of SQL Database (hosted or cloud) management.
- Experience with Cloud Networking (load balancing, WAF, FW, etc).
- Experience with AI tooling (Microsoft Foundry, Github Copilot, etc).
Benefits
- Medical, dental, and vision
- FSA or HSA with company contributions
- Employer paid STD, LTD, AD&D and life insurance
- 401k with 4% employer match
- Work-life balance, flexible time off, and remote work options
- Parental leave
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
• Software Engineering: Python, Django / Flask, SQL DB - Postgres or MySQL • DevOps: Docker, Kubernetes, Kafka pipelines
Software Engineer – Backend, DevOps, Frontend
Our Team SynergyThe UK's most accessible team effectiveness measurement and reporting solution.
• Build an application that brings the best features and practices of software developers to the world of knowledge workers • Backend: Java and Spring Boot, Node.js. NoSQL databases. Microservices architecture • Frontend: TypeScript, React, React-native, Electron. Experience developing Windows native applications is a plus • DevOps: Kubernetes, object storage, and automations. All public clouds (Azure, Google Cloud, AWS) and on-premise (MinIO)
• Collaborate with other Engineering teams to support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews. • Innovate relentlessly: Identify pain points, propose creative solutions, and drive initiatives that simplify, scale, and strengthen the platform. • Maintain services once they are live by measuring and monitoring availability, latency and overall system health. • Own observability: Enhance and expand monitoring and alerting using Datadog; define SLOs/SLIs and create actionable dashboards that drive reliability outcomes. • Drive automation: Develop and improve internal tooling, IaC frameworks, and pipelines (Terraform, GitLab CI/CD) to reduce manual intervention and enable self-healing systems. • Scale systems sustainably through automation and evolve systems by pushing for changes that improve reliability and velocity. • Act as an agent orchestrator using Amazon Kiro: run multiple activities in parallel by leveraging AI agents to accelerate execution, while personally validating results and completing selected tasks manually when needed. • Be on-call. • Practice sustainable incident response and blameless postmortems. Lead post-incident reviews (RCAs) and identify long-term fixes that improve stability, reliability, and developer experience. • Implement monitoring, Logging, alerting, and SLA Reporting. • Create and maintain technical documentation. • Implement, maintain and mature SRE best practices. • Lead incidents: Act as Incident Commander for Incidents; coordinate cross-team response, manage communications, and ensure rapid service restoration. • Provide support for our planning and deployment teams to enable stability, predictability, and scale in our continued growth. • Collaborate with members of the Platform Engineering team to implement and support far-reaching strategic efforts, provide constructive feedback, and foster a collaborative environment. • Work cross-functionally with internal teams and vendors to manage our growth around the globe, with a strong focus on maintaining the high level of performance, availability, and reliability for our users.
Principal Site Reliability Engineer - Observability
ElasticSelf-described as the leading platform for search-powered solutions, Elastic helps organizations, their customers, and their employees find what they need faste
Elastic, the Search AI Company, enables everyone to find the answers they need in real time, using all their data, at scale — unleashing the potential of businesses and people. The Elastic Search AI Platform, used by more than 50% of the Fortune 500, brings together the precision of search and the intelligence of AI to enable everyone to accelerate the results that matter. By taking advantage of all structured and unstructured data — securing and protecting private information more effectively — Elastic’s complete, cloud-based solutions for search, security, and observability help organizations deliver on the promise of AI. What Is The Role : We're looking for a Principal Site Reliability Engineer to join the Observability Solution as part of the team building the next generation of Infrastructure Observability experiences leveraging the new Search AI and agentic capabilities. What You Will Be Doing : - Collaborate with product management, product design, customers and multiple teams across Elastic (especially our own SRE teams) in defining and evolving the end-to-end InfraObs experiences that enable both human and agentic users. - Deliver and continually evolve the experiences leveraging the Elastic Platform capabilities and coding agents. - Be a contact point for other teams within Elastic. Examples include helping Support with difficult cases or aligning with the teams providing the foundations for developing integrations or consulting the Elastic Stack engineers with designing new features. - Foster a culture of mutual respect, collaboration and consensus based decision-making. - Be an awesome person to work with, somebody who sincerely empathizes with others. What You Bring : This is a role for practitioners so we are looking for engineers with a SRE background and experience operating large-scale production services with the help of Observability tools. - Proficiency operating production infrastructure in K8s and at least one of the three major CSPs. - Proficiency using Observability tools. - Working with a high level of autonomy, able to tackle projects and guide them from beginning to end. This covers both technical design and working with other engineers to develop needed components. - Ability to use AI coding agents in the delivery workflow. - Excellent verbal and written communication skills. Collaborating on the internet is hard. We try to be supportive, empathetic, and trusting in all of our interactions. And we expect that from everyone too. Bonus Points : - Experience as a user of the Elastic Stack. Additional Information - We Take Care of Our People As a distributed company, diversity drives our identity. Whether you’re looking to launch a new career or grow an existing one, Elastic is the type of company where you can balance great work with great life. Your age is only a number. It doesn’t matter if you’re just out of college or your children are; we need you for what you can do. We strive to have parity of benefits across regions and while regulations differ from place to place, we believe taking care of our people is the right thing to do. - Competitive pay based on the work you do here and not your previous salary - Health coverage for you and your family in many locations - Ability to craft your calendar with flexible locations and schedules for many roles - Generous number of vacation days each year - Increase your impact - We match up to $2000 (or local currency equivalent) for financial donations and service - Up to 40 hours each year to use toward volunteer projects you love - Embracing parenthood with minimum of 16 weeks of parental leave Different people approach problems differently. We need that. Elastic is an equal opportunity employer and is committed to creating an inclusive culture that celebrates different perspectives, experiences, and backgrounds. Qualified applicants will receive consideration for employment without regard to race, ethnicity, color, religion, sex, pregnancy, sexual orientation, gender perception or identity, national origin, age, marital status, protected veteran status, disability status, or any other basis protected by federal, state or local law, ordinance or regulation. We welcome individuals with disabilities and strive to create an accessible and inclusive experience for all individuals. To request an accommodation during the application or the recruiting process, please email candidate_accessibility@elastic.co. We will reply to your request within 24 business hours of submission. Applicants have rights under Federal Employment Laws, view posters linked below: Family and Medical Leave Act (FMLA) Poster; Pay Transparency Nondiscrimination Provision Poster; Employee Polygraph Protection Act (EPPA) Poster and Know Your Rights (Poster) Elasticsearch develops and distributes technology and information that is subject to U.S. and other countries’ export controls and licensing requirements for individuals who are located in or are nationals of the following sanctioned countries and regions: Belarus, Cuba, Iran, North Korea, Syria, or Russia, including the Ukrainian territories annexed by Russia (The Crimea region of Ukraine, The Donetsk People's Republic (DNR), The Luhansk People's Republic (LNR), Kherson or Zaporizhzhia). If you are located in or are a national of one of the listed countries or regions, an export license may be required as a condition of your employment in this role. Please note that national origin and/or nationality do not affect eligibility for employment with Elastic. Please see here for our Privacy Statement.




