Job Closed

This listing is no longer active.

Revelation Pharma LLC logo
Revelation Pharma LLC

Revelation Pharma is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.

DevOps / AgentOps Engineer

Location

United States

Posted

74 days ago

Salary

$115K - $125K / year

Seniority

Mid Level

No structured requirement data.

Job Description

DevOps / AgentOps Engineer

Revelation Pharma LLC

Role Description This is a hybrid infrastructure and AI operations role — part traditional DevOps, part emerging AgentOps discipline. The DevOps/AgentOps Engineer owns the CI/CD pipelines, infrastructure-as-code, observability stack, and deployment automation for both the application platform and the AI agent layer. They work shoulder-to-shoulder with the Full-Stack and Backend engineers, guide and learn from KMS's build process, and own the shared services infrastructure on the Revelation Pharma side. This is a high-autonomy role for someone who wants to define how a modern health-AI platform is built, deployed, and monitored from day one. Qualifications - AWS Infrastructure (Expert): Deep hands-on experience with IAM policy design, VPC networking, KMS encryption, CloudWatch/CloudTrail, and at least 3 of: HealthLake, Bedrock, Lambda, Step Functions, ECS/Fargate, S3, Clean Rooms. - Infrastructure-as-Code (Expert): Production experience with Terraform or AWS CDK. - CI/CD & Deployment (Expert): End-to-end pipeline design (GitHub Actions, CodePipeline, or equivalent). - Observability & Monitoring (Strong): Production experience building observability stacks. - Security & Compliance (Strong): HIPAA technical safeguard implementation experience. - Agent/AI Operations (Developing to Strong): Experience deploying and monitoring LLM-based systems in production. Requirements - 5+ years in DevOps, SRE, or platform engineering roles with at least 2 years at a staff or senior level. - Production infrastructure experience in healthcare, fintech, or another regulated industry with formal compliance requirements. - Experience working in blended team models (internal + outsourced). - Demonstrated ability to work autonomously in early-stage environments. Benefits - Health care insurance (medical, dental, vision) - Life Insurance - Supplemental Insurance - PTO - 401K matching - Sick leave - Phone/internet reimbursement - Remote work - Top of the line machines

Related Categories

Related Job Pages

More DevOps Engineer Jobs

AlphaSense India logo

Staff Site Reliability Engineer

AlphaSense India

AlphaSense is an equal-opportunity employer. We are committed to a work environment that supports, inspires, and respects all individuals. All employees share in the responsibility for fulfilling AlphaSense’s commitment to equal employment opportunity. AlphaSense does not discriminate against any employee or applicant on the basis of race, color, sex (including pregnancy), national origin, age, religion, marital status, sexual orientation, gender identity, gender expression, military or veteran status, disability, or any other non-merit factor. This policy applies to every aspect of employment at AlphaSense, including recruitment, hiring, training, advancement, and termination. In addition, it is the policy of AlphaSense to provide reasonable accommodation to qualified employees who have protected disabilities to the extent required by applicable laws, regulations, and ordinances where a particular employee works.

DevOps Engineer74 days ago
Full TimeRemoteTeam 1,001-5,000

Role Description Our Site Reliability Engineering team is growing, and we are looking for a highly experienced Staff Site Reliability Engineer to help shape the future of reliability, scalability, and performance at AlphaSense. This is a hands-on, high-impact role where you will: - Architect core reliability platforms. - Lead by example in incident response. - Drive cultural adoption of SRE best practices across our global engineering organization. Our mission is to engineer our platform to the reliability standards of mission-critical systems, targeting 99.99% uptime, while continuously enhancing our systems and processes. This role is key to that mission and goes beyond traditional system maintenance; it’s about pioneering the platforms, practices, and culture that enable engineering to scale effectively. You will act as a force multiplier, mentoring fellow engineers, influencing architectural decisions, and setting the technical bar for reliability across the company. Qualifications - 8+ years of experience in Site Reliability Engineering, DevOps, or a similar role, with at least 3+ of those years operating in a Senior+ SRE position. - Strong background in running production SaaS systems at scale. - Proficiency in at least one programming/scripting language (Python, Go, or similar). - Hands-on expertise with cloud platforms (AWS, GCP, or Azure) and Kubernetes. - Deep understanding of networking fundamentals (TCP/IP, DNS, HTTP/S, load balancing). - Experience with monitoring & alerting (Prometheus, Grafana, Datadog, ELK). - Familiarity with advanced observability (OTEL, continuous profiling). - Proven incident management experience, including leading high-severity incidents and postmortems. - Strong troubleshooting skills across the full stack. - Excellent communication and collaboration skills. Requirements - Architect Reliability Paved Paths: Build frameworks and self-service tooling that let teams own the reliability of their services in a “You Build It, You Run It” culture. - Lead AI-Driven Reliability: Drive our AIOps strategy — automating diagnostics, remediation, and proactive failure prevention. - Champion Reliability Culture: Embed SRE practices across engineering via design reviews, production readiness, and operational standards. - Incident Leadership: Act as Incident Commander during critical events, modeling operational excellence, and ensuring blameless postmortems lead to lasting improvements. - Advance Observability: Deliver end-to-end monitoring, tracing, and profiling (Prometheus, Grafana, OTEL, Continuous Profiling) to optimize performance proactively. - Mentor & Multiply: Elevate engineers across SRE and product teams through mentorship, technical guidance, and knowledge sharing. Company Description AlphaSense is an equal-opportunity employer. We are committed to a work environment that supports, inspires, and respects all individuals. All employees share in the responsibility for fulfilling AlphaSense’s commitment to equal employment opportunity. AlphaSense does not discriminate against any employee or applicant on the basis of race, color, sex (including pregnancy), national origin, age, religion, marital status, sexual orientation, gender identity, gender expression, military or veteran status, disability, or any other non-merit factor. This policy applies to every aspect of employment at AlphaSense, including recruitment, hiring, training, advancement, and termination. In addition, it is the policy of AlphaSense to provide reasonable accommodation to qualified employees who have protected disabilities to the extent required by applicable laws, regulations, and ordinances where a particular employee works. We at AlphaSense have been made aware of fraudulent job postings and individuals impersonating AlphaSense recruiters. These scams may involve fake job offers, requests for sensitive personal information, or demands for payment. Please note: - AlphaSense never asks candidates to pay for job applications, equipment, or training. - All official communications will come from an @alpha-sense.com email address. - If you’re unsure about a job posting or recruiter, verify it on our Careers page. - If you believe you’ve been targeted by a scam or have any doubts regarding the authenticity of any job listing purportedly from or on behalf of AlphaSense please contact us. Your security and trust matter to us.

India
voize logo

DevOps Engineer

voize

Doku einfach einsprechen.

DevOps Engineer74 days ago
Full TimeRemoteTeam 11-50Since 2020H1B No Sponsor

• Own and operate our Kubernetes clusters (AWS EKS production, bare-metal K3s for ML training, on-premises appliances running K3s) using GitOps (FluxCD) and Infrastructure as Code (CloudFormation, Ansible) • Manage the lifecycle of on-premises gateway appliances deployed at customer sites — VM image builds, TLS certificate automation, staged rollouts via GitOps, remote monitoring and troubleshooting • Build and maintain monitoring, alerting, and observability infrastructure (Prometheus, Grafana, Loki, Tempo, OpenTelemetry) across cloud and edge environments • Drive compliance automation — security hardening, access controls, audit logging, encrypted secrets management (SOPS/KMS), and evidence collection for C5, HIPAA, and HDS certifications • Support and scale ML training and data processing infrastructure — GPU cluster management, training job orchestration, and data pipeline reliability, working closely with the ML team to power state-of-the-art speech and language models • Incident response — detection, triage, resolution, and post-incident reviews for infrastructure issues.

Germany
Smartsheet logo

Senior Business DevOps Engineer

Smartsheet

Modern work management platform

DevOps Engineer74 days ago
Full TimeRemoteTeam 1,001-5,000Since 2005H1B Sponsor

Role Description Smartsheet is seeking a Senior Business DevOps Engineer to join our Corporate Systems Development team in Bangalore. This role will focus on building and scaling our CI/CD pipelines, infrastructure automation, monitoring frameworks, and deployment processes supporting mission-critical integrations across Finance, People, Sales, Legal, IT, and Engineering systems. You’ll work across a variety of systems and platforms (AWS, GitLab, DataDog, Terraform, Boomi, UiPath) to streamline deployment of backend integrations and automation solutions. If you thrive on optimizing developer velocity, ensuring system reliability, and automating everything from build to deploy, this role is for you. The position reports to the Senior Manager, Systems Development and collaborates closely with global developers, architects, and application administrators to ensure our platform foundations are secure, efficient, and scalable. You Will: - Design, implement and maintain CI/CD pipelines using GitLab and AWS Code Pipeline for integration, automation, and infrastructure projects. - Develop and manage infrastructure-as-code (Terraform, TerraGrunt, CloudFormation) for consistent, versioned deployment environments. - Build observability and monitoring frameworks using DataDog, Cloudwatch and custom dashboards for integration health. - Automate environment provisioning and configuration management. - Review and remediate any vulnerabilities reported by SecEng monitoring tools like WIZ within the agreed upon SLA, ensuring compliance. - Proactively monitor for code obsolescence and upgrade any AWS code bases. - Collaborate with developers to containerize workloads and streamline Lambda deployment packaging and dependency management. - Enhance system reliability and deployment speed through automation of build/test/release workflows. - Support developers in managing staging and production release processes, including rollback plans and change validation. - Partner with architects to define DevOps best practices for security, scaling, and fault tolerance across integration services. - Participate in a production support rotation, triaging environment-level issues and improving alerting and recovery automation. Qualifications - BS in Computer Science, related field, or equivalent industry experience. - 7+ years of experience in DevOps or System Engineering roles. - Proven expertise in AWS services (Lambda, Step Functions, EC2, SQS/SNS, S3). - Hands-on experience with Terraform, TerraGrunt, GitLab CI/CD, AWS CDK and Infrastructure as Code principles. - Solid understanding of networking, security groups, VPC, and API gateway configurations. - Experience with containerization (Docker) and orchestration. - Familiar with Python and/or Javascript/Typescript scripting for automation. - Strong experience implementing monitoring, logging and alerting systems. - Understanding of deployment strategies (blue/green, canary, rolling). - A strong analytical mindset to identify complex issues, troubleshoot effectively, and propose innovative solutions. - The capacity to share knowledge, coach team members, and foster a learning environment. - Excellent communication and collaboration skills with global, distributed teams. - Proven ability to troubleshoot production systems and drive root-cause resolution. - The flexibility to adjust to changing priorities, new technologies, and evolving project requirements. - Comfortable operating in an Agile environment, and familiar with Sprint ceremonies like Daily stand-up, Sprint planning, Sprint Finalization etc. Benefits - At Smartsheet, your ideas are heard, your potential is supported, and your contributions have real impact. - You’ll have the freedom to explore, push boundaries, and grow beyond your role. - We welcome diverse perspectives and nontraditional paths—because we know that impact comes from individuals who care deeply and challenge thoughtfully. - When you’re doing work that stretches you, excites you, and connects you to something bigger, that’s magic at work. Equal Opportunity Employer Smartsheet is an Equal Opportunity (EEO) employer committed to fostering an inclusive environment with the best employees. It is our policy to provide equal employment opportunities to all qualified applicants in accordance with applicable laws in the US, UK, Australia, Germany, Costa Rica, Japan, Bulgaria, and India. All qualified applicants will receive consideration without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, protected veteran or disabled status, or genetic information. If there are preparations we can make to help ensure you have a comfortable and positive interview experience, please let us know.

KA + 1 moreAll locations: KA | India
Job Closed
Talentgrator logo

DevOps Engineer

Talentgrator

An international company operating in the iGaming industry, focused on building scalable operational processes and supporting business growth across multiple markets. The company works with high-volume financial flows, payment infrastructure, and partner operations, ensuring stability, security, and efficiency across all internal processes. With a strong focus on risk control, fraud prevention, and operational optimization, the team continuously improves internal systems and business processes.

DevOps Engineer74 days ago

Role Description We are looking for an experienced Senior DevOps Engineer. The ideal candidate will have a strong grasp of DevOps principles, technical expertise, and a focus on improving system performance and reliability. - Availability & Performance: Administer servers and virtual machines running Debian, configure load balancers, implement DDoS protection, and work with virtualization platforms (Proxmox VE). - Networking & Infrastructure: Configure networks, manage DNS, optimize routing and load balancing, and ensure high availability. - Automation & IaC: Develop and maintain infrastructure using Ansible, AWX, and Terraform. - Monitoring & Incident Management: Set up monitoring and alerting systems (Zabbix, Grafana, Victoria Metrics, Opsgenie), analyze logs, and troubleshoot system issues. - Containerization & Orchestration: Work with Docker and Kubernetes for application management, develop Helm charts, manage clusters, and integrate with CI/CD pipelines. - Data Systems: Configure and optimize HA solutions for Redis, Kafka, RabbitMQ, PostgreSQL, MongoDB, and ClickHouse. - Process Automation: Develop automation scripts using Python or Golang. - Communication & DevOps Culture: Collaborate with development and QA teams, promote DevOps best practices. - Continuous Improvement: Implement and maintain CI/CD pipelines, identify opportunities for process optimization and automation. Qualifications - 5+ years of experience in a similar role. - Strong knowledge of server hardware and experience with bare-metal hosting. - Excellent troubleshooting skills for production servers running Debian/Ubuntu. - Hands-on experience with virtualization: QEMU KVM, libvirtd, Proxmox VE. - Solid understanding of networking: TCP/IP, DNS, L2/L3 networks, routing, IPTables, VPN. - Experience configuring and managing Nginx. - Proficiency with Docker and Kubernetes, including Helm chart development. - Strong experience with IaC tools: Ansible, Terraform, etc. - CI/CD experience with GitLab CI, ArgoCD, Flux. - Experience working with PostgreSQL, Elasticsearch, Redis, MongoDB, ClickHouse. - Practical experience with Kafka and RabbitMQ. - Experience with monitoring tools: Zabbix, Prometheus, Victoria Metrics, Grafana, ELK Stack. Requirements - Experience with AWS or GCP (Nice to Have). - Programming skills in Python or Golang (Nice to Have). Benefits - 20 days of paid vacation, 5 family days and sick leave compensation. - Flexible workday start, allowing you to manage your schedule comfortably. - Support from a professional corporate coach and psychologist. - Regular internal and external activities, workshops, team trips, and corporate events. - Access to our internal knowledge base, meetups, and team-building initiatives. - Continuous learning opportunities: training in new technologies and ongoing support for your professional development.

Czechia