Job Closed

This listing is no longer active.

Konfío logo
Konfío

La plataforma de soluciones financieras para tu empresa.

Site Reliability Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteSeniorTeam 501-1,000Since 2013H1B No SponsorCompany SiteLinkedIn

Location

Mexico

Posted

98 days ago

Salary

0

Seniority

Senior

Bachelor Degree3 yrs expSpanishAWSJavaScriptPythonRustGo

Job Description

Site Reliability Engineer

Konfío

• Garantizar la alta disponibilidad, escalabilidad y rendimiento de los sistemas y aplicaciones, implementando prácticas de ingeniería de confiabilidad y automatización para prevenir y resolver problemas de infraestructura. • Diseñar, implementar y mantener sistemas escalables y fiables, actuando como un enlace entre desarrollo y operaciones, mediante la automatización de tareas operativas y la resolución de problemas tecnológicos. • Realizar el trabajo que tradicionalmente hacían las operaciones, pero utilizando ingenieros con experiencia en software para resolver problemas complejos. • Ejecutar y mejorar el proceso de gestión de incidencias, garantizando el tiempo de actividad de todos los servicios y procesos y tratando exhaustivamente de prevenir las incidencias.

Job Requirements

  • Licenciatura en Ciencias de la Computación, Ingeniería de Software o similar
  • 3+ años de experiencia utilizando herramientas de Observabilidad y Monitorización como: New Relic, Data Dog, CloudWatch, Opsgenie, PagerDuty.
  • Conocimiento profundo del entorno multicuenta de AWS, con estrategia centralizada de observabilidad y monitorización.
  • Capacidad de programación (scripting) utilizando uno o más lenguajes de alto nivel, como Python, Golang, Rust, JavaScript.
  • Experiencia práctica con soluciones de microservicios, incluidos contenedores y cargas de trabajo de funciones.
  • Experiencia en el diseño, la implementación y el mantenimiento de objetivos de nivel de servicio (SLO) para garantizar el tiempo de actividad y el rendimiento del servicio.
  • Comprensión de las técnicas de saneamiento de telemetría y experiencia en la aplicación de estas técnicas de conformidad con los requisitos reglamentarios y de seguridad.

Benefits

  • Un ambiente de trabajo dinámico y colaborativo donde podrás desarrollar tu potencial al máximo.
  • Oportunidades para aprender y crecer profesionalmente utilizando tecnologías de vanguardia.
  • Un equipo apasionado y talentoso con el que podrás compartir conocimientos y experiencias.
  • Paquete de compensación competitivo y beneficios atractivos.
  • La oportunidad de impactar positivamente en la vida de miles de personas y contribuir al desarrollo del país.

Related Categories

Related Job Pages

More DevOps Engineer Jobs

Flock Safety logo

Solution Deployment Engineer

Flock Safety

We are the first public safety operating system empowering over 2500 cities to eliminate crime.

DevOps Engineer98 days ago
Full TimeRemoteTeam 501-1,000Since 2017H1B Sponsor

• Execute technical solution design and deployment for key customer engagements across the U.S. • Serve as a trusted technical advisor, ensuring deployments align with customer needs and product capabilities. • Troubleshoot and resolve complex technical challenges in real time. • Contribute to team knowledge by documenting best practices and sharing lessons learned. • Partner cross-functionally with Sales, Product, and Customer Success to support customer goals. • Maintain deep technical knowledge of Flock Safety’s products and competitive solutions. • Represent the company at customer meetings, field visits, and industry events. • Provide feedback to improve internal processes, tools, and deployment methodologies.

United States
$70K - $94K / year
Job Closed
Arize AI logo

DevOps Release Manager

Arize AI

Arize AI is a machine learning observability platform for ML practitioners to detect and troubleshoot model issues

DevOps Engineer98 days ago
Full TimeRemoteTeam 51-200Since 2019H1B Sponsor

• Design, build, and maintain scalable backend systems that power the deployment of Arize in customer-managed (on-prem and cloud) environments. • Develop tooling and infrastructure to package, test, and deliver the Arize platform as reliable, production-ready self-hosted releases. • Work across the stack using Go, Java, Python, and Bazel to build reproducible builds and deployment pipelines. • Partner with customers to understand infrastructure constraints and translate them into robust deployment architectures. • Build and optimize services that support high-volume analytics workloads in resource-constrained or isolated environments. • Improve system reliability, observability, and upgradeability for distributed deployments.

United States
$150K - $175K / year
Job Closed
PulsePoint logo

Site Reliability Engineer, K8s

PulsePoint

WebMD and its affiliates is an Equal Opportunity/Affirmative Action employer and does not discriminate on the basis of race, ancestry, color, religion, sex, gender, age, marital status, sexual orientation, gender identity, national origin, medical condition, disability, veterans status, or any other basis protected by law.

DevOps Engineer98 days ago
Full TimeRemoteTeam 201-500

WebMD and its affiliates is an Equal Opportunity/Affirmative Action employer and does not discriminate on the basis of race, ancestry, color, religion, sex, gender, age, marital status, sexual orientation, gender identity, national origin, medical condition, disability, veterans status, or any other basis protected by law. Position Overview Our BI team runs a set of GCP-based APIs and data services that a lot of internal products depend on. As we've grown, keeping things running has increasingly been a side responsibility for engineers who are primarily building features — and that's not sustainable. We're looking for an SRE to own that space: service health, incident response, infrastructure monitoring, and making sure we're not blindly burning cloud budget. The Site Reliability Engineer will ensure the availability, performance, and security of the Business Intelligence team's GCP-hosted APIs and data infrastructure. This role is responsible for proactive monitoring, incident response, and continuous improvement of platform reliability across a cloud-native stack. The engineer will work closely with backend and data engineers to maintain service health and drive operational excellence. This position also carries responsibility for GCP cost visibility, helping the team track and optimize cloud spend through structured monitoring and alerting. Responsibilities - Monitor and maintain uptime of GCP-hosted APIs and services, keeping performance within agreed targets - Lead incident response for BI platform services — triage, resolve, and follow up with post-mortems that actually prevent recurrence - Build and manage observability infrastructure: dashboards, alerts, and logging across GCP services - Track GCP cloud spend and set up cost alerting to flag anomalies before they become problems - Review and fix security gaps — IAP configs, service account permissions, API access controls - Work with data and backend engineers to shore up reliability of data pipelines and BigQuery workflows - Contribute to infrastructure-as-code and help keep deployments documented and reproducible Qualifications - 2+ years in a Site Reliability, DevOps, or Cloud Infrastructure role in a production environment - Bachelor's degree in Computer Science, Engineering, or related field, or equivalent hands-on experience - Practical experience with GCP — Cloud Run, API Gateway, and BigQuery in particular - Experience with monitoring and observability tooling (Cloud Monitoring, Datadog, or similar) - Solid grasp of cloud security fundamentals — IAM, network controls, access management - Proficiency with Git and version control in a team setting Please list the preferred skills here: - CI/CD pipelines and deployment automation (GitHub Actions, Cloud Build, or similar) - Terraform or other infrastructure-as-code tools - Python for scripting or automation - MySQL, Spanner, or BigQuery at any meaningful depth - GCP cost management and spend optimization - Experience with dbt or Looker - Comfortable working across CET/EST hours in a distributed team

United Kingdom
RE Partners logo

DevOps Engineer

RE Partners

We make the Aspirational Attainable. We Do Better Together to Deliver Real Change.

DevOps Engineer98 days ago
Full TimeRemoteTeam 201-500Since 2019H1B No Sponsor

DevOps Engineer We are looking for a DevOps Engineer with basic Golang skills to support a team focused on improving development processes and the product. In this role, you will work closely with a Staff Engineer building with a focus on designing, migrating, implementing, and maintaining the infrastructure and cloud environments that enable these systems to operate reliably and at scale. Your work will involve containerized services, AWS Lambda, Terraform-managed infrastructure, and secure data transfer between AWS accounts, with a strong emphasis on scalability, security, and operational excellence. You will collaborate with engineering, product, and operational teams, while also coordinating technical requirements with the vendor. Strong communication skills are important, as you’ll often act as the bridge between the team and other internal teams, ensuring infrastructure and systems are aligned with internal standards and best practices. What we’re looking for: ● Strong AWS and DevOps experience (infrastructure, deployments, operations, observability) ● Solid experience with infrastructure as code (Terraform preferred) ● Experience with containerized environments (Docker, ECS/EKS, etc.) ● Hands-on experience with AWS services such as Lambda and related cloud tooling ● Golang and Bash skills for scripting, automation, and backend support ● Ability to work independently while collaborating with multiple stakeholders ● Strong communication and coordination skills Nice to have: ● Familiarity with TypeScript or Node.js This role is ideal for engineers who enjoy working towards engineering excellence and helping the team excel. You will collaborate across infrastructure and development teams to deliver impactful solutions. Join Our Global Team: We invite you to apply for the position at RE Partners. Join us in shaping the future of business technology consulting and transforming the way organizations thrive in a digital world. As a diverse, woman-owned global business, we pride ourselves on keeping talent happy – our 7% attrition rate speaks volumes. Bring your talented friends along and earn a referral bonus Equal Opportunity Employer: We are an equal opportunity employer and welcome applications from all qualified individuals regardless of race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, disability, or veteran status.

Brazil + 3 moreAll locations: Brazil | Egypt | Oman | Romania