Job Closed

This listing is no longer active.

Site Reliability Engineer

DevOps EngineerDevOps EngineerFull TimeRemoteMid LevelTeam 10,001

Location

United States

Posted

113 days ago

Salary

0

Seniority

Mid Level

No structured requirement data.

Job Description

Site Reliability Engineer

Cooper Standard

Job Description: Site Reliability Engineer (SRE) About Liveline Liveline enables dramatic improvements in manufacturing performance thorough a unique application of artificial intelligence to provide real-time process control and predictive assistants for plant personnel. Our focus is on automating complex processes, not simply providing dashboards for managers and operators. Our team combines experts in AI with world-class process engineers who can focus on the “last mile” with customers: Extracting data from the process and implementing controls on the shop floor. We speak the language of AI but also industrial controllers. Our hardware and software offerings are scalable and cost-effective whether customers have one production line or hundreds, delivering an ROI that’s attractive to small and medium-sized enterprises. We are passionate about democratizing the power of analytics and advanced automation for manufacturers of almost any size. Through our approach, producers can de-mystify complex processes and free up valuable technicians to focus on more advanced tasks instead of constantly monitoring and adjusting equipment parameters. A Liveline Technologies SRE is responsible for the reliability, performance, observability, and operational excellence of Liveline’s production services. This spans from the factory-floor edge systems to AWS cloud components. You will help build and run resilient infrastructure, automate repetitive work with code (Terraform, Bash, Python), implement monitoring and alerting (Prometheus/Grafana), and participate in incident response/on-call to ensure uptime for mission-critical manufacturing systems. You’ll collaborate closely with controls engineers, data scientists, and software teams to safely deploy changes, define SLIs/SLOs, and continuously improve availability and latency for real-time process control. Primary Responsibilities - Operate Production Systems: Maintain high availability, performance, and security of Liveline’s production stack across AWS and plant/edge environments. - Observability & Monitoring: Stand up, tune, and maintain Prometheus/Grafana dashboards, alerts, recording rules, and runbooks. Implement logs/traces (e.g., OpenTelemetry) and actionable alerting. - Infrastructure as Code: Build and manage reproducible infrastructure with Terraform (VPC, IAM, EC2/EKS/ECS, RDS, S3, CloudWatch, CloudTrail). Apply version control, code reviews, and plan/apply workflows. - Automation & Tooling: Write Bash and Python scripts and small services to automate operational tasks, health checks, failover routines, backup/restore, and environment bootstrapping. - NOC / Incident Response: Participate in a follow-the-sun/on-call rotation; triage and resolve incidents, lead initial comms, and produce blameless postmortems with clear corrective actions. - SLIs/SLOs/Error Budgets: Define and instrument SLIs (availability, latency, error rate, freshness), set SLOs with stakeholders, and manage error budgets to guide release velocity and reliability tradeoffs. - Networking & Connectivity: Support secure, reliable connectivity between factory networks and cloud (site-to-site VPNs, routing, DNS, TLS, private subnets, security groups, network ACLs). - Databases & Storage: Operate and tune PostgreSQL/TimescaleDB, InfluxDB, or similar time-series/relational stores; manage backups, PITR, replication, partitioning, and performance baselining. - CI/CD & Release Engineering: Contribute to build/deploy pipelines (e.g., GitHub Actions/GitLab CI), implement canaries/blue-green strategies, and enforce change management and rollback plans. - Security & Compliance: Enforce least-privilege IAM, secret management (AWS Secrets Manager/SSM), encryption, artifact signing, and basic hardening for Linux and Kubernetes workloads. - Edge & OT Collaboration: Partner with process/controls engineers to ensure reliable data ingestion from PLCs/industrial gateways (e.g., OPC UA/Modbus), and safe deploys to plant edge nodes. - Cost, Capacity & Performance: Right-size compute/storage, set budgets/alerts, forecast capacity, and optimize resource utilization without compromising SLOs. - Documentation & Runbooks: Author and maintain runbooks, architecture diagrams, operational playbooks, and disaster recovery procedures. Education and Qualifications: - Bachelor’s Degree in IT, Computer Science, or Computer Engineering (or equivalent experience). - 5+ years of experience in a corporate IT or startup setting - Familiar with containers (Docker) and orchestration (Kubernetes or ECS). - Experience running production workloads, participating in on-call, and writing postmortems. - Strong communication skills with the ability to explain tradeoffs to non-SRE stakeholders. - Intellectual curiosity, ownership mindset, and bias for automation. - Willingness and ability to travel to customer sites and plants, as necessary. Nice to Have - Kubernetes (EKS), Helm, Kustomize. - Service Mesh/Ingress (Envoy, NGINX, ALB). - Logging/Tracing: OpenSearch/ELK, Loki, OpenTelemetry. - Config Management: Ansible. - Secrets & PKI: HashiCorp Vault, mTLS. - Edge/Industrial Protocols: OPC UA, Modbus, MQTT; experience with industrial gateways. - Compliance exposure (SOC 2, ISO 27001) and change management (ITIL). Position Type: Regular Additional Locations: Additional Information: Cooper Standard is proud of its diverse workforce and committed to providing equal employment opportunities to applicants and employees without regard to race, color, religion, sex, national origin, genetic information, physical or mental disability, age, veteran or military status, or any other characteristic protected by applicable law. We are dedicated to creating an environment at work that not only values diversity but also encourages inclusion and a sense of belonging. We firmly believe that a diverse workplace fosters an environment where our employees can flourish and provide superior service to our customers. Because we recognize and value the range of ways in which people acquire experiences, whether personal, professional, or via education or volunteerism, we invite interested applicants to evaluate the key duties and requirements and apply for any opportunities that fit your experience and qualifications. Applicants with disabilities may be entitled to reasonable accommodations under the Americans with Disabilities Act, as well as certain state and/or local laws. If you believe you require such assistance to complete our online application or to participate in an interview, you (or someone on your behalf) may request assistance by emailing recruitment@cooperstandard.com with a description of the accommodation you seek. Application materials submitted to this email address will not be considered. Remote Status: Remote

Related Categories

Related Job Pages

More DevOps Engineer Jobs

ORAEX CLOUD CONSULTING logo

DevOps Analyst, IaC

ORAEX CLOUD CONSULTING

Data Management • Cloud • DevOps • Observability

DevOps Engineer113 days ago
Full TimeRemoteTeam 51-200Since 2012H1B No Sponsor

• Develop, maintain, and enhance infrastructure automation using Terraform. • Create and manage playbooks, roles, and automation routines with Ansible. • Standardize provisioning and configuration of servers, cloud environments, and workloads. • Implement and continuously improve CI/CD pipelines. • Ensure versioning, traceability, and reuse of infrastructure code. • Support technical teams in adopting automation, DevOps best practices, and governance. • Troubleshoot automated environments, pipelines, and provisioned resources. • Contribute to security, compliance, and change-control practices in production and non-production environments. • Document processes, architectures, and automation workflows.

Brazil
Incentive.me logo

SRE/DevOps Analyst

Incentive.me

Incentives, engagement, loyalty tecnology

DevOps Engineer113 days ago
Full TimeRemoteTeam 51-200Since 2017H1B No Sponsor

• Support the maintenance and evolution of environments and services • Work on monitoring, observability and troubleshooting • Assist with incident response, troubleshooting, disaster recovery and root cause analysis (RCA) • Work with CI/CD pipelines and operational automation • Support the operation of Kubernetes workloads and cloud services • Collaborate with technical teams to improve reliability, availability and operational efficiency • Create and maintain technical documentation for environments, processes and changes • Support security best practices across infrastructure, access controls, pipelines and applications • Participate in root cause analyses (RCA)

Brazil
Job Closed
Keyfactor, Inc. logo

Senior DevOps Engineer

Keyfactor, Inc.

Our mission is to build a connected society, rooted in trust, with identity-first security for every machine and human. Keyfactor helps organizations move fast to establish digital trust at scale — and then maintain it. With decades of cybersecurity experience, Keyfactor is trusted by more than 1,500 companies across the globe. We are proud to continually earn recognition as a Best Place to Work, and we achieve that through our amazing people who cultivate our culture as we grow. We hope you will trust your future with Keyfactor!

DevOps Engineer113 days ago
Full TimeRemoteTeam 501-1,000

Role Description Responsible for building and maintaining the tools, processes, and environments that enable efficient software delivery. This role emphasizes a very hands-on approach towards automating workflows, containerizing applications, implementing CI/CD pipelines, and Infrastructure as Code, working closely with cross-functional teams. Job Responsibilities - Lead design, performance tuning, and debugging of scalable data systems with focus on Search and Analytics (OpenSearch), Event Streaming (Kafka), and NoSQL databases (MongoDB). - Lead the design and implementation of infrastructure automation and deployment processes across multiple cloud providers spanning both containers and Virtual Machines. - Architect and manage scalable containerization and orchestration solutions for applications. - Drive continuous integration and continuous delivery (CI/CD) improvements to optimize workflows. - Monitor and enhance system performance, availability, and security through proactive measures. - Collaborate closely with cross-functional teams to improve automation and deployment strategies. - Implement and maintain Infrastructure as Code (IaC) for consistent and reproducible environments. - Mentor junior engineers, providing technical guidance and fostering skill development. Qualifications - Degree in Computer Science, Engineering, or a related field (or equivalent experience). - 4-6 years of relevant work experience in DevOps or IT infrastructure. - Extensive knowledge of industry trends, company strategy, and cross-functional processes. - Deep understanding of scripting, version control, and infrastructure best practices. - Comprehensive knowledge of system performance, security, and reliability. - Strategic thinking, exceptional problem-solving abilities, high-level proficiency in relevant tools and technologies. - Advanced skills in infrastructure automation and deployment. - Expertise in containerization and orchestration. - Experience with optimizing and scaling CI/CD pipelines. - Ability to lead complex projects, drive strategic initiatives, and influence decision-making. - Strong problem-solving and troubleshooting skills. - Ability to mentor junior team members and lead initiatives. - Any open-source contributions relevant to data and infrastructure projects are a nice-to-have. Travel Requirements - Up to 5% Compensation Salary will be commensurate with experience. Benefits - Second Fridays (a company-wide day off on the second Friday of every month minus November and December of 2025 due to the Holiday schedule). Please note that this benefit is subject to change. - Comprehensive benefit coverage globally. - Generous paid parental leave globally. - Competitive time off globally. - Dedicated employee-focused ambassadors via Key Contributors & Culture Committees. - DIVERSE Commitment, a call to action for a more inclusive and diverse future in business, society, and technology. - The Keyfactor Alliance Program to support DEIB efforts. - Wellbeing resources, wellness allowance, mindfulness app free membership, Wellness Wednesdays. - Global Volunteer Day, company non-profit matching, and 3 volunteer days off. - Monthly Talent development and Cross Functional meetings to support professional development. - Regular All Hands meetings – followed by group gatherings. Our Core Values - Trust is paramount. - Customers are core. - Innovation never stops, it only accelerates. - We deliver with agility. - United by respect. - Teams make “it” happen. Equal Opportunity Employment Keyfactor is a proud equal opportunity employer including but not limited to veterans and individuals with disabilities. Reasonable Accommodation Applicants with disabilities may contact a member of Keyfactor’s People team via people@keyfactor.com and/or telephone at 1.216.785.2990 to request and arrange for accommodations at any time.

United States
Job Closed
MAIA logo

Lead DevOps - Platform Engineer

MAIA

Empowering the Mittelstand through AI-powered SaaS Solutions.

DevOps Engineer113 days ago
Full TimeRemoteTeam 1-10Since 2021H1B Sponsor

• Make MAIA's platform reliable, secure, auditable, and developer-friendly. • Take full ownership of our infrastructure and establish the standards that will govern how it grows. • Own and evolve our CI/CD pipelines (GitHub Actions) and deployment workflows. • Build out and mature our observability stack (Grafana, Loki, Sentry, PostHog). • Implement and own our security fundamentals: IAM, secrets management, TLS, vulnerability scanning, and patch management. • Drive the technical controls required for our ISO 27001 certification and build the systems that produce auditable evidence continuously.

Germany
€70K - €80K / year
Job Closed