Solar technology solutions that help you design, estimate and optimize commercial and utility scale solar assets.
Platform DevOps Engineer
Location
California + 8 moreAll locations: California | Colorado | Florida | New Jersey | New York | North Carolina | Massachusetts | Texas | Utah
Posted
52 days ago
Salary
$126.6K - $180K / year
Seniority
Senior
Job Description
Platform DevOps Engineer
PVcase
• Direct the AWS infrastructure strategy for PVcase Prospect, ensuring the application meets rigorous availability, performance, and security benchmarks. • Collaborate with the Global Platform team to implement unified architectural standards, contributing to organization-wide IaC and security initiatives. • Architect and maintain resilient cloud environments using Terraform and AWS, prioritizing modularity and reuse. • Support the transition toward a self-service enablement model, providing product developers with the tools and guardrails necessary for autonomous deployments. • Manage and refine monitoring, logging, and alerting stacks (Grafana, ELK, Prometheus, Checkly) to ensure proactive incident detection. • Identify and implement opportunities to leverage agentic workflows and AI-assisted tooling to automate complex operational tasks and improve incident response times.
Job Requirements
- Extensive hands-on experience managing complex AWS ecosystems (including VPC, RDS, IAM, EKS, Route53, S3, EFS, Firebase).
- Proven proficiency with Terraform for infrastructure automation and Kubernetes/Docker for container orchestration.
- Strong command of cloud networking (subnets, load balancing, routing) and DevSecOps principles (RBAC, encryption, secret management).
- A pragmatic approach to engineering, with a track record of delivering incremental improvements in high-growth SaaS environments.
- Excellent communication skills, with the ability to act as a technical liaison between US-based product teams and our global infrastructure team.
- Professional familiarity with, or a strong aptitude for, implementing AI-driven agentic workflows to optimize DevOps processes. This includes the use of LLMs and autonomous agents for task automation, documentation, and infrastructure maintenance.
- Comfort utilizing AI-assisted development tools (e.g., GitHub Copilot, Claude Code) to accelerate code generation and infrastructure troubleshooting.
- A proactive interest in evaluating emerging technologies that reduce cognitive load and enhance developer velocity.
Benefits
- Security for your future with our 401(K) plan, where we match 100% on your first 4% of contributions.
- Health, dental, and vision coverage.
- Flexible vacation policy, with a minimum of 3 weeks off.
- Full training and onboarding program for a seamless start.
- Flexible working hours, harmonizing your personal and professional life.
- Half-day Summer Fridays.
- Unlimited remote work policy.
- Internal transparency with company results and salary system, promoting a culture of trust and collaboration.
- Additional paid vacation days, including birthdays, volunteering, and other occasions.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
AI DevOps, Reliability Engineer
BranchBranch, founded in 2014, is a trailblazer in mobile linking and measurement, powering seamless digital experiences for over 100,000 apps and reaching more than
• Design and expand deployment automation • Establish release practices and standards • Extend automation deeper into production paths • Enable verification through automation • Own CI/CD standards across teams • Build pipeline tooling for safe engineering paths • Design and build out environments that mirror production • Bring AI tooling into operations • Champion Infrastructure as Code for provisioning • Operate and tune high-volume data infrastructure • Embed with an assigned engineering team day-to-day • Stand up DORA metrics and use them for improvements
• Developing and managing cloud-based infrastructure on AWS. • Creating and maintaining deployment architectures and continuous delivery pipelines. • Designing high-availability and fault-tolerant solutions for applications. • Implementing monitoring frameworks, including dashboards, alerts, and escalation processes. • Automating infrastructure provisioning and management using Infrastructure as Code (IaC) tools such as Terraform or CloudFormation. • Managing containerized applications and orchestrating deployments using Kubernetes. • Ensuring security best practices are applied across CI/CD pipelines, cloud infrastructure, and microservices. • Optimizing system performance and scalability through observability and proactive monitoring. • Collaborating with development teams to streamline deployment workflows and improve DevOps processes. • Advising clients on best practices for cloud infrastructure, deployment automation, and system security. • Engaging in technical discussions with stakeholders and supporting project execution to ensure timely delivery. • Assisting with the analysis of client requirements. • Working with and supporting Technical Leaders in project execution and timely delivery. • Collaborating with client teams.
Infrastructure / Site Reliability Engineer, SRE
Solvd, Inc.Get things Solvd. | Software Development & QA
• Design, provision, and maintain secure, scalable, and highly available cloud infrastructure (primarily AWS, GCP, or Azure) • Write and maintain modular, clean Terraform or OpenTofu scripts • Manage and optimize containerized environments using Docker and Kubernetes (EKS/GKE) • Build, maintain, and secure robust CI/CD pipelines • Implement modern GitOps workflows to automate application delivery • Design and implement comprehensive observability stacks using tools like Prometheus, Grafana, Datadog, or New Relic • Participate in an engineering on-call rotation, driving root-cause analysis
DevOps Engineer, II
EncouraWe empower students & institutions to create meaningful connections to achieve their goals.
• Own and maintain the reliability, performance, and availability of large-scale production systems — monitoring dashboards, reviewing alerts, and resolving incidents as they arise • Serve as primary on-call and incident responder for customer-impacting issues, triaging, coordinating resolution, and leading post-mortems • Design, build, and improve CI/CD pipelines using Azure DevOps, GitHub Actions, Jenkins, and Octopus Deploy • Automate Azure infrastructure and services (Web Apps, Functions, SQL, Storage, Key Vaults, Entra ID) using IaC tooling • Collaborate closely with engineering, product, and security teams to support deployments, migrations, and compliance initiatives • Drive cloud cost optimization, scalability, and auto-scaling initiatives across hosted environments • Implement and modernize incident management tooling and enforce security best practices and system hardening standards



