EasyLlama

EasyLlama, founded in 2019 and headquartered in Covina, California, is a compliance training and learning management platform that helps organizations create sa

Staff Software Engineer – Platform

Location

California + 19 moreAll locations: California | Colorado | Connecticut | Florida | Illinois | Louisiana | New Jersey | New York | North Carolina | Ohio | Oregon | Massachusetts | Michigan | Minnesota | Pennsylvania | Tennessee | Texas | Virginia | Washington | Wisconsin

Posted

8 days ago

Salary

$163K - $185K / year

Seniority

Lead

Job Description

Staff Software Engineer – Platform

EasyLlama

• Own the technical roadmap for Platform Engineering. • Architect, build, and operate highly scalable backend services using Ruby on Rails. • Design and operate our Kubernetes infrastructure, cloud architecture, and deployment pipelines. • Improve developer productivity through CI/CD, automation, internal tooling, and engineering best practices. • Build and evolve the platform powering our AI products, including LLM orchestration, agent infrastructure, prompt management, evaluation pipelines, model routing, and AI observability. • Own our billing platform, including subscriptions, licensing, payments, usage metering, invoicing, and Stripe integrations. • Design and maintain secure, scalable integrations with HRIS platforms, identity providers, and enterprise customers using REST APIs, OAuth, SAML, SCIM, and webhooks. • Build best-in-class observability through logging, metrics, tracing, alerting, and performance monitoring. • Improve the scalability, security, reliability, and operational excellence of our platform. • Lead architectural decisions that impact multiple engineering teams. • Mentor engineers through technical leadership, design reviews, and engineering best practices. • Diagnose and resolve complex production issues across infrastructure, applications, and distributed systems. • Help establish and grow the Platform Engineering function as EasyLlama scales.

Job Requirements

  • 10+ years of professional software engineering experience
  • Expert-level Ruby on Rails experience building production SaaS applications
  • Strong React experience and the ability to work across the full stack
  • Extensive experience designing distributed systems and service-oriented architectures
  • Deep experience with Kubernetes, Docker, AWS, and cloud-native infrastructure
  • Strong experience building and operating CI/CD pipelines and modern developer tooling
  • Experience owning production infrastructure with high availability and reliability requirements
  • Experience designing and maintaining APIs and enterprise integrations
  • Strong understanding of SQL, database design, caching, background processing, and notification systems like email and SMS
  • Experience leading complex technical initiatives across multiple teams without direct people management
  • Excellent debugging, communication, and architectural design skills
  • A builder’s mindset with strong ownership and a passion for creating platforms that enable other engineers to move faster.

Benefits

  • Competitive employer-sponsored health insurance
  • 401(k) + company matching
  • Professional development reimbursements
  • Flexible, fully remote environment
  • Quarterly remote work stipend

Related Categories

Related Job Pages

More Platform Engineer Jobs

Senior Fullstack Engineer – Developer Platform

JTL-Software GmbH

JTL ist einer der führenden Anbieter von E-Commerce-Software im deutschsprachigen Raum – mit ca. 450 gruppenweiten Mitarbeiter:innen und 50.000 Kunden aus verschiedensten Branchen. Wir entwickeln skalierbare, flexible Lösungen für den Onlinehandel der Zukunft – von der Warenwirtschaft bis zur Shop- und Marktplatzanbindung. Fairness und Respekt sind bei uns gelebte Praxis.

Role Description Your mission is to work on the platform topics that matter most right now. Our focus shifts each quarter depending on where the company needs it most — sometimes frontend, sometimes backend, sometimes infrastructure. - You develop standards, guidelines, and best practices for frontend and backend at JTL. - You create and maintain templates, boilerplates, and libraries that help other teams work more productively and efficiently. - You support the bootstrapping of new projects, so our standards and toolchains are applied from the start. - You regularly evaluate new technologies and assess their value for our landscape. Qualifications - You have several years of professional experience in fullstack development. - You are strong in one stack and solid in the other: React and TypeScript on the frontend, .NET 8 / C# on the backend. - Frontend: component-based development and design systems (Storybook, TailwindCSS), state management and data fetching (e.g. TanStack Query), modern build tools (Vite, turborepo, pnpm), and testing (Vitest, Jest, Playwright). - Backend: CQRS and Domain-Driven Design (DDD), API development with FastEndpoints, and familiarity with GraphQL — ideally having designed a GraphQL API. - You are confident with Docker and common containerisation workflows, and with Git and collaborative workflows (pull requests, code reviews). - You communicate clearly in writing and speech — a big part of the role is driving standards and best practices across teams. - You have a solid understanding of and genuine interest in AI development, since most of the platform features we ship are designed AI-first. - You have business-fluent English (written and spoken) and good German skills. Requirements - Nice to have: Terraform and hyperscalers, preferably Azure. - Micro-frontends. - CI/CD pipelines (GitHub Actions). - Basic Kubernetes and Helm. - Infrastructure as Code (IaC), GitOps, and experience with Argo. Benefits - Remote-first within Germany, with the option to work remotely from eligible countries for up to 180 days per year. - Meal allowance of up to €115 net per month. - Ergonomic workspace allowance for your home office setup. - Regular team events, company-wide gatherings, and summer and Christmas parties to stay connected as a remote-first company. - EGYM Wellpass and JobRad subsidy. - Financial benefits including capital-forming payments (Vermögenswirksame Leistungen) and a company pension scheme.

Germany

Role Description Own the entire technology platform of a lean, profitable, high-traffic performance marketing business. You’ll be responsible for the applications, data platform, and infrastructure that generate revenue every day, with the autonomy to make it faster, more reliable, and easier to scale. What You’ll Own - You’ll be the technical owner of our production platform from end to end. Applications - Our consumer survey and offer web applications running on AWS EC2. - You’ll build new features while keeping the platform fast under real production traffic, including: - Real-time offer serving - Lead capture - Event tracking - Customer-facing web experiences Data Platform - Own our high-traffic PostgreSQL database (Supabase, currently ~50GB and growing). - Responsibilities include: - Schema design - Query optimization - Index strategy - Performance tuning - Database reliability - Performance matters. A single missing database index recently cost us thousands of dollars per day. We want someone who proactively owns performance before it becomes a business problem. Infrastructure - Own the production environment. This includes: - AWS EC2 - Linux servers - Deployments - Monitoring - Logging - Observability - Performance - Uptime - Your mindset should be: - Is it up? - Is it fast? - Will we know about problems before our customers do? Integrations - Maintain and improve integrations with systems including: - Everflow - TrustedForm - IPQS - Google APIs - Tracking pixels and postbacks - Internal AI tooling (our OpenClaw-based AI agent, Zelda) Your First 90 Days - You’ll be expected to quickly become the technical owner of the platform. Success looks like: - Learning the complete survey → offer → revenue pipeline - Becoming the go-to technical owner of the production platform - Implementing meaningful monitoring and alerting for database and application health - Shipping at least one customer-facing feature - Delivering measurable improvements in reliability or performance Qualifications - 7+ years building and owning production web applications - Extensive PostgreSQL experience including indexing, query optimization, schema design, and troubleshooting under load (required) - Strong AWS, EC2, Linux, and production operations experience - Experience with Node.js, TypeScript, Python, and modern JavaScript front-end frameworks - Strong API and systems integration experience - Experience building observable, reliable production systems - Comfortable owning systems independently in a small remote company - Excellent communication skills and a high level of accountability Bonus Points - Experience with any of the following is a plus: - Performance marketing - Affiliate marketing - AdTech - Lead generation - Offer serving - Attribution systems - TrustedForm or consent platforms - High-volume web applications - AI-assisted software development - Modern CI/CD - Observability and production monitoring - Taking legacy codebases from “working” to “well engineered” Why This Role Is Different This isn’t a ticket-taking engineering job. You’ll own a real production platform that generates revenue every day. You’ll have the freedom to make architectural decisions, improve reliability, and build new capabilities without navigating layers of management. The scope is similar to what many engineers only experience as a technical founder, but with the stability of an established, profitable business. If you’re looking for ownership, autonomy, and the opportunity to build systems that have an immediate business impact, we’d love to talk.

Worldwide
$155K - $175K / year
Famedly GmbH logo

Platform Engineer

Famedly GmbH

Famedly is a complete medical collaboration platform delivered as a single decentralized application.

Full TimeRemoteTeam 11-50Since 2019H1B No Sponsor

• Design, build, and operate our Kubernetes platform across multiple clusters • Develop and maintain our GitOps workflows with Argo CD (app-of-apps model, Helm-based deployments) • Operate and evolve core platform services such as Harbor, OpenBao, Zitadel • Build and improve observability across the platform (Grafana, Mimir, Loki, Tempo, Alloy) and contribute to actionable alerting • Manage secrets, backups, and data services (e.g. CNPG/PostgreSQL, VolSync, S3 storage) • Partner with product engineering teams to migrate services from Ansible/VM-based deployments to Kubernetes and define reusable deployment standards • Improve developer experience through documentation, templates, golden paths, and self-service platform capabilities • Troubleshoot platform issues across infrastructure, networking, and application layers (2nd-level support) • Participate in on-call duties to ensure stable platform operations • Contribute to platform architecture decisions, runbooks, and operational documentation

Germany
€60K - €70K / year

OpenShift Platform Engineer

Bright Vision Technologies

Bright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications.

Role Description We are seeking an experienced OpenShift Platform Engineer to design, deploy, and operate enterprise-grade Red Hat OpenShift container platforms running mission-critical workloads across on-premises and cloud environments. In this role you will own the end-to-end platform lifecycle — from cluster provisioning and hardening, through workload onboarding and developer enablement, to ongoing operations and continuous improvement. The ideal candidate will bring deep expertise in OpenShift and Kubernetes internals, strong Linux fundamentals, and a security-first mindset, and will partner with application teams, security engineers, and infrastructure stakeholders to deliver a stable, secure, and developer-friendly container platform. Key Responsibilities - Architect, provision, and operate Red Hat OpenShift clusters across on-premises, virtualized, and major cloud environments, applying best practices for high availability, scalability, and security. - Design and implement multi-tenant OpenShift platforms with appropriate project isolation, network policies, quotas, RBAC, and security context constraints. - Automate cluster lifecycle operations — installation, upgrades, scaling, and decommissioning — using OpenShift installer, GitOps tooling (Argo CD, OpenShift GitOps), and IaC frameworks. - Onboard application development teams to the platform, providing clear self-service patterns, project templates, CI/CD integrations, and supporting documentation. - Operate OpenShift networking, including SDN/OVN-Kubernetes, ingress controllers, routes, network policies, and integration with enterprise firewalls and load balancers. - Manage OpenShift storage strategies (block, file, object), including dynamic provisioning, OpenShift Data Foundation, and integration with external storage platforms. - Implement comprehensive monitoring, logging, and tracing using the OpenShift observability stack (Prometheus, Grafana, EFK, Tempo) and integrate with enterprise SIEM and APM tools. - Harden the platform in collaboration with security teams, including image scanning (Quay/Clair), runtime protection, secret management, and continuous compliance with regulatory frameworks. - Design and maintain robust CI/CD pipelines on OpenShift using Tekton, Jenkins, GitLab CI, or Argo Workflows, enabling safe and frequent application deployments. - Develop disaster-recovery and backup strategies for cluster state, etcd, and persistent workloads, validated through regular failover and restore drills. - Plan, test, and execute cluster upgrades across minor and major versions with minimal disruption to workloads, including pre-upgrade validation and rollback strategies. - Partner with application teams to right-size workloads, manage resource quotas, and tune cluster autoscaling for predictable performance and cost efficiency. - Continuously evolve platform tooling and developer experience, capturing feedback from internal users and translating it into roadmap improvements. - Mentor junior engineers and act as an OpenShift subject-matter expert across the broader organization. Qualifications - Bachelor’s degree in Computer Science, Engineering, or a related technical discipline. - Five or more years of experience operating container platforms in production, with at least three years on Red Hat OpenShift. - Deep, hands-on knowledge of Kubernetes and OpenShift internals, including operators, CRDs, RBAC, and networking. - Strong Linux administration skills, including networking, performance tuning, and troubleshooting. - Hands-on experience with infrastructure-as-code tools such as Ansible, Terraform, or Helm. - Solid experience implementing CI/CD pipelines on OpenShift or Kubernetes using Tekton, Jenkins, or Argo CD. - Strong scripting skills in Bash, Python, or Go. - Working knowledge of cluster monitoring, logging, and tracing tools. - Familiarity with container image security and supply chain hardening. - Excellent troubleshooting, communication, and documentation skills. Preferred Qualifications - Red Hat Certified Specialist in OpenShift Administration or Architect-level Red Hat certification. - Experience operating OpenShift on at least one major public cloud (AWS, Azure, GCP). - Exposure to service mesh implementations (OpenShift Service Mesh, Istio, Linkerd). - Familiarity with regulated environments such as PCI-DSS, HIPAA, or SOC 2. - Experience with GitOps workflows using Argo CD or Flux. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at www.bvteck.com .

United States
100K - 150K / year
Job Closed