Job Closed
This listing is no longer active.
Leading MDR provider trusted by some of the world’s top brands to expel adversaries, minimize risk, & build resilience.
Senior AI Platform Engineer
Location
United States
Posted
75 days ago
Salary
$142.9K - $207.2K / year
Seniority
Senior
Job Description
Senior AI Platform Engineer
Expel
• Architect and maintain end-to-end machine learning training pipelines on AWS (SageMaker, EKS, Step Functions) to ensure reliable and reproducible model development and deployment • Build and maintain infrastructure for production agentic applications using Amazon Bedrock and Bedrock AgentCore — including agent runtimes, memory, secure gateways, and observability at scale • Contribute to the architectural evolution of our ML platform, including evaluating MLOps tooling and participating in buy vs. build decisions • Implement AI/ML governance best practices for model versioning, testing, validation, maintenance, and security • Integrate MLOps best practices with Expel's SDLC, security, and infrastructure standards, working alongside SRE, Platform Engineering, and Security teams • Drive quality, reliability, and scalability improvements through thoughtful engineering and monitoring • Partner with data scientists, software engineers, and stakeholders to operationalize ML models reliably and at scale • Mentor and support junior engineers; foster a culture of engineering excellence • Create and maintain documentation, internal tooling, and enablement resources so practitioners across Expel can work effectively with ML systems • Stay current with the MLOps landscape and bring relevant innovations back to the team
Job Requirements
- 5+ years of relevant software engineering experience with meaningful focus on ML operations and infrastructure
- Degree in Computer Science, Mathematics, Statistics, Engineering, or a related technical field preferred (or a compelling story)
- Strong Python proficiency; familiarity with other languages (Go, JS) is a plus
- Solid experience with CI/CD pipelines, infrastructure-as-code, and containerization for ML workloads
- Hands-on experience with cloud-based ML platforms — AWS (SageMaker, Bedrock, Bedrock AgentCore) strongly preferred; GCP (Vertex AI) experience also valued
- Proven experience operationalizing LLMs and building infrastructure for complex agentic applications — agent orchestration, memory, tool calling, RAG architectures
- Familiarity with ML frameworks including Scikit-Learn, PyTorch, Spark, and TensorFlow
- Working knowledge of continuous retraining, concept drift monitoring, and data drift detection in production
Benefits
- Offer unlimited PTO (that leadership models and encourages)
- Up to 24 weeks of parental leave
- Excellent health benefits
- Pay you a monthly fitness and cell phone stipends — no receipts required
- Support your professional growth with a conference benefit and continuous learning opportunities
- Offer full remote flexibility — work from wherever you do your best work
Related Guides
Related Categories
Related Job Pages
More Platform Engineer Jobs
• As a Senior Data Platform Engineer, you’ll be a hands-on engineer on a small, high-ownership data team. • You’ll work across the full data platform - relational, warehouse, and lakehouse systems - building and operating the pipelines that power compliance, analytics, and reporting workloads. • Design and build pipelines that move data across systems - supporting data lake ingestion, compliance workloads, and cross-domain data flows. • Own pipeline operations end to end: monitoring, incident resolution, data quality, and documentation that lets any team member respond independently. • Identify technical debt and reliability risks and raise them with clear context and proposed next steps. • Design and maintain schemas across relational, warehouse, and lakehouse layers, working with application engineers and product to get data models right. • Build out the platform’s service layer, infrastructure-as-code, and data quality frameworks - this role spans design and implementation. • Keep platform documentation at a level where any team member can understand what exists, how it works, and where the risks are. • Over time, contribute to the analytics engineering layer, including modeling practices and semantic layer development. • Contribute to evaluations of the current platform against emerging architectures and tooling, helping produce trade-off analyses and recommendations. • Bring what you see day to day in the systems you operate into the team’s improvement roadmap and technical direction. • Track and report on platform health metrics: pipeline uptime, failure rates, data freshness, and cost trends. • Mentor peers and junior engineers through code review, pairing, and technical guidance.
Mobile Platform Developer
TailscaleSimple, secure networks for teams of any scale. Built on WireGuard.
• Develop the Tailscale product, contributing to client code and backend services. The client code is a mix of modern Swift, Kotlin and Go. Prior Go expertise is not a requirement. • Bring a special focus on our mobile platforms, iOS and Android, while contributing to common code that supports macOS, Windows and other core client platforms.
Role Description Gestalte die technologische Grundlage einer skalierenden AI-SaaS-Plattform aktiv mit. Wir suchen eine erfahrene Persönlichkeit mit strategischem Blick auf Plattformarchitektur, Infrastruktur und technische Skalierung. In dieser Rolle verantwortest du die Weiterentwicklung unserer technischen Plattform- und Infrastrukturstrategie und stellst sicher, dass unsere Systeme den Anforderungen eines wachsenden AI-SaaS-Unternehmens nachhaltig gerecht werden. Du verbindest technologisches Verständnis mit wirtschaftlichem Denken und sorgst dafür, dass Skalierbarkeit, Stabilität, Security und Effizienz in einer modernen Plattformarchitektur zusammengeführt werden. Dabei arbeitest du eng mit Geschäftsführung, Produktmanagement und Engineering zusammen und gestaltest zentrale Infrastrukturentscheidungen aktiv mit. Aufgaben - Entwicklung und Weiterentwicklung der Zielarchitektur unserer AI- und Voice-basierten SaaS-Plattform - Definition und Umsetzung einer skalierbaren, stabilen und wirtschaftlich effizienten Infrastrukturstrategie - Sicherstellung der Skalierbarkeit bei steigender Nutzung und wachsender Systemlast - Verantwortung für Cloud-, Hosting- und Infrastrukturkonzepte auf Basis moderner Plattformtechnologien - Weiterentwicklung von Security-, Governance- und Compliance-Strukturen - Einführung und Optimierung von Observability-, Monitoring- und Performance-Standards - Optimierung der Infrastrukturkosten mit Blick auf Skalierung, Stabilität und Wirtschaftlichkeit - Aufbau klarer Infrastruktur-Roadmaps sowie Priorisierung technischer Initiativen - Steuerung externer Technologie- und Infrastrukturpartner - Enge Zusammenarbeit mit Engineering, Produktmanagement und Geschäftsführung bei Architektur- und Plattformentscheidungen - Sicherstellung klarer Entscheidungs- und Priorisierungsprozesse zur Vermeidung von Verzögerungen in Infrastruktur- und Plattforminitiativen Qualifications - Mehrjährige Erfahrung in technischen Architektur-, Plattform- oder Infrastrukturrollen - Erfahrung im Aufbau oder in der Weiterentwicklung skalierbarer SaaS-Plattformen - Sehr gutes Verständnis moderner Cloud-Infrastrukturen, idealerweise Azure, AWS oder GCP - Sehr gutes Verständnis von Linux-basierten Systemen und deren Betrieb in produktiven Cloud-Umgebungen - Erfahrung mit AI-/LLM-basierten Systemen oder datenintensiven Plattformumgebungen - Fundierte Kenntnisse in Infrastruktur-Architektur, Plattformdesign und DevOps-nahen Strukturen - Erfahrung mit Security-, Governance- und Compliance-Anforderungen in modernen Cloud-Umgebungen - Fähigkeit, technische Architekturentscheidungen unter Berücksichtigung von Skalierbarkeit, Kosten und Umsetzbarkeit unter Zeitdruck und mit klarer Priorisierung zu treffen - Erfahrung in der Zusammenarbeit mit Engineering-Teams sowie in der Abstimmung mit Business-Stakeholdern - Fähigkeit, komplexe technische Themen verständlich und adressatengerecht zu kommunizieren - Strukturierte, analytische und lösungsorientierte Arbeitsweise - Hohes Maß an Eigenverantwortung, Entscheidungsstärke und Umsetzungsorientierung Benefits - Eine strategisch wichtige Rolle mit direktem Einfluss auf die Weiterentwicklung unserer Plattform - Gestaltungsspielraum in einem dynamischen, technologiegetriebenen Umfeld - Direkte Zusammenarbeit mit der Geschäftsführung und kurzen Entscheidungswegen - Möglichkeit, zentrale Infrastruktur- und Architekturthemen nachhaltig mitzugestalten - Moderne AI- und SaaS-Technologien in einem wachstumsorientierten Umfeld - Hoher Impact auf Skalierbarkeit, Stabilität und Zukunftsfähigkeit unserer Plattform - Flexible und hybride Arbeitsweise mit deutschlandweitem Remote-Setup - Zusammenarbeit mit erfahrenen Produkt-, Engineering- und Technologie-Teams
• Develop, implement, and maintain governance frameworks across the Power Platform ecosystem. • Administer and support Power Platform environments, including Power Apps. • Design and deliver hands-on application solutions, including automation workflows and data architecture. • Maintain a comprehensive understanding of broader platform capabilities. • Design, build, and deploy scalable solutions using Power Pages, Power Apps, Power Automate, and Dataverse.



