Be smart. Smart this!
Lead AI Platform Engineer, LLM, Databricks
Location
Brazil
Posted
5 days ago
Salary
0
Seniority
Senior
Job Description
Lead AI Platform Engineer, LLM, Databricks
Smarthis
• Liderar tecnicamente a evolução da plataforma corporativa de IA e suas capacidades compartilhadas. • Definir padrões de governança para utilização de LLMs, AI Gateway, Skills, MCPs e agentes inteligentes. • Projetar arquiteturas e frameworks reutilizáveis para acelerar iniciativas de IA em toda a organização. • Conduzir avaliações de novas tecnologias e definir a estratégia de evolução da plataforma. • Estabelecer padrões de MLOps, MLflow, versionamento e governança para soluções de IA. • Atuar como referência técnica para Analytics Engineers e demais times consumidores da plataforma. • Trabalhar em conjunto com Cybersecurity, Arquitetura e Platform Engineering para garantir segurança, escalabilidade e conformidade das soluções. • Mentorar engenheiros e disseminar boas práticas de AI Platform Engineering.
Job Requirements
- Experiência sólida com plataformas de IA e Large Language Models (LLMs).
- Experiência construindo frameworks, bibliotecas ou plataformas compartilhadas de IA.
- Domínio de Python para desenvolvimento de soluções de IA.
- Conhecimento em AI Gateway, MLflow, MLOps e Databricks AI.
- Experiência com governança, segurança e boas práticas para IA corporativa.
- Capacidade de definir padrões técnicos e influenciar decisões de arquitetura.
- Inglês para atuação em ambiente global.
- Excelente se tiver
- Experiência com LangChain, LangGraph ou frameworks para agentes inteligentes.
- Vivência com Model Context Protocol (MCP).
- Experiência desenvolvendo AI Agents para ambientes corporativos.
- Conhecimento em Unity Catalog e Data Governance.
- Certificações Microsoft Azure ou Databricks.
- Participação na definição de roadmaps e estratégias para plataformas de IA.
Benefits
- Participação nos lucros e resultados
- Plano de saúde e odontológico
- Cartão de benefícios flexível
- TotalPass
- Acesso ao Zenklub
- Programa de idiomas
- Parceria com a FIAP
- Trabalho híbrido
- Treinamentos internos e plano de carreira estruturado
- Horário flexível
- Programa de mentoria
- Licença parental estendida
- Auxílio-creche
Related Guides
Related Categories
Related Job Pages
More Platform Engineer Jobs
• Configure and support enterprise AI platforms, including OpenAI, Anthropic, Google, and related AI tooling. • Define approved access patterns for LLMs, APIs, MCP, agents, connectors, RAG, automation, and human-in-the-loop workflows. • Partner with business teams to assess AI use cases, clarify outcomes, evaluate feasibility, and recommend the right solution pattern. • Lead or co-lead lighthouse AI projects that prove reusable enterprise patterns and demonstrate measurable business value. • Create solution blueprints for priority AI use cases, including architecture, data needs, access patterns, risk, cost, support model, and production readiness. • Establish cost management, usage tracking, agent visibility, and oversight infrastructure. • Define evaluation approaches for AI solutions, including quality checks, test cases, usage metrics, human review points, failure handling, and improvement loops. • Package successful pilots into reusable reference architectures, implementation templates, and transition plans for production support. • Partner with security, privacy, legal, architecture, data, and IT teams on governance, risk review, and standards. • Advise business leads and citizen developers on when AI solutions can remain team-owned versus when they require enterprise ownership, controls, monitoring, and support. • Work with the AI Enablement Analyst to convert platform capabilities into training, documentation, FAQs, playbooks, and support processes. • Enable Help Desk readiness for setup, initial troubleshooting, common questions, and escalation paths.
Role Description Wir suchen für einen unserer Kunden einen erfahrenen Platform Engineer. Das kleine Team verantwortet das Design und den Betrieb der Container Plattform der ganzen SBB. In Zahlen: Sie haben aktuell 62 Cluster mit total 18'148 CPUs, 80TB Memory und viel Storage (~500TB). Der Automatisierungsgrad ist extrem hoch und sie haben ständig neue spannende Ideen wie sie weiter optimieren könnten. Auch KI hält Einzug im Umfeld und sie sehen viel Potential um in dem Bereich. Es wartet ein Team mit exorbitantem Kubernetes Knowhow. - Du übernimmst zusammen mit dem Team Konzeption, Engineering und Betrieb unserer richtig fetten Container Plattformen auf Basis von Kubernetes/Red Hat OpenShift; - Du verwaltest die Server-Flotte und die Cloud-Infrastruktur nach dem "Infrastructure as Code" Prinzip; - Dabei interagierst Du mit den internen Benutzerinnen, erkennst neue Anforderungen an die Cloud Infrastruktur, erarbeitest neue Lösungen und implementierst diese; - Du kümmerst dich periodisch um Supportanfragen, pflegst Dokumentation und FAQs und verbesserst die Plattform fortlaufend; - Du leistest für die selbst aufgebauten Container Plattform alle sieben Wochen für eine Woche Pikettdienst (SRE Prinzip)! Qualifications - Erfahrung bei der Administration von Kubernetes/OpenShift Clustern und deren Automatisierung - Erfahrung mit Cloud Plattformen (AWS, Azure, OpenStack) - Sehr gute Linux und Netzwerk Troubleshooting Skills (Kabel-Hai) - Hintergrund als Softwareentwicklerin und Kenntnisse in den Programmiersprachen Python/Go wäre ein Pluspunkt. - Fähigkeit, in komplexen und stressigen Situationen den Überblick zu behalten und systematisch an Herausforderungen heran zu gehen - Absoluter Teamplayer der sich mit Herzblut fürs Team und die Plattform einsetzt - Gute Deutsch- und Englischkenntnisse - BizDevOps, SAFe und Scrum/Kanban sind dir geläufig und die Arbeit in einem agilen Team macht dir Spass - Du hast dich mit KI Themen auseinandergesetzt und weisst was ein Modell, Agent, MCP und Skill ist. - Pikett-Dienst 7/24h, d.h. alle 7 Wochen für eine Woche Pikett Company Description
IS Security GRC Platform Engineer
Ochsner HealthOchsner Health is committed to serving, healing, leading, educating, and innovating since 1942. We believe that every award earned, every record broken, and every patient helped is due to the dedicated employees who fill our hallways. At Ochsner, whether you work with patients every day or support those who do, you are making a difference, and that matters.
Role Description The Cybersecurity GRC Engineer supports Ochsner Health’s Cybersecurity Governance, Risk, and Compliance program by serving as a technical and functional resource for the organization’s GRC application, Onspring, and related GRC functions. This role is responsible for: - Administering, maintaining, and improving Onspring workflows, forms, dashboards, reports, risk records, findings, control mappings, assessment processes, evidence repositories, security exception workflows, and user support activities. - Supporting integration and coordination between Onspring and related enterprise applications, reporting tools, workflow systems, and business processes. - Working closely with Information Services, Cybersecurity, Compliance, Privacy, Audit, Legal, third-party risk, and business stakeholders. - Supporting audit readiness, risk management, compliance tracking, remediation management, executive reporting, and continuous improvement of the Cybersecurity GRC program. Qualifications - Required: High school diploma or equivalent. - Preferred: Bachelor's degree in Information Technology, Cybersecurity, Information Systems, Risk Management, Business Administration, Healthcare Administration, or a related field. Requirements - Required: - 2 years information technology experience with master’s degree; OR - 4 years information technology experience with bachelor’s degree; OR - 6 years information technology experience with associate’s degree; OR - 8 years of information technology experience. - Preferred: - 5 to 10 years of related information technology, cybersecurity, governance, risk, compliance, audit, security operations, application administration, or platform administration experience in a regulated industry. - Experience supporting cybersecurity GRC functions in healthcare, federal, government, commercial, or other highly regulated environments. - Experience with GRC and workflow platforms such as Onspring, Archer, MetricStream, Xacta, CSAM, ServiceNow, JIRA, Confluence, Remedy, or similar platforms. Benefits - Competitive salary and benefits package. - Opportunities for professional development and growth. - Supportive work environment focused on making a difference. Company Description Ochsner Health is committed to serving, healing, leading, educating, and innovating since 1942. We believe that every award earned, every record broken, and every patient helped is due to the dedicated employees who fill our hallways. At Ochsner, whether you work with patients every day or support those who do, you are making a difference, and that matters.
Staff AI Platform Engineer
GenesysOrchestrating billions of remarkable experiences in more than 100 countries – through cloud, digital and AI technology.
• Own the end-to-end architecture and evolution of the enterprise AI platform ecosystem, ensuring scalability, security, and performance across tools and integrations • Drive platform strategy decisions through build versus buy analysis, cost modeling, and risk evaluation that guide leadership investment and consolidation priorities • Design and implement self-service platform capabilities that reduce onboarding time and accelerate enterprise-wide adoption of AI tools • Build and standardize internal APIs, SDKs, and infrastructure-as-code solutions that enable consistent, reusable, and scalable integrations • Lead the transition from manual administration to automated, API-driven platform operations to improve efficiency and reduce operational friction • Establish and scale identity and access management, audit logging, and governance frameworks that strengthen security posture and compliance readiness • Translate platform usage, cost, and adoption telemetry into clear, actionable insights that inform procurement, roadmap planning, and enterprise strategy • Lead incident response strategy and implement monitoring and alerting systems that improve reliability and reduce platform disruption



