Inspiring people through data.
Senior Data Engineer (Focus on Data Quality)
Location
Worldwide
Posted
9 days ago
Salary
0
Seniority
Senior
Job Description
Senior Data Engineer (Focus on Data Quality)
Five Acts
Role Description Estamos em busca de uma pessoa Engenheira de Dados Sênior com foco em Qualidade de Dados para atuar na construção, validação e evolução de processos que garantam a confiabilidade dos dados entregues às áreas de negócio. Essa pessoa será responsável por desenvolver mecanismos de qualidade de dados, realizar testes, validações e homologações, além de contribuir para a evolução das práticas de engenharia de dados, assegurando que as informações consumidas por analistas, cientistas de dados, sistemas, dashboards e demais consumidores sejam consistentes, confiáveis e de alto valor para o negócio. - Desenvolver e implementar processos de qualidade de dados ao longo do ciclo de engenharia de dados. - Documentar, testar, validar e homologar pipelines e soluções desenvolvidas. - Garantir a qualidade, consistência e confiabilidade dos dados entregues aos consumidores finais. - Desenvolver soluções utilizando Python, SQL e PySpark em ambiente Databricks. - Construir e manter notebooks e pipelines de processamento de dados. - Participar da análise de requisitos junto às áreas de negócio. - Realizar testes unitários e apoiar os processos de homologação das soluções. - Pesquisar, desenhar e desenvolver novas soluções de engenharia de dados. - Atuar continuamente na melhoria da confiabilidade, eficiência e qualidade dos processos e dos dados. - Disseminar conhecimento técnico e boas práticas de engenharia e qualidade de dados para o time. - Contribuir para que os dados gerem valor para as áreas de negócio e apoiem a tomada de decisão. Qualifications - Tenha pensamento Lean e mentalidade de melhoria contínua. - Seja colaborativa e engajada com as entregas da squad. - Atue com senso de urgência em situações críticas. - Tenha facilidade para trabalhar com experimentação e aprendizado contínuo. - Encare erros como oportunidades de evolução, propondo melhorias de forma constante. - Seja protagonista na resolução de problemas, apoiando o time na busca por soluções. - Valorize a troca de conhecimento e a cultura de feedback. Requirements - Experiência mínima de 4 anos atuando com Engenharia de Dados. - Conhecimento sólido em Python e SQL para transformação, análise e tratamento de dados. - Experiência com PySpark. - Experiência em Databricks, incluindo desenvolvimento de notebooks em Python e PySpark. - Vivência com práticas de testes, validação e homologação de soluções de dados. - Experiência na análise de requisitos e desenvolvimento de soluções de engenharia de dados. - Conhecimento em ferramentas e práticas de DevOps, incluindo: Git, Docker, Jira, Bitbucket, Bamboo, Nexus, Microsserviços. - Compromisso com a qualidade das entregas e boas práticas de desenvolvimento. Benefits - Vales Alimentação e Refeição (Swile) - Flexibilidade para crédito em Auxílio Home-Office (Swile) - Cobertura de até 100% em Plano de Saúde e Odontológico - Seguro de Vida em grupo - Trabalho remoto - Convênio Saúde Mental - psicoterapia online e presencial - Incentivo a certificações e cursos - Convênio para cursos de pós-graduação e MBA (Esalq/USP) - Parceria com escolas de idiomas - Parceria com academias e apps de bem-estar (Wellhub) - Palestras e rodas de conversa internas - Bônus por indicação - Happy hours - Mimos em datas comemorativas
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Senior Data Engineer
TripadvisorTripadvisor, founded in 2000, is an award-winning network for travel information that features real advice from global travelers. The world’s largest travel s
• Providing the organization’s data consumers high quality data sets by data curation, consolidation, and manipulation from a wide variety of large scale (terabyte and growing) sources. • Building high quality data pipelines and ETL processes that interact with terabytes of data on leading platforms such as Snowflake and BigQuery. • Developing and improving our enterprise data marts by creating efficient and scalable data models to be used across the organization. • Partnering with our analytics, data science, crm, and machine learning teams for data solutions. • Responsible for an enterprise data mart integrity, validation, and documentation. • Responsible for the data pipelines’ SLA and dependency management. • Writing technical documentation for data solutions, and presenting at design reviews • Solving data pipeline failure events and implementing sound anomaly detection • Working with various teams from analytics to product owners and front end developers on tracking solutions and solving technical challenges • Leading medium to large size projects in terms of writing technical specs and project planning • Mentoring junior members of the team.
Business Informatics Support Data Engineer
TRILLIUM HEALTH RESOURCESTrillium Health Resources is a Tailored Plan and Managed Care Organization (MCO) serving 46 counties across North Carolina. We manage services for individuals with serious mental health needs, substance use disorders, traumatic brain injuries, and intellectual/development (IDD) disabilities. Our mission is to help individuals and families build strong foundations for healthy, fulfilling lives.
Role Description Trillium Health Resources has a career opening for a Business Informatics Support Data Engineer to join our team! The Business Informatics Support Data Engineer is a highly skilled, collaborative, and technical position responsible for the development and maintenance of reporting systems for Trillium. This role is dedicated to supporting business operations through the development of scalable data solutions using SQL, SSRS, and Power BI. This role will be a primary in the design and support of the Data Warehouse. In addition to technical development, this position will be responsible for maintaining up-to-date documentation in the team wiki and providing mentorship and support to other Business Informatics and Business reporting team members, as needed. On a typical day, you might: - Participate in the design and support of the Data Warehouse. - Collaborate with business units to understand data needs and translate them into effective reporting solutions by gathering requirements and delivering high-quality data solutions that support operational and strategic goals. - Develop, maintain, and optimize SSRS reports and Power BI dashboards to provide actionable insights. - Write and troubleshoot SQL queries to extract, transform, and analyze data from various sources. - Ensure data accuracy, consistency, and integrity across all reporting platforms. - Write and optimize SQL queries for data extraction, transformation, and analysis. Qualifications - Associate’s degree and four (4) years of experience in any of the following areas: Information Technology/MIS, Mathematics (Actuarial/Statistics), Data Analytics, Engineering Sciences, Business, Healthcare Administration, Human Service field, healthcare claims environment and reporting development, n-tier and web-based system development and support with strong technical knowledge in the specialized areas of application system programming, including software and tools including; SQL Server 2016 or above, SQL, My SQL, Azure DevOps, Source Control, SSRS, SSIS, SSMS, Visual Studio 2016 or above, C# and Microsoft.NET Framework 4.0 or above. Requires certification. Applicable certification(s) may be substituted to equivalent degree and experience requirements; - OR Equivalent combination of education/experience. - Two-year degrees require certification. - Must have a valid driver’s license. - Must reside within the United States. Requirements - Degree in Information Technology/MIS, Mathematics (Actuarial/Statistics), Data Analytics, Engineering Sciences, Business, Healthcare Administration, or Human Service field (preferred). - Recent experience with SQL database management and development (i.e. Power BI, Analytical, R, Python, Visual Studio, and/or SSRS) (preferred). - Knowledge or experience with one or more of the following technologies: NoSQL, Snowflake, Databricks, Mongo, or Hadoop (preferred). - Applicable certification(s) including Microsoft data systems certifications, CSTP, ISTQB, ASTQB, MTA, MCSA, MCSD, MCSE, ITIL v3, Power BI, as well as INFORMS, IIBA, AWS, Azure, or equivalent certifications will be accepted (preferred). Benefits - Typical working hours: 8:30 am – 5:00 pm; flexible work schedules with some roles with management approval. - Work-from-home options available for most positions. - Health Insurance with no premium for employee coverage. - Flexible Spending Accounts. - 24 days of Paid Time Off (PTO) plus 12 paid holidays in your first year. - NC Local Government Retirement Pension (defined-benefit plan). - 401k with 5% employer match and immediate vesting. - Public Service Loan Forgiveness (PSLF) qualifying employer. - Quarterly stipend for remote work supplies.
Senior Platform Product Manager, Data & AI
CaseWorthy, Inc.Better Data & Client Management For Whole Person Care
• Lead the Platform Pod. • Own the Cara roadmap across all phases. • Design the Assistants and write the specs. • Drive the infrastructure conversation in partnership. • Own the MCP and API contract standards. • Build design patterns that scale. • Partner across engineering and prioritize in the open. • Support go-to-market efforts.
Role Description We're building out data as a first-class practice at YLD, and we need someone in London who can be the credible technical voice in front of clients and prospects, not just describing what we could do, but backing it with real delivery experience. This is a senior role, reporting to our Head of Engineering. You'll sit alongside our commercial and client services leadership in meetings where technical credibility matters, help shape and win new engagements, and stay close enough to client work to guide it and step in yourself when an engagement needs you. Alongside, we expect you to build our external reputation in data through events, writing, and the wider community. Responsibilities - Hands-on delivery (core, from day one): - Doing hands-on data engineering work directly, pipelines, transformation layers, models, and the infrastructure around them, not just guiding others - Adapting to different client contexts: legacy warehouse migrations, greenfield lakehouses, transformation layers that need rescuing - Guiding the wider team and unblocking hard technical problems as your remit grows beyond your own delivery - Defining and evolving our internal standards for data work: testing, documentation, project structure, code review - Mentoring and growing data engineers across the company - Client and commercial (core, from day one): - Joining client and prospect meetings as the technical authority on data, alongside our client services and commercial leads - Shaping proposals and pitches, translating client problems into a credible data approach and a defensible scope - Acting as the escalation point on live data engagements when technical judgement calls need weight behind them - Travelling to client sites for workshops, discovery sessions, and in-person relationship building, and representing YLD at external data events - Practice, talent, and external profile (growth, over the first year): - Owning the growth of YLD's data practice: capability, reputation, and headcount over time - Contributing to attracting and hiring strong data engineers into YLD - Representing YLD at data events and in the wider community as your standing in the market builds - Supporting our sales team so that data becomes something we can proactively sell, not just respond to Qualifications - Consulting or agency experience is a strong plus, but real client or stakeholder-facing exposure elsewhere is acceptable - A track record of technical leadership: leading a team or function - Comfortable presenting to senior client stakeholders and holding your own under commercial pressure - Cross-functional fluency: you translate engineering constraints into business terms and negotiate realistic commitments - Appetite to build a public presence in the data community over time, even if you don't have one yet Requirements - Not theoretical, evidenced by things you've actually shipped, and deep enough to set the bar for a team and answer a sceptical client: - SQL as engineering: read and review SQL that performs at scale, and understand query planning, engine quirks, and how materialisation choices affect cost and performance well enough to guide a team's decisions - Python for data: built and maintained production data systems in typed, testable Python, not just notebooks, and can set that standard for others - Data modelling: made deliberate choices between dimensional, Data Vault, normalised and denormalised designs, and can explain the trade-offs in flexibility, query performance, and maintainability to both engineers and clients - Transformation architecture: design transformations that are idempotent, incremental, and dependency-aware. You think in DAGs, not scripts - Pipeline design: weigh batch, streaming, and micro-batch trade-offs against latency, complexity, cost, and reprocessability, and pick the right approach for the problem - Data testing: know what to catch at build time (schema contracts, assertions, transformation logic) versus defer to observability, and can make that call for a team - Data observability: treat data reliability like site reliability, with measurable indicators, alerting, incident response, and root cause analysis - CI/CD for data: version, test, and deploy pipelines like software, with environment promotion, rollback strategies, and infrastructure as code - Data governance: implemented lineage, cataloguing, sensitive data classification, or access control, and can hold a client to account on making data auditable and secure - Cost-conscious: optimised warehouse spend, storage strategies, or job efficiency, and treat compute as a resource to manage, not ignore - Platform fluency: worked across orchestration (Airflow, Dagster), warehouses (Snowflake, BigQuery, Databricks), ingestion (Fivetran, Airbyte, custom), and transformation (dbt, Spark) - AI Native: experienced in agentic coding workflows (context, harness, loop, and graph engineering), and able to embed AI Engineering constructs into pipelines (RAG, semantic search, evals, guardrails, etc) Benefits - Company Private Health care (currently provided by Vitality) - Enhanced fully paid maternity and paternity leave for up to 6 months - Company's Pension Scheme - 25 days annual holiday (excluding Public Holidays) - £2000 annual allowance for Training/Conferences - £300 annual allowance for additional hardware - Mental Health support - Bonus (depending on Company performance and results) - Company laptop - Generous referral schemes



