Job Closed

This listing is no longer active.

Cogniify

We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other protected characteristic.

Senior Data Engineer / Analyst

Location

United States

Posted

96 days ago

Salary

$121.7K - $162.2K / year

Seniority

Senior

Job Description

Senior Data Engineer / Analyst

Cogniify

Role Description We’re looking for an experienced Senior Data Engineer/Analyst to lead the design and delivery of production-grade data platforms, pipelines, and analytics solutions that drive business intelligence and AI/ML capabilities across the organization. In this role, you will own critical data workstreams end to end, make key architectural decisions on data modeling, pipeline design, and platform selection, and collaborate with clients and stakeholders to translate business requirements into scalable data solutions. You will champion modern data stack practices using Snowflake, Databricks, dbt, and cloud-native services, while mentoring other engineers and driving the maturity of our data engineering, analytics, and DataOps capabilities. What You’ll Do - Lead the design, development, and optimization of enterprise-grade data pipelines and transformation layers using dbt, Apache Spark, Airflow, Dagster, and cloud-native orchestration services. - Architect and manage data platforms on Snowflake and/or Databricks, including warehouse design, lakehouse architecture, compute optimization, access governance, and cost management. - Define and enforce data modeling standards across the organization using Kimball dimensional modeling, Data Vault 2.0, Activity Schema, or hybrid approaches. - Design and implement end-to-end data ingestion strategies from diverse sources: transactional databases (CDC via Debezium, Fivetran), APIs, event streams (Kafka, Kinesis, Spark Structured Streaming), SaaS platforms, and unstructured data sources. - Build and maintain curated data products, metrics layers, and semantic models that enable self-service analytics across the organization. - Implement comprehensive data quality, observability, and lineage frameworks using dbt tests, Great Expectations, Monte Carlo, Datafold, Elementary, or Soda. - Collaborate with clients and business stakeholders to gather requirements, translate them into data architecture decisions, and drive technical strategy from ideation to production. - Lead the adoption of DataOps practices including CI/CD for data pipelines (GitHub Actions, dbt Cloud, Databricks Asset Bundles), automated testing, and environment promotion workflows. - Design and optimize analytics solutions and advanced dashboards using Tableau, Looker, Power BI, or Sigma Computing, ensuring performance and usability at scale. - Support AI/ML initiatives by designing and maintaining feature stores, training datasets, and model input/output data pipelines in collaboration with ML engineers. - Mentor junior and mid-level data engineers and analysts, conduct architecture and code reviews, and drive knowledge sharing across the team. - Drive data governance initiatives including cataloging (Alation, Atlan, DataHub, Unity Catalog), lineage tracking, PII management, and regulatory compliance (GDPR, CCPA, HIPAA). Qualifications - Bachelor’s or Master’s degree in Computer Science, Data Science, Statistics, Mathematics, or a related field. - 5–8+ years of professional experience in data engineering, analytics engineering, or a closely related role with significant production delivery. - Deep expertise in SQL and advanced data transformation techniques across analytical databases and data warehouses. - Extensive hands-on experience with Snowflake and/or Databricks in production environments, including architecture design, performance tuning, and cost optimization. - Strong experience with dbt (data build tool) for production-grade transformation, testing, documentation, and CI/CD workflows. - Proficiency in Python (Pandas, PySpark, Polars) and Apache Spark for large-scale data processing. - Experience designing and operating workflow orchestration using Airflow, Dagster, Prefect, or cloud-native equivalents. - Strong knowledge of data ingestion patterns and tools: Fivetran, Airbyte, CDC (Debezium), streaming (Kafka, Kinesis, Flink). - Experience with data visualization and BI platforms: Tableau, Looker, Power BI, or Sigma at enterprise scale. - Proven ability to work directly with clients and stakeholders and translate ambiguous business requirements into concrete data solutions. - Demonstrated ability to mentor engineers, lead design reviews, and influence data architecture direction. - Strong understanding of data governance, cataloging, data quality, and compliance frameworks. Preferred Qualifications - Snowflake SnowPro Advanced or Architect, Databricks Data Engineer Professional, or AWS Data Analytics Specialty certification. - Experience with lakehouse architectures using Delta Lake, Apache Iceberg, or Apache Hudi. - Familiarity with real-time analytics and streaming architectures: Kafka, Spark Structured Streaming, Flink, Materialize, or ksqlDB. - Experience with metrics/semantic layers at scale: dbt Semantic Layer, Cube, MetricFlow, or LookML. - Experience designing data platforms that serve AI/ML workloads, including feature stores (Feast, Tecton), vector databases (Pinecone, Weaviate), and RAG data pipelines. - Familiarity with data mesh or data product architecture patterns in large organizations. - Experience with cloud data infrastructure across AWS (S3, Glue, Athena, Lake Formation, Redshift Serverless), Azure (ADLS, Synapse, Fabric, Purview), or GCP (BigQuery, Dataflow, Dataplex). - Knowledge of cost optimization and FinOps for data platforms (Snowflake credit management, Databricks cluster policies, spot instances). Benefits - Unlimited PTO. - Very generous parental leave, much above industry standards. - Entrepreneurial culture where pushing limits and taking risks is everyday business. - Open communication with management and company leadership. - Small, dynamic teams = massive impact. - Medical, Dental and Vision coverage for employees. - Access to Disability & Life insurance. - Mental health and wellbeing support. - Annual bonus program. - Employer Stock Purchase Program (ESPP). - Yearly Team building experiences. - Mentorship and sponsorship opportunities. - Manager resources and support. Salary Range US East/West Coast: $121,700 - $162,200 Disclaimer: These salary ranges are estimates based on market data. Actual compensation may vary depending on factors including work experience, education, skills, and specific geographic location. Company Description We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other protected characteristic.

Related Categories

Related Job Pages

More Data Engineer Jobs

Verisian logo

Data Engineer

Verisian

Accelerate drug time to market through real study traceability and unparalleled trial integrity

Data Engineer96 days ago
Full TimeRemoteTeam 1-10H1B No Sponsor

About Us At Verisian, we build deep tech and AI solutions that enable groundbreaking medical therapies to enter the market faster and safely. New medicines and devices are judged using all available evidence in the proper context and backed by automated analyses that validate and make data and results transparent to companies and regulators alike. Regulators will make the right decisions on novel therapies faster and with greater confidence, protecting patients from harm and making breakthrough treatments available as soon as is possible and safe. In times of increasing trial numbers and complexity, we are building technical innovation to remove crucial bottlenecks for pharmaceutical and medical device companies, as well as public health authorities, directly affecting clinical trial planning, analyses, validation, submission, and market approval. By joining us, you will commit yourself to building software and AI tools that directly contribute to increasing the rate at which medical innovation improves human health and wellbeing. Disclaimer: We welcome candidates of diverse experience levels, and meeting every requirement is not mandatory. Research indicates that underrepresented groups tend to apply only if they meet all qualifications. If you're enthusiastic about the role, please apply and let our recruiters evaluate your application. Culture At Verisian, our mission is to build the future infrastructure of medical innovation. To help us succeed, we're creating a unique employee culture. We always put the mission first. We're fanatically customer obsessed, crafting world-class products that customers love with every interaction. We take extreme ownership and accountability of our work, seeing whatever we do through to completion. We communicate candidly and directly with each other, even when it's uncomfortable. We're innately curious, open to alternative perspectives and invest passionately in our own continuous growth. Role Description As a Data Engineer, you will join our world-class engineering team in building the Verisian Platform. You will work on an application that exposes our clinical trial insights to data managers, statistical programmers, statisticians, medical experts/writers, and regulatory authorities.  The Verisian Platform brings value to a set of highly regulated processes crucial for medical progress and innovation. Your work will support the Planner, Builder, Explorer, Validator, Submitter, and related supporting modules. These core modules of the platform target the planning, exploration/onboarding, building, validation, submission, and review of clinical trials and their results. They enable data managers, statistical programmers, statisticians, medical writers, and regulators to deliver their work faster, at higher quality, lower cost, and in greater confidence. Our pipelines analyze clinical trial documentation, code, logs, data, and results to build a knowledge graph through code traceability. We harness the resulting dataset, column-level and logic lineage to turn clinical trials into Information Infrastructure that can be used by experts and consumed by AI to revolutionize how therapies are evaluated and enter the market. We capture complex processes in fully- and semi-automated workflows that place experts in control and AI automation at their fingertips. We build visualizations to provide our customers with a maximum of insight as fast as possible. Our application stack is based on Next.js and deployed via Docker/Kubernetes in the cloud. The data analysis pipelines run in Argo Workflows. We analyze code based on Antlr4 and Java. AI agents are developed in Python. The data analysis engine is developed in Python. Git is where our code lives, and Github Actions is how it gets out into the world. In tandem, you will create data validation rules (custom DSL) and develop our data analysis engine (Python) that is used to automatically detect inconsistencies and analysis errors in data while enforcing regulatory data standard adherence. These crucial pipelines are an integral part of our platform to expose our game-changing functionality to users and consumed by our AI agents to automate the planning, analysis, validation, and submission of clinical trials. You will be expected to lead the analysis, design, building and testing of components of the engine and data validation rules. As part of our core team, you will join us in designing, prioritising, building and testing new functionality, troubleshooting customer issues, finding root causes, and deploying required fixes to ensure maximal user impact and performance.

United Kingdom
Job Closed
Full TimeRemoteTeam 501-1,000

Role Description Are you passionate about leading a team of skilled engineers in a fast-paced, innovative FinTech environment? Green Street is seeking an experienced Lead Data Engineer to manage and mentor a team dedicated to designing and building robust data pipelines and architecture for the Commercial Real Estate (CRE) industry. You’ll oversee the development of advanced data solutions, integrating both public and proprietary data from diverse sources to create comprehensive data products and insights for our clients and internal research teams. Why join Green Street? We’re a highly collaborative, Agile-driven team committed to engineering excellence, leveraging the latest technologies, and continuously optimizing our data infrastructure. As a leader on our team, you’ll have the opportunity to guide data engineers and QA engineers, shape Green Street's data solutions, and work closely with cross-functional teams to deliver meaningful impact within the CRE space. We love great engineers, and we are excited to get to know you better. Please share your resume today! Responsibilities - Lead and manage a team of data engineers and data QA engineers, fostering a collaborative and high-performance culture - Design and develop robust data ingestion, ETL processes, and ensure data integrity and high availability - Document and manage data models, schema designs, and ER diagrams, ensuring a well-architected and scalable data structure - Oversee the development and maintenance of data architecture, optimizing for performance and reliability - Collaborate with research and product teams to enhance our analytical capabilities and streamline workflows - Conduct analysis on large, aggregated datasets (e.g., demographic, geographic, GIS/spatial) - Coordinate effectively with both onshore and offshore development teams, ensuring alignment and meeting project objectives Qualifications - 10+ years of experience in data engineering, with at least 3 years in a leadership role managing engineering teams - Advanced proficiency in Python (8+ years) and experience with relational databases (MySQL, PostgreSQL, etc.) - Prior experience with Git and Agile methodologies - Strong expertise in data architecture, ETL processes, and data pipeline management - Proven ability to design and optimize database schemas for complex data domains - Experience with cloud platforms, particularly AWS (Amazon RDS, Containers) - Willingness and desire to keep up with new cloud technologies, languages, standards, and practices - Strong problem-solving skills, excellent communication abilities, attention to detail, and a commitment to continuous learning and improvement - Ability to work independently in a fast-paced, agile environment - Ability to collaborate effectively with a remote team - Ability to overlap 3-4 hours with the U.S. West Coast to participate in team discussions and lead cross-time-zone collaboration Nice-to-Haves - Prior knowledge of finance, Real Estate, mathematics, GIS / spatial data is a plus - Experience with data quality best practices and automated data validation processes

PST (UTC-8)
Job Closed
Dotdigital logo

Principal Data Engineer

Dotdigital

Go beyond the expected.

Data Engineer96 days ago
Full TimeRemoteTeam 201-500Since 1999H1B No Sponsor

• Lead the design and implementation of scalable, secure and resilient data systems across streaming, batch and real-time use cases. • Architect data pipelines, model and storage solutions that power analytical and product use cases; using primarily Python and SQL via orchestration tooling that run workloads in the cloud. • Leverage AI to automate both data processing and engineering processes. • Assure and drive best practices relating to data infrastructure, governance, security and observability. • Work with technologists across multiple teams to deliver coherent features and data outcomes. • Support the data team to help adopt data engineering principles. • Identify, validate and promote new tools and technologies that improve the performance and stability of data services.

South Africa
Job Closed
GFT Technologies SE logo

Senior Data Platform Engineer (m/w/d) Senior Data Platform Engineer (m/w/d)

GFT Technologies SE

Procuramos uma pessoa que: Goste de trabalhar em equipe e seja colaborativa em suas atribuições; Tenha coragem para se desafiar e ir além, abraçando novas oportunidades de crescimento; Transforme ideias em soluções criativas e busque qualidade em toda sua rotina; Tenha habilidades de resolução de problemas; Possua habilidade e se sinta confortável para trabalhar de forma independente e gerenciar o próprio tempo; Tenha interesse em lidar com situações adversas e inovadoras no âmbito tecnológico. Big enough to deliver – small enough to care. #VempraGFT #VamosVoarJuntos #ProudToBeGFT

Data Engineer96 days ago
Full TimeRemoteTeam 10,001

GFT Technologies ist ein verantwortungsvolles, KI-zentriertes globales Unternehmen im Bereich der digitalen Transformation. Wir konzipieren fortschrittliche Lösungen für die Daten- und KI-Transformation, modernisieren Technologie-Infrastrukturen und entwickeln Kernsysteme der nächsten Generation für führende Banken, Versicherungen, Industrie- und Robotik-Unternehmen. In enger Zusammenarbeit mit unseren Kunden verschieben wir Grenzen, um ihr volles Potenzial auszuschöpfen. Mit fundierter Branchenexpertise, modernsten Technologien und einem starken Partnernetzwerk bietet GFT verantwortungsvolle, KI-zentrierte Lösungen, die technologische Exzellenz mit hoher Liefer- und Kosteneffizienz vereinen. Das macht uns zu einem verlässlichen Partner für nachhaltigen Geschäftserfolg. Mit über 12.000 Technologie-Expertinnen und -Experten sind wir in mehr als 20 Ländern weltweit tätig und bieten Karrieremöglichkeiten im Bereich führender Software-Innovationen. Die GFT Technologies SE (GFT-XE) ist im SDAX der Deutschen Börse notiert. Let’s Go Beyond! Im Center of Excellence Data gestalten wir das Datenfundament für die KI-Transformation unserer Kunden. Als Senior Data Platform Engineer übernimmst du eine fachlich führende Rolle in Transformationsprojekten – mit Schwerpunkt auf regulierten Branchen, insbesondere Financial Services. Du verbindest strategisches Governance-Design mit operativer Umsetzung und sorgst dafür, dass Datenplattformen nicht nur leistungsfähig, sondern auch regulatorisch compliant, auditierbar und AI-ready sind. Wir setzen alles daran, uns gemeinsam mit dir weiterzuentwickeln, um noch besser zu werden! Unser motiviertes Team freut sich auf dich und deine Erfahrung als Senior Data Platform Engineer (m/w/d)! Deine Aufgaben - Du designst und implementierst skalierbare Datenplattformen in Cloud-Umgebungen (Azure, AWS oder GCP). - Du entwickelst moderne ELT/ETL- und Streaming-Pipelines für Batch- und Near-Real-Time-Verarbeitung. - Du setzt Lakehouse-Architekturen unter Nutzung offener Storage- und Table-Formate um. - Du modellierst Datenstrukturen nach Data Vault 2.0 oder Kimball – abhängig vom fachlichen und architektonischen Kontext. - Du baust konsumfertige Datenprodukte für Analytics-, Reporting- und KI-Anwendungsfälle auf. - Du bereitest Daten für Machine-Learning-Use-Cases vor (inkl. Feature-Engineering-Logiken), ohne selbst ML-Modelle zu trainieren. - Du stellst Datenqualität, Performance und Kostenoptimierung in der Cloud sicher. - Du implementierst Orchestrierung, Automatisierung und CI/CD-Prozesse für stabile, produktionsreife Plattformen. - Du arbeitest eng mit Data Architects, Governance-Teams und Data Scientists zusammen. - Du übernimmst technische Verantwortung innerhalb konkreter Umsetzungsprojekte. Das bringst du mit - Mindestens 3–5 Jahre Erfahrung im Aufbau moderner Datenplattformen - Fundierte Kenntnisse in mindestens einer Cloud-Plattform (Azure, AWS oder GCP) - Sehr gute SQL-Kenntnisse sowie Erfahrung mit Python (oder vergleichbar) - Erfahrung mit Streaming-Technologien (z. B. Kafka, Pub/Sub, Kinesis oder vergleichbar) - Praxis in der Datenmodellierung nach Data Vault 2.0 und/oder Kimball - Verständnis moderner Lakehouse-Architekturen und offener Speicherformate - Erfahrung mit Orchestrierung und Automatisierung von Datenpipelines - Know-how in Performance- und Kostenoptimierung in Cloud-Umgebungen - Erfahrung in der Zusammenarbeit mit Analytics- und Data-Science-Teams - Sehr gute Deutsch- und Englischkenntnisse. Das bieten wir dir - Flexible Arbeitszeiten: Um Familie und Beruf optimal zu vereinbaren, kannst du deinen Arbeitstag nach deinen individuellen Bedürfnissen gestalten. Profitiere darüber hinaus von individuellen Modellen, Workation und Sabbaticals. - Homeoffice: Egal, ob aus dem Büro oder von einem anderen Ort – mobiles Arbeiten gehört für uns zum Alltag. - Mindset: Open Door, Teamspirit und flache Hierarchien sind im #teamGFT keine Buzzwords, sondern gelebte Praxis. - 12.000 Talente weltweit: Profitiere von dem globalen Austausch mit Experten aus über 20 Ländern auf deinem Gebiet. - Weiterbildung & Zertifizierungen: Nimm an Fortbildungen, Konferenzen und Zertifizierungen teil. Wir gehen auf deine individuellen Bedürfnisse ein. - Standortbezogene Extras: Profitiere von weiteren Zusatzleistungen, wie Job Rad, Betrieblicher Altersvorsorge und vielem mehr. - Neueste Technologien: Durch die Arbeit mit international führenden Konzernen und den Einsatz interdisziplinärer Teams arbeiten wir am Puls der Zeit und setzen uns ständig mit den neuesten Methoden und Technologien auseinander.

Germany