Unlocking the prosperity of cross-border communities through finance and technology.
Senior Data Engineer
Location
United Kingdom
Posted
6 days ago
Salary
0
Seniority
Senior
Job Description
Senior Data Engineer
Zepz
• Lead the development and management of our data platform. • Support critical data needs and developing foundational data models essential for advancing cross-border remittances business. • Leverage expertise in data warehousing, transformation, and modeling to build robust and scalable data solutions. • Manage all aspects of the build and deployment lifecycle of data movement and transformation processes. • Oversee user access and resource allocation for compute infrastructure, ensuring efficient scaling and utilization. • Ensure effective use of platform tools and resources for efficient ETL/ELT processing. • Provide technical leadership and mentorship to other data engineers within the team. • Evaluate and integrate new data technologies and tools to enhance the capabilities and efficiency of the data platform.
Job Requirements
- At least 7 years of experience in data engineering
- In-depth, hands-on experience with technologies for data ingestion, job orchestration, data warehousing, and reporting.
- Proven ability to design, implement, and manage robust ETL/ELT pipelines using various tools (e.g., Spark, dbt)
- Strong experience with cloud data platforms (e.g., AWS, GCP, Azure), including compute, storage, and database services.
- Proficient in SQL and advanced programming in at least one language (e.g., Python, Scala, Java).
- Experience with workflow orchestration tools (e.g., Astronomer, Apache Airflow)
- Solid understanding and practical experience with data governance principles, data cataloging, and metadata management.
- Demonstrated ability to optimize data platform performance and manage compute resources efficiently.
- Excellent analytical and problem-solving skills with meticulous attention to detail.
- Strong communication and interpersonal skills, capable of articulating complex technical concepts to both technical and non-technical audiences.
Benefits
- unlimited annual leave
- great healthcare benefits
- employee discounts
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
• Supporting and evolving the Tealium Customer Data Platform • Working closely with IT, Product, Marketing, and Analytics teams to deliver reliable customer data integrations, audience activation, and platform performance • Operational platform management — configuring audiences, push notifications, and website personalisation • Technical integration work, including APIs, webhooks, and maintaining data flows across systems
• CTG is seeking a Senior Data Engineer to design, build, and maintain scalable, efficient data pipelines and systems following modern data engineering best practices. • Design, build, and maintain scalable data pipelines, ETL/ELT workflows, and data models using Python, Apache Spark (PySpark), Databricks, dbt, SQL (PostgreSQL), and AWS Glue. • Develop and optimize AWS-native data platforms leveraging AWS Glue, Amazon EMR, Amazon MWAA (Apache Airflow), Lambda, Step Functions, Amazon S3, Redshift, RDS, DMS, and CloudWatch. • Build high-performance ingestion, transformation, and orchestration workflows for structured and semi-structured data using Apache Iceberg, Parquet, ORC, and Avro. • Design and optimize analytical data platforms using Amazon Athena, Trino, Hive, OpenSearch, and enterprise data catalog technologies. • Integrate enterprise and external data sources across relational and NoSQL platforms including PostgreSQL, Oracle, Redshift, GraphDB, and other NoSQL databases. • Build AI-enabled data solutions using Amazon Bedrock, RAG pipelines, and vector search technologies including Amazon S3 Vector and OpenSearch vector indexes. • Develop cloud infrastructure using CloudFormation (Infrastructure as Code), GitHub, Harness, and enterprise CI/CD pipelines while leveraging SNS, SQS, and EventBridge for event-driven architectures. • Improve the reliability, scalability, performance, and maintainability of enterprise data platforms through monitoring, troubleshooting, automation, and continuous optimization. • Support mission-critical analytics and reporting solutions within large-scale AWS-based federal data environments, implementing solutions that comply with FedRAMP and NIST 800-53 security controls. • Lead modernization initiatives migrating legacy platforms including IBM DataStage, Hadoop, RunDeck, and shell-based workflows to cloud-native AWS services. • Mentor junior engineers through technical guidance, architecture discussions, and code reviews while promoting engineering best practices. • Collaborate with cross-functional teams in an Agile environment to define requirements, deliver high-quality data solutions, and communicate technical concepts effectively to technical and non-technical stakeholders.
• Design, build, and maintain scalable data pipelines, ETL/ELT workflows, and data models using Databricks, dbt, Apache Spark (PySpark), SQL (PostgreSQL), and Python. • Develop and optimize AWS-native data solutions leveraging services including AWS Glue, Amazon EMR, Amazon MWAA (Apache Airflow), Lambda, Step Functions, Amazon S3, Redshift, RDS, DMS, and CloudWatch. • Build high-performance data ingestion, transformation, and orchestration workflows across structured and semi-structured data using Parquet, ORC, Avro, and Apache Iceberg. • Integrate data from enterprise and external sources, including relational and NoSQL databases such as PostgreSQL, Oracle, Redshift, GraphDB, and other NoSQL platforms. • Improve the reliability, scalability, performance, and maintainability of enterprise data platforms through monitoring, troubleshooting, and continuous optimization. • Develop and maintain automated deployment pipelines using Harness and collaborate on cloud infrastructure and platform improvements. • Support data engineering efforts powering mission-critical analytics, reporting, and decision-making across large-scale federal data environments. • Collaborate with cross-functional teams in an Agile environment to define requirements, deliver high-quality data solutions, and continuously improve engineering processes.
• Design, build, and maintain scalable data pipelines, ETL/ELT workflows, and data models using Python, Apache Spark (PySpark), Databricks, dbt, SQL (PostgreSQL), and AWS Glue. • Develop and optimize AWS-native data platforms leveraging AWS Glue, Amazon EMR, Amazon MWAA (Apache Airflow), Lambda, Step Functions, Amazon S3, Redshift, RDS, DMS, and CloudWatch. • Build high-performance ingestion, transformation, and orchestration workflows for structured and semi-structured data using Apache Iceberg, Parquet, ORC, and Avro. • Design and optimize analytical data platforms using Amazon Athena, Trino, Hive, OpenSearch, and enterprise data catalog technologies. • Integrate enterprise and external data sources across relational and NoSQL platforms including PostgreSQL, Oracle, Redshift, GraphDB, and other NoSQL databases. • Build AI-enabled data solutions using Amazon Bedrock, RAG pipelines, and vector search technologies including Amazon S3 Vector and OpenSearch vector indexes. • Develop cloud infrastructure using CloudFormation (Infrastructure as Code), GitHub, Harness, and enterprise CI/CD pipelines while leveraging SNS, SQS, and EventBridge for event-driven architectures. • Improve the reliability, scalability, performance, and maintainability of enterprise data platforms through monitoring, troubleshooting, automation, and continuous optimization. • Support mission-critical analytics and reporting solutions within large-scale AWS-based federal data environments, implementing solutions that comply with FedRAMP and NIST 800-53 security controls. • Lead modernization initiatives migrating legacy platforms including IBM DataStage, Hadoop, RunDeck, and shell-based workflows to cloud-native AWS services. • Mentor junior engineers through technical guidance, architecture discussions, and code reviews while promoting engineering best practices. • Collaborate with cross-functional teams in an Agile environment to define requirements, deliver high-quality data solutions, and communicate technical concepts effectively to technical and non-technical stakeholders.


