The go-to marketplace for the Everyday Worker. We help employers meet job seekers where they are.
Principal Data Engineer
Location
United States
Posted
2 days ago
Salary
0
Seniority
Lead
Job Description
Principal Data Engineer
JobGet
• Lead the technical direction for how data flows, scales, and powers decisions at JobGet • Own the data architecture behind the platform • Drive architectural decisions • Champion modern data stack adoption • Ensure platform reliability • Build production-grade pipelines • Establish data modeling standards • Solve complex data integration and performance challenges • Architect and evolve streaming data infrastructure • Enable machine learning at scale • Establish data governance standards • Implement data validation frameworks • Raise the technical bar through mentoring
Job Requirements
- 10+ years of data engineering experience
- 3–5 years operating at a principal or staff IC level in a start-up environment
- Production-grade experience with Snowflake and dbt
- Experience with real-time streaming technologies such as KSQL or Apache Flink is strongly preferred
- Hands-on AWS experience at scale
- DataOps experience: familiarity with CI/CD for data pipelines
- Significant, demonstrated AI usage in your day-to-day engineering work
- Expert-level proficiency in Python and SQL
- Strong data modeling background
- Experience building data infrastructure that supports ML/AI systems
- Demonstrated track record of setting technical standards across teams
- Strong grasp of data governance fundamentals
- Exceptional communication and stakeholder management skills
Benefits
- Purpose-driven organization
- Flexible time off
- Remote-first
- Flexible work hours - our employees are in multiple time zones
- Medical & dental plans
- Parental leave
- Employee stock options
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
• Collaborates with cross-functional teams to understand data and analytical requirements and translates them into effective Microsoft Fabric–based data solutions. • Mentors and provides technical leadership to Data Engineers through code reviews, design reviews, knowledge-sharing sessions, and engineering best practices. • Establishes and maintains data engineering standards, reusable frameworks, development patterns, and operational best practices to improve consistency and scalability across the data platform. • Leads root cause analysis and resolution efforts for critical production issues impacting enterprise data pipelines, reporting solutions, and platform operations. • Reviews and provides guidance on solution designs, architecture decisions, and implementation approaches to ensure alignment with enterprise standards and best practices. • Designs, develops, and maintains end-to-end data ingestion and transformation processes using Microsoft Fabric OneLake, Lakehouses, Data Pipelines, Dataflows Gen2, and Notebooks. • Formats, cleanses, and stores data in a structured manner to ensure data quality and accessibility for reporting and analysis, following Medallion Architecture. • Works closely with reporting and analytics teams to ensure datasets are accurate, well-modeled, and optimized for Power BI and other analytical workloads. • Creates comprehensive documentation for data pipelines, processes, and solutions. • Participates in data governance activities to ensure data integrity, security, and compliance. • Utilizes source control to manage and track changes to data pipelines and codebase. • Provides level II/III support to troubleshoot and resolve data-related issues and inquiries. • Provides ancillary support for Machine Learning (ML) and Artificial Intelligence (AI) processes. • Evaluates emerging technologies, industry trends, and architectural patterns related to data engineering, analytics, artificial intelligence, and cloud platforms and provides recommendations for adoption. • Manages and optimizes Microsoft Fabric platform performance by monitoring capacity utilization, query performance, Lakehouse and Warehouse workloads, Delta table maintenance (OPTIMIZE/VACUUM), data partitioning, storage consumption, and pipeline execution to ensure scalable, reliable, and cost-effective data operations. • Develops and maintains monitoring, alerting, and observability solutions for Microsoft Fabric pipelines, notebooks, Lakehouses, Warehouses, semantic models, and related data platform components to ensure operational reliability and rapid issue resolution. • Performs other duties as assigned.
• Collaborate with cross-functional teams to understand data and analytical requirements and translate them into effective Microsoft Fabric–based data solutions. • Design, develop and maintain end-to-end data ingestion and transformation processes using Microsoft Fabric Data Pipelines, Dataflows Gen2, and Notebooks. • Format, cleanse, and store data in a structured manner to ensure data quality and accessibility for reporting and analysis, following Medallion Architecture • Work closely with reporting and analytics teams to ensure datasets are accurate, well-modeled, and optimized for Power BI and other analytical workloads. • Create comprehensive developer documentation for data pipelines, processes, and solutions. • Participate in data governance activities to ensure data integrity, security, and compliance. • Utilize source control to manage and track changes to data pipelines and codebase. • Provide level II/III support to troubleshoot and resolve data-related issues and inquiries. • Provide ancillary support for Machine learning (ML) and Artificial Intelligence (AI) processes • Stay updated with industry trends and best practices related to data warehousing, ETL, and data engineering.
Data & Analytics Engineer
Eltropy Inc.Eltropy is on a mission to disrupt the way people access financial services. Eltropy enables financial institutions to digitally engage in a secure and compliant way. Using our world-class digital communications platform, community financial institutions can improve operations, engagement, and productivity. CFIs (Community Banks and Credit Unions) use Eltropy to communicate with consumers via Text, Video, Secure Chat, co-browsing, screen sharing, and chatbot technology — all integrated in a single platform bolstered by AI, skill-based routing, and other contact center capabilities. Customers are our North Star No Fear - Tell the truth Team of Owners Eltropy is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.
Role Description Eltropy is a digital conversations platform for credit unions and community financial institutions in the US. The Data Engineering & Analytics team builds the AWS data pipelines and customer-facing dashboards that power analytics across the platform. We are looking for a Data & Analytics Engineer with 3-4 years of experience who can own dashboard delivery end to end along with the pipelines behind it. The ideal candidate learns fast, builds product context quickly, listens well, and collaborates effectively across product, engineering, DevOps, and customer-facing teams. Key Responsibilities - Own dashboard changes end to end in QuickSight and ThoughtSpot - new metrics and filters, SPICE refresh management, internal-to-production promotion, and post-release validation. - Build and maintain batch and streaming ETL pipelines on AWS using Glue (PySpark), S3, Redshift, and Airflow (MWAA) DAGs. - Write and optimize Redshift SQL; debug query performance, connection contention, and data mismatches across sources. - Support near-real-time ingestion (Kafka/MSK CDC → Glue Streaming → S3 → Redshift). - Investigate customer-reported analytics discrepancies (Jira/support tickets), root-cause them in the data, and communicate findings clearly to support, product, and engineering. - Set up and respond to pipeline monitoring - CloudWatch metrics and alarms, monitoring DAGs, refresh health - and participate in incident triage and RCA. - Develop deep product knowledge: understand what each metric means to our credit union customers and translate product changes into data model and dashboard updates. - Ensure data quality, validation, and consistency across systems. Qualifications - 3-4 years of experience in data engineering and/or analytics engineering. - Strong SQL on Redshift (or a similar MPP warehouse) and solid Python/PySpark. - Hands-on experience with the AWS data stack: S3, Glue, Redshift, CloudWatch. - Workflow orchestration with Apache Airflow — authoring, debugging, and deploying DAGs. - Data modeling and warehousing fundamentals. - BI dashboarding experience with QuickSight, ThoughtSpot, Tableau, or Power BI. - Quick learner — able to grasp an unfamiliar product and data model fast and work independently. - Strong listening and collaboration skills; comfortable coordinating across product, engineering, DevOps, and customer-facing teams. Requirements - Streaming/CDC experience: Kafka (MSK), Debezium, Spark Structured Streaming. - AWS infrastructure awareness: IAM roles, security groups, Secrets Manager, Kinesis/Firehose. - Fintech or B2B SaaS analytics exposure. Security Responsibilities - Adhere to Eltropy's policies on security, confidentiality, availability, and privacy; handle customer and financial-institution data responsibly, protect access credentials, and report security events promptly through Eltropy's channels.
Data Scientist / Machine Learning Engineer
Sequoia ConnectOur core expertise lies in connecting Top Technologists with Top Companies through unparalleled IT headhunting solutions
Role Description We are currently searching for a Data Scientist / Machine Learning Engineer: - Build and calibrate change-detection and anomaly models on multi-temporal Sentinel-1/2 imagery over pipeline corridors. - Learn per-site "normal terrain" baselines and validate detections against a ground-truth event log (detection rate, lead time, false-positive rate, AUC). - Fine-tune geospatial foundation models (Prithvi-EO or similar) with LoRA/PEFT on limited labelled data. - Implement SAR techniques for displacement: amplitude change, coherence, and pixel-offset tracking to measure pipe and dune movement. - Develop dune-migration tracking (optical flow / feature tracking), migration direction, and mobility indices. - Engineer robust ingestion from Copernicus (CDSE / Sentinel Hub / STAC) and fuse optical, SAR, DEM, and ERA5 wind data. - Design labelling strategy (encroachment masks, severity) and a train/validation split that avoids leakage. - Communicate results and limitations honestly to technical and business stakeholders. Qualifications - 7+ years of applied data science / ML experience, with hands-on geospatial remote sensing. - Strong Python programming skills, including numpy, rasterio/GDAL, xarray, scikit-image, and geopandas/shapely. - Working knowledge of optical and SAR data (spectral indices, backscatter/dB, resolution trade-offs, revisit). - Deep learning expertise with PyTorch, including experience fine-tuning models (transfer learning, LoRA/PEFT). - Proven ability in model validation and calibration: ROC/AUC, thresholding, cross-validation, and handling weak/few labels. - Experience with time-series / change-detection methods and coordinate reference systems (UTM, reprojection). - High-Performance Mindset: Resilience, emotional intelligence, and a focus on agile delivery. - Technologist DNA: A deep understanding of the difference between "coding" and "engineering." Requirements - Desired experience with InSAR / SAR offset tracking (SNAP, ISCE, or equivalent) for surface/structure displacement. - Familiarity with geospatial foundation models (Prithvi-EO, TerraTorch, HLS) and segmentation. - Knowledge of Copernicus/CDSE, Sentinel Hub, STAC, and Planetary Computer. - Exposure to Aeolian geomorphology, dune dynamics, or the oil & gas / pipeline-integrity domain. - Experience with MLOps and cloud environments (containerisation, scheduled inference, geospatial data pipelines). - MSc/PhD in Remote Sensing, Geospatial Science, Earth Observation, CS/ML, Physics, or equivalent experience. - Familiarity with cloud-native foundations or AI coding assistants. Benefits - Work Arrangement: We value flexibility to support your lifestyle. This position is available as Remote. Languages - Advanced Oral English: For seamless collaboration with global teams. - Advanced Spanish. Special Notes - Preference for candidates with Space Tech experience, though not mandatory.


