Prodege logo
Prodege

Founded in 2005, Prodege, LLC is a private American company headquartered in El Segundo, California, specializing in online marketing, consumer surveys, and mar

Staff Data Engineer

Location

United States

Posted

2 days ago

Salary

$185K - $220K / year

Seniority

Lead

Job Description

Staff Data Engineer

Prodege

Role Description This is a role for an engineer who wants to own core components of a modern data platform. We are looking for a Staff Data Engineer to architect, build, and operate production data systems across Prodege. This is a deeply hands-on technical role. You lead by building production-grade systems, setting engineering standards, and delivering scalable data architecture, not by working strictly at an abstract planning level. If you prefer delegating execution or working solely on isolated pipelines, this is not the role for you. But if you are a technical lead who owns data platforms end to end, from ingestion, streaming, and Medallion modeling to observability, governance, feature store foundations, and deployment, keep reading. You will build and evolve platform capabilities for a business serving over 120 million registered users. Operating within a high-scale data environment featuring: - 400 Terabyte footprint - 100 Terabyte Iceberg lake - 50 million daily events - 500 million pipeline records You will deliver the data foundations that power analytics, experimentation, machine learning, and AI-driven decision making across all Prodege products. If you enjoy building distributed systems at scale, working closely with cross-functional technical teams, and driving an AI-first engineering strategy, this role is for you. What you will own - Architecture, implementation, and operational reliability of major data platform domains, including pipelines, modeling layers, and data services - High-scale batch, ELT, and near-real-time streaming pipelines powering business intelligence, machine learning, experimentation, and product analytics - Platform standards for data governance, schema evolution, data contracts, lineage, and end-to-end observability - Data infrastructure and feature pipelines supporting machine learning, experimentation frameworks, and AI-driven applications - Technical quality and engineering standards across the team through direct code contributions, architecture design reviews, and technical mentorship What makes this role exciting - You will directly shape core data platform capabilities that drive analytics, machine learning, and business intelligence across the enterprise - You will own major technical domains from initial system design through production deployment and lifecycle management - You will build data foundations across consumer rewards, performance marketing, customer experience, and multiple owned digital properties - You will operate at true engineering scale, managing a 400 Terabyte data footprint, 50 million daily events, 500 million daily pipeline records, 50 plus Kafka topics, and 300 thousand daily queries - You will establish platform patterns that accelerate Prodege transition to an AI-first software engineering model What you will do - Architect, build, and operate high-capacity batch, ELT, event-driven, and near-real-time streaming data pipelines - Construct production-grade data platform components utilizing Snowflake, dbt, Iceberg, Trino, Kafka, and modern lakehouse technologies - Design scalable data models adhering to Medallion architecture principles to support business intelligence, advanced analytics, and machine learning - Enforce platform discipline around data contracts, schema evolution, lineage tracking, access governance, and system observability - Optimize data infrastructure for performance, query speed, system reliability, availability, and cost efficiency - Partner cross-functionally with Engineering, Product, Analytics, Business Intelligence, and Machine Learning teams to deliver trusted data foundations - Apply AI-assisted engineering tools to accelerate development velocity, automated testing, system debugging, and technical documentation What you will bring - Five to eight or more years of hands-on data engineering experience building large-scale data systems, ideally in advertising technology, marketing technology, consumer internet, or high-volume marketplace environments - Advanced technical expertise with SQL, Python, Snowflake, and dbt - Demonstrated background designing, deploying, and operating production-grade batch, ELT, and streaming pipeline architectures - In-depth understanding of modern data architecture paradigms, including Medallion architecture, data contracts, schema evolution, event-driven systems, and data modeling - Practical experience building data systems that directly support analytics, experimentation platforms, and machine learning workloads - Proven ability to optimize data pipelines and storage systems for scale, reliability, throughput, and cost performance - Strong communication and technical leadership skills, with a track record of driving technical decisions through design reviews, code reviews, and cross-functional alignment Bonus points - Experience with Iceberg, Trino, Kafka, Flink, Kinesis, Spark, or related lakehouse and distributed streaming frameworks - Experience building ML feature pipelines, feature stores, or model training data infrastructure - Experience architecting self-service analytics frameworks or experimentation platforms - Experience with modern DataOps methodologies, workflow orchestration, and data observability tools - Experience leveraging modern AI-assisted software development tools to enhance engineering productivity Pay Transparency The anticipated base salary range for this position is $185,000 to $220,000. The final salary offered to a successful candidate will be dependent on several factors that may include, but are not limited to; the type and length of experience within the job, type and length of experience within the industry, the type and length of knowledge and skills for the position, education, training, etc. Prodege is a multi-state employer and final compensation within this range could be impacted by work location. Please note that the compensation details listed in US role postings reflect the base salary only, and do not include bonus, equity, or benefits. Benefits - Comprehensive benefits package including medical, dental, vision, STD, LTD, and basic life insurance - Flexible PTO and paid sick leave prorated based on hire date - Eight paid holidays throughout the calendar year Equal Employment Opportunity Statement At Prodege, we are committed to creating a diverse and inclusive environment. We are proud to be an Equal Opportunity Employer and do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, veteran status, or any other characteristic protected by law. We encourage individuals of all backgrounds to apply. FCIHO Employers will consider for employment qualified applicants with criminal histories in a manner consistent with the requirements of FCIHO.

Related Categories

Related Job Pages

More Data Engineer Jobs

DraftKings Inc. logo

Senior Data Science Engineer, Personalization

DraftKings Inc.

Defining what it means to build and deliver the most extraordinary sports & entertainment experiences.The Crown is Yours

Data Engineer2 days ago
Full TimeRemoteTeam 1,001-5,000Since 2012H1B No Sponsor

• Lead end-to-end data science initiatives, from problem definition and modeling to production deployment and performance monitoring • Design and implement statistical models and machine learning algorithms that solve complex business challenges and enhance user engagement • Partner with Engineering, Product, and Analytics teams to integrate scalable data science solutions into live production systems • Drive innovation in personalized game recommendations and page layout optimization by experimenting with new algorithms and data-driven approaches • Translate complex technical findings into clear, actionable insights that empower stakeholders to make informed decisions • Deploy and maintain production-grade Python software • Mentor junior data scientists, champion best practices, and elevate the team’s technical rigor and impact

Canada
Sigma Software Group logo

Data Engineer

Sigma Software Group

We support enterprises, product houses, and startups with custom software solutions development and IT consulting.

Data Engineer2 days ago
Full TimeRemoteTeam 1,001-5,000Since 2002H1B No Sponsor

• Design, develop, and maintain scalable ETL/ELT pipelines • Build and optimize data processing solutions using Python and Apache Spark • Develop and support real-time data streaming applications using Kafka • Design and implement enterprise data solutions leveraging Snowflake • Build cloud-native data architectures and services on AWS • Ensure data quality, reliability, scalability, and performance across the data ecosystem • Troubleshoot, monitor, and optimize production data pipelines • Participate in technical discussions, code reviews, and engineering best practice initiatives • Collaborate with cross-functional teams to support data-driven business initiatives

Brazil
Fora Financial logo

Staff Data Engineer

Fora Financial

Leading Business Financing Lender

Data Engineer2 days ago
Full TimeRemoteTeam 51-200Since 2008H1B Sponsor

Role Description Fora Financial is at an inflection point, modernizing our legacy stack to build the foundation for AI-native analytics. To lead this effort, we are hiring a Staff Data Engineer to build and own the platform backbone for governed reporting and trusted AI workflows. This is a hands-on Staff IC role on a small Data & AI team. You will make strategic architecture calls—from ingestion patterns and Snowflake design to data contracts and SLAs—and then get into the weeds to build, harden, or rebuild pipelines. We are looking for a systems thinker who understands business impact and operational burden, and who can partner closely with Analytics, Engineering, and vendors to turn fragmented source systems into trustworthy data products. What you will own - Architecture & Strategy - Data platform architecture: ingestion patterns, warehouse design, environment strategy, orchestration, access governance, and reliability standards. - Freshness strategy: deciding which data needs real-time, near-real-time, daily, or ad hoc refreshes — and designing accordingly. - Streaming vs. batch decisions: making pragmatic tradeoffs across business value, cost, complexity, failure modes, and operational burden. - Execution & Reliability - Source ingestion: batch, incremental, API-based, file-based, CDC, and streaming patterns where they make sense. - Pipeline reliability: dependencies, retries, alerts, backfills, incident response, runbooks, monitoring, and support expectations. - New source onboarding: requirements → source profiling → ingestion design → QA → documentation → support ownership. - Legacy migration: helping retire brittle reporting paths such as Azure Data Factory, SQL backup workflows, TRS Daily, and other duplicate pipelines. - Governance & Quality - Snowflake platform operations: roles, permissions, service accounts, connector ownership, environment separation, performance, cost, and governance. - Data contracts: schema-change handling, new-field availability, upstream SLAs, source defects, and escalation paths. - Data quality and observability: freshness, volume movement, nulls, duplicates, reconciliation, anomaly detection, and critical business-rule checks. - AI-enabled leverage: using AI and automation to improve debugging, documentation, pipeline scaffolding, testing, monitoring, and operational workflows. Qualifications - Deep data engineering judgment: You have designed, built, and operated production platforms, not just individual pipelines. - Hands-on depth: You seamlessly move from architecture discussions to Python, SQL, deployment scripts, and production debugging. - Strong ingestion fundamentals: APIs, CDC, backfills, idempotency, schema drift, and failure recovery. - Snowflake fluency: Warehouse design, RBAC, performance tuning, and cost controls. - Data quality discipline: You know which checks matter and make quality visible before users find issues. - Independent ownership & communication: You can sequence ambiguous work, write useful design docs, align technical decisions with business outcomes, and carry problems to resolution. - AI-native leverage: You actively use LLMs and agents to accelerate engineering work without outsourcing judgment. Nice to have - Lending, fintech, or financial-services data experience. - CDC, Debezium, Fivetran, Airbyte, Azure Data Factory, dbt Cloud, Dagster, Airflow, Prefect, or equivalent tooling. - Snowflake performance tuning, RBAC, data sharing, warehouse cost optimization, or Iceberg. - Data observability with Monte Carlo, Elementary, dbt tests, custom monitors, or similar. - Data contracts, source SLAs, or schema-change processes with Engineering teams. - AI-native analytics, semantic layers, MCP servers, agent QA, or governed context retrieval. - Lightweight internal tools, scripts, or agents that reduce repetitive platform work. Compensation and logistics - Base salary: $175,000–$200,000 - Fully remote within the US; Eastern or Central time zones preferred. - Reports to the VP of Data & AI. - Final compensation is based on scope of past ownership, technical judgment, and ability to set direction independently. Benefits - Company-subsidized medical, dental, and vision plans. - 401(k) plan with company match. - Life insurance at no cost to employees. - Generous time off plan, including rollover vacation days. - Health care and dependent care flexible spending accounts. - Commuter benefits. - Remote working model. - Weekly breakfast, snacks, and Friday lunches provided onsite.

United States
$175K - $200K / year
INFUSE logo

Data-Focused Project Coordinator

INFUSE

Demand Performance Delivered

Data Engineer2 days ago
ContractRemoteTeam 1,001-5,000H1B No Sponsor

Role Description We are looking for a detail-oriented and organized Data-Focused Project Coordinator to support the planning, execution, and completion of key projects. This role requires proficiency in Business Intelligence tools , strong data management skills, and the ability to work effectively in a fast-paced environment. - Support project management tasks, ensuring timely delivery. - Use Business Intelligence tools to analyze and visualize data for actionable insights. - Maintain accurate project documentation and validate data accuracy. - Collaborate with teams to create detailed reports and support decision-making. Qualifications - Proficiency in Business Intelligence tools . - Strong attention to detail, organizational, and analytical skills. - A basic understanding of database systems and validation processes is a plus. - Experience using and improving AI-driven workflows is highly valued. - Proficiency in Russian and/or Ukrainian languages. Company Description INFUSE is a demand generation company providing solutions to B2B organizations. We help clients and partners to deliver audience, buyers, and account engagement that meets their goals. With a global team operating across more than 60 countries, the company has been recognized multiple times on the prestigious Inc. 5000 list, and has received over 60 other industry awards, including Inc. Best Workplaces.

Worldwide