Job Closed

This listing is no longer active.

Effectual logo
Effectual

Cloud Confidently®

Data Engineer

Data EngineerData EngineerFull TimeRemoteSeniorTeam 201-500H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

94 days ago

Salary

$132K - $166K / year

Seniority

Senior

Job Description

Data Engineer

Effectual

• Building and delivering high quality data architectures and pipelines that support clients, business analysts, and data scientists. • Interfaces with other technology teams to extract, transform, and load [ETL] data from a wide variety of data sources. • Continually improve ongoing reporting and processes, as well as automate or simplify self-service for clients. • Develop, code, and deploy scripts written in the Python programming language, as Python is the language of Data. • Assemble large, complex datasets to meet both functional and non-functional requirements. • Build high-performance algorithms, predictive models, and proofs of concept while using modern programming languages and tools to integrate systems and manage complex data workflows. • Ensure the data strategies and architectures are in compliance with regulatory requirements. • Work with business executives to ensure the data architecture aligns with business requirements. • Optimize data retrieval, develop dashboards, reports and perform database maintenance tasks. • Collaborate closely with data scientists, helping to ensure that the organization's data infrastructure meets the requirements of complex data analytics.

Job Requirements

  • Bachelor’s or master’s degree in Computer Science, Engineering or a related field
  • 4-7 years experience working as a Data Engineer, preferably in a professional services or consulting environment
  • Strong proficiency in programming languages such as Python, Java, or Scala, with expertise in data processing frameworks and libraries (e.g., Spark, Hadoop, SQL, etc.)
  • In-depth knowledge of database systems (relational and NoSQL), data modeling, and data warehousing concepts
  • Experience with cloud-based data platforms and services (e.g., AWS, Azure, Google Cloud) including familiarity with relevant tools and technologies (e.g., S3, Redshift, BigQuery, etc.)
  • Computer Vision/Intelligent Document Processing.
  • Proficiency in designing and implementing ETL processes and data integration workflows using tools like Apache Airflow, Informatica, or Talend
  • Familiarity with data governance practices, data quality frameworks, and data security principles
  • Strong analytical and problem-solving skills, with the ability to translate business requirements into technical solutions
  • Excellent communication and collaboration skills, with the ability to effectively work with clients and cross-functional teams
  • Self-motivated and proactive, with a passion for learning and staying updated with the latest trends and advancements in the field of data engineering
  • Able to work with ambiguity and turn client wants and needs into working stories, epics which can be executed upon during a sprint.
  • A firm understanding of the SDLC process
  • An understanding of object-oriented programming
  • Needs minimal direction
  • AWS background, AWS Cloud Formation & Data Migration Services (DMS)
  • Solution Engineer mindset

Benefits

  • Medical, dental, and vision health insurances,
  • Short term disability, long term disability and life insurances,
  • 401k with Company match
  • Paid time off (PTO) (120 hours PTO that accrue over one year)
  • Paid time off for major holidays (14 days per year)
  • These and any other employee benefit offerings are subject to management’s discretion and may change at any time.

Related Categories

Related Job Pages

More Data Engineer Jobs

Jumio Corporation logo

Data Intern

Jumio Corporation

Identity verification through informed AI.

Data Engineer94 days ago
InternshipRemoteTeam 201-500Since 2010H1B Sponsor

Role Description The Intern will play a vital role in advancing the company's capabilities in computer vision and fraud detection projects, as well as contributing to research and development initiatives. The incumbent will assist in various tasks related to biometric data collection, algorithm testing, and model validation. Proficiency in Python, along with knowledge of machine learning and MATLAB, is essential for success in this role. Familiarity with AWS and SageMaker is advantageous. - Assist in computer vision and fraud detection projects by contributing to algorithm testing, model validation, and biometric data collection. - Utilize Python, MATLAB, and deep learning libraries such as PyTorch and TensorFlow for image processing and analysis tasks. - Collaborate with the research team to implement benchmarking metrics and perform ROC analysis. - Ensure accuracy in data collection, labeling, and algorithmic testing by paying meticulous attention to detail. - Develop and optimize SQL queries to extract and analyze data for machine learning model training. - Communicate effectively with team members through written and verbal channels to provide updates on project progress and findings. Qualifications - Availability to start immediately is required. - Enrollment in a graduate program in computer science, computer/electrical engineering, or related fields. - Proficiency in Python with working knowledge of image processing libraries. - Understanding of benchmarking metrics. - Attention to detail and ability to adapt and learn quickly in a fast-paced environment. Requirements - Knowledge of different biometric modalities, eKYC, and presentation attacks. - Understanding of databases, cloud computing, and storage, particularly AWS data and ML pipelines such as SageMaker and S3 buckets. - The compensation is $23 per hour. Jumio Values - Integrity - Diversity - Empowerment - Accountability - Leading Innovation Equal Opportunities Jumio is a collaboration of people with different ideas, strengths, interests, and cultures. We welcome applications and colleagues from all backgrounds and of all statuses.

United States
$23 / hour
Job Closed
ReWorks Solutions logo

Data Entry Specialist (AR/AP)

ReWorks Solutions

Building quality global teams that drive efficiency and results

Data Engineer94 days ago
Full TimeRemoteTeam 201-500Since 2024H1B No Sponsor

Position: Data Entry Specialist (AR/AP) Working Hours: US Hours (9am-5pm EST) Full-Time, Remote Work. Key Responsibilities - Accurately enter and manage Accounts Receivable (AR) and Accounts Payable (AP) data in company financial systems. - Verify and reconcile invoices, payment records, and financial documents to ensure correctness. - Assist with generating financial reports related to AR/AP activities. - Communicate with vendors, clients, and internal departments to resolve discrepancies and clarify financial records. - Maintain organized and up-to-date records of all AR/AP transactions. - Support finance and accounting teams with data entry and administrative tasks as required.

Philippines
Job Closed
ZoomInfo Technologies LLC logo

Senior Data Engineer

ZoomInfo Technologies LLC

ZoomInfo (NASDAQ: GTM) is the Go-To-Market Intelligence Platform that empowers businesses to grow faster with AI-ready insights, trusted data, and advanced automation. Its solutions provide more than 35,000 companies worldwide with a complete view of their customers, making every seller their best seller.

Data Engineer94 days ago
Full TimeRemoteTeam 1,001-5,000

Role Description We are looking for a highly skilled Senior Data Engineer to become part of our core Data & AI Engineering team. In this pivotal role, you will be responsible for designing and expanding enterprise-level data infrastructure that enables ZoomInfo's internal teams to interact with data comprehensively—extracting, exploring, analyzing, and generating insights—through various platforms using ZI's internal chat agent. The ideal candidate has a strong background in big data processing, pipeline orchestration, and data modeling, with a proven track record of delivering scalable and high-quality data solutions in fast-paced, data-centric product environments. Given the dynamic nature of emerging technologies, this role requires an individual who excels at exploration and embraces continuous learning as core responsibilities. You'll constantly research and implement innovative solutions while integrating vast, diverse data sources into our AI applications, including our industry-leading LLM-powered systems. What you’ll do: - Design, develop, and maintain high-performance, product-centric data pipelines using Airflow, DBT, and Python. - Architect and optimize the massive-scale data warehouse and lakehouse that serves as our single source of truth for all customer data, primarily using Snowflake. - Lead the integration of diverse structured and unstructured data sources (e.g., web data, third-party APIs) into our data ecosystem, ensuring high-quality and reliable ingestion. - Implement and enforce Model Context Protocol (MCP) or similar architectures to feed accurate and contextual data into our LLM-powered products for applications like Retrieval Augmented Generation (RAG) and advanced search. - Collaborate with ML engineers, data scientists, and product managers to translate business needs into scalable data solutions that directly enhance customer value. - Define, monitor, and enforce data quality SLAs across all pipelines and products, ensuring data accuracy and lineage are a top priority. - Mentor and coach junior engineers, promoting best practices in code quality, data architecture, and operational excellence. - Participate in architectural decisions and long-term strategy planning for our enterprise-wide data infrastructure, with a focus on cost, performance, and reliability. Qualifications - Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field. - 8+ years of progressive experience in data engineering, with a track record of leadership and impact. - Demonstrated experience in implementing or scaling data infrastructure for a data-centric product company. Requirements - Expert-level SQL for building performant, scalable queries and transformations on massive datasets. - Strong Python programming skills with a focus on distributed computing, data manipulation, and building robust APIs. - Production-level experience for large-scale batch and streaming data processing. - Hands-on experience with DBT (Data Build Tool) for advanced data modeling and transformations in a modern data stack. - Deep knowledge of Snowflake data warehouse design, optimization, and cost modeling. - Experience implementing Model Context Protocol (MCP) or similar architectures to feed structured and unstructured data into LLM-powered systems. - Strong understanding of data architecture concepts, including data lakes, event-driven architectures (e.g., Kafka), ETL/ELT, and data mesh. - Proficiency with cloud platforms (GCP and/or AWS) and infrastructure as code (e.g., Terraform). Benefits - Excellent communication skills – ability to explain complex technical concepts to both engineering teams and non-technical stakeholders. - Strategic & Product-Oriented Thinking – can translate business objectives and customer needs into scalable, high-impact data solutions. - Leadership & Mentorship – experience guiding and uplifting engineering teams to achieve their full potential. - Stakeholder Management – able to collaborate effectively across departments (Product, Engineering, Sales, Compliance). - Agility & Adaptability – thrives in ambiguous, evolving environments and can rapidly prototype and iterate on solutions. - Strong documentation habits and ability to evangelize best practices across the organization.

Worldwide
Job Closed
Full TimeRemoteTeam 11-50H1B No Sponsor

At Murmuration, we believe that America’s promise is shaped and reshaped by the best ideas and ideals of its communities, and the dreams of the people who believe in a better life for themselves, their families, and each other.  We help organizations build power in their communities in four key ways: we organize a network of values-aligned partners; we provide deep, data-driven insights into people, places, and perspectives; we develop tools that make organizing and engagement easy and more effective; and we offer services that strengthen our partners’ capacity to lead change in their communities. We envision an America where every community has what it needs to help people lead healthy, free, and dignified lives. We work to redesign the systems and structures we all depend on — how we learn, live, govern, and solve problems — so that they are just, equitable, resilient, and rooted in shared responsibility. By strengthening the ties that hold communities together, we aim for civic life defined by collective action and care, with effective leadership that truly represents everyone.  We are a collaborative, curious, and creative team of organizers, scientists, teachers, technologists, campaign veterans, and more who share the unwavering belief that we can use our gifts in service of transforming America — together. We’ve built our team guided by the belief that the whole is greater than the sum of its parts. And so we support each other relentlessly — rallying together to face challenges the same way we celebrate each other’s wins. About the Position You're a mission-driven data engineer with a track record of building infrastructure that delivers real impact. You're deeply curious not just about how systems work but about the messy, real-world data they must support. You've done this before: you understand what "good" looks like, you deliver high quality solutions, and you help the teams around you operate more effectively. At Murmuration, you'll work with voter files from 50 states (each with its own format and quirks), census data, polling results, geographic boundaries, election returns, and more. These data sources arrive at different cadences, anywhere from multiple times a day during early voting to once a year from the census. Your role is to bring order to this complexity by building the pipelines, data contracts, and governance that transform these disparate inputs into Atlas, our unified representation of American civic life. This role is for someone who's energized by complex data problems, comfortable with ambiguity, and motivated by the understanding that strong infrastructure and data foundations is a core part of what makes research, product development, and meaningful civic impact possible. You'll partner closely with data scientists, researchers, and product teams who rely on what you build. You also stay curious about emergent technologies, particularly AI, and apply them thoughtfully to amplify your impact and the effectiveness of the broader team. Job Level P4 What You'll Do - Own data pipelines and infrastructure: Design, implement, and evolve scalable, production-grade systems using tools such as Dagster, Airflow, Snowflake, AWS, MongoDB, and dbt. Apply a cloud-native and DevOps mindset using CI/CD, infrastructure-as-code, monitoring, and automated testing to build reliable systems. Partner with cross-functional teams to deliver solutions that meet both immediate product needs and long-term organizational strategy. - Lead data ingestion and integration: Bring in complex, high-volume datasets while ensuring strong data contracts, freshness, quality, integrity, and lineage, and build systems that empower domain experts to contribute to and maintain their own data pipelines. - Transform raw data into trusted data products: Convert raw inputs into structured, usable datasets  that empower our analytical and product teams. Collaborate closely with operational data managers to ensure data models and intuitive, reliable alignment with how data is consumed in practice. - Leverage AI: Make informed judgement calls about how AI can be a force-multiplier for both your own work and the team’s and how it can’t. - Elevate the team: Mentor engineers, actively shape technical direction through architectural reviews and roadmap planning, and build team culture through documentation and knowledge sharing.

United States
Job Closed