Founded in 2003, First Advantage provides comprehensive background-check insights and solutions, enabling employers and housing providers to make confident choi
Data Engineer (US Remote)
Location
United States
Posted
133 days ago
Salary
$75K - $90K / year
Seniority
Mid Level
No structured requirement data.
Job Description
Data Engineer (US Remote)
First Advantage
At First Advantage (Nasdaq: FA), people are at the heart of everything we do. From our customers and partners to our greatest advantage — our team members. Operating with empathy and compassion, First Advantage fosters a global inclusive workforce devoted to the diverse voices that make up our talent and products. Our team members empower each other to be their authentic selves and treat all with respect, integrity, and fairness. Say hello to a rewarding career and come join a leading provider of mission-critical background screening solutions to some of the most recognized Fortune 100 and Global 500 brands. What We Do: We are on the frontline of recruitment enabling organizations to Hire Smarter. Onboard Faster™ First Advantage is an HR Tech company delivering innovative solutions and insights to enable our clients to manage risk and hire the best talent. Leveraging an advanced technology platform, First Advantage builds fully scalable, configurable screening programs that meet the unique needs of over 30,000 clients. Headquartered in Atlanta, GA and with an internationally distributed workforce spanning 17 countries with about 5,000 employees, First Advantage performs over 100 million screens in over 200 countries and territories annually. We are seeking a Data Engineer with solid Python/PySpark programming skills to join the Data Engineering Team and help us build the Data Analytics Platform in Azure cloud. Who You Are: You are self-motivated and ready to “roll up your sleeves." While you are an independent contributor, you are also collaborative. You can spearhead a project and see it through from start to completion. As a team player, you navigate cross-functional teams and work well with team members in other business units and departments toward a common goal. An Innovator — you see gaps in current processes or workflows as an opportunity to improve and try something new. A lifelong learner and always seeking out opportunities to learn and upskill, you understand the importance of thorough and secure screenings and are interested in the Human Capital sector and the confluence of people, process, and technology. What You'll Do: - Develop reusable, metadata-driven data pipelines - Automate and optimize any data platform related processes - Build integrations with data sources and data consumers - Add data transformation methods to shared ETL libraries - Write unit tests - Develop solutions for the Databricks data platform monitoring - Proactively resolve any performance or quality issues in ETL processes - Cooperate with infrastructure engineering team to set up cloud resources - Contribute to data platform wiki / documentation - Perform code reviews and ensures code quality - Initiate and implements improvements to the data platform architecture What You Need to be Successful: - Programming: Python/PySpark, SQL - Proficient in building robust data pipelines using Databricks Spark - Experienced in dealing with large and complex datasets - Knowledgeable about building data transformations modules organized as libraries (Python packages) - Familiar with Databricks Delta optimization techniques (partitioning, z-ordering, compaction, etc.) - Experienced in developing CI/CD pipelines - Experienced in leveraging event brokers (Kafka /Event Hubs / Kinesis) to integrate with data sources and data consumers - Understanding of basic networking concepts - Familiar with Agile Software Development methodologies (Scrum) Nice to Have Skills: - Understanding of stream processing challenges and familiarity with Spark Structured Streaming - Experience with IaC (Terraform, Bicep or other) - Experience running containerized applications (Azure Container Apps, Kubernetes) - Experience building event sourcing solutions - Familiarity with platforms for change data capture (e.g. Debezium) - Knowledge of Azure cloud native solutions (e.g. Azure Data Factory, Azure Function App, Azure Container Instances) Why First Advantage is Your Next Big Career Move First Advantage is going through a technology transformation! We are looking for experts who are excited to work with advanced technologies and provide best-in-class user experiences, drive the development and deployment of scalable solutions, and smoothly guide our agile teams and clients through meaningful changes as we continue to expand our impact. Additional benefits offered to our eligible people include: - Competitive benefits package, including health care, life insurance and Multisport, - Challenging projects. - Spacious, modern, and fully equipped office space in the heart of Krakow. - Flexibility and possibility to work remotely. - Superior co-working and personal development experience. What Are You Waiting For? Apply Today! You have learned a little about us today – we want to learn about you! If you think this position and our company are a great fit for your areas of interest and expertise, tell us about you by applying now! The base salary range for this position is approximately $75,000-90,000 annually plus there is additional opportunity for Variable Compensation. This range reflects our good faith estimate to pay fairly as to what our ideal candidates are likely to expect, and we tailor our offers within the range based on the selected candidate’s experience, industry knowledge, technical and communication skills, and other factors that may prove relevant during the interview process. United States Equal Opportunity Employment: First Advantage is proud to be a global leader in removing barriers and supporting our community members to ensure the changing demographics of the workforce are reflected in our hiring and employment practices. We value all of our candidates, employees, and clients, and place great emphasis on hiring and supporting qualified individuals in each role. We are an equal opportunity employer. We do not discriminate on the basis of race, color, ethnicity, ancestry, religion, sex, national origin, sexual orientation, age, citizenship status, marital status, disability, gender identity, gender expression, veteran status, genetic information, or any other area protected by applicable law.
Job Requirements
- Programming: Python/PySpark, SQL.
- Proficient in building robust data pipelines using Databricks Spark.
- Experienced in dealing with large and complex datasets.
- Knowledgeable about building data transformations modules organized as libraries (Python packages).
- Familiar with Databricks Delta optimization techniques (partitioning, z-ordering, compaction, etc.).
- Experienced in developing CI/CD pipelines.
- Experienced in leveraging event brokers (Kafka/Event Hubs/Kinesis) to integrate with data sources and data consumers.
- Understanding of basic networking concepts.
- Familiar with Agile Software Development methodologies (Scrum).
- Understanding of stream processing challenges and familiarity with Spark Structured Streaming.
- Experience with IaC (Terraform, Bicep or other).
- Experience running containerized applications (Azure Container Apps, Kubernetes).
- Experience building event sourcing solutions.
- Familiarity with platforms for change data capture (e.g. Debezium).
- Knowledge of Azure cloud native solutions (e.g. Azure Data Factory, Azure Function App, Azure Container Instances).
Benefits
- Competitive benefits package, including health care, life insurance and Multisport.
- Challenging projects.
- Spacious, modern, and fully equipped office space in the heart of Krakow.
- Flexibility and possibility to work remotely.
- Superior co-working and personal development experience.
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Software Engineer III, Data Engineering
GitHub, Inc.GitHub is the world’s leading AI-powered developer platform with 150 million developers and counting. We’re also home to the biggest open-source community on earth (and 99% of the world’s software has open-source code in its DNA). Many of the apps and programs you use every day are built on GitHub. Our teams are dreamers, doers, and pioneers, leading the way in AI, driving humanitarian efforts around the globe, and even sending open source to Mars (and beyond!). At GitHub, our goal is to create the space you need to do your best work. We’re remote-first and offer competitive pay, generous learning and growth opportunities, and excellent benefits to support you, wherever you are—because we know that people flourish when they can work on their own terms. Join us, and let’s change the world, together.
About GitHub GitHub is the world’s leading platform for agentic software development — powered by Copilot to build, scale, and deliver secure software. Over 180 million developers, including more than 90% of the Fortune 100 companies, use GitHub to collaborate, and more than 77,000 organisations have adopted GitHub Copilot. Locations In this role you can work from Remote, United States Overview As a software engineer at GitHub, you will enhance the collaboration experience at GitHub by working closely with a community of engineers and designers with a distributed, diverse and passionate team delivering the services that millions of developers depend on. In this role you will design, prototype, implement, ship and support highly performant and inspiring user experiences with your team. We are looking for creative problem solvers and diverse thinkers, people who care about culture as well as customers and features. We believe that how we do things is as important as what we do. Big vision, a common purpose, passion for quality, curiosity, dedication, and investment in fun and collaboration are what lead to great results. Great products reflect the teams that build them. Responsibilities - Design, develop, test and ship high-quality technical solutions that scale across multiple GitHub services and become intimately familiar with the systems you build and take pride in writing maintainable code. - Provide technical leadership, mentorship, pairing opportunities, and code reviews to encourage the growth of others; support teams in producing extensible and maintainable code, ensuring integration with downstream dependencies and adherence to quality standards. - Own and advocate for the health and quality of the systems that the team builds, including participating in on-call for first responder rotations and live incidents. - Write architecture briefs and proposals and carry out code experiments. - Design and implement APIs to facilitate seamless integration between software components. - Utilize CI/CD tools to set up automated pipelines for continuous integration and delivery. - Collaborate with cross-functional teams and partner with stakeholders and lead discussions for technical solutions, including design and cost considerations. - Create and guide others in 1) developing clear testing plans to assure solution quality, reliability, and performance; 2) defining success metrics; and 3) integrating customer feedback for continuous improvement - all while ensuring system architecture meets security and compliance standards. - Maintain executional and operational excellence within and potentially across teams/organizations. - Apply debugging tools and telemetry to verify assumptions, proactively resolve issues, and optimize code performance and maintainability. Qualifications Required Qualifications: - 4+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software languages including, but not limited to, C, C++, C#, JavaScript, Go, Ruby, Rust, or Python - OR Associate’s Degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 3+ years experience - OR Bachelor's Degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 2+ years experience in Computer Science, or related technical discipline with proven experience coding in languages including, but not limited to, C, C++, C#, JavaScript, Go, Ruby, Rust, or Python - OR Master's Degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field - OR equivalent experience. Preferred Qualifications: - Demonstrated experience with large-scale system architecture and design, particularly in cloud-based environments, with a strong understanding of distributed systems and microservices. - Experience working closely with product management, design, and other engineering teams to drive cross-functional projects and deliver high-quality products - Experience in one or more scripting languages (e.g., Bash, Python, or a similar language) - Experience with cloud environments and/or Cloud Native Compute Foundation (CNCF) concepts - Experience working with both relational (e.g. mysql) and most importantly non-relational datastores (e.g. Cosmos) - Experience working with Azure resource such as Azure Storage (blob and table particularly), Azure Redis Cache, Azure Data Explorer Clusters. - Experience operating Cosmos DB clusters at scale. Compensation Range The base salary range for this job is USD $107,700.00 - USD $285,900.00 /Yr. These pay ranges are intended to cover roles based across the United States. An individual's base pay depends on various factors including geographical location and review of experience, knowledge, skills, abilities of the applicant. At GitHub certain roles are eligible for benefits and additional rewards, including annual bonus and stock. These rewards are allocated based on individual impact in role. In addition, certain roles also have the opportunity to earn sales incentives based on revenue or utilization, depending on the terms of the plan and the employee's role. GitHub values - Customer-obsessed - Ship to learn - Growth mindset - Own the outcome - Better together - Diverse and inclusive Manager fundamentals - Model - Coach - Care Leadership principles - Create clarity - Generate energy - Deliver success Who We Are GitHub is the world’s leading AI-powered developer platform with 150 million developers and counting. We’re also home to the biggest open-source community on earth (and 99% of the world’s software has open-source code in its DNA). Many of the apps and programs you use every day are built on GitHub. Our teams are dreamers, doers, and pioneers, leading the way in AI, driving humanitarian efforts around the globe, and even sending open source to Mars (and beyond!). At GitHub, our goal is to create the space you need to do your best work. We’re remote-first and offer competitive pay, generous learning and growth opportunities, and excellent benefits to support you, wherever you are—because we know that people flourish when they can work on their own terms. Join us, and let’s change the world, together. EEO Statement GitHub is made up of people from a wide variety of backgrounds and lifestyles. We embrace diversity and invite applications from people of all walks of life. We don't discriminate against employees or applicants based on gender identity or expression, sexual orientation, race, religion, age, national origin, citizenship, disability, pregnancy status, veteran status, or any other differences. Also, if you have a disability, please let us know if there's any way we can make the interview process better for you; we're happy to accommodate!
Imagine being an employee-owner of a company guided by engaged and empowered team members like yourself. Where a culture of respect, flexibility, and accountability aren't just ideals - they're our foundation, and diverse backgrounds and perspectives are valued as drivers of innovation and growth. Join us, as together, we are Building a Better World for All of Us®. You belong at SEH SEH is currenting searching for a Reality Capture Manager to join our talented Geospatial Data Services team! Why our employee-owners love SEH: - "I was on vacation last week and had zero concerns that my colleagues would help out with anything that came into my inbox!" – GIS Analyst - "What company has a CEO who cares enough to seek out one-on-one conversations ranging from 'How are you?' to 'What do you think would help the company?' SEH, that's who. " – Civil Engineering Technician - "Having the feeling that my voice matters and believing that SEH truly cares about the employees is so satisfying!" – Sr Financial Analyst - "It feels good having colleagues and supervisors that provide support and resources for growth and learning!" – Civil Engineer - "This is the first company I've worked for with a true entrepreneurial spirit." – Sr Mechanical Engineer Why you’ll love SEH: - Collaborate on amazing projects of varying size and complexity that positively impact communities - Being 100% employee-owned means we all share in the company’s success - Career development through continued education, licensure/certification, skills, and technical training - Work arrangements that promote work/life balance - Flexible holidays enable individuals to tailor their festivities - Paid Family Leave provides time to care for loved ones, whether family by birth or family by choice This Opportunity: Field Operations & Data Capture - Oversee the use of UAVs and terrestrial LiDAR, SLAM systems, and photogrammetry to collect high-resolution spatial data across diverse project sites. - Coordinate with project managers and reality capture team, to scope, plan, and execute reality capture missions. - Ensure compliance with insurance, FAA regulations and airspace restrictions and SEH and site specific safety requirements. Data Processing & Modeling - Manage post-processing workflows to generate georectified 3D models and point clouds. - Integrate reality capture data with BIM, GIS, and CAD platforms. - Collaborate with survey teams to merge drone and ground-based data for enhanced accuracy. Technology & Innovation - Evaluate and implement emerging reality capture tools and software. - Maintain SEH’s fleet of drones and scanning equipment for suitability and compliance with federal, state and local regulations. - Develop SOPs and best practices for reality capture across the organization. Team Leadership & Training - Mentor and train drone pilots and reality capture specialists. - Support onboarding and certification (e.g., FAA Part 107) for new team members. - Promote a culture of safety, innovation, and continuous learning. Client Engagement & Strategy - Communicate the value of reality capture to internal and external stakeholders. - Provide scope and fee estimates for reality capture services. - Assist and advise Lead Practices with incorporation of high density data sets into design workflows. - Support marketing and visualization efforts with aerial imagery and 3D assets. Essential Qualifications: - Proficiency in tools such as: Civil3d, Autodesk ReCap, Revit, Navisworks, Faro Scene, Leica Cyclone, DroneDeploy, Pix4D, OpenDroneMap, Trimble Business Center. - Strong understanding of industrial plant operations, engineering and architectural design and construction workflows. - Experience with GPS/GNSS systems and ground control integration. Preferred Qualifications: - FAA Part 107 Drone Certification Who We Are Better Places. Clean Water. Renewing Infrastructure. Improving Mobility. SEH is an employee-owned engineering, architectural, planning, and environmental company, offering a wide variety of services. We've been helping government, industrial, and commercial clients find solutions to complex challenges since 1927. Our 900+ employee-owners across the US unite behind our core purpose of Building a Better World for All of Us®. Base compensation is expected to be in the range of $135,000 and $145,000 based on skill set and experience. Check out our full benefits package at SEH Hiring Journey. Due to current business and operational considerations, unable to hire employees residing in the following states at this time: AK, AR, CA, CT, DE, HI, KY, MA, RI, VT, and PR. Candidates willing to relocate should indicate this in their application. The selected candidate must be authorized to work for any employer in the U.S. without requiring visa sponsorship now or in the future. SEH is an Equal Opportunity Employer, committed to providing equal employment opportunities to all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, disability, or veteran status. We take affirmative action to ensure that all employment decisions are based on merit, qualifications, and abilities. Women and Minorities are encouraged to apply. Notice to Third Party Agencies: SEH does not accept unsolicited resumes from third party recruiting firms. Absent a signed Service Agreement by SEH’s Talent Director, SEH reserves the right to pursue and hire these candidates without financial obligation to recruiters or agencies. #LI-KS1
• Streamlining Data Integration You’ll design and automate scalable systems to ingest and transform data from diverse sources, ensuring seamless and efficient data flow across the organization. • Safeguarding Data Quality By implementing robust monitoring frameworks, you’ll proactively detect and resolve data integrity issues, maintaining trust in analytics and reporting. • Establishing Data Governance You’ll lead the development of governance processes to manage metadata, access, and retention, ensuring compliance and secure data usage for internal and external stakeholders. • Building Scalable Data Pipelines You’ll architect reliable and high-performance ETL/ELT pipelines with built-in monitoring and alerts, enabling timely and accurate data delivery for business needs. • Optimizing Database Design and Performance Through thoughtful physical data modeling and indexing strategies, you’ll enhance database efficiency and scalability for large-scale operations. • Modernizing Data Infrastructure You’ll develop and operate advanced storage and processing solutions using distributed and cloud platforms, supporting big data initiatives and analytics. • Automating Data Workflows By leveraging modern tools and techniques, you’ll reduce manual data preparation tasks, boosting productivity and minimizing errors. • Mentoring and Agile Collaboration You’ll coach junior team members and contribute to agile practices like DevOps and Scrum, accelerating delivery of critical analytics projects and fostering team growth.
Principal Data Architect
UnqorkUsing CaaS (Codeless-as-a-Service) to accelerate time-to-market & eliminate legacy code for the enterprise 🚀
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description We're seeking an experienced, visionary Principal Data Architect and Leader to drive the strategy, design, and operations of our enterprise-grade data architecture and infrastructure. This critical role leads our data functions, ensuring our platform provides a robust, scalable, secure, and highly available foundation for customers developing, deploying, and hosting applications using this robust data layer. The ideal candidate will possess deep expertise in data architecture, cloud-native technologies, and operational excellence, directly impacting our ability to serve thousands of customer applications. Report to our Engineering Manager Architecture & Strategy - Define and own the long-term data architecture strategy for Unqork's core platform, covering data modeling, query design, storage topology, and access patterns across a complex data environment that includes MongoDB as well as integration with Relational and Columnar database models. - Ensure the data layer meets the security, performance, scalability, and resilience requirements of enterprise-grade, mission-critical applications. - Evaluate and recommend the right database technologies, indexing strategies, caching layers, ETL and search infrastructure for each class of workload — including when to use MongoDB Atlas Search, Caching (e.g., Redis), Querying (e.g., Kafka) and Streaming components. - Own the end-to-end solutions around data transfers using solutions like ETL for customers. - Own the data architecture for Unqork's AI-driven development layer — defining persistence, versioning, and query standards for AI-generated configurations. - Lead the design of declarative data models and schemas that enable non-technical users to build complex logic while maintaining strict data integrity. - Define the architectural boundary between database-layer computation (aggregation pipelines, indexing) and application-layer computation (Node.js post-processing, in-memory caching), and establish standards for which work belongs where. - Create and maintain comprehensive documentation including data architecture blueprints, indexing governance policies, query standards, and migration playbooks. - Own capacity planning and cost modeling for data infrastructure resources as Unqork scales. Leadership & Data Operations - Mentor and grow a team of data engineers and database engineers responsible for the health and performance of Unqork's data platform. - Establish data modeling best practices and enforce standardization across all environments — including schema conventions, index lifecycle management, and pagination contract design. - Oversee the design and operation of Unqork's database infrastructure — defining thresholds, coverage policies, write amplification limits, and manual override processes. - Drive data operational excellence by implementing and refining query performance monitoring, slow query alerting, explain plan review processes, and incident response playbooks for database degradation events. - Define and enforce data access governance — including RBAC data model standards, cache TTL policies, and the rules under which eventual consistency is acceptable vs. when strong consistency is required. - Comply with security regulations while working on data designs and patterns for Unqork platform. - Partner with Product to translate product requirements into data model decisions, and identify where relaxing a product constraint unlocks a disproportionate architectural improvement. Qualifications - 10+ years of progressive experience in data architecture, database engineering, or a related field, with at least 3 years in a principal or architect-level role. - Extensive experience designing and managing enterprise-grade, multi-tenant data infrastructure for SaaS platforms. - Expert-level proficiency with MongoDB — including aggregation pipeline design, index strategy (B-tree, text, vector), replica sets, sharding, and query execution plan analysis (IXSCAN vs. COLLSCAN). - Deep, hands-on expertise with our core data technology stack: - MongoDB / MongoDB Atlas (aggregation pipelines, Atlas Search, Atlas Vector Search, sharding) - Relational/SQL Databases (Operational and Business Intelligence schema and query partners) - Redis (caching strategy, TTL design, cache invalidation, pub/sub) - Node.js (application/database boundary, worker threads, event loop awareness) - RBAC and access control data patterns (denormalization, write-time materialization, owner list caching) - AI/ML data infrastructure (semantic search, LLM-friendly schema design, columnar database design) - Proven ability to write architectural decision records that hold up over time — capturing not just the recommendation but the alternatives considered and the conditions under which the decision should be revisited. - Proven ability to lead technical teams, manage complex data migration projects, and influence cross-functional stakeholders including Product and Engineering leadership. - Strong understanding of data security principles, multi-tenant isolation patterns, and enterprise compliance requirements (SOC 2, ISO 27001). Benefits - 💻 Work from home with a remote-first community - 🏝 Unlimited PTO (and the encouragement to use it) - 📝 Student loan payback program - 🏥 100% employer-covered medical, dental, and vision options available to you and your dependents - 💸 Flexible Spending Account (FSA) - 🏠 Monthly stipend toward your WFH setup, vacation, development and more - 💰 Employer-sponsored 401(k) with contribution match - 🏋🏻♀️ Subsidized ClassPass Membership - 🍼 Generous Paid Parental Leave Hiring Ranges - Tier 1: $229,000 - $286,200 - Tier 2: $215,100 - $268,900 Company Description Unqork embraces a culture of security and privacy awareness by consistently safeguarding sensitive information, adhering to company policies, and actively participating in training and initiatives to protect our data and the privacy of our stakeholders. Unqork is an equal opportunity employer. We will consider all qualified applicants without regard to race, color, nationality, gender, gender identity or expression, sexual orientation, religion, disability or age.




