Job Closed

This listing is no longer active.

Big Data Engineer

Location

United States

Posted

78 days ago

Salary

$104.6K - $156.8K / year

Seniority

Mid Level

Job Description

Big Data Engineer

AdvanSix

Role Description AdvanSix is seeking a Big Data Engineer to build and operate our enterprise Unified Data Layer (UDL) - spanning IT and OT - to deliver trustworthy, performant data products that power Finance, Operations, Supply Chain & Logistics, HSE, Commercial, and corporate analytics. You’ll engineer batch/CDC/streaming pipelines, model curated/semantic layers, and harden run-state with testing, CI/CD, security, and observability. You’ll partner closely with the data team and larger IT organization. Mission - Design and deliver scalable, secure data pipelines and data models that safely connect operational systems to analytics. - Ensure trusted and well‑governed data. - Enable repeatable delivery of BI, ML, AI, and automation solutions. Data Engineering & Modeling - Build ingestion pipelines (batch, CDC, streaming) from S/4HANA/DataSphere, PHD/historian, LIMS, TMS, HSE, and other sources into landing → curated → semantic layers. - Implement data contracts, schema/versioning, SCD handling, partitioning, and performance tuning (file formats, clustering, caching). - Develop dimensional/semantic models that back certified Power BI datasets and APIs for apps/agents. OT/IT Integration & Safety - Integrate OT data via OPC UA/MQTT, broker/DMZ patterns, read-only historian feeds, and event/batch frames—no control-net reads. - Collaborate with plant controls on change control, signal quality, and downtime windows. Quality, Security & Observability - Embed data quality rules, unit/integration tests, and validation checks (freshness, completeness, drift/PSI). - Instrument lineage and end-to-end monitoring; build alerting and on-call runbooks to minimize MTTR. - Enforce RBAC, secrets management, PII/HSE classifications, and retention aligned to Governance/MDM policies. CI/CD, Cost & Reliability - Automate build/test/deploy with Git-based CI/CD (environments, approvals, blue/green). - Track and optimize cost/performance (cluster sizing, autoscaling, cache strategy); contribute to FinOps reviews. Collaboration & Documentation - Partner with Reporting & BI on semantic model contracts, RLS, and performance SLAs; avoid direct system scraping. - Produce “readme” docs, data dictionaries, runbooks, and post-incident reviews; support knowledge transfer with vendors. Qualifications - Minimum 5 years' in data engineering building production pipelines at scale (batch/CDC/streaming). - Hands-on with Azure data stack: Databricks or Fabric/Synapse, ADF/Pipelines, ADLS/OneLake, Azure SQL/SQL MI, Key Vault. - Strong SQL and Python/PySpark; comfort with Spark Structured Streaming and performance tuning. - Experience implementing tests/observability (freshness, schema, expectations), and Git-based CI/CD. - Familiarity with SAP S/4HANA structures and SAP DataSphere semantic modeling. - OT concepts: historians (PHD/PI), OPC UA/MQTT, event/batch frames, ISA-95/99 basics. - Understanding of Power BI consumption (semantic models, RLS) and APIs for downstream AI/ML apps/agents. Preferred Qualifications - Time-series/data-quality tooling (e.g., Great Expectations or equivalent patterns), feature/metric stores. - MDM concepts (keys, survivorship), lineage/catalog tooling. - TMS/WMS, LIMS, Historian, HSE domain exposure; Lean/Six Sigma mindset; FinOps awareness. Benefits - We provide benefits that are industry competitive and focused on employee well-being. - Total Rewards program includes a competitive compensation, health, dental, vision & wellness programs, paid vacation, 401K with company matching, health savings programs, disability & life insurance, employee assistance program. - Tuition reimbursement for continued education, certifications, training, and development. - Work within a fast paced and innovative company, meeting passionate colleagues and partners with diverse backgrounds and experiences.

Related Categories

Related Job Pages

More Data Engineer Jobs

Full TimeRemoteTeam 10,001+Since 2006H1B Sponsor

• Partner with Technical Product Managers to understand product vision, roadmap priorities, and business objectives • Analyze and refine incoming development requests to clarify problem statements, scope, and success criteria • Analyze system interactions, integrations, APIs, and data flows to define expected behavior and outcomes • Create and maintain documentation such as process flows, data mappings, interface behavior descriptions, and business rules • Support backlog refinement by ensuring work items are well-defined, prioritized appropriately, and ready for delivery • Define acceptance criteria, test scenarios, and expected outcomes for validation activities • Support the TPM in stakeholder discussions by providing detailed analysis, examples, and clarifications, along with note taking and action item tracking

United States
$63K - $64K / year
Job Closed
Full TimeRemoteTeam 11-50H1B No Sponsor

• Design and build scalable, secure, and cost-effective data solutions in Snowflake • Develop and optimize data pipelines using tools such as dbt, Python, CloverDX, and cloud-native services • Participate in discovery sessions with clients to gather requirements and translate them into solution designs and project plans • Collaborate with engagement managers and account teams to help scope work and provide technical input for Statements of Work (SOWs) • Serve as a Snowflake subject matter expert, guiding best practices in performance tuning, cost optimization, access control, and workload management • Lead modernization and migration initiatives to move clients from legacy systems into Snowflake • Integrate Snowflake with BI tools, governance platforms, and AI/ML frameworks • Contribute to internal accelerators, frameworks, and proofs of concept • Mentor junior engineers and support knowledge sharing across the team

United States
Careerflow.ai logo

Norwegian Data Annotator

Careerflow.ai

Commitments Required: 40 hours per week with overlap of 6 hours with PST. Engagement type: Contractor (no medical/paid leave). Duration of contract: 6 months with opportunity to extend; expected start date is 1st week of Jun-2026. Location: North America and LATAM.

Data Engineer78 days ago
Full TimeRemoteTeam 11-50

Role Description We are looking for native or fluent language speakers to help train AI systems by reviewing and annotating content in their language (Norwegian). This is a long-term, remote gig. No prior experience is required. On the job training will be provided. What You Will Do - Review content in your native language (e.g., newspaper articles, website screenshots, advertisements) - Draw bounding boxes around specific elements as instructed (e.g., highlight the commercial aspect of an article) - Answer structured questions about the content - Write short summaries when required - Follow clear annotation guidelines to ensure quality and consistency Qualifications - Native or highly fluent speakers of the target language - Students, freelancers, part-time workers, or anyone seeking flexible short-term work - People comfortable following written instructions on a computer Requirements - Strong reading comprehension in your language - Basic computer skills (no technical background needed) - Attention to detail - Ability to follow structured instructions Role Details - Duration: Minimum 6 months (Long-term Project) - Pay Rate: $8 USD / hour - Hours: 30–40 hours/week - Location: Fully remote - Work: Not voice-based — text and image annotation only - No Interview, only assessment: 15–20 minute language skills assessment You will be contracted by Careerflow on behalf of one of our clients.

United States
$8 / hour

Role Description GoVivid Co. is seeking a detail-oriented Data Entry Admin to support our growing digital media and video production operations. This remote position plays a critical role in maintaining accurate records, organizing production data, and ensuring smooth workflow across internal systems. The ideal candidate is highly organized, tech-savvy, and comfortable working in a fast-paced, creative environment. - Accurately input, update, and maintain large volumes of data across internal systems and databases in a remote work environment - Organize and manage digital files related to video production, client projects, and content assets - Review data for errors, inconsistencies, and missing information, ensuring high levels of accuracy - Assist with tracking project timelines, deliverables, and production workflows - Maintain and update spreadsheets, CRM systems, and cloud-based platforms - Support internal teams by preparing reports and organizing operational data - Ensure proper documentation and filing of client and project-related information - Collaborate with cross-functional teams to streamline data handling processes in a remote setting Qualifications - Proven experience in data entry, administrative support, or similar roles - Strong typing skills with a high level of accuracy and attention to detail - Proficiency in Microsoft Excel, Google Sheets, and database systems - Ability to manage multiple tasks and meet deadlines in a remote work environment - Strong organizational and time management skills - Excellent communication skills and ability to work independently Requirements - Experience working with digital media, content production, or creative agencies - Familiarity with CRM tools and cloud storage platforms - Basic understanding of video production workflows is a plus Benefits - Flexible remote work environment - Opportunity to work with a fast-growing creative and video production company - Career growth and skill development opportunities - Collaborative and innovative team culture - Salary: 28 - 35 USD Per hour

Worldwide
$28 - $35 / hour