Arcadia logo
Arcadia

We transform data into powerful insights that deliver results.

Analytics Engineer – Life Sciences Delivery Operations

Analytics EngineerAnalytics EngineerFull TimeRemoteSeniorTeam 201-500H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

3 days ago

Salary

$175K - $200K / year

Seniority

Senior

Bachelor Degree5 yrs expExperience acceptedEnglishAWSCloudPySparkPythonSDLCSparkSQL

Job Description

Analytics Engineer – Life Sciences Delivery Operations

Arcadia

• Author and maintain dbt models and PySpark transformation jobs, replacing ad-hoc Snowflake scripts with governed, version-controlled, tested code • Design and implement delivery endpoint configurations as code-customer, delivery target (Snowflake, S3), cadence, cohort filters, incremental and full historical refresh methods • Write production-grade Python and PySpark for data transformation, validation automation, and delivery pipeline components, including customer-specific data models and schema validation logic • Coordinate and execute monthly RWD deliveries across all active channel partners: delivery job execution, manifest generation and validation, tokenization workflows, and QC • Own the channel partner data inquiry queue-triage, investigate, resolve, and communicate on data questions and discrepancies; you are the primary research contact for channel partners • Follow SDLC best practices: author requirements, write test plans, manage releases, and maintain operating documentation in Confluence • Leverage AI tools (including Claude Code) to accelerate development, automate documentation, generate and verify code, and improve operational throughput

Job Requirements

  • Bachelor's or Master's degree in Computer Science, Data Science, Statistics, or a related field (or equivalent professional experience)
  • 5+ years of hands-on data engineering experience (production pipelines, dbt, Spark/PySpark, cloud data infrastructure) AND 5+ years of direct experience with life sciences RWD data (claims, EHR, clinical); these disciplines can overlap-5 years total is sufficient if you bring meaningful depth in both
  • Production-grade SQL proficiency in Snowflake or a comparable columnar warehouse: complex joins, CTEs, window functions, incremental patterns – you write this fluently
  • Python and/or PySpark for data transformation: you have written and debugged production Spark jobs, not just automation scripts
  • dbt: hands-on experience authoring models, tests, macros, and yml documentation; familiarity with incremental strategies and model validation
  • AWS S3: practical experience with file staging, delivery paths, bucket structure, and lifecycle management in a data engineering context
  • HIPAA de-identification: working knowledge of Safe Harbor requirements and how they are applied in data pipelines before data leaves your custody
  • SDLC fundamentals: you write requirements, author test plans, manage releases, and document your work – this is not new to you
  • CI/CD and source control: Git/GitHub, PR-based review workflows, branching strategies
  • External customer experience: you have led (or actively presented within) technical data discussions with partner analytics or science teams and can communicate complex data concepts clearly in writing and verbally
  • Self-starter who operates independently in ambiguous, high-growth environments and a natural collaborator when the work calls for it.

Benefits

  • Be at the center of a high-stakes, high-impact engineering RWD delivery pipeline you help create will define how Arcadia delivers RWD to life science partners at scale
  • Become the definitive internal expert on one of the most complex and valuable real-world healthcare datasets in the market, with the autonomy to shape how it is engineered, measured, and delivered
  • Be on the front lines of AI adoption-use cutting-edge tools to accelerate your work and shape how the team operates in an AI-first environment
  • Flexible, fully remote work environment, with resources and support to do your best work
  • Exposure to senior leaders across the entire life science and corporate engineering teams
  • A clear path to grow into a player/manager role as Arcadia's life sciences delivery team scales
  • Become a member of the talented, energized, diverse, and purpose-driven Arcadian community

Related Categories

Related Job Pages

More Analytics Engineer Jobs

Indigenous Pact PBC, Inc. logo

Data Platform & Analytics Engineer

Indigenous Pact PBC, Inc.

Increased Funding. Increased Access. Better Care.

Full TimeRemoteTeam 11-50H1B No Sponsor

• Design, develop, and maintain data pipelines, integrations, and automated data workflows • Build and support data lake, data warehouse, Databricks, and analytics platform solutions • Develop, optimize, and maintain ETL/ELT processes to ingest, transform, and organize data from internal and external systems • Create and maintain analytics-ready datasets, semantic models, and data structures optimized for Power BI consumption, in collaboration with BI Analysts on semantic model design • Develop and maintain data transformation workflows using SQL, Python, Databricks, and other modern data platform tools • Own operational health of the data platform, including monitoring, alerting, incident response, and continuous reliability improvements • Implement data validation, testing, reconciliation, and monitoring processes to ensure data accuracy and integrity • Develop and maintain technical documentation, data dictionaries, lineage documentation, and operational procedures • Collaborate with application developers, BI analysts, and business stakeholders to define data requirements and deliver data solutions • Establish and promote data engineering standards, best practices, and governance processes • Evaluate and recommend improvements to data architecture, tools, processes, and technologies • Administer database and data platform security, including access management, role-based permissions, and auditing of data access • Ensure data solutions comply with privacy, regulatory, and data governance requirements • Manage database and platform operations including backup, recovery, disaster recovery planning, and environment management • Monitor and optimize platform compute utilization and cost, including Databricks cluster configuration and workload efficiency • Implement data retention, archival, and cataloging policies • Participate in solution design, architecture reviews, and strategic data initiatives

California
$140K - $170K / year
Cribl logo

Staff Analytics Engineer

Cribl

Cribl, the Data Engine for IT and Security, empowers organizations to transform their data strategy.

Full TimeRemoteTeam 501-1,000Since 2017H1B Sponsor

• own the long-term architecture and evolution of Cribl's analytics engineering platform • design, build, and maintain certified dbt models as the authoritative source for business-critical metrics • establish and enforce analytics engineering standards for modeling, testing, documentation, and code review • design and maintain semantic and metadata layers that enable reliable AI-powered analytics and self-service • partner with analysts to migrate high-value business logic from Omni into governed warehouse models • partner with Data Engineering to improve source reliability, warehouse architecture, and Snowflake performance and cost efficiency • mentor analysts and analytics engineers on dbt development, data modeling, and analytics engineering best practices • work may happen across many time zones

California
$145K - $190K / year
Rackner logo

Senior Business Analyst – Data Quality Assurance Engineer

Rackner

DevSecOps and AI from Cloud to Mission Edge | Kubernetes Partner | Multicloud | 8(a) | HUBZone

Full TimeRemoteTeam 11-50H1B No Sponsor

• Own testing activities from planning through defect resolution and release validation. • Build technical depth across SQL, databases, ETL processes, APIs, and cloud applications. • Strengthen automated testing, data profiling, traceability, and quality practices. • Support healthcare systems where accurate, secure, and dependable data matters. • Review business and technical requirements and translate them into clear test conditions, acceptance criteria, and validation strategies. • Support requirements development through stakeholder discussions, user stories, process analysis, and backlog refinement. • Build, execute, and maintain automated and manual tests using established tools and frameworks. • Develop test plans and test cases for healthcare data, applications, reports, APIs, and user interfaces. • Validate data-ingestion and ETL processes by comparing source data, transformation rules, and target outputs. • Write SQL queries and joins to investigate data quality, confirm business rules, and identify missing, duplicate, inconsistent, or unexpected records. • Perform functional, system, integration, regression, component, and performance testing. • Maintain traceability between requirements, test coverage, defects, and delivery outcomes. • Support user-acceptance testing, defect triage, root-cause analysis, and resolution. • Partner with developers, analysts, product owners, and subject-matter experts to resolve complex application and data issues. • Document test results, risks, defects, and recommendations for technical and non-technical stakeholders. • Help improve testing tools, methodologies, and automation strategies.

United States
Full TimeRemoteTeam 5,001-10,000Since 1995H1B No Sponsor

• Build and maintain DBT models for cross-domain analytics use cases • Create reusable facts, dimensions, and marts for analytics consumption • Support cross-domain data products such as Churn, LTV, Segmentation, and Customer Intelligence • Implement dbt tests, source freshness checks, and reconciliation logic • Translate business requirements into data models and metric logic • Support data quality checks and model validation • Document model definitions, business rules, assumptions, and lineage • Collaborate closely with analytics and domain teams to ensure datasets are fit for reporting and decision-making

Brazil