Stellus Rx logo
Stellus Rx

Trusted, pharmacist-led health support in every moment that matters.

Senior Data Engineer

Location

Peru

Posted

2 days ago

Salary

0

Seniority

Senior

Job Description

Senior Data Engineer

Stellus Rx

• Develop, construct, and maintain large-scale data processing systems that collect data from a variety of structured and unstructured sources — using AI code generation tools to accelerate pipeline authoring, reduce boilerplate, and improve code quality. • Build and optimize ELT pipelines using AI-assisted tooling to identify bottlenecks, suggest optimizations, and automate routine pipeline maintenance tasks. • Identify, design, and implement internal process improvements: use AI to automate manual processes, optimize data delivery, and re-design infrastructure for greater scalability — replacing manual analysis with AI-driven discovery of improvement opportunities. • Build the infrastructure required for optimal extraction, transformation, and loading of data from various sources; use AI to accelerate infrastructure-as-code authoring and configuration. • Prepare data for data scientist exploration and discovery using AI-assisted data profiling and quality assessment tools — surfacing anomalies, schema drift, and data gaps faster than manual inspection allows. • Perform data wrangling and munging for downstream analytics and machine learning; leverage AI tools to generate and validate transformation logic against business rules. • Assemble large, complex datasets that meet functional and non-functional business requirements; use AI to rapidly evaluate dimensional modeling approaches and ontology alignment strategies. • Enable large-scale machine learning by designing and maintaining annotated datasets, elastic search approaches, and scalable data lake structures that support AI/ML workloads. • Create and maintain analytics pipelines that generate data and insight to power business decision-making; use AI-assisted analysis to proactively surface trends, anomalies, and opportunities within pipeline outputs. • Collaborate with data scientists, analysts, and business stakeholders on requirements for dimensional modeling, distributed ETL pipelines, and cross-repository data migration. • Evaluate, compare, and improve design patterns, data lifecycle approaches, and data ontology alignment — using AI to model trade-offs and accelerate proof-of-concept validation. • Work with data and analytics experts to continuously improve the functionality, reliability, and intelligence of data systems. • Perform root cause analysis on internal and external data and processes using AI-assisted investigation tools — replacing slow, manual log and lineage review with faster, AI-accelerated diagnostics. • Develop and maintain data quality frameworks; use AI to automate anomaly detection, schema validation, and data contract enforcement across pipelines. • Develop a strong understanding of company domains, strategic direction, and user needs to ensure data systems are aligned to business outcomes, not just technical requirements.

Job Requirements

  • 4+ years of experience in a Data Engineer role.
  • Graduate degree in Computer Science, Statistics, Informatics, Information Systems, or another quantitative field.
  • Advanced SQL knowledge and experience with relational databases and query authoring.
  • Demonstrated, hands-on experience using AI tools to accelerate data engineering tasks — pipeline development, data quality automation, code generation, or root cause analysis — with specific examples you can speak to.
  • Experience building and optimizing data pipelines, architectures, and datasets.
  • Strong analytic skills working with unstructured and disconnected datasets.
  • Experience with big data tools: Hadoop, Spark, Kafka, etc.
  • Experience with relational and NoSQL databases including Postgres and Cassandra.
  • Experience with pipeline and workflow management tools: Airflow, Luigi, Azkaban, or similar.
  • Experience with AWS cloud services: EC2, EMR, RDS, Redshift.
  • Experience with stream-processing systems: Storm, Spark Streaming, or similar.
  • Working knowledge of message queuing, stream processing, and highly scalable data stores.
  • Proficiency in object-oriented/scripting languages: Python, Java, Scala, C++, or similar.
  • Experience supporting cross-functional teams in dynamic, agile environments.
  • Familiarity with AI-assisted data quality or observability platforms (e.g., Monte Carlo, Soda, or similar).
  • Experience with LLM-based data processing pipelines or retrieval-augmented generation (RAG) architectures.
  • Healthcare data experience; familiarity with FHIR/HL7 standards a plus.

Benefits

  • Health insurance
  • Professional development opportunities

Related Categories

Related Job Pages

More Data Engineer Jobs

Stellus Rx logo

Senior Data Architect

Stellus Rx

Trusted, pharmacist-led health support in every moment that matters.

Data Engineer2 days ago
Full TimeRemoteTeam 201-500Since 2022

• Define and maintain enterprise data architecture standards across structured, semi-structured, and unstructured data domains. • Use AI-assisted modeling tools to accelerate data model design and validate designs against business requirements. • Design and govern the organization's cloud data lake, data warehouse, and lakehouse architectures on AWS. • Establish data ontology, taxonomy, and semantic layer standards that enable AI systems to reason over organizational data accurately. • Evaluate emerging data architecture patterns and build a roadmap for their adoption across Stellus Rx. • Design scalable data models and ELT/ETL pipeline architectures that support analytics and AI/ML workloads. • Define standards for data partitioning, indexing, caching, and storage optimization. • Partner with Data Engineers to translate architectural blueprints into production-ready pipelines. • Define and enforce data governance frameworks, data quality standards, and data contracts across the enterprise. • Develop and maintain a master data management (MDM) strategy.

Peru
interVal logo

Data Engineer

interVal

The Visibility Engine to Grow AUM in Less Time

Data Engineer2 days ago
Full TimeRemoteTeam 11-50Since 2019

• Design, develop, and maintain scalable data pipelines for ingestion, transformation, and delivery of large datasets across diverse industries. • Implement and ensure data privacy and security best practices, supporting data sovereignty and compliance with regulatory requirements. • Collaborate closely with AI/ML engineers, Data Scientists, and Platform engineers to enable advanced analytics and AI capabilities while retaining strict data control. • Optimize data platforms and systems for performance, reliability, and cost efficiency. • Build tools and frameworks for secure, privacy-preserving data processing and orchestration. • Develop and maintain documentation, data models, and technical workflows. • Partner with cross-functional teams to launch new data-driven product features and solutions.

United States
Full TimeRemoteTeam 501-1,000

• Utilize consulting and technical skills to be able to work in a client-facing project environment independently. • Be responsible for your own execution and sometimes lead individual work streams on client engagements as assigned and under supervision of engagement lead. • Collaborate with other team members to successfully deliver on projects. • Work effectively and directly communicate with both internal and client and/or partner teams. • Develop full ownership of your execution on client engagements, you'll become involved in the project planning and solution stages of engagements as well. • Design and implement complex ETL/ELT pipelines with evidence of improved data processing times. • Successfully lead small data warehousing projects with measurable performance enhancements under the management of an engagement lead. • Contribute to real-time data processing solutions and managed streaming data. • Implement security and compliance measures for data pipelines. • Design and implement version control and branching strategies and integrate them into CI/CD for promoting and testing in higher environments.

Illinois + 1 moreAll locations: Illinois | Virginia
$120K - $150K / year
Clover Health logo

Director of Engineering, Data Products

Clover Health

Clover is a healthcare technology company helping members live their healthiest lives with our Medicare Advantage plans.

Data Engineer2 days ago
Full TimeRemoteTeam 501-1,000H1B Sponsor

• Own the Data Products technical roadmap, canonical data modeling, and ingestion scalability. • Manage and optimize the clinical data ingestion pipeline. • Oversee the scaling of customer data onboarding processes and the continuous improvement of data hygiene, canonical storage standards, and observability and monitoring of the data pipelines. • Enable shared funnel visibility across the organization and support the development of a robust enterprise reporting framework. • Drive the adoption of an AI-powered Software Development Life Cycle (SDLC), leveraging AI tools to accelerate coding, testing, and deployment processes while maintaining stringent code quality and security standards.

California
$223.1K - $290K / year