TetraScience logo
TetraScience

Open | Cloud-Native | Purpose-Built for Science

Principal Architect, Platform & Data Lake

Data EngineerData EngineerFull TimeRemoteLeadTeam 51-200Since 2015H1B SponsorCompany SiteLinkedIn

Location

United States

Posted

5 days ago

Salary

0

Seniority

Lead

Job Description

Principal Architect, Platform & Data Lake

TetraScience

Role Description TetraScience is building the scientific data and AI cloud for biopharma. The platform is the innermost foundation for developing, delivering and operating our enterprise grade, secure, compliant Scientific Data and AI capabilities customers rely on. In this role, you will own the platform architecture, evolution and growth scaling across: - Enterprise Platform - Scientific Search - AI/ML Ops - Developer Platform - Developer Productivity - Lakehouse Platform - Partner Integrations - Cloud Infrastructure This is a senior IC leadership role. You set technical direction, own the decisions that cross team boundaries, and close architectural gaps before they become business risks. After achieving strong product-market fit and traction, we are entering a growth scaling phase where we are expanding our industry partnerships and developer experience to rapidly build the foundations of AI-native scientific data and workflows in production. The scope of this role is intentionally broad. We are looking for experienced candidates who cover a majority of these areas. Strong candidates bring deep fingerprints in one of two architectural profiles, with meaningful range across both: - Enterprise Data & AI Platforms - Data, Knowledge, and Developer Products Qualifications - 12+ years in software engineering, with at least 5 at staff or principal level in a SaaS platform or data infrastructure context. - Deep architecture ownership in at least one of the two fingerprint profiles above, with meaningful range across the other. - Demonstrated ownership of enterprise authentication and authorization systems at scale: SAML, OIDC, fine-grained RBAC across a multi-tenant SaaS product. - Hands-on experience with AI/ML serving infrastructure: you have built and operated model inference pipelines under production load. - Search architecture experience: you have designed and operated a search platform that handles diverse query types (keyword, semantic, or hybrid) across large structured or semi-structured datasets. - Hands-on experience with data lake architectures at scale: Delta Lake or Apache Iceberg, schema evolution patterns, partition pruning, and the trade-offs between query performance and storage cost. - Infrastructure fluency on AWS with Kubernetes or ECS. - Ability to write and defend architecture decisions: RFCs, trade-off documents, design reviews. - Strong cross-team communication. - Comfort operating across strategy, architecture, and operations in the same week. Requirements - Authn/Authz architecture is documented, consistent across services, and passing enterprise security reviews without heroics from a single engineer. - AI/ML infrastructure has a clear architecture and roadmap for MLE inference and training use cases, with strong operational telemetry and cost visibility. - The developer platform has clear SDKs and a set of standard templates for scientific use cases to start from, with adoption and delivery by multiple scientific use case teams. - Operational excellence based on a clear O11y architecture rolled out, with every production service having SLOs defined, monitored and managed. - Cost governance with customer chargeback attribution architecture and operationalized with the finance and field teams. - Lakehouse platform architecture and operational buildout as a Data Products Platform with strong DX and operational scaling. - Evolve IDS to open standards based schema and encoding with strongly typed data models and schema-on-write enforcement. - Published reference architecture for each partner class (lab instrument manufacturers and AI models), with one partner successfully onboarded against each without bespoke engineering support. Benefits - Competitive compensation with equity - Unlimited PTO - Company-paid Life Insurance, LTD/STD - 401(k)

Related Categories

Related Job Pages

More Data Engineer Jobs

Role Description We are seeking an experienced Senior Data Architect to design, develop, and manage enterprise data warehouse solutions. The ideal candidate will have strong expertise in: - SQL Server - Data modeling - Database architecture - ETL - Data governance Experience with Azure DevOps, .NET/C#, or Java is preferred. Candidates should have 8+ years of IT experience, including at least 5 years in data architecture and data warehouse design. Qualifications - Data Warehouse & Data Modeling - SQL Server & SQL - Database Architecture - Data Governance & Master Data - Azure DevOps - .NET/C# or Java - Strong analytical and communication skills Requirements - 8+ years of IT experience - At least 5 years in data architecture and data warehouse design Company Description

United States
$45 - $50 / hour

Senior Data Engineer

UnitedHealth Group

UnitedHealth Group is a healthcare and well-being company that’s dedicated to improving the health outcomes of millions around the world. We are comprised of

Data Engineer5 days ago

Role Description The Senior Data Engineer – Exadata Platform Engineering is responsible for the administration, maintenance, performance optimization, and lifecycle management of Oracle Exadata engineered systems, including both On-Premises Exadata and Exadata Cloud Services (ExaCS). This role focuses on Exadata infrastructure and cloud environments. This position ensures the availability, performance, scalability, and security of mission-critical, large-scale platform environments while providing operational support, capacity planning, automation, monitoring, and troubleshooting across hybrid cloud and on-premises environments. Responsibilities include: - Administer and maintain Oracle Exadata Database Machine, ZDLRA, and Exadata Cloud environments, including compute nodes, storage cells, operating systems, networking, and supporting platform components. - Manage Exadata infrastructure performance, capacity planning, storage utilization, and growth forecasting to ensure platform scalability and reliability. - Configure, monitor, and optimize Exadata-specific infrastructure technologies, including Smart Scan, Storage Indexes, Smart Flash Cache, and IORM. - Administer and support Exadata networking, backup infrastructure, and disaster recovery platform components. - Monitor system health, performance, and availability using Oracle Enterprise Manager (OEM), including alert management, reporting, and operational support. - Support Exadata deployments, migrations, and platform operations across on-premises, OCI, Exadata Cloud Service (ExaCS), KVM, OCI@Azure, and OCI@AWS environments. - Develop and maintain automation solutions for platform administration, monitoring, patching, and operational processes using Ansible, Python, and Shell scripting. - Troubleshoot and resolve complex Exadata platform issues. - Participate in incidents, problems, change, and on-call support processes to ensure platform stability, availability, resiliency, and operational excellence. You will enjoy the flexibility to telecommute from anywhere within the U.S. as you take on some tough challenges. Qualifications - High School Diploma/GED - 5+ years of hands-on experience administering Oracle Exadata engineered systems in production environments - 5+ years of experience supporting Exadata infrastructure, including storage cells, compute nodes, operating systems, and networking - 5+ years of experience monitoring and optimizing Exadata environments using features such as Smart Scan, Storage Indexes, Smart Flash Cache, and IORM - 5+ years of experience with Exadata capacity planning, storage management, backup infrastructure, and high availability/disaster recovery platforms - 4+ years of experience troubleshooting complex Exadata infrastructure issues - 4+ years of experience automating administrative processes using Ansible, Python, Shell scripting, or similar tools - 4+ years of experience with Incident, Problem, and Change Management processes in enterprise production environments Requirements - Undergraduate degree or equivalent experience - 1+ years of experience with ZDLRA administration would be an added benefit - Experience supporting Exadata Cloud Service (ExaCS), OCI, Oracle Database@Azure, or Oracle Database@AWS environments strongly preferred - Experience with OEM Administration would be an added benefit - Experience in ticketing tools such as ServiceNow Benefits - Comprehensive benefits package - Incentive and recognition programs - Equity stock purchase - 401k contribution (all benefits are subject to eligibility requirements) Application Deadline This will be posted for a minimum of 2 business days or until a sufficient candidate pool has been collected. Job posting may come down early due to volume of applicants.

United States
$91.7K - $163.7K / year
Full TimeRemoteTeam 1,001-5,000H1B No Sponsor

• Lead the modernisation and transformation of Newsquest's data platform and supporting architecture, creating a scalable, secure and sustainable foundation for future growth. • Define and deliver a roadmap for data migration, integration, governance, reporting and analytics capabilities. • Lead the implementation of a modern data platform, including the development of a data warehouse, data lake, lakehouse or equivalent cloud-based solution. • Work with stakeholders across the business to understand requirements and translate them into practical data solutions that deliver measurable business value. • Collaborate closely with IT teams, data engineers, system owners and third-party suppliers to ensure successful project delivery. • Manage programme plans, budgets, risks, dependencies and resources across multiple workstreams. • Establish and embed robust data governance, ownership and quality standards across the organisation, increasing confidence and trust in business data. • Lead and support internal data engineering resources, ensuring successful delivery of platform objectives. • Drive platform adoption across the business, improving accessibility to trusted data and reducing reliance on manual processes. • Oversee post-implementation stabilisation, optimisation and continuous improvement activities, ensuring the platform continues to evolve alongside business requirements. • Identify future opportunities to enhance reporting, analytics, automation and data-driven decision-making.

United Kingdom
Keyrus logo

Data Engineer – GCP (Mid/Senior)

Keyrus

#MakeDataMatter #HumanizingTheFuture

Data Engineer5 days ago
Full TimeRemoteTeam 1,001-5,000Since 1996H1B Sponsor

• Modern, scalable, and secure data architectures on Google Cloud Platform. • High-performance data pipelines for ingestion, transformation, and data delivery. • Analytical solutions for Data Lake, Data Warehouse, and Data Mart. • Robust and optimized data models for analytical consumption and decision-making. • Environments with governance, data quality, security, and compliance across the entire data lifecycle.

Brazil