We're partners in transformation. We help clients activate ideas and solutions to take advantage of a new world of opportunity. We are a team of 80,000 strong, working with over 6,000 clients, including 80% of the Fortune 500, across North America, Europe and Asia.
Principal Data Scientist
Location
United States
Posted
5 days ago
Salary
$75 - $90 / hour
Seniority
Lead
Job Description
Principal Data Scientist
TEKsystems
Role Description TEKsystems is partnered with a software company in Raleigh that needs to hire a Senior Data Scientist for their flagship product, a leading global provider of information and analytics. Recently our client has focused on the general availability of their product for U.S. customers, a generative AI solution designed to transform legal work. Their product delivers trusted results in a familiar, easy-to-use interface with linked hallucination-free legal citations that combine the power of generative AI with proprietary search technology, Shepard’s® Citations functionality, and authoritative content. This role leads the design and development of an advanced multi-agent AI platform that powers intelligent research, drafting, and reasoning capabilities for large-scale enterprise knowledge environments. You will architect agent frameworks, optimize retrieval-augmented generation pipelines, fine-tune language models, and build the infrastructure that enables AI systems to collaborate, plan, and execute complex tasks reliably. The work directly shapes the next generation of AI-driven professional tools used by experts in high-stakes domains. Core Responsibilities - Architect and implement multi-agent systems capable of planning, tool use, and coordinated task execution. - Design and optimize RAG pipelines including embeddings, hybrid retrieval, reranking, and context-window strategies. - Fine-tune and evaluate small, medium, and large language models for domain-specific reasoning and summarization. - Develop prompt engineering frameworks, guardrails, and automated evaluation suites for agent reliability. - Build scalable ML services and APIs for production deployment in distributed environments. - Collaborate with product, engineering, and domain experts to translate complex workflows into agentic AI solutions. - Establish best practices for model evaluation, observability, safety, and compliance. - Mentor DS/ML engineers and contribute to long-term AI strategy and architecture. Qualifications - 6–12+ years in Data Science / ML Engineering, with deep experience in LLM-based systems. - Proven experience building agentic architectures (planner-executor, tool-use agents, ReAct-style reasoning). - Strong background in RAG, embeddings, retrieval optimization, and evaluation. - Expertise in NLP, transformers, deep learning, and model fine-tuning. - Proficiency with PyTorch, HuggingFace, LangChain/LlamaIndex, Ray, Kubernetes, and vector databases. - Experience designing production-grade ML systems with monitoring, evaluation, and observability. - Strong communication skills and ability to lead technical direction. Preferred Qualifications - Experience in enterprise search, knowledge management, or high-compliance domains. - Experience with model distillation, LoRA/QLoRA, PEFT, and model compression. - Experience building evaluation frameworks for hallucination, grounding, and agent reliability. - Familiarity with knowledge graphs, symbolic reasoning, or hybrid neuro-symbolic systems. - Publications, patents, or open-source contributions in LLMs or agent systems. - Strong coding skills in Python (7+ years). - Natural problem solver, able to take a lead in collaborating to resolve issues. - Proficiency in IDE debugging: VSCODE and PYCHARM. - 5+ years of experience in AI and machine learning. - Deep understanding of machine learning algorithms, classification models, diagnostic testing of models. - Experience working directly with Transformer-based architectures including BERT, RoBERTa, T5, etc., and familiarity with large language models and fine-tuning. - Experience with conversational search / semantic search, reinforcement learning, prompt engineering, hallucination mitigation. - Working understanding of the business risks associated with applying LLM (LangChain) in a business. - Experience working with AWS, RAG, SageMaker, SQL. Requirements - Expert Level experience. - Contract to Hire position based out of Raleigh, NC. Benefits - Medical, dental & vision. - Critical Illness, Accident, and Hospital. - 401(k) Retirement Plan – Pre-tax and Roth post-tax contributions available. - Life Insurance (Voluntary Life & AD&D for the employee and dependents). - Short and long-term disability. - Health Spending Account (HSA). - Transportation benefits. - Employee Assistance Program. - Time Off/Leave (PTO, Vacation or Sick Leave). Workplace Type This is a fully remote position. Application Deadline This position is anticipated to close on Jul 31, 2026.
Related Guides
Related Categories
Related Job Pages
More Data Scientist Jobs
• Leads and manages the entire lifecycle of data science projects, from conceptualization and design to development, deployment, and ongoing optimization. • Collaborates with cross-functional teams to define project scope, objectives, and success metrics. • Ensures projects align with organizational goals and deliver measurable impact on healthcare outcomes. • Leverages deep understanding of machine learning algorithms (decision trees, neural networks, graphical models, etc.) to build sophisticated predictive models for diverse healthcare applications. • Utilizes clustering, dimension reduction, and deep generative models to uncover hidden patterns and insights within large, complex healthcare datasets. • Applies rigorous validation techniques to ensure model accuracy, reliability, and fairness. • Oversees the deployment of models into production environments, ensuring seamless integration with existing systems. • Extracts insights from clinical and operational data sources (Epic Clarity, HL7, DICOM, and other enterprise data sources) to inform decision-making and guide project direction. • Translates complex technical findings into compelling narratives that resonate with non-technical stakeholders. • Facilitates data-driven decision-making by effectively communicating the value and impact of AI models. • Mentors and guides junior data scientists, fostering their professional growth and technical expertise. • Promotes a culture of collaboration, knowledge sharing, and continuous learning within the data science team. • Contributes to developing best practices and standards for data science and machine learning within the organization. • Stays abreast of the latest advancements in machine learning and healthcare research to identify opportunities for improvement and innovation. • Experiments with new approaches and technologies to enhance model performance and expand the organization's data science capabilities.
Role Description ClinChoice is searching for a Clinical R Programmer/Principal Data Scientist Consultant to join one of our clients. We are seeking a Principal Data Scientist Consultant with strong experience in R and a solid background in clinical programming. The ideal candidate will have hands-on experience developing SDTM and ADaM datasets using R, along with working knowledge of SAS. This role requires someone who can support clinical trial deliverables, ensure regulatory compliance, and collaborate closely with biostatistics and clinical data teams. - Develop, validate, and maintain SDTM and ADaM datasets using R following CDISC standards. - Support TLF (Tables, Listings, Figures) generation in R or SAS as needed. - Write efficient, reproducible, and well-structured R scripts for clinical data analysis and reporting. - Collaborate with statisticians, data managers, and clinical teams to understand programming requirements. - Perform QC checks, reconcile data issues, and ensure deliverables meet regulatory expectations (e.g., FDA, EMA). - Contribute to programming workflows, documentation, and version control best practices. - Support automation initiatives and R-based pipeline development. - Utilize SAS for legacy studies or where SAS support is required. Qualifications - Bachelor’s or Master’s degree in Statistics, Computer Science, Mathematics, Life Sciences, or related field. - 4–6+ years of experience in clinical programming, with a strong focus on R. - Proven experience in creating SDTM and ADaM datasets using R. - Working knowledge of SAS programming. - Solid understanding of CDISC standards (SDTM, ADaM). - Experience with clinical trial data, regulatory submissions, and QC processes. - Strong analytical, problem-solving, and documentation skills. Requirements - Experience with R packages such as tidyverse, haven, pharmaverse (e.g., admiral, tidyCDISC), or other clinical programming toolkits. - Understanding of R Markdown, Shiny apps, or reproducible reporting tools. - Exposure to GxP validation, version control (Git), and automated workflows. - Experience working in a CRO or pharmaceutical environment. Company Description ClinChoice is a global full-service CRO specializing in clinical development and functional solutions for pharmaceutical, biotechnology, medical device, and consumer health companies. We have over 28 years of proven high-quality delivery and results across all our services, with over 4,000 professionals in more than 20 countries across the Americas, Europe, and Asia-Pacific. Our mission drives our culture: to contribute to a healthier and safer world by accelerating the development and commercialization of innovative drugs and devices. Our employees are the most valuable company asset, and they are the fulcrum around which all ClinChoice activities are built, and close management and training are the core instruments to develop and maintain highly-qualified personnel. The continuous training keeps the resources qualified in terms of competence and expertise and gives all personnel the clear tools needed to manage both internal and client processes with the same methodology. The success of these core values is evidenced by our low industry-average turnover rates. ClinChoice is an equal opportunity employer. We have based our success on attracting, developing, and promoting talent, guided by a commitment to diversity and inclusivity. Our employees come from very diverse backgrounds: gender, race, beliefs, and ethnicities. We recognize this is our strength and celebrate it.
Role Description As a Senior Data Scientist at the Experian Innovation Lab, your role will be to research and develop analytical solutions, prototype new products, and evaluate data assets. You must bring an experience in predictive modeling, machine learning, and deep learning to this position. You will report into the Sr. Principal Data Scientist. You'll have the opportunity to: - Create advanced machine learning analytical solutions to extract insights from diverse structured and unstructured data sources. - Unearth data value by selecting and applying the right machine learning, deep learning and processing techniques. - Refine data manipulation and retrieval through the design of efficient data structures and storage solutions. - Innovate with tools designed for data processing and information retrieval. - Dissect and document vast datasets, analyzing them to highlight patterns and insights. - Solve complex challenges by developing impactful algorithms. - Ensure model excellence by validating performance scores and analyzing Return on investment and benefits. Qualifications - PhD degree in Machine Learning, Data Science, AI, Computer Science, or a related quantitative field. - 2+ years of experience in AI, data science, or predictive modeling. - Proficiency in multiple programming languages, including Python. - Experience in deep learning (CNN, RNN, LSTM, attention models), machine learning methodologies (SVM, GLM, boosting, random forest), graph models, or reinforcement learning. - Experience with open-source tools for deep learning and machine learning technology such as pytorch, Keras, tensorflow, scikit-learn, pandas. - Experience with large data analysis using pySpark. - Experience with LLMs and the relevant tools in the Generative AI domain. - Experience with Hadoop and NoSQL related technologies such as Map Reduce, Hive, HBase, mongoDB, Cassandra. - Experience modifying and applying advanced algorithms to address practical problems. - Experience with online, mobile marketing analytics. - Experience with GPU programming. Benefits - Great compensation package and bonus plan. - Core benefits including medical, dental, vision, and matching 401K. - Flexible work environment, ability to work remote, hybrid or in-office. - Flexible time off including volunteer time off, vacation, sick and 12-paid holidays.
• Lead a data science team and cross-functional projects • Innovate and improve machine learning models • Partner with business leaders and technical experts across the company • Develop new data sources and improve our modeling methodology • Apply models with sound risk management


