wagey.ggwagey.gg
29,762  jobs29,762  jobs
Browse Tech JobsCompaniesFeaturesPricingFAQs
Log InGet Started Free
Jobs(29,762)/Senior Data Scientist Role(355)/paradigm-health (4) - Senior Data Scientist - LLM Evaluation & Business Intelligence
Pro members applied to this job 36 hours before you saw itGet Pro ›
paradigm-health

paradigm-health - Senior Data Scientist - LLM Evaluation & Business Intelligence

Remote - US-Remote5d ago
RemoteSeniorNABiotechnologyLife SciencesSenior Data ScientistData ScientistReportingPythonDocumentationBusiness IntelligenceSQLdbtHexDatabrickshypothesisCloseMentoringData Governance

Requirements

• You are a rigorous, versatile data scientist who is energized by making AI systems measurable and trustworthy. You care about getting the number right and being able to defend it. You're equally comfortable designing an evaluation framework for an LLM pipeline and writing the SQL and dbt models that power reliable reporting. • You bring a strong interest in clinical research and a passion for advancing healthcare through data innovation. You're autonomous and driven — you can take an ambiguous problem and run with it, structuring the approach yourself rather than waiting for a fully specified spec. You're curious, collaborative, and hold yourself to data science best practices as a default, not an afterthought. • Bachelor's degree or higher in data science, statistics, biostatistics, computer science, mathematics, epidemiology, or a related field. • 4+ years of experience as a data scientist, analytics engineer, ML/data professional, or other highly analytical role in life sciences, biotech, healthcare, or a related industry. • Solid understanding of how production LLM pipelines work, and hands-on experience evaluating LLM or NLP systems — accuracy metrics, error analysis, RAG/retrieval quality, prompt evaluation, or similar. • Strong proficiency in SQL, with experience building production data models in dbt (or a comparable transformation framework). • Proficiency in at least one programming language (Python preferred; PySpark a plus). • Experience with modern data platforms such as Databricks, and BI/analytics tools such as Hex, or similar. • Fluency with data science best practices: reusability, version control, documentation, testing, code review, and reproducibility. • Excellent written and verbal communication, with the ability to translate technical concepts into actionable insights for non-technical stakeholders and to develop reporting logic collaboratively with them. • Comfort with AI tooling and agentic workflows in day-to-day work. • Comfort with ambiguity and adaptability to a mission-driven, fast-paced startup environment. • Master's degree or higher in a quantitative field. • Experience establishing or working within a semantic layer for organization-wide reporting. • Experience integrating LLM-based workflows into production systems, and working closely with ML/AI engineering teams. • Familiarity with clinical data elements (oncology a plus) and the clinical trial industry or research operations. • Strong statistical knowledge, including regression, classification, hypothesis testing, causal inference, or Bayesian methods (e.g., continuous toxicity monitoring for trial protocols). • Prior experience mentoring other data scientists or leading cross-functional projects (for senior candidates). • At Paradigm Health, we are committed to providing equal employment opportunities to all qualified individuals. We encourage and welcome candidates from all backgrounds and perspectives to apply for our open positions. We are interested in all qualified individuals and ensure that all employment decisions are based on job-related factors such as skills, experience, and qualifications.

Responsibilities

• Design and run evaluations of production LLM pipelines, developing accuracy and quality metrics that tell us how well our systems perform on real clinical tasks. • Support evaluation of our Patient Trial Evaluation (PTE) pipeline — assessing prompts and trial-matching logic for accuracy in reporting, and surfacing where and why they fail. • Measure and improve the effectiveness of our retrieval-augmented generation (RAG) systems, from retrieval quality through final output. • Use accuracy metrics to inform cost/quality tradeoffs across trials, prompts, and models, giving the team a clear basis for which approaches to ship. • Establish and maintain a library of "best prompts" for recurring clinical concepts, backed by evidence rather than intuition. • Partner with AI Engineering to close the loop between evaluation findings and production improvements. • Build and maintain production-grade data models using SQL and dbt on Databricks, ensuring analytics logic is reliable, maintainable, and production-ready. • Help establish and grow a semantic layer that enables consistent, trusted reporting across the organization. • Advance data governance practices across dbt, Databricks, and Hex — documentation, testing, versioning, and clear ownership of metrics. • Collaborate with non-technical internal stakeholders to develop reporting logic, answer data questions, and translate business needs into well-defined data products. • Support both internal and external reporting needs with accurate, well-tested datasets. • For senior applicants: mentor and support junior data scientists — providing guidance on evaluation and analytics approaches, technical best practices, and professional growth.

Apply in one click

Upload My Resume

Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT

Apply in One Click
Apply in One Click

Similar roles

RedditReddit - Senior Staff Data Scientist - Consumer Experimentation1mo ago
·Remote - Ontario, Canada
RemoteNAStaffData AnalyticsData ScientistSenior Data ScientistSQLPythonhypothesisMentoringDocumentation
omnidianomnidian - Senior Data Scientist1w ago
·Remote - USA·Equity
RemoteNASeniorDiagnosticsCloud ComputingData ScientistSenior Data ScientistLearning & DevelopmentReportingPythonscikit-learnSQLPandasAWSAzureClaudeNumPyMLflowDatabricksMentoringCloseCAIA
kariusdxkariusdx - Senior Clinical Data Scientist1w ago
·Redwood City, CA (Hybrid) or Remote (USA) - Hybrid
In OfficeNASeniorGenomicsClinical ResearchData ScientistSenior Data ScientistSQLReportingACCADocumentationData QualityData VisualizationData Governance
HandshakeHandshake - Senior Data Scientist, Product Analytics1w ago
·Remote - San Francisco, California, United States·$162k - $203k/year + Equity
RemoteNASeniorArtificial IntelligenceSenior Data ScientistData ScientistReportingModeGoal SettingSQLDecision MakingJupyterPythonPandasHex
RedditReddit - Senior Staff Data Scientist - Consumer Relevance1mo ago
·Remote - USA·$233k - $233k/year + Equity
RemoteNAStaffData AnalyticsData ScientistSenior Data ScientistSQLPythonMentoring
RevenueCatRevenueCat - Senior Data Scientist1w ago
·Remote - Americas; EMEA - USA *·$220k - $220k/year + Equity
RemoteNASeniorData AnalyticsData ScientistSenior Data ScientistSQLPythonRampDecision Making
MozillaMozilla - Senior Staff Data Scientist1w ago
·Remote - Canada; Remote US·Equity
RemoteNAStaffData AnalyticsNonprofitData ScientistSenior Data ScientistSQLPython
MozillaMozilla - Senior Staff Data Scientist1w ago
·Remote - Remote; Remote Canada; Remote US·Equity
RemoteNAStaffData AnalyticsNonprofitData ScientistSenior Data ScientistSQLPython

Browse more by category

Show 355 moreSenior Data ScientistShow 276 moreData ScientistShow 6,158 moreReportingShow 4,582 morePythonShow 4,589 moreDocumentationShow 226 moreBusiness IntelligenceShow 2,477 moreSQLShow 311 moredbtShow 60 moreHexShow 300 moreDatabricks
Privacy·Terms··Contact·FAQ·Wagey on X