Publications
35 works across longitudinal EHR modeling, clinical NLP, multi-agent LLM systems, and biomedical representation learning, grouped by type and listed newest first.
Journal articles
5- 2026
Phenotypic prediction of missense variants via deep contrastive learning
Nature Biomedical Engineering
PheMART connects missense variants to 4,179 clinical phenotypes in a shared metric space, combining protein language models, interaction networks, medical knowledge graphs, and EHR-derived signal to support rare disease diagnosis.
- 2026
Genomic classification to predict survival in metastatic prostate cancer: development of Somatic Tumor Risk Assessment for Overall Survival–Prostate (STRATOS-P)
JCO Precision Oncology
A prognostic classifier built from comprehensive genomic profiling of 7,201 veterans with metastatic prostate cancer in the VHA National Precision Oncology Program.
- 2025
The role of Whole Health in enhancing tobacco cessation outcomes for veterans: a retrospective cohort study
Journal of General Internal Medicine
Links participation in Whole Health services to measurable tobacco cessation outcomes across the Veterans Health Administration.
- 2024
CoRTEx: contrastive learning for representing terms via explanations with applications on constructing biomedical knowledge graphs
Journal of the American Medical Informatics Association 31(9), 1912–1920
Explanation-augmented contrastive learning that sharpens biomedical term representations for large-scale knowledge graph construction.
- 2021
A feasibility study of 2-D microwave thorax imaging based on the supervised descent method
Electronics 10(3), 352
Learning-based 2-D microwave thorax imaging via the supervised descent method for fast structural reconstruction.
Conference and workshop papers
11- 2026
CONTEXTOR: contextualized high-order contrastive learning
International Conference on Machine Learning (ICML)
Recasts high-order relation inference as a dynamic query–response process, contextualizing candidate entities through asymmetric conditional modulation.
- 2026
MARTI: a framework for multi-agent LLM systems reinforced training and inference
International Conference on Learning Representations (ICLR)
An open framework unifying reinforced training and inference for multi-agent LLM systems.
- 2025
ReviewRL: towards automated scientific review with reinforcement learning
Conference on Empirical Methods in Natural Language Processing (EMNLP) Main conference
Trains review generation with verifiable rewards so that automated reviews stay grounded in the paper rather than drifting into fluent generality.
- 2025
TrajSurv: learning continuous latent trajectories from electronic health records for trustworthy survival prediction
Machine Learning for Healthcare Conference (MLHC) PMLR 298
Learns a continuous-time latent trajectory per patient, so survival predictions come with an inspectable account of how risk accumulated.
- 2025
Traj-CoA: patient trajectory modeling via chain-of-agents for lung cancer risk prediction
NeurIPS GenAI4Health Workshop
A chain of LLM agents reads a patient's notes in order and passes forward a compressed timeline, making zero-shot lung cancer risk prediction traceable to source evidence.
- 2025
UW-BioNLP at ChemoTimelines 2025: thinking, fine-tuning, and dictionary-enhanced LLM systems for chemotherapy timeline extraction
Clinical NLP Workshop (ACL) pp. 40–56 1st place, subtask 2
Winning system for generating patient chemotherapy timelines from raw clinical notes, combining fine-tuning, chain-of-thought, and dictionary-enhanced retrieval.
- 2025
Scalability of LLM-based multi-agent systems for scientific code generation: a preliminary study
MathNLP Workshop (EMNLP)
Shows a minimalist actor–critic pair can beat a reasoning model at equal compute while cutting token cost by 75%.
- 2024
UltraMedical: building specialized generalists in biomedicine
NeurIPS Datasets and Benchmarks Track Spotlight
A large biomedical instruction dataset and model suite that reaches domain specialization without giving up general capability.
- 2024
Large language models as biomedical hypothesis generators: a comprehensive evaluation
Conference on Language Modeling (COLM)
Evaluates biomedical hypothesis generation on date-partitioned background–hypothesis pairs to control for contamination.
- 2023
Large language models are zero shot hypothesis proposers
NeurIPS Instruction Tuning Workshop
Tests whether language models can propose scientifically meaningful biomedical hypotheses without task-specific training.
- 2022
Automatic biomedical term clustering by learning fine-grained term representations
BioNLP Workshop (ACL) pp. 91–96
Fine-grained term representations that separate near-synonyms well enough to cluster biomedical terminology automatically.
Preprints and working papers
9- 2026
Traj-Evolve: a self-evolving multi-agent system for patient trajectory modeling in lung cancer early detection
arXiv:2606.02812
Pairs an experience pool of retrieved similar patients with RL-optimized agent coordination, beating nine baselines over five years of patient history.
- 2026
NatureBench: can coding agents match the published SOTA of Nature-family papers?
arXiv:2606.24530
Ninety tasks drawn from peer-reviewed Nature papers. The best agent exceeds published results on 17.8% of them, mostly by reducing science to familiar supervised learning.
- 2026
EpiEvolve: self-evolving agents for streaming pandemic forecasting under regime shifts
arXiv:2606.05513
Keeps a fixed LLM forecaster but adapts through episodic memory, halving the time to recover after a variant regime shift.
- 2026
TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection
arXiv:2604.10386
A training-free multi-agent LLM framework with memory that matches or beats supervised baselines across 15 cancer types (AUC 0.64–0.80) while producing reviewable evidence.
- 2026
Self-improving agents in the era of experience: a survey of self- to meta-evolution
Preprint
Surveys agents that improve themselves, and the meta-layer that decides which parts of the improvement process should change.
- 2025
A survey of reinforcement learning for large reasoning models
arXiv:2509.08827
A broad survey of RL for reasoning models: foundational components, open problems, training resources, and where scaling runs into walls.
- 2025
CaSBRE: causality-inspired semi-supervised biomedical relation extraction
Under review
Disentangles causal from spurious features so relation extraction stays robust with scarce labels and unseen entities.
- 2023
Hierarchical pretraining for biomedical term embeddings
arXiv:2307.00266
Hierarchy-aware pretraining that encodes graded semantic relatedness between biomedical terms.
- 2022
BIOS: an algorithmically generated biomedical knowledge graph
arXiv:2203.09975
A biomedical knowledge graph built by machine learning rather than manual curation, at a scale manual curation cannot reach.
Conference abstracts
9- 2026
Predicting multi-cancer risk from EHR data using multi-agent LLMs
Journal of Clinical Oncology ASCO Annual Meeting
- 2026
STRATOS-P clinical model for prognosis and selection of patients with metastatic hormone-sensitive prostate cancer for intermittent therapy
Journal of Clinical Oncology ASCO Annual Meeting
- 2026
Development of an oncology generative AI foundation model trained on more than a million longitudinal patient journeys across the United States
Journal of Clinical Oncology ASCO Annual Meeting
- 2026
MSR67 — Zero-shot lung cancer risk prediction from longitudinal electronic health records with a chain-of-agents framework
Value in Health ISPOR
- 2026
MSR68 — Characterizing oncology patient journeys and health state transitions using a data-driven Markov transition matrix in large-scale electronic health records
Value in Health ISPOR
- 2026
MSR172 — Can a generative patient journey foundation model alleviate the burden of cancer screening?
Value in Health ISPOR
- 2026
RWD140 — Patient journey foundational model for scalable imputation of missing units of measurement in electronic health records data
Value in Health ISPOR
- 2026
Adapting the Global Burden of Disease Healthcare Access and Quality Index for the Veterans Health Administration: a feasibility study
AcademyHealth Annual Research Meeting
- 2025
Population-level tobacco cessation outcomes associated with implementing Whole Health at the Veterans Health Administration
AcademyHealth Annual Research Meeting
Thesis
1- 2026
Towards trustworthy modeling of patient trajectory with longitudinal electronic health records
PhD dissertation, University of Washington
Advised by Ruth Etzioni and Meliha Yetisgen. Degree conferred June 2026.