| Aug 24, 2026 | Two of our papers have been accepted to Findings of EMNLP 2026! Does Accuracy Equal Evidence? studies reasoning faithfulness under KV-cache compression, while LONGQAEVAL develops reliable, resource-conscious evaluation methods for long-form clinical QA. Congratulations to the students and collaborators who made this work possible! 🎉 |
| Jul 20, 2026 | A clean sweep at COLM 2026: 3/3 papers accepted! ReLay studies the benefits and risks of personalized health summaries; MedConceal tests how well models uncover and address patients’ hidden concerns; and SEA learns reusable diagnostic rules through jointly optimized reasoning and dual memory. I’ll be in San Francisco for the conference—please say hello! 🌉🎉 |
| May 26, 2026 | MedGuideX turns clinical practice guidelines into executable decision logic for LLM training, improving average clinical-reasoning accuracy by 10.28% and producing physician-preferred rationales. 🩺 |
| Apr 30, 2026 | Two papers are now published in the Journal of Biomedical Informatics: PlainQAFact introduces a sentence-aware, retrieval-augmented metric for checking factual consistency in biomedical plain-language summaries, while our crowdsourced study finds that LLM summaries can feel as clear as human-written ones yet produce worse actual comprehension. Reliable health communication must measure both factuality and what readers truly understand. 🧭 |
| Mar 29, 2026 | Our new counterfactual multi-agent framework tests how changing individual clinical findings shifts competing diagnoses, improving diagnostic accuracy and making agent reasoning more evidence-grounded on complex cases. 🔄 |
| Jan 29, 2026 | EHR-RAG combines event- and time-aware retrieval with iterative and counterfactual evidence reasoning for long-horizon health records, improving Macro-F1 by 10.76% over strong LLM baselines. 🏥 |
| Oct 12, 2025 | LONGQAEVAL shows that long-form clinical QA evaluation should mix sentence- and answer-level judgments by criterion, and that sampling a small subset of sentences can retain reliability while reducing expert annotation cost. 📋 |
| Sep 20, 2024 | Excited to share that our paper for plain language summarization evaluation has been accepted to #EMNLP2024! 🌟 I’ll be in Miami for the conference—feel free to reach out if you’d like to chat! 📩 |
| Jan 15, 2024 | Check out my new publication: Retrieval Augmentation of Large Language Models for Lay Language Generation |