Publications

(2025). MARVEL: Modular Abstention for Reliable and Versatile Expert LLMs. ICML 2025.
(2025). AutoScale-Automatic Prediction of Compute-optimal Data Composition for Training LLMs. COLM 2025.
(2025). Do Language Models Mirror Human Confidence? Exploring Psychological Insights to Address Overconfidence in LLMs. ACL 2025.
(2025). Know Your Limits: A Survey of Abstention in Large Language Models. TACL 2025.
(2024). Mitigating Overconfidence in Large Language Models: A Behavioral Lens on Confidence Estimation and Calibration. NeurIPS 2024 BML.
(2024). Characterizing LLM Abstention Behavior in Science QA with Context Perturbations. EMNLP 2024.
(2024). Laboratory-Scale AI: Open-Weight Models are Competitive with ChatGPT Even in Low-Resource Settings. FAccT 2024.
(2024). OmniMotionGPT: Animal Motion Generation with Limited Data. CVPR 2024.
(2023). InfoVisDial: An Informative Visual Dialogue Dataset by Bridging Large Multimodal and Language Models. Microsoft Research.
(2023). CCQ: cross-class query network for partially labeled organ segmentation. AAAI 2023.