Web IR NLP Group @ NUS
Web IR NLP Group @ NUS
Tour
News
People
Events
Publications
Projects
Contact
Light
Dark
Automatic
Paper-Conference
Discursive Circuits: How Do Language Models Understand Discourse Relations?
We introduce a new task, Completion under Discourse Relation (CuDR), to reveal the underlying transformer circuits responsible for discourse understanding.
Yisong Miao
,
Min-Yen Kan
PDF
Cite
Code
Project
Beyond In-Context Learning: Aligning Long-form Generation of Large Language Models via Task-Inherent Attribute Guidelines
LongGuide generates task-inherent metric and output-constraint guidelines to improve long-form LLM generation beyond demonstrations alone.
Xuan Long Do
,
Duong Ngoc Yen
,
Trong Xuan Do
,
Anh Tuan Luu
,
Kenji Kawaguchi
,
Shafiq Joty
,
Min-Yen Kan
,
Nancy F. Chen
PDF
Cite
What Makes a Good Natural Language Prompt?
A meta-analysis of prompting research that organizes natural-language prompt quality into 21 properties across six dimensions.
Xuan Long Do
,
Duy C. Dinh
,
Hai N. Nguyen
,
Kenji Kawaguchi
,
Nancy F. Chen
,
Shafiq Joty
,
Min-Yen Kan
PDF
Cite
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
A systematic evaluation of output-format bias in LLMs, with prompting and fine-tuning strategies for reducing format sensitivity.
Xuan Long Do
,
Hai N. Nguyen
,
Tiviatis Sim
,
Hieu Dao
,
Shafiq Joty
,
Kenji Kawaguchi
,
Nancy F. Chen
,
Min-Yen Kan
PDF
Cite
Aligning Large Language Models with Human Opinions through Persona Selection and Value-Belief-Norm Reasoning
Chain-of-Opinion selects and reasons over explicit and implicit personae to improve LLM opinion prediction and alignment.
Xuan Long Do
,
Kenji Kawaguchi
,
Min-Yen Kan
,
Nancy F. Chen
Cite
URL
DnA-Eval: Enhancing Large Language Model Evaluation through Decomposition and Aggregation
DnA-Eval decomposes and aggregates LLM-as-judge evaluation to improve interpretability and agreement with human preferences.
Minzhi Li (Ella)
,
Zhengyuan Liu
,
Shumin Deng
,
Shafiq Joty
,
Nancy F. Chen
,
Min-Yen Kan
Cite
URL
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
A NeurIPS 2024 workshop paper on boosting reasoning with MCTS-driven iterative preference learning.
Yuxi Xie
,
Anirudh Goyal
,
Wenyue Zheng
,
Min-Yen Kan
,
Timothy P. Lillicrap
,
Kenji Kawaguchi
,
Michael Qizhe Shieh
PDF
Cite
Code
OpenReview
arXiv
Multi-expert Prompting Improves Reliability, Safety and Usefulness of Large Language Models
An EMNLP 2024 prompting method that simulates and aggregates multiple expert responses to improve LLM reliability, safety, and usefulness.
Xuan Long Do
,
Duong Ngoc Yen
,
Anh Tuan Luu
,
Kenji Kawaguchi
,
Min-Yen Kan
,
Nancy F. Chen
PDF
Cite
Advancing Adversarial Suffix Transfer Learning on Aligned Large Language Models
An EMNLP 2024 study of adversarial suffix transfer learning on aligned large language models.
Hongfu Liu
,
Yuxi Xie
,
Ye Wang
,
Michael Qizhe Shieh
PDF
Cite
Code
DOI
ACL Anthology
DataTales: A Benchmark for Real-World Intelligent Data Narration
An EMNLP 2024 benchmark of 4.9k financial reports and market data for evaluating data narration by language models.
Yajing Yang
,
Qian Liu
,
Min-Yen Kan
Cite
DOI
URL
«
»
Cite
×