상세 보기
AdvisorQA: Towards Helpful and Harmless Advice-seeking Question Answering with Collective Intelligence
- Kim, Minbeom;
- Lee, Hwanhee;
- Park, Joonsuk;
- Lee, Hwaran;
- Jung, Kyomin
WEB OF SCIENCE
1SCOPUS
0초록
As the integration of large language models into daily life is on the rise, there is still a lack of dataset for advising on subjective and personal dilemmas. To address this gap, we introduce AdvisorQA, which aims to improve LLMs' capability to offer advice for deeply subjective concerns, utilizing the LifeProTips Reddit forum. This forum features a dynamic interaction where users post advice-seeking questions, receiving an average of 8.9 advice per query, with 164.2 upvotes from hundreds of users, embodying a collective intelligence. Therefore, we've completed a dataset encompassing daily life questions, diverse corresponding responses, and majority vote ranking, which we use to train a helpfulness metric. In baseline experiments, models aligned with AdvisorQA dataset demonstrated improved helpfulness through our automatic metric, as well as GPT-4 and human evaluations. Additionally, we expanded the independent evaluation axis to include harmlessness. AdvisorQA marks a significant leap in enhancing QA systems to provide subjective, helpful, and harmless advice, showcasing LLMs' improved understanding of human subjectivity.
- 제목
- AdvisorQA: Towards Helpful and Harmless Advice-seeking Question Answering with Collective Intelligence
- 저자
- Kim, Minbeom; Lee, Hwanhee; Park, Joonsuk; Lee, Hwaran; Jung, Kyomin
- 발행일
- 2025
- 유형
- Proceedings Paper
- 저널명
- PROCEEDINGS OF THE 2025 CONFERENCE OF THE NATIONS OF THE AMERICAS CHAPTER OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: HUMAN LANGUAGE TECHNOLOGIES, VOL 1: LONG PAPERS
- 권
- 1
- 페이지
- 6545 ~ 6565