数据集 / YIRONGCHEN/SoulChatCorpus

YIRONGCHEN/SoulChatCorpus 已完整同步

SoulChatCorpus

该项目为项目https://github.com/scutcyr/SoulChat的数据集部分。

最近更新

数据样例

{
    "id": 0,
    "topic": "成长",
    "messages": [
        {
            "role": "user",
            "content": "最近感觉很焦虑,不知道如何缓解。"
        },
        {
            "role": "assistant",
            "content": "我明白你的感受。首先,要知道焦虑是一种正常的情绪。但是,当它影响到你的生活质量时,就需要采取行动了。"
        },
        {
            "role": "user",
            "content": "可是我总觉得自己做什么都不够好,很害怕失败。"
        },
        {
            "role": "assistant",
            "content": "这可能是完美主义导致的。完美主义者总是把难度设置得很高,并在失败时很难接受自己。其实,你是一个有自尊和价值的人,无论成败。"
        },
        {
            "role": "user",
            "content": "但是我总是无法集中注意力,会经常拖延。"
        },
        {
            "role": "assistant",
            "content": "拖延是一种很常见的自我保护机制,但长期下来对心理健康负面影响很大。你可以尝试分段任务,每段任务完成后休息片刻,然后继续下一个任务。"
        },
        {
            "role": "user",
            "content": "但是每天都很疲惫,很难找到动力去改变。"
        },
        {
            "role": "assistant",
            "content": "情绪低落,疲惫和动力不足是相互作用的。你可以尝试建立自己的日常规划,合理分配时间,保证充足的睡眠和营养,以此建立更加健康的自己。"
        },
        {
            "role": "user",
            "content": "其实我很害怕自己无法排解焦虑情绪。"
        },
        {
            "role": "assistant",
            "content": "我会一直在这里支持你,并分享一些我的应对技巧。你可以与朋友交流,寻求专业帮助,做一些自己感兴趣的事情,放松身心。"
        },
        {
            "role": "user",
            "content": "我会尝试这些方法,谢谢你的支持和理解。"
        },
        {
            "role": "assistant",
            "content": "不要忘记,焦虑并不是终点,而是一个可以克服的过程。你是勇敢的,你可以做到。任何时候,都欢迎你来和我交流。"
        }
    ]
}

声明

  • 本项目使用了ChatGLM-6B 模型的权重,需要遵循其MODEL_LICENSE,因此,本项目仅可用于您的非商业研究目的
  • 本项目提供的SoulChat模型致力于提升大模型的共情对话与倾听能力,然而,模型的输出文本具有一定的随机性,当其作为一个倾听者的时候,是合适的,但是不建议将SoulChat模型的输出文本替代心理医生等的诊断、建议。本项目不保证模型输出的文本完全适合于用户,用户在使用本模型时需要承担其带来的所有风险!
  • 您不得出于任何商业、军事或非法目的使用、复制、修改、合并、发布、分发、复制或创建SoulChat模型的全部或部分衍生作品。
  • 您不得利用SoulChat模型从事任何危害国家安全和国家统一、危害社会公共利益、侵犯人身权益的行为。
  • 您在使用SoulChat模型时应知悉,其不能替代医生、心理医生等专业人士,不应过度依赖、服从、相信模型的输出,不能长期沉迷于与SoulChat模型聊天。

致谢

本项目由华南理工大学未来技术学院 广东省数字孪生人重点实验室发起,得到了华南理工大学信息网络工程研究中心、电子与信息学院等学院部门的支撑,同时致谢广东省妇幼保健院、广州市妇女儿童医疗中心、中山大学附属第三医院、合肥综合性国家科学中心人工智能研究院等合作单位。

同时,我们感谢以下媒体或公众号对本项目的报道(排名不分先后):

引用

@inproceedings{chen-etal-2023-soulchat,
    title = "{S}oul{C}hat: Improving {LLM}s{'} Empathy, Listening, and Comfort Abilities through Fine-tuning with Multi-turn Empathy Conversations",
    author = "Chen, Yirong  and
      Xing, Xiaofen  and
      Lin, Jingkai  and
      Zheng, Huimin  and
      Wang, Zhenyu  and
      Liu, Qi  and
      Xu, Xiangmin",
    editor = "Bouamor, Houda  and
      Pino, Juan  and
      Bali, Kalika",
    booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2023",
    month = dec,
    year = "2023",
    address = "Singapore",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2023.findings-emnlp.83",
    pages = "1170--1183",
    abstract = "Large language models (LLMs) have been widely applied in various fields due to their excellent capability for memorizing knowledge and chain of thought (CoT). When these language models are applied in the field of psychological counseling, they often rush to provide universal advice. However, when users seek psychological support, they need to gain empathy, trust, understanding and comfort, rather than just reasonable advice. To this end, we constructed a multi-turn empathetic conversation dataset of more than 2 million samples, in which the input is the multi-turn conversation context, and the target is empathetic responses that cover expressions such as questioning, comfort, recognition, listening, trust, emotional support, etc. Experiments have shown that the empathy ability of LLMs can be significantly enhanced when finetuning by using multi-turn dialogue history and responses that are closer to the expression of a psychological consultant.",
}
}

Star History

Star History Chart

4 个文件

浏览文件