数据集 / Robin021/DISC-Law-SFT

Robin021/DISC-Law-SFT 已完整同步

DISC-Law-SFT Dataset

Legal Intelligent systems in Chinese require a combination of various abilities, including legal text understanding and generation. To achieve this, we have constructed a high-quality supervised fine-tuning dataset called DISC-Law-SFT, which covers different legal scenarios such as legal information extraction, legal judgment prediction, legal document summarization, and legal question answering. DISC-Law-SFT comprises two subsets, DISC-Law-SFT-Pair and DISC-Law-SFT-Triplet. The former aims to introduce legal reasoning abilities to the LLM, while the latter helps enhance the model's capability to utilize external legal knowledge. For more detailed information, please refer to our technical report. The distribution of the dataset is:

We currently open-source most of the DISC-Law-SFT Dataset.

More detail and news check our homepage !

3 个文件

浏览文件