Skip to main navigation Skip to search Skip to main content

基于文本挖掘和自动分类的法院裁判决策支持系统设计

Translated title of the contribution: Count Judgment Decision Support System Based on Text-mining and Machine Learning
  • 朱青
  • , 卫柯臻
  • , 丁兰琳
  • , 黎建强

    Research output: Journal Publications and ReviewsRGC 21 - Publication in refereed journalpeer-review

    Abstract

    In many other countries with the continental legal system, the constant generation of new legal relationships makes, the defect of statute law which is unable to be timely formulate and modify gradually become obvious. As the number of dispute lawsuit rapidly grows,many countries in the world face the problem how to improve the efficiency of the judicial system under the premise of guaranteeing the quality of the trial. Therefore, in addition to reforming the system,the decision support system will effectively improve judicial decisions.

    In this paper, medical damage judgment documents in China are taken as example, and a court judgment decision support system (CJ-DSS) is proposed based on text mining and the automatic classification technology. The system can predict the trail results of the new lawsuit texts according to the previous cases verdict: rejected and no rejected. By combining different feature extraction methods (DF, Chi-square and DF-CHI feature combination extraction method)and classifiers (SVM,ANN and KNN), multiple combinations that meet the expected performance as the base learning machines are selected. Based on the theory of Delphi Method, integrated learning is used to predict new cases. Integrated learning refers to constructing a new model and using the prediction result of base learning machines that have met expectations as input after proper training,and finally outputting aprediction result with maximum probability through linear or non-linear calculations.

    At the same time, by combining with real cases,it is found that the combination feature extraction method can indeed improve the classifier’s performance, especially for SVM, ANN and KNN classifiers.In addition,the system classification performance became more consistent after integrated learning. The best performance reached 93.3%, which significantly increased system accuracy.

    This paper’s data source is the "BeiDaFaBao" legal database. "Medical malpractice" is used as the keyword and more than 300 court verdict and mediation documents from 2013 are retrieved. Due to the short format of mediation documents and its brief case explanations, they are eliminated from the study. The rest of the documents are trained and tested after preprocessing.

    In previous studies, the accuracy of text classification system has been greatly influenced by the training set size: the larger the training set data, the better the performance. This paper has a reference value for constructing structured high-performance system based on a small sample training set in the future. Meanwhile, since the process of labelling documents is costly, therefore, the study and model construction for unlabeled text should be the focus of future research for data scientists.
    Translated title of the contributionCount Judgment Decision Support System Based on Text-mining and Machine Learning
    Original languageChinese (Simplified)
    Pages (from-to)170-178
    Journal中国管理科学
    Issue number1
    DOIs
    Publication statusPublished - Jan 2018

    Research Keywords

    • 文本挖掘
    • 自动分类
    • 决策支持系统
    • CJ-DSS
    • text-mining
    • automatic text classification
    • decision support system

    Fingerprint

    Dive into the research topics of 'Count Judgment Decision Support System Based on Text-mining and Machine Learning'. Together they form a unique fingerprint.

    Cite this