TIPS:Tailored Information Extraction in Public Security Using Domain-Enhanced Large Language Model  

在线阅读下载全文

作  者:Yue Liu Qinglang Guo Chunyao Yang Yong Liao 

机构地区:[1]School of Cyber Science and Technology,University of Science and Technology of China,Hefei,230026,China [2]National Engineering Research Center for Public Safety Risk Perception and Control by Big Data,China Academy of Electronics and Information Technology,Beijing,100041,China

出  处:《Computers, Materials & Continua》2025年第5期2555-2572,共18页计算机、材料和连续体(英文)

摘  要:Processing police incident data in public security involves complex natural language processing(NLP)tasks,including information extraction.This data contains extensive entity information—such as people,locations,and events—while also involving reasoning tasks like personnel classification,relationship judgment,and implicit inference.Moreover,utilizing models for extracting information from police incident data poses a significant challenge—data scarcity,which limits the effectiveness of traditional rule-based and machine-learning methods.To address these,we propose TIPS.In collaboration with public security experts,we used de-identified police incident data to create templates that enable large language models(LLMs)to populate data slots and generate simulated data,enhancing data density and diversity.We then designed schemas to efficiently manage complex extraction and reasoning tasks,constructing a high-quality dataset and fine-tuning multiple open-source LLMs.Experiments showed that the fine-tuned ChatGLM-4-9B model achieved an F1 score of 87.14%,nearly 30%higher than the base model,significantly reducing error rates.Manual corrections further improved performance by 9.39%.This study demonstrates that combining largescale pre-trained models with limited high-quality domain-specific data can greatly enhance information extraction in low-resource environments,offering a new approach for intelligent public security applications.

关 键 词:Public security information extraction large language model prompt engineering 

分 类 号:TP391[自动化与计算机技术—计算机应用技术]

 

参考文献:

正在载入数据...

 

二级参考文献:

正在载入数据...

 

耦合文献:

正在载入数据...

 

引证文献:

正在载入数据...

 

二级引证文献:

正在载入数据...

 

同被引文献:

正在载入数据...

 

相关期刊文献:

正在载入数据...

相关的主题
相关的作者对象
相关的机构对象