الملخص
Large Language Models (LLMs) are increasingly applied to legal and criminal justice tasks, yet existing work focuses almost exclusively on post-arrest scenarios where the suspect's identity is already known, leaving the critical pre-arrest challenge of inferring suspect characteristics from incomplete evidence largely unexplored. To fill this gap, we introduce the Profiling, Investigation, and Judgment (PIJ), comprising 2,500 real homicide cases from five countries. PIJ evaluates LLMs across three tasks that span the entire criminal investigation pipeline: criminal profiling, which requires abductive reasoning to infer suspect attributes from fragmentary scene evidence, crime process reconstruction, which tests structured information extraction, and sentence prediction, which demands legal deductive reasoning. We evaluate 9 powerful LLMs and find that performance degrades systematically as tasks shift from explicit fact extraction to implicit reasoning over unknown suspect profiles. Categories requiring inferential reasoning, such as motivation and victim-offender relationships, remain the primary bottlenecks. Further analysis reveals substantial gaps between LLMs and human experts, along with pervasive biases in gender, age, and motive attribution. Our findings indicate that pre-arrest inference from incomplete evidence remains an open challenge.
الكلمات المفتاحية
الموضوع
بيانات النشر
- المجلة
- غير متاح
- وصول مفتوح
- وصول مفتوح أخضر
اقتبس هذه المقالة
APA 7
Yao, Y., Cao, Y., Chen, G., Yang, X., Wu, J., Wu, Z., Chao, L. S., & Wong, D. F. (2026). Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence. https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence
MLA 9
Yao, Yutong, et al. "Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence." https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence.
شيكاغو (المؤلف–التاريخ)
Yao, Yutong, Yanjie Cao, Guanhua Chen, Xu Yang, Junchao Wu, Zeyu Wu, Lidia S. Chao, and Derek F. Wong. 2026. "Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence." https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence.
هارفارد
Yao, Y., Cao, Y., Chen, G., Yang, X., Wu, J., Wu, Z., Chao, L. S. and Wong, D. F. (2026) 'Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence', Available at: https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence.
فانكوفر
Yao Y, Cao Y, Chen G, Yang X, Wu J, Wu Z, et al. Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence. https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence
IEEE
Y. Yao, Y. Cao, G. Chen, X. Yang, J. Wu, Z. Wu, L. S. Chao, and D. F. Wong, "Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence," https://omanscience.com/ar/articles/before-the-arrest-benchmarking-llms-on-criminal-profiling-from-incomplete-evidence.