Aug 2026· Applied and Computational Engineering· Vol 255, pp. 42-48· 0 citations
TL;DR
This study applies a Random Forest classifier to the FBI's national Uniform Crime Reporting (UCR) dataset containing 218,069 single-bias hate crimes from 1991to 2020 to demonstrate the predictability of indicators within official records.
Abstract
This study applies a Random Forest classifier to the FBI's national Uniform Crime Reporting (UCR) dataset containing 218,069 single-bias hate crimes from 1991to 2020. Within the dataset, crimes were predicted based on bias-driven motives: race, sexual orientation, gender, religion, or disability. Thirteen incident features were analyzed to form the prediction. The primary model chosen for this study oversampled the two smallest categories and tuned hyperparameters. Ultimately, an accuracy of 68.7% with F1 = 0.80 and recall = 0.93 was achieved for the main prediction target: racial bias. Five trials were conducted with different configurations of the train/test model, compared across two metrics. It can be seen that there is a significant trade-off between total balance within all 5 categories versus overall accuracy. It was determined the incident year was the dominant predictive feature, likely due to FBI changes in definition and classification across the 1991-2020 period. Results demonstrate the predictability of indicators within official records and are limited to UCR metrics of data collection.
The increase in crime can be caused by several factors. As growth in population and density, criminal behavior, domestic crime growth, and so on. This causes difficulties for the law enforcement agencies to control and monitor on a regular crime basis since it requires forecasting and probabilities, significant progres...
This study presents a machine learning–based framework that integrates spatially linked real estate transaction data and crime records to predict residential property prices in the New York City (NYC) housing market over the period 2010–2024. The motivation stems from the limitation of traditional valuation approaches,...
Özgür Özgenç, Esin Ayşe Zaimoğlu· Bitlis Eren Üniversitesi Fen...· 0 citations
This study proposes a Machine Learning (ML_-based framework for the automatic classification of accident narratives into Risk and Non-Risk categories to support the risk identification phase of occupational risk management. A manually labeled dataset containing 278 accident narratives derived from the publicly availabl...
Key predictors of high-risk age at childbirth included maternal age, early marriage, current contraceptive use, husband occupation, number of children, husband age and education, respondent age, and regional disparities are identified.
A. Siam, Tanzila Tanjim Tusra, M. Hossain· PLoS ONE· 0 citations
The pipeline approach ensured no data leakage in cross-validation, and the findings support ensemble ML models with SMOTE as a preprocessing step for imbalanced CVD datasets.
M. Maindarkar· Journal of Intelligent Decis...· 0 citations
Objectivity in military criminal judgments is crucial for judicial legitimacy but is frequently compromised by semantic bias. To the best of our knowledge, this is the first study to specifically address automated bias detection within the Indonesian military legal domain, bridging a significant gap in the literature t...
Bayu Ardiyansyah, Lutfi Indra Nur Praditya, Galih Wasis Wicaksono et al.· Jurnal Teknik Informatika (J...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.