Comparing Crime Rates in Various Communities Using K-Means
CSEF · 2015 Behavioral & Social Sciences
Overview
Objectives/Goals The purpose of this experiment was to compare the extent to which various factors affected the crime rates in communities throughout the country, using the K-Means clustering algorithm. Methods/Materials The experiment used a combination of the K-Means clustering algorithm and a regression analysis to account for all outliers. A sample code for K-Means, found online, was edited and compiled to suit the needs of the analysis. A data set was used for this experiment, where only the variables required were used and the communities with unknown data points were eliminated. The data was formatted to an excel file, the program was run on that data and the cluster centroids results were graphed. The best type (polynomial/logarithmic) line of fit was chosen depending on the data set trend. Results It was found that low number of police units per 100,000 people was the biggest contributor to crime rates, followed by high poverty rates, low high school graduation rates, high homelessness rates, and high unemployment rates. The relationship between the number of police units per 100,000 people and the crime rate was that the crime rate went down as the number of police units per 100,000 people went up. However, the relationship between the unemployment rate and the crime rate showed that the crime rate only went down marginally as the unemployment rate went down. Conclusions/Discussion The hypothesis was partially supported; the order in which the factors affected crime rates was only partially right. For example, unemployment rates were not the second largest contributor to high crime rates; they were the smallest contributors. A possible explanation for this was that the use of the K-means clustering algorithm produced results different from those of previous experiments.
Summary statement
This project compared the extent to which various factors affected the crime rates in communities throughout the country, using the K-Means clustering algorithm.
Help received
None
Competition history
- CSEF 2015
Resources
Related projects
CSEF · 2017
Developing a Predictive Model for On-Campus Crime Using Machine Learning Algorithms and Reporting via Mobile App
ISEF · 2017
A Novel Supervised Machine Learning Approach to Predicting Crime Information
ISEF · 2017
Developing a Predictive Model for On-Campus Crime Using Machine Learning Algorithms and Reporting via Mobile App
CSEF · 2011
An Application of Geographic Profiling to Graffiti Crimes
CSEF · 2019
Factors Affecting the California Science and Engineering Fair Results
CSEF · 2002
Law and Order: Benford's and Zipf's Laws
CSEF · 2018
Illegal Immigrants and Crime: The Hard Statistics
CSEF · 2019
Investigating the Correlation between Demographics and Attitudes towards Contemporary Social Movements
Closest projects by meaning, across every fair and year in the corpus.
Browse more like this
Source: California Science & Engineering Fair public projects