Predicting crimes in urban environments is crucial for public safety, yet existing prediction methods often struggle to align the knowledge across diverse cities that vary dramatically in data availability of specific crime types. We propose HYpernetwork-enhanced Spatial Temporal Learning (HYSTL), a framework that can effectively train a unified, stronger crime predictor without assuming identical crime types in different cities' records. In HYSTL, instead of parameterising a dedicated predictor per crime type, a hypernetwork is designed to dynamically generate parameters for the prediction function conditioned on the crime type of interest. To bridge the semantic gap between different crime types, a structured crime knowledge graph is built, where the learned representations of crimes are used as the input to the hypernetwork to facilitate parameter generation. As such, when making predictions for each crime type, the predictor is additionally guided by its intricate association with other relevant crime types. Extensive experiments are performed on two cities with non-overlapping crime types, and the results demonstrate HYSTL outperforms state-of-the-art baselines.
Predicting crime hotspots in a city is a complex and critical task with significant societal implications. Numerous spatiotemporal correlations and irregularities pose substantial challenges to this endeavor. Existing methods commonly employ fixed-time granularities and sequence prediction models. However, determining appropriate time granularities is difficult, leading to inaccurate predictions for specific time windows. For example, users might ask: What are the crime hotspots during 12:00-20:00? To address this issue, we introduce FlexiCrime, a novel event-centric framework for predicting crime hotspots with flexible time intervals. FlexiCrime incorporates a continuous-time attention network to capture correlations between crime events, which learns crime context features, representing general crime patterns across time points and locations. Furthermore, we introduce a type-aware spatiotemporal point process that learns crime-evolving features, measuring the risk of specific crime types at a given time and location by considering the frequency of past crime events. The crime context and evolving features together allow us to predict whether an urban area is a crime hotspot given
This article investigates crime patterns across European countries in 2022 using Compositional Data Analysis (CoDA) to address limitations of traditional statistical approaches in dealing with the relative nature of crime data. Recognizing crime types as components of a whole, we employ CoDA to explore relationships between different crime categories while respecting their inherent interdependencies. The study utilizes k-means clustering to group countries based on their crime profiles, identifying three distinct clusters largely aligning with geographical locations. This clustering is visualized through t-SNE and geographic mapping, revealing regional similarities. Further analysis using Robust Principal Component Analysis on identified crime clusters reveals insightful relationships between specific crime types, such as homicide, smuggling, and financial crimes, and how their prevalence varies across countries. The findings reveals distinct crime patterns across Europe, highlighting regional commonalities while also highlighting divergences like Norway and Latvia that deviate from their expected geographical classifications. Moreover, the study identifies specific crime groups; f
Crime remains one of the significant problems that countries are grappling with globally. With shrinking economies and increasing poverty, crime has been on the rise in many countries. In this paper, we propose a system of non-linear ordinary differential equations to model crime dynamics in the presence of imitation. The model consists of four independent compartments: individuals who are not at risk of committing a crime, individuals at risk of committing a crime, individuals committing a crime, and individuals convicted and jailed for a crime. The model is analyzed using the basic reproduction number. The analysis shows the system has a locally asymptotically stable crime-free equilibrium when the basic reproduction number is less than unity. The model exhibits a backward bifurcation in which two endemic equilibria coexist with the crime-free equilibrium. When the basic reproduction number exceeds unity, the system has a locally asymptotically stable endemic equilibrium, and the crime-free becomes unstable. Numerical simulations are carried out to verify the analytical results. The sensitivity analysis shows that the relapse rate highly influences the basic reproduction number o
The classification of crime into discrete categories entails a massive loss of information. Crimes emerge out of a complex mix of behaviors and situations, yet most of these details cannot be captured by singular crime type labels. This information loss impacts our ability to not only understand the causes of crime, but also how to develop optimal crime prevention strategies. We apply machine learning methods to short narrative text descriptions accompanying crime records with the goal of discovering ecologically more meaningful latent crime classes. We term these latent classes "crime topics" in reference to text-based topic modeling methods that produce them. We use topic distributions to measure clustering among formally recognized crime types. Crime topics replicate broad distinctions between violent and property crime, but also reveal nuances linked to target characteristics, situational conditions and the tools and methods of attack. Formal crime types are not discrete in topic space. Rather, crime types are distributed across a range of crime topics. Similarly, individual crime topics are distributed across a range of formal crime types. Key ecological groups include identit
The purpose of this study was to determine the association of location and types of crimes in the Philippines and understand the impact of COVID-19 lockdowns by comparing the crime incidence and associations before and during the pandemic. A document review was used as the main method of data collection using the datasets from the Philippine Statistics Authority- Annual Statistical Yearbook (PSA-ASY). The dataset contained the volume of index crimes in the Philippines from 2016 to 2020. The index crimes were broken down into two major categories: crimes against persons and crimes against property. Incidence of crime-by-crime type was available for different administrative regions in the Philippines. Chi-square test and correlation plot of chi-square residual were used to determine the associations between the locations and types of index crimes. A correlation plot of the chisquare residual was used to investigate the patterns of associations. Results suggest that the continuing effort of the Philippine government to fight against criminality has resulted in a steady decline in the incidence of index crimes in the Philippines. The pandemic too contributed to the decline of crime inc
Self-exciting point processes are widely used to model the contagious effects of crime events living within continuous geographic space, using their occurrence time and locations. However, in urban environments, most events are naturally constrained within the city's street network structure, and the contagious effects of crime are governed by such a network geography. Meanwhile, the complex distribution of urban infrastructures also plays an important role in shaping crime patterns across space. We introduce a novel spatio-temporal-network point process framework for crime modeling that integrates these urban environmental characteristics by incorporating self-attention graph neural networks. Our framework incorporates the street network structure as the underlying event space, where crime events can occur at random locations on the network edges. To realistically capture criminal movement patterns, distances between events are measured using street network distances. We then propose a new mark for a crime event by concatenating the event's crime category with the type of its nearby landmark, aiming to capture how the urban design influences the mixing structures of various crime
Crime prediction is a widely studied research problem due to its importance in ensuring safety of city dwellers. Starting from statistical and classical machine learning based crime prediction methods, in recent years researchers have focused on exploiting deep learning based models for crime prediction. Deep learning based crime prediction models use complex architectures to capture the latent features in the crime data, and outperform the statistical and classical machine learning based crime prediction methods. However, there is a significant research gap in existing research on the applicability of different models in different real-life scenarios as no longitudinal study exists comparing all these approaches in a unified setting. In this paper, we conduct a comprehensive experimental evaluation of all major state-of-the-art deep learning based crime prediction models. Our evaluation provides several key insights on the pros and cons of these models, which enables us to select the most suitable models for different application scenarios. Based on the findings, we further recommend certain design practices that should be taken into account while building future deep learning bas
Objectives: To develop a deep learning framework to evaluate if and how incorporating micro-level mobility features, alongside historical crime and sociodemographic data, enhances predictive performance in crime forecasting at fine-grained spatial and temporal resolutions. Methods: We advance the literature on computational methods and crime forecasting by focusing on four U.S. cities (i.e., Baltimore, Chicago, Los Angeles, and Philadelphia). We employ crime incident data obtained from each city's police department, combined with sociodemographic data from the American Community Survey and human mobility data from Advan, collected from 2019 to 2023. This data is aggregated into grids with equally sized cells of 0.077 sq. miles (0.2 sq. kms) and used to train our deep learning forecasting model, a Convolutional Long Short-Term Memory (ConvLSTM) network, which predicts crime occurrences 12 hours ahead using 14-day and 2-day input sequences. We also compare its performance against three baseline models: logistic regression, random forest, and standard LSTM. Results: Incorporating mobility features improves predictive performance, especially when using shorter input sequences. Notewort
We use high resolution data to investigate the association between crime incidence and proximity to different types of public schools over the past fifteen years in the city of Philadelphia. We employ two statistical methods, regression modeling and propensity score matching, in order to better isolate the association between crime and school proximity while controlling for the demographic, economic, land use and disorder characteristics of the surrounding neighborhood. With both of these approaches, we find significantly increased crime incidence near to public schools regardless of crime outcome, educational level and time period. The effect of school proximity on crime varies substantially depending on whether or not school is in session, as well as between different types of crime and educational levels of the school. We see the largest effects of school proximity on crime for violent crimes near to high schools during their in-session time periods. Our results support several theories which suggest that crime should be elevated near to schools, as well as finding significant associations between crime and other aspects of the built environment.
Predicting crime using machine learning and deep learning techniques has gained considerable attention from researchers in recent years, focusing on identifying patterns and trends in crime occurrences. This review paper examines over 150 articles to explore the various machine learning and deep learning algorithms applied to predict crime. The study provides access to the datasets used for crime prediction by researchers and analyzes prominent approaches applied in machine learning and deep learning algorithms to predict crime, offering insights into different trends and factors related to criminal activities. Additionally, the paper highlights potential gaps and future directions that can enhance the accuracy of crime prediction. Finally, the comprehensive overview of research discussed in this paper on crime prediction using machine learning and deep learning approaches serves as a valuable reference for researchers in this field. By gaining a deeper understanding of crime prediction techniques, law enforcement agencies can develop strategies to prevent and respond to criminal activities more effectively.
Philadelphia's problem with high crime rates continues to be exacerbated as Philadelphia's residents, community leaders, and law enforcement officials struggle to address the root causes of the problem and make the city safer for all. In this work, we deeply understand crime in Philadelphia and offer novel insights for crime mitigation within the city. Open source crime data from 2012-2022 was obtained from OpenDataPhilly. Density-Based Spatial Clustering of Applications with Noise (DBSCAN) was used to cluster geographic locations of crimes. Clustering of crimes within each of 21 police districts was performed, and temporal changes in cluster distributions were analyzed to develop a Non-Systemic Index (NSI). Home Owners' Loan Corporation (HOLC) grades were tested for associations with clusters in police districts labeled `systemic.' Crimes within each district were highly clusterable, according to Hopkins' Mean Statistics. NSI proved to be a good measure of differentiating systemic ($<$ 0.06) and non-systemic ($\geq$ 0.06) districts. Two systemic districts, 19 and 25, were found to be significantly correlated with HOLC grade (p $=2.02 \times 10^{-19}$, p $=1.52 \times 10^{-13}$)
Crime is an unlawful act that carries legal repercussions. Bangladesh has a high crime rate due to poverty, population growth, and many other socio-economic issues. For law enforcement agencies, understanding crime patterns is essential for preventing future criminal activity. For this purpose, these agencies need structured crime database. This paper introduces a novel crime dataset that contains temporal, geographic, weather, and demographic data about 6574 crime incidents of Bangladesh. We manually gather crime news articles of a seven year time span from a daily newspaper archive. We extract basic features from these raw text. Using these basic features, we then consult standard service-providers of geo-location and weather data in order to garner these information related to the collected crime incidents. Furthermore, we collect demographic information from Bangladesh National Census data. All these information are combined that results in a standard machine learning dataset. Together, 36 features are engineered for the crime prediction task. Five supervised machine learning classification algorithms are then evaluated on this newly built dataset and satisfactory results are a
Crime is one of the greatest threats to urban security. Around 80 percent of the world's population lives in countries with high levels of criminality. Most of the crimes committed in the cities take place in their urban environments. This paper presents the development and validation of a digital shadow platform for modeling and simulating urban crime. This digital shadow has been constructed using data-driven agent-based modeling and simulation techniques, which are suitable for capturing dynamic interactions among individuals and with their environment. Our approach transforms and integrates well-known criminological theories and the expert knowledge of law enforcement agencies (LEA), policy makers, and other stakeholders under a theoretical model, which is in turn combined with real crime, spatial (cartographic) and socio-economic data into an urban model characterizing the daily behavior of citizens. The digital shadow has also been instantiated for the city of Malaga, for which we had over 300,000 complaints available. This instance has been calibrated with those complaints and other geographic and socio-economic information of the city. To the best of our knowledge, our digi
Adoption of AI by criminal entities across traditional and emerging financial crime paradigms has been a disturbing recent trend. Particularly concerning is the proliferation of generative AI, which has empowered criminal activities ranging from sophisticated phishing schemes to the creation of hard-to-detect deep fakes, and to advanced spoofing attacks to biometric authentication systems. The exploitation of AI by criminal purposes continues to escalate, presenting an unprecedented challenge. AI adoption causes an increasingly complex landscape of fraud typologies intertwined with cybersecurity vulnerabilities. Overall, GenAI has a transformative effect on financial crimes and fraud. According to some estimates, GenAI will quadruple the fraud losses by 2027 with a staggering annual growth rate of over 30% [27]. As crime patterns become more intricate, personalized, and elusive, deploying effective defensive AI strategies becomes indispensable. However, several challenges hinder the necessary progress of AI-based fincrime detection systems. This paper examines the latest trends in AI/ML-driven financial crimes and detection systems. It underscores the urgent need for developing agi
Urban safety and security play a crucial role in improving life quality of citizen and the sustainable development of urban. Traditional urban crime research focused on leveraging demographic data, which is insufficient to capture the complexity and dynamics of urban crimes. In the era of big data, we have witnessed advanced ways to collect and integrate fine-grained urban, mobile, and public service data that contains various crime-related sources as well as rich environmental and social information. The availability of big urban data provides unprecedented opportunities, which enable us to conduct advanced urban crime research. Meanwhile, environmental and social crime theories from criminology provide better understandings about the behaviors of offenders and complex patterns of crime in urban. They can not only help bridge the gap from what we have (big urban data) to what we want to understand about urban crime (urban crime analysis); but also guide us to build computational models for crime. In this article, we give an overview to key theories from criminology, summarize crime analysis on urban data, review state-of-the-art algorithms for various types of computational crime
Crime is highly concentrated in a few places, is committed by a few offenders and is suffered by a few victims. In recent decades, the concentration of crime has become an accepted fact, yet, little is known in terms of how to measure this concentration of crime such that the metric takes into account the fact that crime has, in general, a low frequency, it fluctuates, it is highly concentrated and has a certain degree of randomness. Here, the most frequently used metrics for concentration of crimes are reviewed. A null model with complete randomness is used for comparing between different concentration metrics, which allows constructing a sensitivity analysis for every metric against varying crime rates. Results show that most ways of measuring the concentration of crime are in fact, showing only that crime is rare or that it has fluctuations, but fail to work as a method to compare the concentration of crime between different regions, types of crime or across time.
There are challenges faced in today's world in terms of crime analysis when it comes to graphical visualization of crime patterns. Geographical representation of crime scenes and crime types become very important in gathering intelligence about crimes. This provides a very dynamic and easy way of monitoring criminal activities and analyzing them as well as producing effective countermeasures and preventive measures in solving them. But we need effective computer tools and intelligent systems that are automated to analyze and interpret criminal data in real time effectively and efficiently. These current computer systems should have the capability of providing intelligence from raw data and creating a visual graph which will make it easy for new concepts to be built and generated from crime data in order to solve, understand and analyze crime patterns easily. This paper proposes a new method of visualizing and analyzing crime patterns based on geographical crime data by using Formal Concept Analysis, or Galois Lattices, a data analysis technique grounded on Lattice Theory and Propositional Calculus. This method considered the set of common and distinct attributes of crimes in such a
Exposure to crime and violence can harm individuals' quality of life and the economic growth of communities. In light of the rapid development in machine learning, there is a rise in the need to explore automated solutions to prevent crimes. With the increasing availability of both fine-grained urban and public service data, there is a recent surge in fusing such cross-domain information to facilitate crime prediction. By capturing the information about social structure, environment, and crime trends, existing machine learning predictive models have explored the dynamic crime patterns from different views. However, these approaches mostly convert such multi-source knowledge into implicit and latent representations (e.g., learned embeddings of districts), making it still a challenge to investigate the impacts of explicit factors for the occurrences of crimes behind the scenes. In this paper, we present a Spatial-Temporal Metapath guided Explainable Crime prediction (STMEC) framework to capture dynamic patterns of crime behaviours and explicitly characterize how the environmental and social factors mutually interact to produce the forecasts. Extensive experiments show the superiority
In this work, the spread of crime dynamics in the US is analyzed from a mathematical scope, an epidemiological model is established, including five compartments: Susceptible (S), Latent 1 (E1), Latent 2 (E2), Incarcerated (I), and Recovered (R). A system of differential equations is used to model the spread of crime. A result to show the positivity of the solutions for the system is included. The basic reproduction number and the stability for the disease-free equilibrium results are calculated following epidemiological theories. Numerical simulations are performed with US parameter values. Understanding the dynamics of the spread of crime helps to determine what factors may work best together to reduce violent crime.