Hot-Spot Analysis with Simple Features

Identify and understand clusters of points (typically representing the locations of places or events) stored in simple-features (SF) objects. This is useful for analysing, for example, hot-spots of crime events. The package emphasises producing results from point SF data in a single step using reasonable default values for all other arguments, to aid rapid data analysis by users who are starting out. Functions available include kernel density estimation (for details, see Yip (2020) ), analysis of spatial association (Getis and Ord (1992) ) and hot-spot classification (Chainey (2020) ISBN:158948584X).


sfhotspot

CRANstatus CRANchecks CRAN RStudio mirrordownloads Lifecycle:stable Codecov testcoverage

sfhotspot provides functions to identify and understand clusters of points (typically representing the locations of places or events). All the functions in the package work on and produce simple features (SF) objects, which means they can be used as part of modern spatial analysis in R.

Installation

You can install the development version of sfhotspot from GitHub with:

# install.packages("remotes")
remotes::install_github("mpjashby/sfhotspot")

Functions

sfhotspot has the following functions. All can be used by just supplying an SF object containing points, or can be configured using the optional arguments to each function.

name use
hotspot_count() Count the number of points in each cell of a regular grid. Cell size can be set by the user or chosen automatically.
hotspot_change() Measure the change in the count of points in each cell between two periods of time.
hotspot_kde() Estimate kernel density for each cell in a regular grid. Cell size and bandwidth can be set by the user or chosen automatically.
hotspot_dual_kde() Compare the kernel density of two layers of points, e.g. to estimate the local risk of an event occurring relative to local population.
hotspot_gistar() Calculate the Getis–Ord $G_i^*$ statistic for each cell in a regular grid, while optionally estimating kernel density. Cell size, bandwidth and neighbour distance can be set by the user or chosen automatically.
hotspot_classify() Classify grid cells according to whether they have had significant clusters of points at different time periods. All parameters can be chosen automatically or be set by the user using the hotspot_classify_params() helper function.
hotspot_dbscan() Identify clusters of points using the DBSCAN algorithm.

The results produced by hotspot_count(), hotspot_change(), hotspot_kde(), hotspot_dual_kde(), hotspot_gistar(), hotspot_classify() and hotspot_dbscan() can be easily plotted using hotspot_map(), or combined with other ggplot2 layers using hotspot_layer().

There are also included datasets:

  • memphis_robberies, containing records of 2,245 robberies in Memphis, TN, in 2019.
  • memphis_robberies_jan, containing the same data but only for the 206 robberies recorded in January 2019.
  • memphis_population, containing population counts for the centroids of 10,393 census blocks in Memphis, TN, in 2020.

Example

We can use the hotspot_gistar() function to identify cells in a regular grid in which there are more/fewer points than would be expected if the points were distributed randomly. In this example, the points represent the locations of personal robberies in Memphis, which is a dataset included with the package.

# Load packages
library(sf)
Linking to GEOS 3.13.0, GDAL 3.8.5, PROJ 9.5.1; sf_use_s2() is TRUE
library(sfhotspot)
library(tidyverse)
── Attaching core tidyverse packages ──────────────────────── tidyverse 2.0.0 ──
✔ dplyr     1.2.1     ✔ readr     2.2.0
✔ forcats   1.0.1     ✔ stringr   1.6.0
✔ ggplot2   4.0.3     ✔ tibble    3.3.1
✔ lubridate 1.9.5     ✔ tidyr     1.3.2
✔ purrr     1.2.2     

── Conflicts ────────────────────────────────────────── tidyverse_conflicts() ──
✖ dplyr::filter() masks stats::filter()
✖ dplyr::lag()    masks stats::lag()
ℹ Use the conflicted package (<http://conflicted.r-lib.org/>) to force all conflicts to become errors
# Transform data to UTM zone 15N so that we can think in metres, not decimal 
# degrees
memphis_robberies_utm <- st_transform(memphis_robberies, "EPSG:32615")


# Identify hotspots, set all the parameters automatically by not specifying cell 
# size, bandwidth, etc.
memphis_robberies_htspt <- hotspot_gistar(memphis_robberies_utm, quiet = TRUE)


# Visualise the hotspots by showing only those cells that have significantly
# more points than expected by chance. For those cells, show the estimated
# density of robberies.
autoplot(memphis_robberies_htspt)

Reference manual

It appears you don't have a PDF plugin for this browser. You can click here to download the reference manual.

install.packages("sfhotspot")

1.1.1 by Matt Ashby, 8 days ago


https://pkgs.lesscrime.info/sfhotspot/


Report a bug at https://github.com/mpjashby/sfhotspot/issues


Browse source code at https://github.com/cran/sfhotspot


Authors: Matt Ashby [aut, cre]


Documentation:   PDF Manual  


MIT + file LICENSE license


Imports classInt, cli, dbscan, ggplot2, ggspatial, isoband, rlang, sf, SpatialKDE, spdep, tibble

Suggests testthat, knitr, lubridate, rmarkdown, quarto


See at CRAN