Bache, Richard, Crestani, Fabio, Canter, David V. and Youngs, Donna E. (2010) A Language Modelling approach to linking criminal styles with offender characteristics. Data & Knowledge Engineering, 69 (3). pp. 303-315. ISSN 0169-023X

The ability to infer the characteristics of offenders from their criminal behaviour (‘offender profiling’) has only been partially successful since it has relied on subjective judgments based on limited data. Words and structured data used in crime descriptions recorded by the police relate to behavioural features. Thus Language Modelling was applied to an existing police archive to link behavioural features with significant characteristics of offenders. Both multinomial and multiple Bernoulli models were used. Although categories selected are gender, age group, ethnic appearance and broad occupation (employed or not), in principle this can be applied to any characteristic recorded. Results indicate that statistically significant relationships exist between all characteristics for many types of crime. Bernoulli models tend to perform better than multinomial ones. It is also possible to identify automatically specific terms which when taken together give insight into the style of offending related to a particular group.

CanterLanguage.pdf - Submitted Version

Download (239kB) | Preview


Downloads per month over past year

Add to AnyAdd to TwitterAdd to FacebookAdd to LinkedinAdd to PinterestAdd to Email