arrow
Return

Augmenting natural hazard exposure modelling using natural language processing

delete2024-02-01
delete0
delete
OA
AI
J
Justin Schembri
R
Roberto Gentile *
DOI:10.1016/j.ijdrr.2023.104202delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Natural hazard exposure modelling involves constructing databases that describe the elements (people and built environment) exposed to some hazard in a selected location. These databases are often constructed using information from censuses, cadastral data, or satellite imagery. In this work, we suggest complementing hazard exposure modelling using an alternative and unconventional data source: the text components of building permits. The proposed methodology, Natural Language Processing for the Global Exposure Database (NLP4GED), adopts natural language processing techniques to extract building -by -building exposure attributes in line with the GED4ALL taxonomy (Global Exposure Database for ALL). This three -step methodology involves using: a classifier to filter permits potentially containing exposure information; a clustering algorithm to identify semantically similar permits; and regular expressions (or regex) to extract exposure -attributes. As an illustrative application, we apply NLP4GED to wrangle an unstructured real -world dataset of 100,989 building permits in Malta. We effectively provide relevant exposure attributes (i.e., year of construction, building height, and occupancy) for 23,076 buildings presented in a geographic information system (GIS) environment.
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

International Journal of Disaster Risk Reduction cover
International Journal of Disaster Risk Reduction
IF:
4.5
Papers:
6.0K
Citations:
2.1W

Organization

U
university of london
Scholars:
21.5W
Papers: 19.7W
Citations: 305