Zero-Shot Building Age Classification from Facade Image Using GPT-4

Z. Zeng; J. M. Goo; X. Wang; B. Chi; M. Wang; J. Boehm

doi:10.5194/isprs-archives-XLVIII-2-2024-457-2024

The International Archives of the Photogrammetry, Remote Sensing and Spatial Information Sciences (Jun 2024)

Zero-Shot Building Age Classification from Facade Image Using GPT-4

Z. Zeng,
J. M. Goo,
X. Wang,
B. Chi,
M. Wang,
J. Boehm

Affiliations

Z. Zeng: Department of Civil, Environmental and Geomatic Engineering, University College London, Gower Street, London, WC1E 6BT, UK
J. M. Goo: Department of Civil, Environmental and Geomatic Engineering, University College London, Gower Street, London, WC1E 6BT, UK
X. Wang: Department of Civil, Environmental and Geomatic Engineering, University College London, Gower Street, London, WC1E 6BT, UK
B. Chi: Department of Geography, University College London, Gower Street, London, WC1E 6BT, UK
M. Wang: Department of Civil, Environmental and Geomatic Engineering, University College London, Gower Street, London, WC1E 6BT, UK
J. Boehm: Department of Civil, Environmental and Geomatic Engineering, University College London, Gower Street, London, WC1E 6BT, UK

DOI: https://doi.org/10.5194/isprs-archives-XLVIII-2-2024-457-2024
Journal volume & issue: Vol. XLVIII-2-2024
pp. 457 – 464

Abstract

Read online

A building’s age of construction is crucial for supporting many geospatial applications. Much current research focuses on estimating building age from facade images using deep learning. However, building an accurate deep learning model requires a considerable amount of labelled training data, and the trained models often have geographical constraints. Recently, large pre-trained vision language models (VLMs) such as GPT-4 Vision, which demonstrate significant generalisation capabilities, have emerged as potential training-free tools for dealing with specific vision tasks, but their applicability and reliability for building information remain unexplored. In this study, a zero-shot building age classifier for facade images is developed using prompts that include logical instructions. Taking London as a test case, we introduce a new dataset, FI-London, comprising facade images and building age epochs. Although the training-free classifier achieved a modest accuracy of 39.69%, the mean absolute error of 0.85 decades indicates that the model can predict building age epochs successfully albeit with a small bias. The ensuing discussion reveals that the classifier struggles to predict the age of very old buildings and is challenged by fine-grained predictions within 2 decades. Overall, the classifier utilising GPT-4 Vision is capable of predicting the rough age epoch of a building from a single facade image without any training. The code and dataset are available at https://zichaozeng.github.io/ba_classifier.

Published in The International Archives of the Photogrammetry, Remote Sensing and Spatial Information Sciences

ISSN: 1682-1750 (Print); 2194-9034 (Online)
Publisher: Copernicus Publications
Country of publisher: Germany
LCC subjects: Technology: Engineering (General). Civil engineering (General): Applied optics. Photonics
Website: http://www.isprs.org/publications/archives.aspx

About the journal