<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "https://jats.nlm.nih.gov/nlm-dtd/publishing/3.0/journalpublishing3.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article" dtd-version="3.0" xml:lang="en">
<front>
<journal-meta>
<journal-id journal-id-type="publisher">ISPRS-Annals</journal-id>
<journal-title-group>
<journal-title>ISPRS Annals of the Photogrammetry, Remote Sensing and Spatial Information Sciences</journal-title>
<abbrev-journal-title abbrev-type="publisher">ISPRS-Annals</abbrev-journal-title>
<abbrev-journal-title abbrev-type="nlm-ta">ISPRS Ann. Photogramm. Remote Sens. Spatial Inf. Sci.</abbrev-journal-title>
</journal-title-group>
<issn pub-type="epub">2194-9050</issn>
<publisher><publisher-name>Copernicus Publications</publisher-name>
<publisher-loc>Göttingen, Germany</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5194/isprs-annals-XII-4-W1-2026-187-2026</article-id>
<title-group>
<article-title>Data-Aware 3D Model Enrichment: A Multimodal Pipeline Combining LiDAR Point Clouds and Geotagged Imagery</article-title>
</title-group>
<contrib-group><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Haffadi</surname>
<given-names>Rashid</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
</contrib>
<contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Billen</surname>
<given-names>Roland</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
</contrib>
</contrib-group><aff id="aff1">
<label>1</label>
<addr-line>GeoScITY, Spheres Research Unit, University of Liège, 4000 Liège, Belgium</addr-line>
</aff>
<pub-date pub-type="epub">
<day>28</day>
<month>09</month>
<year>2026</year>
</pub-date>
<volume>XII-4/W1-2026</volume>
<fpage>187</fpage>
<lpage>194</lpage>
<permissions>
<copyright-statement>Copyright: &#x000a9; 2026 Rashid Haffadi</copyright-statement>
<copyright-year>2026</copyright-year>
<license license-type="open-access">
<license-p>This work is licensed under the Creative Commons Attribution 4.0 International License. To view a copy of this licence, visit <ext-link ext-link-type="uri"  xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</ext-link></license-p>
</license>
</permissions>
<self-uri xlink:href="https://isprs-annals.copernicus.org/articles/XII-4-W1-2026/187/2026/isprs-annals-XII-4-W1-2026-187-2026.html">This article is available from https://isprs-annals.copernicus.org/articles/XII-4-W1-2026/187/2026/isprs-annals-XII-4-W1-2026-187-2026.html</self-uri>
<self-uri xlink:href="https://isprs-annals.copernicus.org/articles/XII-4-W1-2026/187/2026/isprs-annals-XII-4-W1-2026-187-2026.pdf">The full text article is available as a PDF file from https://isprs-annals.copernicus.org/articles/XII-4-W1-2026/187/2026/isprs-annals-XII-4-W1-2026-187-2026.pdf</self-uri>
<abstract>
<p>Enhancing 3D city models with facade-level semantics is essential for urban analyses such as flood damage estimation, energy performance modelling, and accessibility assessment. Many existing enrichment methods rely on a single data source and do not systematically produce standardised CityJSON geometry. This study presents a data-aware pipeline that enriches LoD2 building models with windows, doors, and floor levels through two interchangeable workflows, one operating on Mobile Laser Scanning (MLS) point clouds and one on geotagged street-level imagery, that converge to a shared floor estimation and CityJSON enrichment backend. Both workflows employ the Grounding DINO open vocabulary detector in zero shot mode, eliminating the need for task specific training data. Evaluated on data from the Vesdre valley (Belgium), the LiDAR workflow processed 206 buildings and achieved 95.5% and 97.0% within one count accuracy for window and door counts, with 90.8% exact match floor count accuracy. The image workflow processed 338 buildings from a separate area, reaching 93.5% and 95.0% within one count accuracy for windows and doors, with comparable floor estimation, the lower opening accuracy reflecting street-level occlusions. The resulting enriched CityJSON models contain Window, Door, and FloorSurface semantics at LoD3, demonstrating a scalable approach to urban model enrichment that adapts to locally available data.</p>
</abstract>
<counts><page-count count="8"/></counts>
</article-meta>
</front>
<body/>
<back>
</back>
</article>