arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:2607.26421v1 [eess.SP] 29 Jul 2026

MVLA-GR: A Phase-Free Multipath-Based Geometry Reconstruction Method via Multi-View Likelihood Accumulation for ISAC

Bowei Xing, , Yuxiang Zhang, , Jianhua Zhang, , Yifeng Xiong, , Hongbo Xing, , Li Yu, , and Guangyi Liu This work was supported in part by National Natural Science Foundation of China (62401084, 62525101), National Key R&D Program of China (2023YFB2904803, 2023YFB2904805), National Science and Technology Major Project (2025ZD1301700), Guangdong Major Project of Basic and Applied Basic Research (2023B0303000001), Natural Science Foundation of Beijing Xiaomi Innovation Joint Foundation (L243002), and the Beijing University of Posts and Telecommunications-China Mobile Communications Group Co., Ltd. Joint Institute. (Corresponding author: Jianhua Zhang.)B. Xing, Y. Zhang, J. Zhang, H. Xing, and L. Yu are with the State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications, Beijing 100876, China (e-mail: bwxing@bupt.edu.cn; zhangyx@bupt.edu.cn; jhzhang@bupt.edu.cn; hbxing@bupt.edu.cn; li.yu@bupt.edu.cn).Y. Xiong is with the School of Information and Electronic Engineering, Beijing University of Posts and Telecommunications, Beijing 100876, China (e-mail: yifengxiong@bupt.edu.cn).G. Liu is with the Future Research Laboratory, China Mobile Research Institute, Beijing 100053, China (e-mail: liuguangyi@chinamobile.com).
Abstract

Integrated sensing and communication (ISAC) enables wireless systems to reuse communication signals for environmental sensing, where reconstructing the geometry of surrounding objects is a representative sensing task. However, many conventional methods rely on coherent processing and require accurate phase information, which is often hard to guarantee in practical communication systems, particularly at high carrier frequencies. To address this problem, this paper proposes a Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) method based on channel impulse response (CIR) measurements, which uses only delay and power observations without requiring phase information. The method extracts dominant multipath components from each observation, and for each candidate spatial location, accumulates components across views whose propagation distances match the location as supporting evidence. A soft distance-matching kernel is introduced to tolerate range estimation errors and viewpoint-dependent scattering migration, and the received power of each component is used as a reliability weight. A joint thresholding strategy combining response magnitude and angular support continuity then converts the continuous support map into a binary geometry estimate. Ray-tracing simulations on canonical and complex targets, as well as real-world vehicle measurements at 36 GHz, demonstrate that MVLA-GR can effectively recover target geometry, providing a low-complexity phase-free solution for ISAC.

I Introduction

Integrated sensing and communication (ISAC) enables wireless systems to reuse communication signals for sensing the surrounding environment [13, 12, 5]. Among various sensing tasks such as target detection, parameter estimation, and contour characterization [23], reconstructing the geometry of surrounding objects, such as their contour, shape, and spatial layout, is of particular interest because it provides a structural description of the propagation environment rather than isolated target parameters. Such geometric information is a key enabler for emerging concepts such as the digital twin channel, where a synchronized geometric model of the environment is used to characterize and predict wireless propagation [22, 21]. It also benefits various communication tasks in 6G networks, including environment-aware beamforming, blockage prediction, and proactive beam management.

In ISAC systems, pilot signals embedded in communication waveforms are commonly used for channel estimation, from which the channel impulse response (CIR) can be obtained as a basic observation. A wideband CIR describes the multipath structure of the wireless channel in the delay domain, where each propagation delay corresponds to a propagation distance and each multipath component reflects a scattering or reflection event in the environment, whose strength further depends on the radar cross section of the underlying scatterer [25]. As a result, the CIR inherently carries geometric information about the surrounding scatterers [10], and its acquisition is naturally aligned with the communication-native operation of ISAC. This makes CIR a particularly suitable observation for environmental geometry reconstruction. Along this direction, the wireless environmental information theory has recently been proposed as a new paradigm for 6G environment intelligence communication [24].

A common way to reconstruct environmental geometry from wave-domain observations is to exploit the complex-valued wave response through coherent imaging. Representative examples include synthetic aperture radar (SAR) imaging [4], microwave tomography [17], and inverse-scattering reconstructions [8, 7, 14, 11], which jointly use the magnitude and phase of the measured field to reconstruct target shape or contrast. These methods can provide accurate reconstructions when reliable magnitude and phase measurements are available. However, in practical ISAC scenarios, although the magnitude of the CIR can be reliably extracted from communication pilots, the absolute phase reference of each CIR measurement and its consistency across observation positions remain challenging to maintain. Such challenges arise from two distinct sources. From the hardware perspective, oscillator drift, timing offsets, and the lack of cross-view phase synchronization at high carrier frequencies introduce per-observation phase offsets that are not easily predictable from one observation to the next [3, 15]. From the physical perspective, the reflection phase introduced by surface scattering depends on both material composition and incidence geometry through the complex-valued Fresnel reflection coefficient, and varies non-trivially with viewpoint, especially for materials and surface conditions that are not known a priori. These factors together limit the practical applicability of coherent imaging in ISAC scenarios. From this perspective, phase-free approaches that rely on the magnitude information in the CIR are intrinsically robust to such phase uncertainties and are therefore better suited for ISAC scenarios.

Phase-free reconstruction has been studied along two main directions. One direction is phaseless inverse scattering, which reconstructs object shape or material distribution from the magnitude or intensity of scattered fields by solving a wave-equation-based nonlinear inverse problem [2, 6, 16]. Such methods typically require a well-defined forward electromagnetic model and dense spatial or frequency sampling, and involve nonlinear iterative optimization, which makes them less suitable for lightweight CIR-based sensing in ISAC. The other direction is direct spatial accumulation, which projects measured delay-domain responses such as CIRs or PDPs back into the spatial domain and accumulates them over multiple observation positions [19, 26]. Incoherent backprojection is a representative example of this direction, and CIR-based heatmap construction follows a similar principle.

Among these two directions, direct spatial accumulation is more compatible with the lightweight, communication-native nature of ISAC, but it still has clear limitations when applied to CIR-based geometry reconstruction. First, direct spatial accumulation projects the entire PDP back into space, so that not only the dominant peaks but also sidelobes, noise floor, and other non-peak portions contribute to the spatial support. Since these non-peak portions of the PDP do not correspond to physical scattering events, their accumulation tends to blur the reconstruction and make the result sensitive to the detailed shape of the measured PDP. Second, direct spatial accumulation is essentially a heuristic projection scheme that does not explicitly model the underlying observation process, such as the statistical behavior of distance estimation, the viewpoint dependence of dominant scattering, or the angular continuity of true scattering structures. These unmodeled factors directly affect the reliability of the accumulated support, but cannot be addressed within the projection framework itself.

To address these gaps, a phase-free environmental geometry reconstruction method based on multi-view likelihood accumulation is proposed in this paper, with CIR-derived delay and power observations as the only input. The main contributions are summarized as follows.

  • A phase-free framework for environmental geometry reconstruction in ISAC is established, in which the problem is formulated as voxel-wise support estimation from CIR measurements, and only delay and power observations extracted from the CIR are used as the basic input.

  • Within this framework, the proposed method takes individual multipath components as the basic processing units rather than the entire PDP, and converts each detected component into a distance-based geometric constraint through multi-view likelihood accumulation. A soft distance-matching kernel is introduced to tolerate range estimation errors and viewpoint-dependent scattering migration, the received power of each component serves as a reliability weight during accumulation, and a joint thresholding strategy combining response magnitude and angular support continuity is further applied to suppress artifacts and convert the continuous support map into a binary geometry estimate.

  • The proposed method is validated through ray-tracing simulations on canonical and complex targets under various angular sampling intervals and SNR conditions, as well as real-world vehicle measurements at 36 GHz. Experimental results demonstrate that MVLA-GR can effectively recover target geometry and outperforms a standard incoherent backprojection baseline in the majority of tested conditions, confirming its practical applicability for ISAC scenarios.

The remainder of this paper is organized as follows. Section II introduces the system model and formulates the phase-free environmental geometry reconstruction problem from CIR measurements. Section III presents the proposed MVLA-GR method, including multi-view likelihood accumulation, parameter design, and binary reconstruction. Section IV provides simulation and measurement results to evaluate the proposed method.

II System Model and Problem Formulation

Refer to caption
Figure 1: System model for multi-view geometry reconstruction. {𝐓m}\{\mathbf{T}_{m}\} denote the transceiver positions relative to the target.

As illustrated in Fig. 1, we consider a multi-view geometry reconstruction scenario in which a single omnidirectional transceiver collects echo signals at MM positions {𝐓m}\{\mathbf{T}_{m}\} relative to the target. As discussed in Section IV, this multi-view observation structure can be physically realized either by a moving transceiver in a static environment, or by a stationary transceiver observing a moving target.

II-A Voxel Representation

Under the Born approximation [1], the electromagnetic interaction between the incident field and the target is treated as a single-scattering process in which each scattering element contributes independently, without accounting for mutual coupling or multiple reflections between elements. This assumption allows the target response to be represented by a set of spatially distributed scattering coefficients. Accordingly, the region of interest is discretized into NN voxels with centers 𝐩n3\mathbf{p}_{n}\in\mathbb{R}^{3}, n=1,,Nn=1,\ldots,N, each associated with a nonnegative effective scattering intensity xn0x_{n}\geq 0. The voxel intensity vector is defined as

𝐱=[x1,x2,,xN]T+N.\mathbf{x}=[x_{1},x_{2},\ldots,x_{N}]^{T}\in\mathbb{R}_{+}^{N}. (1)

Here, xnx_{n} is interpreted as an effective scattering intensity used for geometry reconstruction. Specifically, xn=0x_{n}=0 indicates the absence of a scatterer, while xn>0x_{n}>0 indicates the presence of a scatterer. The support of 𝐱\mathbf{x}, i.e., the set of voxels for which xn>0x_{n}>0, encodes the geometric structure and contour of the target. Therefore, the objective of the considered geometry reconstruction problem is not to recover the exact value of every xnx_{n}, but to determine which voxels satisfy xn>0x_{n}>0.

II-B Electromagnetic Propagation and Received Signal Model

To establish the relationship between the voxel scattering coefficients {xn}\{x_{n}\} and the received signal at each observation position, we characterize the round-trip propagation between the transceiver and each voxel. Starting from the free-space Green’s function

G(𝐫,𝐫)=ejk𝐫𝐫4π𝐫𝐫,G(\mathbf{r},\mathbf{r}^{\prime})=\frac{e^{\,jk\|\mathbf{r}-\mathbf{r}^{\prime}\|}}{4\pi\|\mathbf{r}-\mathbf{r}^{\prime}\|}, (2)

where k=2π/λk=2\pi/\lambda is the wavenumber and λ=c/fc\lambda=c/f_{c} is the carrier wavelength, the round-trip propagation between the mm-th transceiver position 𝐓m\mathbf{T}_{m} and voxel nn is modeled as the product of two such Green’s functions, leading to an effective complex channel gain

gm,n=ejϕm,n(4πrm,n)2,g_{m,n}=\frac{e^{\,j\phi_{m,n}}}{(4\pi r_{m,n})^{2}}, (3)

where rm,n=𝐩n𝐓mr_{m,n}=\|\mathbf{p}_{n}-\mathbf{T}_{m}\| is the geometric distance between voxel nn and the mm-th transceiver position. The phase term ϕm,n\phi_{m,n} accounts for the round-trip propagation phase 2krm,n2kr_{m,n} together with hardware-induced contributions such as oscillator drift and timing offsets, which are not directly available in practice. In this model, gm,ng_{m,n} captures the round-trip propagation effect, while xnx_{n} represents the effective scattering intensity associated with voxel nn.

A wideband pilot signal s(t)s(t) of duration TT and bandwidth BB is assumed to be transmitted at each observation position and reused for sensing. The specific waveform of s(t)s(t) is not restricted in this work; any standard wideband pilot whose autocorrelation provides a main-lobe range resolution on the order of c/(2B)c/(2B) is applicable, such as OFDM training sequences, linear frequency-modulated chirps, or pseudo-random sequences commonly used in communication systems. The corresponding round-trip delay from the mm-th transceiver position to voxel nn is

τm,n=2rm,nc,\tau_{m,n}=\frac{2r_{m,n}}{c}, (4)

where cc denotes the speed of light. The received signal at the mm-th transceiver position consists of target echoes, reflections from the surrounding environment, and additive noise:

ym(t)=n=1Nxngm,ns(tτm,n)+senv,m(t)+ηm(t),y_{m}(t)=\sum_{n=1}^{N}x_{n}\,g_{m,n}\,s(t-\tau_{m,n})+s_{\mathrm{env},m}(t)+\eta_{m}(t), (5)

where senv,m(t)s_{\mathrm{env},m}(t) denotes reflections from objects outside the imaging region and ηm(t)\eta_{m}(t) denotes the receiver noise.

Since the imaging region occupies a bounded spatial area, the target echoes correspond to a well-defined interval of round-trip delays determined by the geometry of the imaging region and the transceiver trajectory. Reflections from objects outside this region arrive at delays outside this interval and can be suppressed by range gating. After retaining only the delay interval [τmin,τmax][\tau_{\min},\tau_{\max}] corresponding to the imaging region, the received signal is modeled as

ym(t)=n=1Nxngm,ns(tτm,n)+ηm(t).y_{m}(t)=\sum_{n=1}^{N}x_{n}\,g_{m,n}\,s(t-\tau_{m,n})+\eta_{m}(t). (6)

II-C CIR Measurement Model

The CIR is obtained by correlating the received signal ym(t)y_{m}(t) with the known pilot s(t)s(t) over the observation interval of duration TT:

hm(τ)=0Tym(t)s(tτ)dt.h_{m}(\tau)=\int_{0}^{T}y_{m}(t)\,s^{*}(t-\tau)\,\mathrm{d}t. (7)

Substituting (6) into (7) yields

hm(τ)=n=1Nxngm,np(ττm,n)+η~m(τ),h_{m}(\tau)=\sum_{n=1}^{N}x_{n}\,g_{m,n}\,p(\tau-\tau_{m,n})+\tilde{\eta}_{m}(\tau), (8)

where

p(τ)=0Ts(t)s(tτ)dtp(\tau)=\int_{0}^{T}s(t)\,s^{*}(t-\tau)\,\mathrm{d}t (9)

is the autocorrelation of the pilot waveform, serving as the effective point-spread function of the system, and η~m(τ)\tilde{\eta}_{m}(\tau) is the filtered noise term. For a pilot signal occupying bandwidth BB, the range resolution is

ΔR=c2B,\Delta R=\frac{c}{2B}, (10)

which determines the minimum resolvable separation in the range dimension.

Discretizing the delay axis into LL taps {τl=l/B}l=0L1\{\tau_{l}=l/B\}_{l=0}^{L-1} and evaluating (8) at each tap, the CIR can be written in matrix-vector form as

𝐡m=𝐆m𝐱+𝜼m,\mathbf{h}_{m}=\mathbf{G}_{m}\mathbf{x}+\boldsymbol{\eta}_{m}, (11)

where 𝐡m=[hm(τ0),,hm(τL1)]TL\mathbf{h}_{m}=[h_{m}(\tau_{0}),\ldots,h_{m}(\tau_{L-1})]^{T}\in\mathbb{C}^{L}, 𝜼mL\boldsymbol{\eta}_{m}\in\mathbb{C}^{L} is the noise vector, and 𝐆mL×N\mathbf{G}_{m}\in\mathbb{C}^{L\times N} is the observation matrix with (l,n)(l,n)-th element

[𝐆m]l,n=gm,np(τlτm,n).[\mathbf{G}_{m}]_{l,n}=g_{m,n}\,p(\tau_{l}-\tau_{m,n}). (12)

In (12), the delay τm,n\tau_{m,n} is geometrically determined by the known positions 𝐓m\mathbf{T}_{m} and 𝐩n\mathbf{p}_{n} via (4). However, fully specifying (12) also requires phase information that is reliable and mutually consistent across different observation positions. In the considered mobile single-transceiver system, this condition is difficult to guarantee due to the high phase sensitivity at high frequencies, hardware-dependent phase offsets, and the lack of precise cross-view phase synchronization. Coherent inversion based on 𝐆m\mathbf{G}_{m} is therefore not pursued in this work, and the subsequent development relies only on phase-insensitive features extracted from the CIR.

It is worth noting that, although the CIR hm(τ)h_{m}(\tau) is itself complex-valued, the per-observation phase factors arising from oscillator drift, sampling timing offsets, and transceiver localization errors appear as global multiplicative unit-modulus factors on hm(τ)h_{m}(\tau), which are eliminated by the squared-magnitude operation |hm(τ)|2|h_{m}(\tau)|^{2}. The squared-magnitude operation therefore yields a stable representation that does not depend on these uncertain phase quantities, while remaining sensitive to the underlying multipath structure through the magnitude information.

Accordingly, instead of relying on phase, this work adopts the power delay profile (PDP), defined as

Pm(τ)=|hm(τ)|2,P_{m}(\tau)=|h_{m}(\tau)|^{2}, (13)

as a phase-insensitive representation for subsequent geometry reconstruction.

II-D Phase-Free Observations

Based on the phase-insensitive representation in (13), this work extracts phase-free observations from the PDP at each observation position. Specifically, a standard peak detection procedure is applied to Pm(τ)P_{m}(\tau) to identify KmK_{m} detected multipath components, each characterized by its detected delay and received power. The detection is based on local maxima of Pm(τ)P_{m}(\tau), with a noise-relative magnitude threshold to suppress spurious detections, and a minimum inter-peak delay separation on the order of 1/B1/B to avoid splitting a single main lobe into multiple peaks.

Due to the finite range resolution ΔR\Delta R, multiple voxel contributions with sufficiently close round-trip delays may be mapped to the same detected multipath component in the PDP. Let 𝒩m,i\mathcal{N}_{m,i} denote the index set of voxels contributing to the ii-th detected component at the mm-th view. Its complex amplitude can be written as

αm,i=n𝒩m,ixngm,np(τm,iτm,n)+η~m,i,\alpha_{m,i}=\sum_{n\in\mathcal{N}_{m,i}}x_{n}\,g_{m,n}\,p(\tau_{m,i}-\tau_{m,n})+\tilde{\eta}_{m,i}, (14)

where η~m,i\tilde{\eta}_{m,i}\in\mathbb{C} is the effective complex noise term at the corresponding delay bin after matched filtering. From each detected multipath component, two phase-free quantities are extracted:

Rm,i\displaystyle R_{m,i} =cτm,i2,\displaystyle=\frac{c\,\tau_{m,i}}{2}, (15)
pm,i\displaystyle p_{m,i} =|αm,i|2|n𝒩m,ixngm,np(τm,iτm,n)|2+ξm,i,\displaystyle=|\alpha_{m,i}|^{2}\approx\left|\sum_{n\in\mathcal{N}_{m,i}}x_{n}\,g_{m,n}\,p(\tau_{m,i}-\tau_{m,n})\right|^{2}+\xi_{m,i}, (16)

where Rm,iR_{m,i} is the propagation distance inferred from the detected delay τm,i\tau_{m,i}, pm,ip_{m,i} is the received power of the ii-th detected multipath component, and ξm,i\xi_{m,i} denotes a power-domain perturbation arising from receiver noise and its cross-coupling with the signal term. The complete phase-free observation set at the mm-th view is then defined as

𝒞m={(Rm,i,pm,i)}i=1Km,\mathcal{C}_{m}=\bigl\{(R_{m,i},\;p_{m,i})\bigr\}_{i=1}^{K_{m}}, (17)

where KmK_{m} is the number of detected multipath components.

In the subsequent development, the distances {Rm,i}\{R_{m,i}\} provide the geometric constraints used for voxel localization, while the powers {pm,i}\{p_{m,i}\} are used as reliability indicators for multi-view fusion.

II-E Problem Formulation

Since the geometry reconstruction objective is to identify the support of 𝐱\mathbf{x}, we introduce the voxel occupancy variable

zn={1,xn>0,0,xn=0,n=1,,N,z_{n}=\begin{cases}1,&x_{n}>0,\\ 0,&x_{n}=0,\end{cases}\qquad n=1,\ldots,N, (18)

and define the binary occupancy vector

𝐳=[z1,,zN]T{0,1}N.\mathbf{z}=[z_{1},\ldots,z_{N}]^{T}\in\{0,1\}^{N}. (19)

The geometry reconstruction problem is therefore formulated as the estimation of the voxel occupancy pattern from the multi-view phase-free observations {𝒞m}m=1M\{\mathcal{C}_{m}\}_{m=1}^{M}:

𝐳^=argmax𝐳{0,1}NP(𝐳{𝒞m}m=1M).\hat{\mathbf{z}}=\arg\max_{\mathbf{z}\in\{0,1\}^{N}}P\!\left(\mathbf{z}\mid\{\mathcal{C}_{m}\}_{m=1}^{M}\right). (20)

Solving (20) exactly requires the likelihood P({𝒞m}m=1M𝐳)P(\{\mathcal{C}_{m}\}_{m=1}^{M}\mid\mathbf{z}), which in turn requires an explicit model relating the observed powers {pm,i}\{p_{m,i}\} to the underlying voxel occupancies. However, as seen from (16), the received power depends on the unknown phases embedded in {gm,n}\{g_{m,n}\}. Even if the occupied voxels were known, the observed powers could not be predicted explicitly without these phases. Consequently, the exact likelihood cannot be written in closed form, and the posterior in (20) is not directly tractable.

Nevertheless, the phase-free observations still contain useful information for support identification. In particular, the distance measurements {Rm,i}\{R_{m,i}\} provide geometric constraints on the possible locations of scatterers, while the received powers {pm,i}\{p_{m,i}\} reflect the relative reliability of the detected multipath components. This suggests that the voxel occupancy can be approximated from multi-view geometric consistency, with received power serving as a reliability weight. The proposed MVLA-GR method in Section III is developed along this line by constructing a tractable approximation of the posterior probability of voxel occupancy from the observation sets {𝒞m}m=1M\{\mathcal{C}_{m}\}_{m=1}^{M}.

III Multi-View Likelihood Accumulation Method for Geometry Reconstruction

Refer to caption
Figure 2: Overview of the proposed Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) pipeline.

This section presents the proposed Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) method. As discussed at the end of Section II, the posterior in (20) is not directly tractable, but the phase-free observations still carry useful geometric information. The proposed method exploits this information by approximating the posterior probability of voxel occupancy through two phase-free quantities that the CIR observations provide reliably, namely the propagation distances {Rm,i}\{R_{m,i}\} and the received powers {pm,i}\{p_{m,i}\}. The distances are used to test whether each voxel is geometrically consistent with the observed multipath components across views, and the powers are used to weight the contribution of each component according to its reliability. By accumulating these weighted geometric matches over all observation views, a voxel-wise support score sns_{n} is obtained as a tractable surrogate for the posterior probability of voxel occupancy. The support map {sn}n=1N\{s_{n}\}_{n=1}^{N} is then converted into the final binary reconstruction 𝐳^\hat{\mathbf{z}} through a joint thresholding step that combines response magnitude and angular support continuity.

These steps are summarized in the processing pipeline shown in Fig. 2. From left to right, multipath components are first extracted from the PDP at each observation position to form the phase-free observations. The region of interest is then discretized into a voxel grid, and the multi-view geometric consistency between each voxel and the extracted distance–power observations is evaluated and accumulated into a support map. Finally, the joint thresholding step is applied to convert the continuous support map into a binary geometry estimate. The remainder of this section describes each of these components in detail.

III-A Multi-View Likelihood Accumulation

III-A1 Hard-Decision Occupancy Support

For voxel nn and observation view mm, consider the event that at least one detected propagation distance is geometrically consistent with the voxel distance, i.e.,

m,n={i{1,,Km}s.t.|Rm,irm,n|ϵR},\mathcal{E}_{m,n}=\left\{\exists\,i\in\{1,\dots,K_{m}\}\ \text{s.t.}\ |R_{m,i}-r_{m,n}|\leq\epsilon_{R}\right\}, (21)

where Rm,iR_{m,i} is the ii-th detected propagation distance at view mm, rm,nr_{m,n} is the geometric distance between voxel nn and the mm-th transceiver position, and ϵR\epsilon_{R} is a distance tolerance related to the range resolution.

If voxel nn corresponds to a true scattering location, the event m,n\mathcal{E}_{m,n} is more likely to occur. In contrast, for an unoccupied voxel, this event can only occur accidentally with a much smaller probability. Based on this observation, we define the hard-decision support indicator

Im,n=𝕀(m,n),I_{m,n}=\mathbb{I}(\mathcal{E}_{m,n}), (22)

where 𝕀()\mathbb{I}(\cdot) denotes the indicator function.

Treating the MM observation views as independent binary trials, the fraction of views in which the event m,n\mathcal{E}_{m,n} occurs provides a coarse approximation to the posterior probability of voxel occupancy:

P~hard(zn=1{𝒞m}m=1M)1Mm=1MIm,n,\tilde{P}_{\mathrm{hard}}\!\left(z_{n}=1\mid\{\mathcal{C}_{m}\}_{m=1}^{M}\right)\approx\frac{1}{M}\sum_{m=1}^{M}I_{m,n}, (23)

where P~hard()\tilde{P}_{\mathrm{hard}}(\cdot) denotes a surrogate posterior probability rather than an exact one. Under this hard-decision model, voxels supported by a larger fraction of observation views are more likely to be occupied.

III-A2 Soft Distance-Matching Kernel

The hard-decision rule in (23) is not sufficiently robust in practice. First, the extracted distances Rm,iR_{m,i} are affected by finite bandwidth, peak detection errors, and measurement noise. Second, under viewpoint variation, the dominant scattering point associated with the same physical structure may migrate along the target surface, especially in specular-dominated scenarios. As a result, a strict binary decision may reject physically meaningful observations that exhibit only a small distance deviation.

To address these issues, we replace the hard-decision indicator by a soft probabilistic kernel derived from the statistics of the distance estimation error.

For each detected multipath component, we assume that its corresponding peak in the power delay profile is sufficiently isolated from adjacent components, and that the measurement SNR is sufficiently high. Under these conditions, the extracted propagation distance can be modeled as

Rm,i=rm,i+εm,i,R_{m,i}=r^{\star}_{m,i}+\varepsilon_{m,i}, (24)

where Rm,iR_{m,i} is the extracted propagation distance, rm,ir^{\star}_{m,i} is the true propagation distance of the ii-th component, and εm,i\varepsilon_{m,i} is the associated distance estimation error. We assume that εm,i\varepsilon_{m,i} approximately follows a zero-mean Gaussian distribution, i.e., εm,i𝒩(0,σε2)\varepsilon_{m,i}\sim\mathcal{N}(0,\sigma_{\varepsilon}^{2}). This assumption is motivated by the asymptotic normality of maximum-likelihood estimators under additive Gaussian noise [20, 9], and is commonly adopted as a tractable approximation in time-delay estimation problems [18]. In practical multipath scenarios where peaks are not fully isolated, this Gaussian model is still used here as an approximate error model for soft geometric matching.

Under this Gaussian error model, the conditional likelihood of observing distance Rm,iR_{m,i} given that voxel nn is occupied and located at geometric distance rm,nr_{m,n} is proportional to

P(Rm,izn=1,rm,n)exp((rm,nRm,i)22σR2),P(R_{m,i}\mid z_{n}=1,\,r_{m,n})\propto\exp\!\left(-\frac{(r_{m,n}-R_{m,i})^{2}}{2\sigma_{R}^{2}}\right), (25)

where σR=σε\sigma_{R}=\sigma_{\varepsilon} is the standard deviation of the range estimation error. Defining the distance mismatch

dm,i(n)=rm,nRm,i,d_{m,i}^{(n)}=r_{m,n}-R_{m,i}, (26)

this conditional likelihood motivates the following truncated Gaussian kernel as the soft replacement for the hard indicator Im,nI_{m,n}:

ϕ(dm,i(n))={exp((dm,i(n))22σR2),|dm,i(n)|RT,0,otherwise,\phi\!\left(d_{m,i}^{(n)}\right)=\begin{cases}\exp\!\left(-\dfrac{(d_{m,i}^{(n)})^{2}}{2\sigma_{R}^{2}}\right),&|d_{m,i}^{(n)}|\leq R_{\mathrm{T}},\\[6.0pt] 0,&\text{otherwise},\end{cases} (27)

where RTR_{\mathrm{T}} is a truncation radius beyond which the voxel is regarded as geometrically incompatible with the observation. The truncation ensures that clearly inconsistent distance observations do not contribute to the voxel support and also improves computational efficiency.

III-A3 Power-Based Weighting

Different detected multipath components do not provide equally reliable geometric evidence and should therefore not contribute equally to voxel occupancy support. The received power pm,ip_{m,i} is adopted as a reliability weight for two main reasons.

First, under matched filtering, the delay estimation error εm,i\varepsilon_{m,i} decreases with increasing SNR. Components with higher received power generally correspond to higher effective SNR, and their extracted distances are therefore more reliable. Weighting by power naturally suppresses weak and noisy components whose geometric constraints are less trustworthy.

Second, the observed power also reflects the strength of the corresponding propagation path. Stronger detected components usually carry more informative evidence about the underlying scattering structure, while weaker components are more likely to be unstable or noise-contaminated. Therefore, assigning larger weights to stronger components helps the multi-view accumulation focus on more credible geometric evidence.

Accordingly, to preserve the relative reliability differences among components within the same view while preventing any single view from dominating the overall fusion, the weight for each detected component is defined through view-wise normalization:

wm,i=pm,i=1Kmpm,.w_{m,i}=\frac{p_{m,i}}{\sum_{\ell=1}^{K_{m}}p_{m,\ell}}. (28)

The view-wise normalization in (28) implicitly assumes that the observation views are of comparable quality, which is typically valid in the considered multi-view sensing settings, where all observations are collected by the same transceiver in a controlled measurement campaign and therefore share similar noise levels and detection conditions. Under this assumption, the relative powers within each view directly reflect the relative reliability of the detected components.

III-A4 Accumulation and Posterior Approximation

With the soft kernel and power-based weights defined above, the contribution from the ii-th detected component at view mm to voxel nn is

m,i(n)=wm,iϕ(dm,i(n)).\ell_{m,i}^{(n)}=w_{m,i}\,\phi\!\left(d_{m,i}^{(n)}\right). (29)

Accumulating all detected components within view mm, we define the total support of voxel nn under the mm-th observation as

am,n=i=1Kmwm,iϕ(dm,i(n)),a_{m,n}=\sum_{i=1}^{K_{m}}w_{m,i}\,\phi\!\left(d_{m,i}^{(n)}\right), (30)

where KmK_{m} is the number of detected multipath components at view mm.

The quantity am,na_{m,n} can be interpreted as a surrogate of the posterior probability of voxel occupancy under the mm-th observation, i.e.,

am,nP~m(zn=1𝒞m),a_{m,n}\approx\tilde{P}_{m}\!\left(z_{n}=1\mid\mathcal{C}_{m}\right), (31)

where P~m()\tilde{P}_{m}(\cdot) denotes a surrogate posterior probability rather than an exact one. In this sense, am,na_{m,n} is the soft generalization of the hard indicator Im,nI_{m,n} in (22).

To aggregate evidence across all observation views, we define the normalized multi-view support score of voxel nn as

sn=1Mm=1Mam,n=1Mm=1Mi=1Kmwm,iϕ(dm,i(n)).s_{n}=\frac{1}{M}\sum_{m=1}^{M}a_{m,n}=\frac{1}{M}\sum_{m=1}^{M}\sum_{i=1}^{K_{m}}w_{m,i}\,\phi\!\left(d_{m,i}^{(n)}\right). (32)

Accordingly, sns_{n} can be interpreted as a tractable approximation to the posterior probability of occupancy for voxel nn under the full multi-view observations, i.e.,

snP~(zn=1{𝒞m}m=1M).s_{n}\approx\tilde{P}\!\left(z_{n}=1\mid\{\mathcal{C}_{m}\}_{m=1}^{M}\right). (33)

A voxel with a larger sns_{n} is supported by more observation views and more reliable multipath components, and is therefore more likely to be occupied. The collection {sn}n=1N\{s_{n}\}_{n=1}^{N} forms the reconstructed occupancy map used in all subsequent processing.

The computational complexity of the proposed accumulation process is dominated by the evaluation of voxel-wise contributions for all detected multipath components across all observation views, resulting in

𝒪(Nm=1MKm)=𝒪(MNK¯),\mathcal{O}\!\left(N\sum_{m=1}^{M}K_{m}\right)=\mathcal{O}(MN\bar{K}), (34)

where NN is the number of voxels, KmK_{m} is the number of detected multipath components at view mm, and K¯\bar{K} denotes the average number of components per view.

III-B Parameter Selection and Binary Reconstruction

The accumulation process in Section III-A yields a continuous voxel-wise support score sns_{n}, which serves as a surrogate of the posterior probability of voxel occupancy. To obtain a reliable binary reconstruction from this continuous support map, two issues must be further addressed. First, the kernel parameters should be chosen such that valid supports can accumulate across views while maintaining sufficient radial discrimination. Second, accidental overlaps among supports from unrelated scattering structures must be suppressed. These two aspects are discussed in the following subsections.

III-B1 Parameter σR\sigma_{R} (Radial Kernel Scale)

Refer to caption
Figure 3: Illustration of viewpoint-dependent migration of dominant scattering locations.

In (27), the parameter σR\sigma_{R} controls how fast the soft support decays as the distance mismatch dm,i(n)d_{m,i}^{(n)} increases. As established above, σR\sigma_{R} should reflect the effective uncertainty of the extracted propagation distance. Although this uncertainty in general depends on multiple factors, including receiver noise, the dominant scale in multipath-rich scenarios is set by the system range resolution ΔR=c/(2B)\Delta R=c/(2B), which determines the main-lobe width of the pulse shaping function p(τ)p(\tau) and thereby the achievable precision of peak detection. Accordingly, σR\sigma_{R} is parameterized as

σR=κΔR,\sigma_{R}=\kappa\,\Delta R, (35)

where κ\kappa is a dimensionless proportionality constant controlling the tolerance of the radial kernel. A smaller κ\kappa yields sharper radial discrimination but weaker robustness to range errors, whereas a larger κ\kappa increases tolerance at the cost of reduced spatial selectivity. In practice, κ\kappa should be chosen according to the trade-off between range uncertainty and discrimination ability, and is determined in the experimental section.

III-B2 Parameter RTR_{\mathrm{T}} (Truncation Radius)

The parameter RTR_{\mathrm{T}} in (27) determines the effective support radius of the radial kernel. Under a purely range-limited model, RTR_{\mathrm{T}} would naturally be chosen on the same order as σR\sigma_{R}. In practical geometry reconstruction scenarios, however, the dominant scattering point associated with the same physical reflecting region may change with the observation viewpoint, especially under specular-dominated scattering conditions. As illustrated in Fig. 3, when the transceiver moves from 𝐓1\mathbf{T}_{1} to 𝐓2\mathbf{T}_{2}, the dominant scattering response may shift from one point on the target surface to another nearby point. Although these responses still originate from the same physical reflecting region, their effective propagation distances may differ due to viewpoint-dependent migration of the dominant scattering location.

If RTR_{\mathrm{T}} is chosen solely according to the range-resolution scale, valid supports associated with the same physical structure may fail to overlap across different views, weakening the subsequent likelihood accumulation. To preserve cross-view overlap among supports corresponding to the same physical scattering region, the truncation radius should also account for viewpoint-induced scattering migration.

To this end, we introduce a migration scale

δmove=ηD,\delta_{\mathrm{move}}=\eta D, (36)

where DD denotes the maximum spatial extent of the target and η\eta is a small dimensionless coefficient. The parameter δmove\delta_{\mathrm{move}} characterizes the typical range variation caused by viewpoint-dependent movement of dominant scattering locations. Its value depends on target geometry, surface properties, and angular sampling density, and should be selected to provide sufficient overlap among valid supports across views without excessively enlarging the kernel support.

Refer to caption
Figure 4: Illustration of reconstruction artifacts caused by accidental intersections of range constraints associated with different scattering structures.

Since the effective support radius must cover both the range-uncertainty scale and the migration scale, the truncation radius is chosen as

RT=max(σR,δmove).R_{\mathrm{T}}=\max(\sigma_{R},\delta_{\mathrm{move}}). (37)

This ensures that the likelihood kernel remains compact enough to suppress clearly incompatible voxels, while still allowing valid supports from different observation views to overlap and accumulate on the same physical scattering structure.

III-B3 Joint Thresholding for Binary Reconstruction

The accumulated support map {sn}n=1N\{s_{n}\}_{n=1}^{N} provides a continuous spatial representation of voxel occupancy support. However, before converting this support map into a binary reconstruction, spurious responses caused by accidental overlaps among unrelated scattering structures must be suppressed. Such responses appear as reconstruction artifacts, as illustrated in Fig. 4.

A true scattering voxel is expected to be supported not only by a large accumulated response, but also over a sufficiently wide continuous range of observation directions. By contrast, an artifact voxel is usually produced by accidental overlaps among unrelated observations and tends to be supported only within a limited or fragmented angular region. Therefore, in addition to the accumulated response magnitude, the continuity of voxel support across observation directions provides an important cue for artifact suppression.

To characterize this persistence, we first define a binary support indicator for voxel nn at view mm as

bm,n={1,am,n>0,0,am,n=0,b_{m,n}=\begin{cases}1,&a_{m,n}>0,\\ 0,&a_{m,n}=0,\end{cases} (38)

where am,n>0a_{m,n}>0 holds if and only if at least one detected component at view mm satisfies |dm,i(n)|RT|d_{m,i}^{(n)}|\leq R_{\mathrm{T}}, as defined in (27).

Let ϕm\phi_{m} denote the observation angle of the mm-th transceiver position with respect to the target center, and assume that {ϕm}\{\phi_{m}\} are ordered increasingly. For voxel nn, the support sequence {bm,n}m=1M\{b_{m,n}\}_{m=1}^{M} generally forms several continuous supporting intervals along the angular axis. Denote these intervals by Ωn,\Omega_{n,\ell}, =1,,Ln\ell=1,\dots,L_{n}. For each interval Ωn,\Omega_{n,\ell}, its angular width is defined as

Θn,=maxmΩn,ϕmminmΩn,ϕm.\Theta_{n,\ell}=\max_{m\in\Omega_{n,\ell}}\phi_{m}-\min_{m\in\Omega_{n,\ell}}\phi_{m}. (39)

The angular support range of voxel nn is then defined as the maximum width among all continuous supporting intervals:

Θn=max=1,,LnΘn,.\Theta_{n}=\max_{\ell=1,\dots,L_{n}}\Theta_{n,\ell}. (40)

To exploit both the response magnitude and the angular persistence, we adopt a joint thresholding strategy. First, a magnitude threshold is applied to the accumulated voxel responses {sn}\{s_{n}\}, and only voxels whose response values lie within the top p%p\% of all voxels are retained. Second, an angular-support threshold TθT_{\theta} is imposed, and only voxels satisfying

ΘnTθ\Theta_{n}\geq T_{\theta} (41)

are preserved. The joint thresholding criterion is therefore written as

voxel n is retained if(snTop-p%,ΘnTθ).\text{voxel }n\text{ is retained if}\quad\Big(s_{n}\in\text{Top-}p\%,\;\Theta_{n}\geq T_{\theta}\Big). (42)

This criterion has a clear physical interpretation: a voxel is regarded as a reliable scattering location only if it exhibits both a high support score and a sufficiently wide continuous support range across observation directions. As a result, spurious voxels caused by accidental support overlap can be effectively suppressed, while voxels associated with true scattering structures are more likely to be retained.

The final binary reconstruction is then obtained by setting z^n=1\hat{z}_{n}=1 for retained voxels and z^n=0\hat{z}_{n}=0 otherwise, thereby yielding an estimate 𝐳^\hat{\mathbf{z}} of the occupancy vector in (20).

IV Simulation and Measurement Results

Before presenting the experimental results, we briefly note that the proposed framework applies to two complementary ISAC scenarios that share the same multi-view CIR observation structure: (i) a mobile transceiver collecting CIRs along its trajectory to reconstruct static surroundings, supporting tasks such as environment-aware beam management and digital twin construction; and (ii) a static base station observing a moving target whose successive positions provide the multi-view observations, supporting tasks such as target classification and behavior recognition. The simulations in this section follow the first scenario with a fixed circular trajectory, while the vehicle measurement in Section IV-C corresponds to the second scenario, where the antenna is fixed and the vehicle rotates on a turntable to emulate the relative aspect-angle variation between a stationary base station and a moving target.

The proposed MVLA-GR method is evaluated as follows. First, the model parameters are determined using several canonical geometric targets, so that all subsequent evaluations are carried out under a unified parameter setting rather than target-specific tuning. Second, under this fixed parameter configuration, the proposed method is quantitatively compared with a baseline method on a more complex star-shaped target under different angular sampling intervals and signal-to-noise ratios (SNRs). Finally, real-world measurement results on a vehicle target are presented to demonstrate the practical applicability of the proposed framework.

All simulation results in this section are generated using ray-tracing data produced by Remcom Wireless InSite. The carrier frequency is set to 36 GHz, and the signal bandwidth is 3 GHz, corresponding to a range resolution of ΔR=c/(2B)=0.05\Delta R=c/(2B)=0.05 m.

Unless otherwise specified, the transceiver positions are distributed along a circular trajectory with a radius of 10 m around the target, and the reconstruction region is discretized into a uniform voxel grid. It should be emphasized that the proposed MVLA-GR method does not rely on a circular trajectory itself, but only on the knowledge of the transceiver positions at which the CIR measurements are collected. The circular observation pattern adopted here is introduced purely as a convenient and reproducible simulation configuration.

In all simulation experiments, additive noise is introduced in the PDP domain after the clean CIR data are converted into PDPs. Multipath components are then re-extracted from the noisy PDPs by thresholding and peak detection, and the resulting phase-free observations are used for geometry reconstruction. The reconstruction performance in this section is evaluated using the Chamfer distance (CD). Let 𝒳={𝐱u}u=1NX\mathcal{X}=\{\mathbf{x}_{u}\}_{u=1}^{N_{X}} denote the set of ground-truth occupied voxel centers, and let 𝒴={𝐲v}v=1NY\mathcal{Y}=\{\mathbf{y}_{v}\}_{v=1}^{N_{Y}} denote the set of reconstructed occupied voxel centers. The Chamfer distance is defined as

CD=10log10(\displaystyle\mathrm{CD}=0\log_{10}\!\Bigg( 1NXu=1NXminv𝐱u𝐲v22\displaystyle\frac{1}{N_{X}}\sum_{u=1}^{N_{X}}\min_{v}\|\mathbf{x}_{u}-\mathbf{y}_{v}\|_{2}^{2} (43)
+1NYv=1NYminu𝐲v𝐱u22).\displaystyle\quad+\frac{1}{N_{Y}}\sum_{v=1}^{N_{Y}}\min_{u}\|\mathbf{y}_{v}-\mathbf{x}_{u}\|_{2}^{2}\Bigg).

The resulting CD is expressed in dB(m2), and a smaller CD indicates a closer geometric match between the reconstructed result and the ground-truth geometry.

IV-A Parameter Determination Using Canonical Geometric Targets

Before conducting comparative evaluations on complex targets, the two model parameters in MVLA-GR, namely the kernel-scale parameter κ\kappa and the migration parameter η\eta, are first determined using several canonical geometric targets. This step is introduced to obtain a unified parameter setting from simple geometries, rather than tuning the method specifically for the star-shaped target used in the subsequent comparisons.

IV-A1 Canonical Target Setup

Three canonical targets are considered for parameter determination: a circle, a square, and an equilateral triangle. Their geometric sizes are chosen as follows: the circle has a radius of 4 m, the square has a side length of 8 m, and the equilateral triangle has a height of 6 m. These three targets represent smooth curved boundaries, straight edges, and sharp-corner structures, respectively, which correspond to three fundamental scattering mechanisms relevant to MVLA-GR: continuous scattering migration along curved surfaces, piecewise-constant reflection along flat segments, and edge diffraction at sharp corners. Since most practical targets are composed of these three basic scattering elements, parameters determined on this basis are expected to generalize across diverse target geometries. For the circle and square, the maximum extent DD is 8 m. For the equilateral triangle, the maximum extent DD is taken as its side length, which is 434\sqrt{3} m.

In this experiment, each target is observed from 72 uniformly distributed views over the full 360360^{\circ} angular range, corresponding to an angular interval of 55^{\circ}. The SNR is fixed to 20 dB, and the remaining system parameters are kept identical across the three targets.

IV-A2 Selection Criterion for κ\kappa and η\eta

Refer to caption
(a)
Refer to caption
(b)
Refer to caption
(c)
Refer to caption
(d)
Figure 5: Parameter sweep results for κ\kappa and η\eta. (a) Averaged CD over the three canonical targets. (b) Circle. (c) Square. (d) Equilateral triangle.

As described in Section III, the parameter κ\kappa determines the radial kernel scale through σR=κΔR\sigma_{R}=\kappa\Delta R, while η\eta determines the migration scale through δmove=ηD\delta_{\mathrm{move}}=\eta D.

To determine a unified parameter pair, a grid search is performed over candidate values of κ\kappa and η\eta. For each parameter pair (κ,η)(\kappa,\eta), the proposed MVLA-GR method is applied to reconstruct the three canonical targets. During post-processing, the final binary reconstruction is obtained by the joint thresholding rule introduced in Section III, where the retained top p%p\% support values and the angular-support threshold TθT_{\theta} are both scanned over their admissible ranges. For each target, the minimum CD obtained over all tested (p,Tθ)(p,T_{\theta}) combinations is taken as the reconstruction error associated with the current parameter pair (κ,η)(\kappa,\eta).

Let CDq(κ,η)\mathrm{CD}_{q}(\kappa,\eta) denote this minimum CD obtained on the qq-th target, where

q{circle,square,triangle}.q\in\{\text{circle},\text{square},\text{triangle}\}.

The aggregated performance of a parameter pair is defined as the average CD over the three targets:

J(κ,η)=13q=13CDq(κ,η).J(\kappa,\eta)=\frac{1}{3}\sum_{q=1}^{3}\mathrm{CD}_{q}(\kappa,\eta). (44)

The final parameter pair is selected as

(κ,η)=argminκ,ηJ(κ,η),(\kappa^{\star},\eta^{\star})=\arg\min_{\kappa,\eta}J(\kappa,\eta), (45)

and is then fixed for all subsequent simulations and measurements in this section.

Refer to caption
Figure 6: Ray-tracing simulation setup for the star-shaped target.

Fig. 5 shows the parameter sweep results. Fig. 5(a) presents the averaged CD map defined in (44), while Fig. 5(b)–(d) show the corresponding CD maps for the circle, square, and triangle, respectively. The minimum of the averaged CD map is attained at κ=0.5\kappa^{\star}=0.5 and η=0.04\eta^{\star}=0.04.

IV-B Comparative Results on the Star-Shaped Target

Refer to caption
(a) Optimal, SNR=0 dB
Refer to caption
(b) Optimal, SNR=5 dB
Refer to caption
(c) Optimal, SNR=10 dB
Refer to caption
(d) Optimal, SNR=15 dB
Refer to caption
(e) Optimal, SNR=20 dB
Refer to caption
(f) Manual, SNR=0 dB
Refer to caption
(g) Manual, SNR=5 dB
Refer to caption
(h) Manual, SNR=10 dB
Refer to caption
(i) Manual, SNR=15 dB
Refer to caption
(j) Manual, SNR=20 dB
Figure 7: CD versus angular interval. Top row: results under optimal threshold selection. Bottom row: results under calibrated manual thresholds.

With the parameter pair (κ,η)=(0.5,0.04)(\kappa^{\star},\eta^{\star})=(0.5,0.04) fixed according to Section IV-A, we next evaluate the proposed MVLA-GR method on a more complex star-shaped target. Compared with the canonical targets used for parameter determination, the star-shaped target contains both sharp corners and concave structures, and therefore provides a more challenging geometry for reconstruction. The simulation configuration of the star-shaped target and the corresponding observation positions are illustrated in Fig. 6.

To benchmark the proposed method, we consider a standard incoherent backprojection (BP) baseline. For each voxel 𝐩n\mathbf{p}_{n}, the baseline accumulates PDP samples at the corresponding round-trip delay across all observation views, i.e.,

InBP=m=1MPm(τm,n),I_{n}^{\mathrm{BP}}=\sum_{m=1}^{M}P_{m}(\tau_{m,n}), (46)

where τm,n=2rm,n/c\tau_{m,n}=2r_{m,n}/c is the round-trip delay between voxel 𝐩n\mathbf{p}_{n} and the mm-th observation position. The final binary reconstruction is obtained by retaining the top-p%p\% voxels of {InBP}\{I_{n}^{\mathrm{BP}}\}. By contrast, for the proposed MVLA-GR method, the final binary reconstruction is obtained through the joint thresholding rule introduced in Section III.

Unless otherwise specified, all simulations in this subsection are conducted under the same carrier frequency, bandwidth, trajectory radius, and voxel resolution as those used in Section IV-A. The original ray-tracing data are generated on a dense 11^{\circ} angular grid. For a given angular interval Δθ\Delta\theta, the corresponding observation set is obtained by uniformly subsampling this dense grid. To reduce the dependence on the starting observation angle, the final reported CD is obtained by averaging the results over all valid angular offset patterns.

IV-B1 Comparison Under Optimal Threshold Selection

We first compare the proposed method with the baseline under the optimal threshold setting for each configuration. For the proposed MVLA-GR method, the post-processing parameters, namely the retained top p%p\% support values and the angular-support threshold TθT_{\theta}, are jointly scanned, and the minimum CD over all tested combinations is used as the final result. For the baseline method, the top-p%p\% threshold is scanned, and the minimum resulting CD is reported.

Fig. 7(a)–(e) shows the reconstruction error as a function of the angular interval under five representative SNR values. In general, the CD increases as the angular interval becomes larger, since fewer observation views are available and the geometric constraints become weaker. Under most tested conditions, the proposed MVLA-GR method achieves lower CD values than the baseline, and the advantage becomes more evident as the angular sampling becomes sparser.

The SNR sensitivity of the proposed method depends strongly on the angular sampling density. When the angular interval is very small, the proposed method is relatively insensitive to SNR, whereas the influence of SNR becomes increasingly visible as the angular interval grows. This indicates that the effects of view density and observation quality are strongly coupled. When the observation views are sufficiently dense, the multi-view geometric constraints are highly redundant, and the accumulated likelihood remains well concentrated around the true contour even if the front-end peak extraction is moderately degraded by noise. As a result, the reconstruction accuracy of the proposed method varies only slightly with SNR in the dense-view regime. In contrast, when the angular interval becomes larger, the number of available views decreases and the geometric constraints become weaker, so the quality of the extracted peaks plays a more critical role, making the reconstruction performance more sensitive to SNR.

Under extremely dense angular sampling, the baseline can become competitive with the proposed method. This is because, in this regime, the observation views are already highly redundant, and the standard backprojection can directly benefit from dense multi-view accumulation. By contrast, the proposed method adopts a unified parameter setting fixed in Section IV-A rather than being re-optimized for this dense-view regime, which may make this unified setting slightly suboptimal when the angular sampling is extremely dense. Nevertheless, as the angular interval increases, the proposed method exhibits a clearer advantage and maintains better reconstruction accuracy under more practically relevant sparse-view conditions.

The baseline appears less sensitive to SNR than the proposed method, because the two methods are affected by noise in different ways. The baseline directly accumulates PDP samples at the corresponding round-trip delays, so noise is partially averaged out during integration. By contrast, MVLA-GR relies on peak extraction before accumulation, so noise additionally affects the front-end observation generation through missed peaks, false peaks, and peak-location perturbations, leading to a clearer dependence on SNR especially when the angular sampling is not sufficiently dense.

Representative reconstruction results of the proposed MVLA-GR method are shown in Fig. 8. The first row illustrates the effect of angular interval when the SNR is fixed at 20 dB, while the second row shows the effect of SNR when the angular interval is fixed at 55^{\circ}. These visual results are consistent with the quantitative CD curves: denser angular sampling and higher SNR both lead to clearer and more complete reconstruction of the star-shaped contour.

Refer to caption
(a) Δθ=1\Delta\theta=1^{\circ}, SNR = 20 dB
Refer to caption
(b) Δθ=5\Delta\theta=5^{\circ}, SNR = 20 dB
Refer to caption
(c) Δθ=10\Delta\theta=10^{\circ}, SNR = 20 dB
Refer to caption
(d) Δθ=30\Delta\theta=30^{\circ}, SNR = 20 dB
Refer to caption
(e) Δθ=5\Delta\theta=5^{\circ}, SNR = 15 dB
Refer to caption
(f) Δθ=5\Delta\theta=5^{\circ}, SNR = 10 dB
Refer to caption
(g) Δθ=5\Delta\theta=5^{\circ}, SNR = 5 dB
Refer to caption
(h) Δθ=5\Delta\theta=5^{\circ}, SNR = 0 dB
Figure 8: Representative reconstruction results of the proposed MVLA-GR method under different sensing conditions.

IV-B2 Comparison Under Calibrated Manual Thresholds

Although the optimal-threshold results reveal the best achievable reconstruction performance under each configuration, such exhaustive tuning is generally unavailable in practice. We therefore further consider a calibrated manual-threshold setting. Specifically, for each SNR, the retained top p%p\% support ratio and the angular-support threshold TθT_{\theta} of the proposed MVLA-GR method are fixed according to the average of their optimal values over different angular intervals at that SNR. For the baseline, the top-p%p\% threshold is calibrated in the same way. The resulting thresholds are then reused for all angular intervals under the corresponding SNR without further tuning. The calibrated threshold values used for each SNR are summarized in Table I.

TABLE I: Calibrated Manual Thresholds
SNR (dB) MVLA-GR: pp MVLA-GR: TθT_{\theta} (deg) BP: pp
0 0.6090 30.4115 0.3000
5 0.6374 52.5926 0.2975
10 0.6814 69.9835 0.3008
15 0.6736 84.2016 0.2938
20 0.7008 96.8477 0.2942

Fig. 7(f)–(j) shows the corresponding CD curves under this calibrated manual-threshold setting. Compared with the optimal-threshold results in Fig. 7(a)–(e), both methods exhibit performance degradation, which is expected since the thresholds are no longer individually optimized for each configuration. However, the relative performance trend between the two methods remains informative. In particular, as the SNR increases, the proposed MVLA-GR method shows a clearer advantage over the baseline under most angular intervals. This indicates that, although the fixed thresholds introduce some mismatch, the proposed method is still able to exploit multi-view geometric consistency more effectively when the observation quality is sufficiently good.

Table I also shows that the calibrated thresholds of the proposed MVLA-GR method exhibit a clear SNR dependence: the retained top-p%p\% ratio increases with SNR because cleaner observations contain fewer artifact responses, and TθT_{\theta} increases monotonically with SNR because more complete and reliable peaks at higher SNR make true scattering voxels remain continuously supported over a wider angular range. By contrast, the calibrated top-p%p\% threshold of the baseline remains relatively stable across SNR, which is consistent with its low SNR sensitivity discussed above.

The SNR-dependent advantage of MVLA-GR can be explained from two perspectives. When the SNR is moderate or high, the extracted peaks are sufficiently reliable, and the geometric-consistency mechanism of MVLA-GR can still be effectively exploited even under fixed thresholds, so its higher reconstruction ceiling translates into a clear advantage over the baseline. When the SNR is low, peak extraction becomes more vulnerable to missed peaks, false peaks, and location perturbations, which makes MVLA-GR more sensitive to threshold mismatch. Overall, the comparison across Fig. 7 suggests that MVLA-GR not only achieves a higher best-case performance, but also preserves its advantage under calibrated manual thresholds when the SNR is moderate or high, indicating meaningful robustness in practical thresholding conditions.

IV-C Vehicle Measurement Validation

To further assess the proposed MVLA-GR framework under practical conditions, real measurement data collected from a vehicle target are considered in this subsection.

Refer to caption
Figure 9: Measurement setup in the anechoic chamber.

IV-C1 Measurement Setup

The measurement is conducted in an anechoic chamber to suppress external interference and unwanted environmental reflections. The measurement platform follows the setup reported in [25], where its calibration accuracy has been verified using a metal sphere as a reference target. A Volkswagen T-CROSS vehicle is used as the target, whose physical dimensions are approximately 4.22×1.76×1.60m4.22\times 1.76\times 1.60~\mathrm{m} (length ×\times width ×\times height).

The measurement system consists of a vector network analyzer (VNA), two horn antennas, and the associated RF front-end components. The two antennas are placed in close proximity to emulate a quasi-monostatic sensing configuration. During the experiment, the vehicle is mounted on a turntable, while the antennas remain fixed and point toward the target, as shown in Fig. 9. By rotating the turntable, the target is observed from multiple aspect angles, which is equivalent to collecting multi-view measurements around the target. The main measurement configuration is summarized in Table II.

TABLE II: Measurement Configuration
Parameters Value / Type
Carrier frequency fcf_{c} / Bandwidth BB 36 GHz / 3 GHz
Range resolution ΔR\Delta R 0.05 m
Number of views MM (Δθ\Delta\theta) 36 (1010^{\circ})
Target–antenna distance 12 m
Antenna height 0.9 m
Antenna type (Beamwidth) Horn antenna (14.3914.39^{\circ})
Target size 4.22×1.76×1.60m4.22\times 1.76\times 1.60~\mathrm{m}

IV-C2 Measured Data Processing and Reconstruction Configuration

For each observation angle, the VNA records the complex frequency response over the operating bandwidth. After calibration, the corresponding CIR is obtained by inverse Fourier transform, and the PDP is computed as the squared magnitude of the CIR. Since the measured data already contain practical noise and system imperfections, no additional synthetic noise is introduced in this experiment. Following the same phase-free reconstruction framework used in the simulation part, dominant multipath components are extracted from the measured PDPs and represented by their peak distances and peak powers.

The extracted peaks are then fed into the proposed MVLA-GR method. The model parameters are directly inherited from Section IV-A, namely κ=0.5\kappa^{\star}=0.5 and η=0.04\eta^{\star}=0.04. Since the antennas have a finite beamwidth rather than omnidirectional radiation patterns, each distance observation is only projected within the corresponding beam-covered angular sector. This beam constraint is incorporated into the likelihood accumulation process to suppress support outside the illuminated region.

The reconstruction region is discretized with a voxel size of 0.005 m. After likelihood accumulation, the final binary result is obtained using manually selected threshold parameters for visualization. Quantitative CD evaluation is omitted here, since accurate voxel-wise ground-truth occupancy labels are not available for the real vehicle target. The reconstruction result is therefore assessed visually.

Refer to caption
(a)
Refer to caption
(b)
Figure 10: Measured reconstruction results of the vehicle target obtained by the proposed MVLA-GR method. (a) Continuous support map before thresholding. (b) Final binary reconstruction result after thresholding.

IV-C3 Measurement Results

The measured reconstruction results of the vehicle target are shown in Fig. 10. Fig. 10(a) presents the continuous support map obtained after multi-view likelihood accumulation, while Fig. 10(b) shows the final binary reconstruction result after thresholding.

It can be seen that the continuous support map already reveals the main horizontal contour of the vehicle, although the support remains spatially diffuse before thresholding. After thresholding, the main contour becomes significantly clearer, and the elongated body structure as well as the overall aspect ratio of the vehicle are well preserved. These results indicate that the proposed likelihood accumulation mechanism remains effective under practical measurement conditions.

Compared with the idealized simulation results, the measured reconstruction is less regular and exhibits local discontinuities and some contour thickening. This is expected because the dominant measured scattering points originate from electromagnetically strong parts such as edges, corners, and metallic structures rather than from the outermost physical boundary, the 3D vehicle geometry is projected onto a horizontal imaging plane, and the measured data inevitably contain residual noise, calibration errors, and other non-ideal scattering effects. The recovered contour should therefore be interpreted as the effective projection of dominant scatterers on the imaging plane rather than the exact physical footprint of the vehicle.

Nevertheless, the reconstructed result still preserves the major geometric characteristics of the vehicle. This provides evidence that the proposed MVLA-GR framework is not limited to idealized ray-tracing data, but can also operate on real measured observations.

V Conclusion

In this paper, we addressed the problem of phase-free environmental geometry reconstruction in ISAC systems, where the absolute phase reference of each CIR measurement and its consistency across observation positions are challenging to maintain, due to both hardware imperfections such as oscillator drift and timing offsets, and physical-layer factors such as material- and viewpoint-dependent reflection phases. The proposed MVLA-GR method reconstructs target geometry from CIR measurements by exploiting the geometric consistency of multi-view propagation paths, using only delay and power information extracted from the power delay profile. A joint thresholding strategy combining response magnitude and angular support continuity is developed to suppress reconstruction artifacts and convert the continuous support map into a binary geometry estimate.

Simulation results on canonical geometric targets and a complex star-shaped target demonstrated that the proposed method achieves lower reconstruction error than the incoherent BP baseline in the majority of tested conditions. Real-world vehicle measurements at 36 GHz further verified the practical applicability of the proposed framework, where the main geometric characteristics of the vehicle were successfully recovered from phase-free CIR observations.

Future work will further investigate the integration of MVLA-GR with ISAC tracking modules, where target trajectory information from cooperative sensing can guide the segmentation of CIR observations and enable real-time geometry reconstruction of moving targets in 6G mobile networks.

References

  • [1] M. Born and E. Wolf (2013) Principles of optics: electromagnetic theory of propagation, interference and diffraction of light. Elsevier. Cited by: §II-A.
  • [2] Z. Chen and G. Huang (2016) A direct imaging method for electromagnetic scattering data without phase information. SIAM Journal on Imaging Sciences 9 (3), pp. 1273–1297. External Links: Document Cited by: §I.
  • [3] M. Chung, L. Liu, and O. Edfors (2022) Phase-noise compensation for OFDM systems exploiting coherence bandwidth: modeling, algorithms, and analysis. IEEE Transactions on Wireless Communications 21 (5), pp. 3040–3056. External Links: Document Cited by: §I.
  • [4] I. G. Cumming and F. H. Wong (2005) Digital processing of synthetic aperture radar data: algorithms and implementation. Artech House, Norwood, MA, USA. Cited by: §I.
  • [5] F. Dong, F. Liu, Y. Cui, W. Wang, K. Han, and Z. Wang (2023) Sensing as a service in 6G perceptive networks: a unified framework for ISAC resource allocation. IEEE Transactions on Wireless Communications 22 (5), pp. 3522–3536. External Links: Document Cited by: §I.
  • [6] X. Ji, X. Liu, and B. Zhang (2019) Phaseless inverse source scattering problem: phase retrieval, uniqueness and direct sampling methods. Journal of Computational Physics: X 1, pp. 100003. External Links: Document Cited by: §I.
  • [7] Y. Jiang, F. Gao, S. Jin, and T. J. Cui (2025) Electromagnetic property sensing in ISAC with multiple base stations: algorithm, pilot design, and performance analysis. IEEE Transactions on Wireless Communications 24 (4), pp. 3400–3416. External Links: Document Cited by: §I.
  • [8] Y. Jiang, F. Gao, and S. Jin (2024) Electromagnetic property sensing: a new paradigm of integrated sensing and communication. IEEE Transactions on Wireless Communications 23 (10), pp. 13471–13483. External Links: Document Cited by: §I.
  • [9] S. M. Kay (1993) Fundamentals of statistical signal processing, volume I: estimation theory. Prentice Hall. Cited by: §III-A2.
  • [10] E. Leitinger, F. Meyer, F. Hlawatsch, K. Witrisal, F. Tufvesson, and M. Z. Win (2019) A belief propagation algorithm for multipath-based SLAM. IEEE Transactions on Wireless Communications 18 (12), pp. 5613–5629. External Links: Document Cited by: §I.
  • [11] B. Lin, C. Zhao, F. Gao, G. Y. Li, J. Qian, and H. Wang (2024) Environment reconstruction based on multi-user selection and multi-modal fusion in ISAC. IEEE Transactions on Wireless Communications 23 (10), pp. 15083–15095. External Links: Document Cited by: §I.
  • [12] F. Liu, Y. Cui, C. Masouros, J. Xu, T. X. Han, Y. C. Eldar, and S. Buzzi (2022) Integrated sensing and communications: toward dual-functional wireless networks for 6G and beyond. IEEE Journal on Selected Areas in Communications 40 (6), pp. 1728–1767. External Links: Document Cited by: §I.
  • [13] G. Liu, R. Xi, Z. Han, L. Han, X. Zhang, L. Ma, Y. Wang, M. Lou, J. Jin, Q. Wang, and J. Wang (2024) Cooperative sensing for 6G mobile cellular networks: feasibility, performance, and field trial. IEEE Journal on Selected Areas in Communications 42 (10), pp. 2863–2876. External Links: Document Cited by: §I.
  • [14] Y. Luo, L. Yu, T. Wu, Y. Zhang, and J. Zhang (2026) A robust CSI-based scatterer geometric reconstruction method for 6G ISAC system. IEEE Wireless Communications Letters 15, pp. 2129–2133. External Links: Document Cited by: §I.
  • [15] N. J. Myers and R. W. Heath (2019) Message passing-based joint CFO and channel estimation in mmWave systems with one-bit ADCs. IEEE Transactions on Wireless Communications 18 (6), pp. 3064–3077. External Links: Document Cited by: §I.
  • [16] J. Ning, F. Han, and J. Zou (2024) A direct sampling-based deep learning approach for inverse medium scattering problems. Inverse Problems 40 (1), pp. 015005. Cited by: §I.
  • [17] M. Pastorino (2010) Microwave imaging. John Wiley & Sons. Cited by: §I.
  • [18] A. H. Quazi (1981) An overview on the time delay estimate in active and passive systems for target localization. IEEE Transactions on Acoustics, Speech, and Signal Processing 29 (3), pp. 527–533. Cited by: §III-A2.
  • [19] G. Sun, F. Zhang, and S. Pan (2020) Photonic-assisted high-resolution incoherent back projection synthetic aperture radar imaging. Optics Communications 466, pp. 125633. External Links: Document Cited by: §I.
  • [20] H. L. V. Trees (2001) Detection, estimation, and modulation theory, part I: detection, estimation, and linear modulation theory. John Wiley & Sons. Cited by: §III-A2.
  • [21] H. Wang, J. Zhang, G. Nie, L. Yu, Z. Yuan, T. Li, J. Wang, and G. Liu (2025) Digital twin channel for 6G: concepts, architectures and potential applications. IEEE Communications Magazine 63 (3), pp. 24–30. External Links: Document Cited by: §I.
  • [22] J. Wang, J. Zhang, Y. Zhang, Y. Sun, G. Nie, L. Shi, P. Zhang, and G. Liu (2025) Radio environment knowledge pool for 6G digital twin channel. IEEE Communications Magazine 63 (5), pp. 158–164. External Links: Document Cited by: §I.
  • [23] Y. Wang, M. Tao, and S. Sun (2024) Cramér-Rao bound analysis and beamforming design for integrated sensing and communication with extended targets. IEEE Transactions on Wireless Communications 23 (11), pp. 15987–16000. External Links: Document Cited by: §I.
  • [24] J. Zhang, L. Yu, S. Liu, Y. Cai, Y. Zhang, H. Xing, and T. Jiang (2026) Wireless environmental information theory: a new paradigm toward 6G online and proactive environment intelligence communication. Engineering 56, pp. 186–200. External Links: Document Cited by: §I.
  • [25] Y. Zhang, J. Zhang, H. Gong, X. Hu, J. Zhang, H. Xing, S. Luo, Y. Xiong, L. Yu, Z. Yuan, G. Liu, and T. Jiang (2026) A unified RCS modeling of typical targets for 3GPP ISAC channel standardization and experimental analysis. IEEE Journal on Selected Areas in Communications 44, pp. 702–716. External Links: Document Cited by: §I, §IV-C1.
  • [26] L. Zhou, C. Huang, and Y. Su (2012) A fast back-projection algorithm based on cross correlation for GPR imaging. IEEE Geoscience and Remote Sensing Letters 9 (2), pp. 228–232. External Links: Document Cited by: §I.