MVLA-GR: A Phase-Free Multipath-Based Geometry Reconstruction Method via Multi-View Likelihood Accumulation for ISAC
Abstract
Integrated sensing and communication (ISAC) enables wireless systems to reuse communication signals for environmental sensing, where reconstructing the geometry of surrounding objects is a representative sensing task. However, many conventional methods rely on coherent processing and require accurate phase information, which is often hard to guarantee in practical communication systems, particularly at high carrier frequencies. To address this problem, this paper proposes a Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) method based on channel impulse response (CIR) measurements, which uses only delay and power observations without requiring phase information. The method extracts dominant multipath components from each observation, and for each candidate spatial location, accumulates components across views whose propagation distances match the location as supporting evidence. A soft distance-matching kernel is introduced to tolerate range estimation errors and viewpoint-dependent scattering migration, and the received power of each component is used as a reliability weight. A joint thresholding strategy combining response magnitude and angular support continuity then converts the continuous support map into a binary geometry estimate. Ray-tracing simulations on canonical and complex targets, as well as real-world vehicle measurements at 36 GHz, demonstrate that MVLA-GR can effectively recover target geometry, providing a low-complexity phase-free solution for ISAC.
I Introduction
Integrated sensing and communication (ISAC) enables wireless systems to reuse communication signals for sensing the surrounding environment [13, 12, 5]. Among various sensing tasks such as target detection, parameter estimation, and contour characterization [23], reconstructing the geometry of surrounding objects, such as their contour, shape, and spatial layout, is of particular interest because it provides a structural description of the propagation environment rather than isolated target parameters. Such geometric information is a key enabler for emerging concepts such as the digital twin channel, where a synchronized geometric model of the environment is used to characterize and predict wireless propagation [22, 21]. It also benefits various communication tasks in 6G networks, including environment-aware beamforming, blockage prediction, and proactive beam management.
In ISAC systems, pilot signals embedded in communication waveforms are commonly used for channel estimation, from which the channel impulse response (CIR) can be obtained as a basic observation. A wideband CIR describes the multipath structure of the wireless channel in the delay domain, where each propagation delay corresponds to a propagation distance and each multipath component reflects a scattering or reflection event in the environment, whose strength further depends on the radar cross section of the underlying scatterer [25]. As a result, the CIR inherently carries geometric information about the surrounding scatterers [10], and its acquisition is naturally aligned with the communication-native operation of ISAC. This makes CIR a particularly suitable observation for environmental geometry reconstruction. Along this direction, the wireless environmental information theory has recently been proposed as a new paradigm for 6G environment intelligence communication [24].
A common way to reconstruct environmental geometry from wave-domain observations is to exploit the complex-valued wave response through coherent imaging. Representative examples include synthetic aperture radar (SAR) imaging [4], microwave tomography [17], and inverse-scattering reconstructions [8, 7, 14, 11], which jointly use the magnitude and phase of the measured field to reconstruct target shape or contrast. These methods can provide accurate reconstructions when reliable magnitude and phase measurements are available. However, in practical ISAC scenarios, although the magnitude of the CIR can be reliably extracted from communication pilots, the absolute phase reference of each CIR measurement and its consistency across observation positions remain challenging to maintain. Such challenges arise from two distinct sources. From the hardware perspective, oscillator drift, timing offsets, and the lack of cross-view phase synchronization at high carrier frequencies introduce per-observation phase offsets that are not easily predictable from one observation to the next [3, 15]. From the physical perspective, the reflection phase introduced by surface scattering depends on both material composition and incidence geometry through the complex-valued Fresnel reflection coefficient, and varies non-trivially with viewpoint, especially for materials and surface conditions that are not known a priori. These factors together limit the practical applicability of coherent imaging in ISAC scenarios. From this perspective, phase-free approaches that rely on the magnitude information in the CIR are intrinsically robust to such phase uncertainties and are therefore better suited for ISAC scenarios.
Phase-free reconstruction has been studied along two main directions. One direction is phaseless inverse scattering, which reconstructs object shape or material distribution from the magnitude or intensity of scattered fields by solving a wave-equation-based nonlinear inverse problem [2, 6, 16]. Such methods typically require a well-defined forward electromagnetic model and dense spatial or frequency sampling, and involve nonlinear iterative optimization, which makes them less suitable for lightweight CIR-based sensing in ISAC. The other direction is direct spatial accumulation, which projects measured delay-domain responses such as CIRs or PDPs back into the spatial domain and accumulates them over multiple observation positions [19, 26]. Incoherent backprojection is a representative example of this direction, and CIR-based heatmap construction follows a similar principle.
Among these two directions, direct spatial accumulation is more compatible with the lightweight, communication-native nature of ISAC, but it still has clear limitations when applied to CIR-based geometry reconstruction. First, direct spatial accumulation projects the entire PDP back into space, so that not only the dominant peaks but also sidelobes, noise floor, and other non-peak portions contribute to the spatial support. Since these non-peak portions of the PDP do not correspond to physical scattering events, their accumulation tends to blur the reconstruction and make the result sensitive to the detailed shape of the measured PDP. Second, direct spatial accumulation is essentially a heuristic projection scheme that does not explicitly model the underlying observation process, such as the statistical behavior of distance estimation, the viewpoint dependence of dominant scattering, or the angular continuity of true scattering structures. These unmodeled factors directly affect the reliability of the accumulated support, but cannot be addressed within the projection framework itself.
To address these gaps, a phase-free environmental geometry reconstruction method based on multi-view likelihood accumulation is proposed in this paper, with CIR-derived delay and power observations as the only input. The main contributions are summarized as follows.
-
•
A phase-free framework for environmental geometry reconstruction in ISAC is established, in which the problem is formulated as voxel-wise support estimation from CIR measurements, and only delay and power observations extracted from the CIR are used as the basic input.
-
•
Within this framework, the proposed method takes individual multipath components as the basic processing units rather than the entire PDP, and converts each detected component into a distance-based geometric constraint through multi-view likelihood accumulation. A soft distance-matching kernel is introduced to tolerate range estimation errors and viewpoint-dependent scattering migration, the received power of each component serves as a reliability weight during accumulation, and a joint thresholding strategy combining response magnitude and angular support continuity is further applied to suppress artifacts and convert the continuous support map into a binary geometry estimate.
-
•
The proposed method is validated through ray-tracing simulations on canonical and complex targets under various angular sampling intervals and SNR conditions, as well as real-world vehicle measurements at 36 GHz. Experimental results demonstrate that MVLA-GR can effectively recover target geometry and outperforms a standard incoherent backprojection baseline in the majority of tested conditions, confirming its practical applicability for ISAC scenarios.
The remainder of this paper is organized as follows. Section II introduces the system model and formulates the phase-free environmental geometry reconstruction problem from CIR measurements. Section III presents the proposed MVLA-GR method, including multi-view likelihood accumulation, parameter design, and binary reconstruction. Section IV provides simulation and measurement results to evaluate the proposed method.
II System Model and Problem Formulation
As illustrated in Fig. 1, we consider a multi-view geometry reconstruction scenario in which a single omnidirectional transceiver collects echo signals at positions relative to the target. As discussed in Section IV, this multi-view observation structure can be physically realized either by a moving transceiver in a static environment, or by a stationary transceiver observing a moving target.
II-A Voxel Representation
Under the Born approximation [1], the electromagnetic interaction between the incident field and the target is treated as a single-scattering process in which each scattering element contributes independently, without accounting for mutual coupling or multiple reflections between elements. This assumption allows the target response to be represented by a set of spatially distributed scattering coefficients. Accordingly, the region of interest is discretized into voxels with centers , , each associated with a nonnegative effective scattering intensity . The voxel intensity vector is defined as
| (1) |
Here, is interpreted as an effective scattering intensity used for geometry reconstruction. Specifically, indicates the absence of a scatterer, while indicates the presence of a scatterer. The support of , i.e., the set of voxels for which , encodes the geometric structure and contour of the target. Therefore, the objective of the considered geometry reconstruction problem is not to recover the exact value of every , but to determine which voxels satisfy .
II-B Electromagnetic Propagation and Received Signal Model
To establish the relationship between the voxel scattering coefficients and the received signal at each observation position, we characterize the round-trip propagation between the transceiver and each voxel. Starting from the free-space Green’s function
| (2) |
where is the wavenumber and is the carrier wavelength, the round-trip propagation between the -th transceiver position and voxel is modeled as the product of two such Green’s functions, leading to an effective complex channel gain
| (3) |
where is the geometric distance between voxel and the -th transceiver position. The phase term accounts for the round-trip propagation phase together with hardware-induced contributions such as oscillator drift and timing offsets, which are not directly available in practice. In this model, captures the round-trip propagation effect, while represents the effective scattering intensity associated with voxel .
A wideband pilot signal of duration and bandwidth is assumed to be transmitted at each observation position and reused for sensing. The specific waveform of is not restricted in this work; any standard wideband pilot whose autocorrelation provides a main-lobe range resolution on the order of is applicable, such as OFDM training sequences, linear frequency-modulated chirps, or pseudo-random sequences commonly used in communication systems. The corresponding round-trip delay from the -th transceiver position to voxel is
| (4) |
where denotes the speed of light. The received signal at the -th transceiver position consists of target echoes, reflections from the surrounding environment, and additive noise:
| (5) |
where denotes reflections from objects outside the imaging region and denotes the receiver noise.
Since the imaging region occupies a bounded spatial area, the target echoes correspond to a well-defined interval of round-trip delays determined by the geometry of the imaging region and the transceiver trajectory. Reflections from objects outside this region arrive at delays outside this interval and can be suppressed by range gating. After retaining only the delay interval corresponding to the imaging region, the received signal is modeled as
| (6) |
II-C CIR Measurement Model
The CIR is obtained by correlating the received signal with the known pilot over the observation interval of duration :
| (7) |
Substituting (6) into (7) yields
| (8) |
where
| (9) |
is the autocorrelation of the pilot waveform, serving as the effective point-spread function of the system, and is the filtered noise term. For a pilot signal occupying bandwidth , the range resolution is
| (10) |
which determines the minimum resolvable separation in the range dimension.
Discretizing the delay axis into taps and evaluating (8) at each tap, the CIR can be written in matrix-vector form as
| (11) |
where , is the noise vector, and is the observation matrix with -th element
| (12) |
In (12), the delay is geometrically determined by the known positions and via (4). However, fully specifying (12) also requires phase information that is reliable and mutually consistent across different observation positions. In the considered mobile single-transceiver system, this condition is difficult to guarantee due to the high phase sensitivity at high frequencies, hardware-dependent phase offsets, and the lack of precise cross-view phase synchronization. Coherent inversion based on is therefore not pursued in this work, and the subsequent development relies only on phase-insensitive features extracted from the CIR.
It is worth noting that, although the CIR is itself complex-valued, the per-observation phase factors arising from oscillator drift, sampling timing offsets, and transceiver localization errors appear as global multiplicative unit-modulus factors on , which are eliminated by the squared-magnitude operation . The squared-magnitude operation therefore yields a stable representation that does not depend on these uncertain phase quantities, while remaining sensitive to the underlying multipath structure through the magnitude information.
Accordingly, instead of relying on phase, this work adopts the power delay profile (PDP), defined as
| (13) |
as a phase-insensitive representation for subsequent geometry reconstruction.
II-D Phase-Free Observations
Based on the phase-insensitive representation in (13), this work extracts phase-free observations from the PDP at each observation position. Specifically, a standard peak detection procedure is applied to to identify detected multipath components, each characterized by its detected delay and received power. The detection is based on local maxima of , with a noise-relative magnitude threshold to suppress spurious detections, and a minimum inter-peak delay separation on the order of to avoid splitting a single main lobe into multiple peaks.
Due to the finite range resolution , multiple voxel contributions with sufficiently close round-trip delays may be mapped to the same detected multipath component in the PDP. Let denote the index set of voxels contributing to the -th detected component at the -th view. Its complex amplitude can be written as
| (14) |
where is the effective complex noise term at the corresponding delay bin after matched filtering. From each detected multipath component, two phase-free quantities are extracted:
| (15) | ||||
| (16) |
where is the propagation distance inferred from the detected delay , is the received power of the -th detected multipath component, and denotes a power-domain perturbation arising from receiver noise and its cross-coupling with the signal term. The complete phase-free observation set at the -th view is then defined as
| (17) |
where is the number of detected multipath components.
In the subsequent development, the distances provide the geometric constraints used for voxel localization, while the powers are used as reliability indicators for multi-view fusion.
II-E Problem Formulation
Since the geometry reconstruction objective is to identify the support of , we introduce the voxel occupancy variable
| (18) |
and define the binary occupancy vector
| (19) |
The geometry reconstruction problem is therefore formulated as the estimation of the voxel occupancy pattern from the multi-view phase-free observations :
| (20) |
Solving (20) exactly requires the likelihood , which in turn requires an explicit model relating the observed powers to the underlying voxel occupancies. However, as seen from (16), the received power depends on the unknown phases embedded in . Even if the occupied voxels were known, the observed powers could not be predicted explicitly without these phases. Consequently, the exact likelihood cannot be written in closed form, and the posterior in (20) is not directly tractable.
Nevertheless, the phase-free observations still contain useful information for support identification. In particular, the distance measurements provide geometric constraints on the possible locations of scatterers, while the received powers reflect the relative reliability of the detected multipath components. This suggests that the voxel occupancy can be approximated from multi-view geometric consistency, with received power serving as a reliability weight. The proposed MVLA-GR method in Section III is developed along this line by constructing a tractable approximation of the posterior probability of voxel occupancy from the observation sets .
III Multi-View Likelihood Accumulation Method for Geometry Reconstruction
This section presents the proposed Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) method. As discussed at the end of Section II, the posterior in (20) is not directly tractable, but the phase-free observations still carry useful geometric information. The proposed method exploits this information by approximating the posterior probability of voxel occupancy through two phase-free quantities that the CIR observations provide reliably, namely the propagation distances and the received powers . The distances are used to test whether each voxel is geometrically consistent with the observed multipath components across views, and the powers are used to weight the contribution of each component according to its reliability. By accumulating these weighted geometric matches over all observation views, a voxel-wise support score is obtained as a tractable surrogate for the posterior probability of voxel occupancy. The support map is then converted into the final binary reconstruction through a joint thresholding step that combines response magnitude and angular support continuity.
These steps are summarized in the processing pipeline shown in Fig. 2. From left to right, multipath components are first extracted from the PDP at each observation position to form the phase-free observations. The region of interest is then discretized into a voxel grid, and the multi-view geometric consistency between each voxel and the extracted distance–power observations is evaluated and accumulated into a support map. Finally, the joint thresholding step is applied to convert the continuous support map into a binary geometry estimate. The remainder of this section describes each of these components in detail.
III-A Multi-View Likelihood Accumulation
III-A1 Hard-Decision Occupancy Support
For voxel and observation view , consider the event that at least one detected propagation distance is geometrically consistent with the voxel distance, i.e.,
| (21) |
where is the -th detected propagation distance at view , is the geometric distance between voxel and the -th transceiver position, and is a distance tolerance related to the range resolution.
If voxel corresponds to a true scattering location, the event is more likely to occur. In contrast, for an unoccupied voxel, this event can only occur accidentally with a much smaller probability. Based on this observation, we define the hard-decision support indicator
| (22) |
where denotes the indicator function.
Treating the observation views as independent binary trials, the fraction of views in which the event occurs provides a coarse approximation to the posterior probability of voxel occupancy:
| (23) |
where denotes a surrogate posterior probability rather than an exact one. Under this hard-decision model, voxels supported by a larger fraction of observation views are more likely to be occupied.
III-A2 Soft Distance-Matching Kernel
The hard-decision rule in (23) is not sufficiently robust in practice. First, the extracted distances are affected by finite bandwidth, peak detection errors, and measurement noise. Second, under viewpoint variation, the dominant scattering point associated with the same physical structure may migrate along the target surface, especially in specular-dominated scenarios. As a result, a strict binary decision may reject physically meaningful observations that exhibit only a small distance deviation.
To address these issues, we replace the hard-decision indicator by a soft probabilistic kernel derived from the statistics of the distance estimation error.
For each detected multipath component, we assume that its corresponding peak in the power delay profile is sufficiently isolated from adjacent components, and that the measurement SNR is sufficiently high. Under these conditions, the extracted propagation distance can be modeled as
| (24) |
where is the extracted propagation distance, is the true propagation distance of the -th component, and is the associated distance estimation error. We assume that approximately follows a zero-mean Gaussian distribution, i.e., . This assumption is motivated by the asymptotic normality of maximum-likelihood estimators under additive Gaussian noise [20, 9], and is commonly adopted as a tractable approximation in time-delay estimation problems [18]. In practical multipath scenarios where peaks are not fully isolated, this Gaussian model is still used here as an approximate error model for soft geometric matching.
Under this Gaussian error model, the conditional likelihood of observing distance given that voxel is occupied and located at geometric distance is proportional to
| (25) |
where is the standard deviation of the range estimation error. Defining the distance mismatch
| (26) |
this conditional likelihood motivates the following truncated Gaussian kernel as the soft replacement for the hard indicator :
| (27) |
where is a truncation radius beyond which the voxel is regarded as geometrically incompatible with the observation. The truncation ensures that clearly inconsistent distance observations do not contribute to the voxel support and also improves computational efficiency.
III-A3 Power-Based Weighting
Different detected multipath components do not provide equally reliable geometric evidence and should therefore not contribute equally to voxel occupancy support. The received power is adopted as a reliability weight for two main reasons.
First, under matched filtering, the delay estimation error decreases with increasing SNR. Components with higher received power generally correspond to higher effective SNR, and their extracted distances are therefore more reliable. Weighting by power naturally suppresses weak and noisy components whose geometric constraints are less trustworthy.
Second, the observed power also reflects the strength of the corresponding propagation path. Stronger detected components usually carry more informative evidence about the underlying scattering structure, while weaker components are more likely to be unstable or noise-contaminated. Therefore, assigning larger weights to stronger components helps the multi-view accumulation focus on more credible geometric evidence.
Accordingly, to preserve the relative reliability differences among components within the same view while preventing any single view from dominating the overall fusion, the weight for each detected component is defined through view-wise normalization:
| (28) |
The view-wise normalization in (28) implicitly assumes that the observation views are of comparable quality, which is typically valid in the considered multi-view sensing settings, where all observations are collected by the same transceiver in a controlled measurement campaign and therefore share similar noise levels and detection conditions. Under this assumption, the relative powers within each view directly reflect the relative reliability of the detected components.
III-A4 Accumulation and Posterior Approximation
With the soft kernel and power-based weights defined above, the contribution from the -th detected component at view to voxel is
| (29) |
Accumulating all detected components within view , we define the total support of voxel under the -th observation as
| (30) |
where is the number of detected multipath components at view .
The quantity can be interpreted as a surrogate of the posterior probability of voxel occupancy under the -th observation, i.e.,
| (31) |
where denotes a surrogate posterior probability rather than an exact one. In this sense, is the soft generalization of the hard indicator in (22).
To aggregate evidence across all observation views, we define the normalized multi-view support score of voxel as
| (32) |
Accordingly, can be interpreted as a tractable approximation to the posterior probability of occupancy for voxel under the full multi-view observations, i.e.,
| (33) |
A voxel with a larger is supported by more observation views and more reliable multipath components, and is therefore more likely to be occupied. The collection forms the reconstructed occupancy map used in all subsequent processing.
The computational complexity of the proposed accumulation process is dominated by the evaluation of voxel-wise contributions for all detected multipath components across all observation views, resulting in
| (34) |
where is the number of voxels, is the number of detected multipath components at view , and denotes the average number of components per view.
III-B Parameter Selection and Binary Reconstruction
The accumulation process in Section III-A yields a continuous voxel-wise support score , which serves as a surrogate of the posterior probability of voxel occupancy. To obtain a reliable binary reconstruction from this continuous support map, two issues must be further addressed. First, the kernel parameters should be chosen such that valid supports can accumulate across views while maintaining sufficient radial discrimination. Second, accidental overlaps among supports from unrelated scattering structures must be suppressed. These two aspects are discussed in the following subsections.
III-B1 Parameter (Radial Kernel Scale)
In (27), the parameter controls how fast the soft support decays as the distance mismatch increases. As established above, should reflect the effective uncertainty of the extracted propagation distance. Although this uncertainty in general depends on multiple factors, including receiver noise, the dominant scale in multipath-rich scenarios is set by the system range resolution , which determines the main-lobe width of the pulse shaping function and thereby the achievable precision of peak detection. Accordingly, is parameterized as
| (35) |
where is a dimensionless proportionality constant controlling the tolerance of the radial kernel. A smaller yields sharper radial discrimination but weaker robustness to range errors, whereas a larger increases tolerance at the cost of reduced spatial selectivity. In practice, should be chosen according to the trade-off between range uncertainty and discrimination ability, and is determined in the experimental section.
III-B2 Parameter (Truncation Radius)
The parameter in (27) determines the effective support radius of the radial kernel. Under a purely range-limited model, would naturally be chosen on the same order as . In practical geometry reconstruction scenarios, however, the dominant scattering point associated with the same physical reflecting region may change with the observation viewpoint, especially under specular-dominated scattering conditions. As illustrated in Fig. 3, when the transceiver moves from to , the dominant scattering response may shift from one point on the target surface to another nearby point. Although these responses still originate from the same physical reflecting region, their effective propagation distances may differ due to viewpoint-dependent migration of the dominant scattering location.
If is chosen solely according to the range-resolution scale, valid supports associated with the same physical structure may fail to overlap across different views, weakening the subsequent likelihood accumulation. To preserve cross-view overlap among supports corresponding to the same physical scattering region, the truncation radius should also account for viewpoint-induced scattering migration.
To this end, we introduce a migration scale
| (36) |
where denotes the maximum spatial extent of the target and is a small dimensionless coefficient. The parameter characterizes the typical range variation caused by viewpoint-dependent movement of dominant scattering locations. Its value depends on target geometry, surface properties, and angular sampling density, and should be selected to provide sufficient overlap among valid supports across views without excessively enlarging the kernel support.
Since the effective support radius must cover both the range-uncertainty scale and the migration scale, the truncation radius is chosen as
| (37) |
This ensures that the likelihood kernel remains compact enough to suppress clearly incompatible voxels, while still allowing valid supports from different observation views to overlap and accumulate on the same physical scattering structure.
III-B3 Joint Thresholding for Binary Reconstruction
The accumulated support map provides a continuous spatial representation of voxel occupancy support. However, before converting this support map into a binary reconstruction, spurious responses caused by accidental overlaps among unrelated scattering structures must be suppressed. Such responses appear as reconstruction artifacts, as illustrated in Fig. 4.
A true scattering voxel is expected to be supported not only by a large accumulated response, but also over a sufficiently wide continuous range of observation directions. By contrast, an artifact voxel is usually produced by accidental overlaps among unrelated observations and tends to be supported only within a limited or fragmented angular region. Therefore, in addition to the accumulated response magnitude, the continuity of voxel support across observation directions provides an important cue for artifact suppression.
To characterize this persistence, we first define a binary support indicator for voxel at view as
| (38) |
where holds if and only if at least one detected component at view satisfies , as defined in (27).
Let denote the observation angle of the -th transceiver position with respect to the target center, and assume that are ordered increasingly. For voxel , the support sequence generally forms several continuous supporting intervals along the angular axis. Denote these intervals by , . For each interval , its angular width is defined as
| (39) |
The angular support range of voxel is then defined as the maximum width among all continuous supporting intervals:
| (40) |
To exploit both the response magnitude and the angular persistence, we adopt a joint thresholding strategy. First, a magnitude threshold is applied to the accumulated voxel responses , and only voxels whose response values lie within the top of all voxels are retained. Second, an angular-support threshold is imposed, and only voxels satisfying
| (41) |
are preserved. The joint thresholding criterion is therefore written as
| (42) |
This criterion has a clear physical interpretation: a voxel is regarded as a reliable scattering location only if it exhibits both a high support score and a sufficiently wide continuous support range across observation directions. As a result, spurious voxels caused by accidental support overlap can be effectively suppressed, while voxels associated with true scattering structures are more likely to be retained.
The final binary reconstruction is then obtained by setting for retained voxels and otherwise, thereby yielding an estimate of the occupancy vector in (20).
IV Simulation and Measurement Results
Before presenting the experimental results, we briefly note that the proposed framework applies to two complementary ISAC scenarios that share the same multi-view CIR observation structure: (i) a mobile transceiver collecting CIRs along its trajectory to reconstruct static surroundings, supporting tasks such as environment-aware beam management and digital twin construction; and (ii) a static base station observing a moving target whose successive positions provide the multi-view observations, supporting tasks such as target classification and behavior recognition. The simulations in this section follow the first scenario with a fixed circular trajectory, while the vehicle measurement in Section IV-C corresponds to the second scenario, where the antenna is fixed and the vehicle rotates on a turntable to emulate the relative aspect-angle variation between a stationary base station and a moving target.
The proposed MVLA-GR method is evaluated as follows. First, the model parameters are determined using several canonical geometric targets, so that all subsequent evaluations are carried out under a unified parameter setting rather than target-specific tuning. Second, under this fixed parameter configuration, the proposed method is quantitatively compared with a baseline method on a more complex star-shaped target under different angular sampling intervals and signal-to-noise ratios (SNRs). Finally, real-world measurement results on a vehicle target are presented to demonstrate the practical applicability of the proposed framework.
All simulation results in this section are generated using ray-tracing data produced by Remcom Wireless InSite. The carrier frequency is set to 36 GHz, and the signal bandwidth is 3 GHz, corresponding to a range resolution of m.
Unless otherwise specified, the transceiver positions are distributed along a circular trajectory with a radius of 10 m around the target, and the reconstruction region is discretized into a uniform voxel grid. It should be emphasized that the proposed MVLA-GR method does not rely on a circular trajectory itself, but only on the knowledge of the transceiver positions at which the CIR measurements are collected. The circular observation pattern adopted here is introduced purely as a convenient and reproducible simulation configuration.
In all simulation experiments, additive noise is introduced in the PDP domain after the clean CIR data are converted into PDPs. Multipath components are then re-extracted from the noisy PDPs by thresholding and peak detection, and the resulting phase-free observations are used for geometry reconstruction. The reconstruction performance in this section is evaluated using the Chamfer distance (CD). Let denote the set of ground-truth occupied voxel centers, and let denote the set of reconstructed occupied voxel centers. The Chamfer distance is defined as
| (43) | ||||
The resulting CD is expressed in dB(m2), and a smaller CD indicates a closer geometric match between the reconstructed result and the ground-truth geometry.
IV-A Parameter Determination Using Canonical Geometric Targets
Before conducting comparative evaluations on complex targets, the two model parameters in MVLA-GR, namely the kernel-scale parameter and the migration parameter , are first determined using several canonical geometric targets. This step is introduced to obtain a unified parameter setting from simple geometries, rather than tuning the method specifically for the star-shaped target used in the subsequent comparisons.
IV-A1 Canonical Target Setup
Three canonical targets are considered for parameter determination: a circle, a square, and an equilateral triangle. Their geometric sizes are chosen as follows: the circle has a radius of 4 m, the square has a side length of 8 m, and the equilateral triangle has a height of 6 m. These three targets represent smooth curved boundaries, straight edges, and sharp-corner structures, respectively, which correspond to three fundamental scattering mechanisms relevant to MVLA-GR: continuous scattering migration along curved surfaces, piecewise-constant reflection along flat segments, and edge diffraction at sharp corners. Since most practical targets are composed of these three basic scattering elements, parameters determined on this basis are expected to generalize across diverse target geometries. For the circle and square, the maximum extent is 8 m. For the equilateral triangle, the maximum extent is taken as its side length, which is m.
In this experiment, each target is observed from 72 uniformly distributed views over the full angular range, corresponding to an angular interval of . The SNR is fixed to 20 dB, and the remaining system parameters are kept identical across the three targets.
IV-A2 Selection Criterion for and
As described in Section III, the parameter determines the radial kernel scale through , while determines the migration scale through .
To determine a unified parameter pair, a grid search is performed over candidate values of and . For each parameter pair , the proposed MVLA-GR method is applied to reconstruct the three canonical targets. During post-processing, the final binary reconstruction is obtained by the joint thresholding rule introduced in Section III, where the retained top support values and the angular-support threshold are both scanned over their admissible ranges. For each target, the minimum CD obtained over all tested combinations is taken as the reconstruction error associated with the current parameter pair .
Let denote this minimum CD obtained on the -th target, where
The aggregated performance of a parameter pair is defined as the average CD over the three targets:
| (44) |
The final parameter pair is selected as
| (45) |
and is then fixed for all subsequent simulations and measurements in this section.
IV-B Comparative Results on the Star-Shaped Target
With the parameter pair fixed according to Section IV-A, we next evaluate the proposed MVLA-GR method on a more complex star-shaped target. Compared with the canonical targets used for parameter determination, the star-shaped target contains both sharp corners and concave structures, and therefore provides a more challenging geometry for reconstruction. The simulation configuration of the star-shaped target and the corresponding observation positions are illustrated in Fig. 6.
To benchmark the proposed method, we consider a standard incoherent backprojection (BP) baseline. For each voxel , the baseline accumulates PDP samples at the corresponding round-trip delay across all observation views, i.e.,
| (46) |
where is the round-trip delay between voxel and the -th observation position. The final binary reconstruction is obtained by retaining the top- voxels of . By contrast, for the proposed MVLA-GR method, the final binary reconstruction is obtained through the joint thresholding rule introduced in Section III.
Unless otherwise specified, all simulations in this subsection are conducted under the same carrier frequency, bandwidth, trajectory radius, and voxel resolution as those used in Section IV-A. The original ray-tracing data are generated on a dense angular grid. For a given angular interval , the corresponding observation set is obtained by uniformly subsampling this dense grid. To reduce the dependence on the starting observation angle, the final reported CD is obtained by averaging the results over all valid angular offset patterns.
IV-B1 Comparison Under Optimal Threshold Selection
We first compare the proposed method with the baseline under the optimal threshold setting for each configuration. For the proposed MVLA-GR method, the post-processing parameters, namely the retained top support values and the angular-support threshold , are jointly scanned, and the minimum CD over all tested combinations is used as the final result. For the baseline method, the top- threshold is scanned, and the minimum resulting CD is reported.
Fig. 7(a)–(e) shows the reconstruction error as a function of the angular interval under five representative SNR values. In general, the CD increases as the angular interval becomes larger, since fewer observation views are available and the geometric constraints become weaker. Under most tested conditions, the proposed MVLA-GR method achieves lower CD values than the baseline, and the advantage becomes more evident as the angular sampling becomes sparser.
The SNR sensitivity of the proposed method depends strongly on the angular sampling density. When the angular interval is very small, the proposed method is relatively insensitive to SNR, whereas the influence of SNR becomes increasingly visible as the angular interval grows. This indicates that the effects of view density and observation quality are strongly coupled. When the observation views are sufficiently dense, the multi-view geometric constraints are highly redundant, and the accumulated likelihood remains well concentrated around the true contour even if the front-end peak extraction is moderately degraded by noise. As a result, the reconstruction accuracy of the proposed method varies only slightly with SNR in the dense-view regime. In contrast, when the angular interval becomes larger, the number of available views decreases and the geometric constraints become weaker, so the quality of the extracted peaks plays a more critical role, making the reconstruction performance more sensitive to SNR.
Under extremely dense angular sampling, the baseline can become competitive with the proposed method. This is because, in this regime, the observation views are already highly redundant, and the standard backprojection can directly benefit from dense multi-view accumulation. By contrast, the proposed method adopts a unified parameter setting fixed in Section IV-A rather than being re-optimized for this dense-view regime, which may make this unified setting slightly suboptimal when the angular sampling is extremely dense. Nevertheless, as the angular interval increases, the proposed method exhibits a clearer advantage and maintains better reconstruction accuracy under more practically relevant sparse-view conditions.
The baseline appears less sensitive to SNR than the proposed method, because the two methods are affected by noise in different ways. The baseline directly accumulates PDP samples at the corresponding round-trip delays, so noise is partially averaged out during integration. By contrast, MVLA-GR relies on peak extraction before accumulation, so noise additionally affects the front-end observation generation through missed peaks, false peaks, and peak-location perturbations, leading to a clearer dependence on SNR especially when the angular sampling is not sufficiently dense.
Representative reconstruction results of the proposed MVLA-GR method are shown in Fig. 8. The first row illustrates the effect of angular interval when the SNR is fixed at 20 dB, while the second row shows the effect of SNR when the angular interval is fixed at . These visual results are consistent with the quantitative CD curves: denser angular sampling and higher SNR both lead to clearer and more complete reconstruction of the star-shaped contour.
IV-B2 Comparison Under Calibrated Manual Thresholds
Although the optimal-threshold results reveal the best achievable reconstruction performance under each configuration, such exhaustive tuning is generally unavailable in practice. We therefore further consider a calibrated manual-threshold setting. Specifically, for each SNR, the retained top support ratio and the angular-support threshold of the proposed MVLA-GR method are fixed according to the average of their optimal values over different angular intervals at that SNR. For the baseline, the top- threshold is calibrated in the same way. The resulting thresholds are then reused for all angular intervals under the corresponding SNR without further tuning. The calibrated threshold values used for each SNR are summarized in Table I.
| SNR (dB) | MVLA-GR: | MVLA-GR: (deg) | BP: |
|---|---|---|---|
| 0 | 0.6090 | 30.4115 | 0.3000 |
| 5 | 0.6374 | 52.5926 | 0.2975 |
| 10 | 0.6814 | 69.9835 | 0.3008 |
| 15 | 0.6736 | 84.2016 | 0.2938 |
| 20 | 0.7008 | 96.8477 | 0.2942 |
Fig. 7(f)–(j) shows the corresponding CD curves under this calibrated manual-threshold setting. Compared with the optimal-threshold results in Fig. 7(a)–(e), both methods exhibit performance degradation, which is expected since the thresholds are no longer individually optimized for each configuration. However, the relative performance trend between the two methods remains informative. In particular, as the SNR increases, the proposed MVLA-GR method shows a clearer advantage over the baseline under most angular intervals. This indicates that, although the fixed thresholds introduce some mismatch, the proposed method is still able to exploit multi-view geometric consistency more effectively when the observation quality is sufficiently good.
Table I also shows that the calibrated thresholds of the proposed MVLA-GR method exhibit a clear SNR dependence: the retained top- ratio increases with SNR because cleaner observations contain fewer artifact responses, and increases monotonically with SNR because more complete and reliable peaks at higher SNR make true scattering voxels remain continuously supported over a wider angular range. By contrast, the calibrated top- threshold of the baseline remains relatively stable across SNR, which is consistent with its low SNR sensitivity discussed above.
The SNR-dependent advantage of MVLA-GR can be explained from two perspectives. When the SNR is moderate or high, the extracted peaks are sufficiently reliable, and the geometric-consistency mechanism of MVLA-GR can still be effectively exploited even under fixed thresholds, so its higher reconstruction ceiling translates into a clear advantage over the baseline. When the SNR is low, peak extraction becomes more vulnerable to missed peaks, false peaks, and location perturbations, which makes MVLA-GR more sensitive to threshold mismatch. Overall, the comparison across Fig. 7 suggests that MVLA-GR not only achieves a higher best-case performance, but also preserves its advantage under calibrated manual thresholds when the SNR is moderate or high, indicating meaningful robustness in practical thresholding conditions.
IV-C Vehicle Measurement Validation
To further assess the proposed MVLA-GR framework under practical conditions, real measurement data collected from a vehicle target are considered in this subsection.
IV-C1 Measurement Setup
The measurement is conducted in an anechoic chamber to suppress external interference and unwanted environmental reflections. The measurement platform follows the setup reported in [25], where its calibration accuracy has been verified using a metal sphere as a reference target. A Volkswagen T-CROSS vehicle is used as the target, whose physical dimensions are approximately (length width height).
The measurement system consists of a vector network analyzer (VNA), two horn antennas, and the associated RF front-end components. The two antennas are placed in close proximity to emulate a quasi-monostatic sensing configuration. During the experiment, the vehicle is mounted on a turntable, while the antennas remain fixed and point toward the target, as shown in Fig. 9. By rotating the turntable, the target is observed from multiple aspect angles, which is equivalent to collecting multi-view measurements around the target. The main measurement configuration is summarized in Table II.
| Parameters | Value / Type |
|---|---|
| Carrier frequency / Bandwidth | 36 GHz / 3 GHz |
| Range resolution | 0.05 m |
| Number of views () | 36 () |
| Target–antenna distance | 12 m |
| Antenna height | 0.9 m |
| Antenna type (Beamwidth) | Horn antenna () |
| Target size |
IV-C2 Measured Data Processing and Reconstruction Configuration
For each observation angle, the VNA records the complex frequency response over the operating bandwidth. After calibration, the corresponding CIR is obtained by inverse Fourier transform, and the PDP is computed as the squared magnitude of the CIR. Since the measured data already contain practical noise and system imperfections, no additional synthetic noise is introduced in this experiment. Following the same phase-free reconstruction framework used in the simulation part, dominant multipath components are extracted from the measured PDPs and represented by their peak distances and peak powers.
The extracted peaks are then fed into the proposed MVLA-GR method. The model parameters are directly inherited from Section IV-A, namely and . Since the antennas have a finite beamwidth rather than omnidirectional radiation patterns, each distance observation is only projected within the corresponding beam-covered angular sector. This beam constraint is incorporated into the likelihood accumulation process to suppress support outside the illuminated region.
The reconstruction region is discretized with a voxel size of 0.005 m. After likelihood accumulation, the final binary result is obtained using manually selected threshold parameters for visualization. Quantitative CD evaluation is omitted here, since accurate voxel-wise ground-truth occupancy labels are not available for the real vehicle target. The reconstruction result is therefore assessed visually.
IV-C3 Measurement Results
The measured reconstruction results of the vehicle target are shown in Fig. 10. Fig. 10(a) presents the continuous support map obtained after multi-view likelihood accumulation, while Fig. 10(b) shows the final binary reconstruction result after thresholding.
It can be seen that the continuous support map already reveals the main horizontal contour of the vehicle, although the support remains spatially diffuse before thresholding. After thresholding, the main contour becomes significantly clearer, and the elongated body structure as well as the overall aspect ratio of the vehicle are well preserved. These results indicate that the proposed likelihood accumulation mechanism remains effective under practical measurement conditions.
Compared with the idealized simulation results, the measured reconstruction is less regular and exhibits local discontinuities and some contour thickening. This is expected because the dominant measured scattering points originate from electromagnetically strong parts such as edges, corners, and metallic structures rather than from the outermost physical boundary, the 3D vehicle geometry is projected onto a horizontal imaging plane, and the measured data inevitably contain residual noise, calibration errors, and other non-ideal scattering effects. The recovered contour should therefore be interpreted as the effective projection of dominant scatterers on the imaging plane rather than the exact physical footprint of the vehicle.
Nevertheless, the reconstructed result still preserves the major geometric characteristics of the vehicle. This provides evidence that the proposed MVLA-GR framework is not limited to idealized ray-tracing data, but can also operate on real measured observations.
V Conclusion
In this paper, we addressed the problem of phase-free environmental geometry reconstruction in ISAC systems, where the absolute phase reference of each CIR measurement and its consistency across observation positions are challenging to maintain, due to both hardware imperfections such as oscillator drift and timing offsets, and physical-layer factors such as material- and viewpoint-dependent reflection phases. The proposed MVLA-GR method reconstructs target geometry from CIR measurements by exploiting the geometric consistency of multi-view propagation paths, using only delay and power information extracted from the power delay profile. A joint thresholding strategy combining response magnitude and angular support continuity is developed to suppress reconstruction artifacts and convert the continuous support map into a binary geometry estimate.
Simulation results on canonical geometric targets and a complex star-shaped target demonstrated that the proposed method achieves lower reconstruction error than the incoherent BP baseline in the majority of tested conditions. Real-world vehicle measurements at 36 GHz further verified the practical applicability of the proposed framework, where the main geometric characteristics of the vehicle were successfully recovered from phase-free CIR observations.
Future work will further investigate the integration of MVLA-GR with ISAC tracking modules, where target trajectory information from cooperative sensing can guide the segmentation of CIR observations and enable real-time geometry reconstruction of moving targets in 6G mobile networks.
References
- [1] (2013) Principles of optics: electromagnetic theory of propagation, interference and diffraction of light. Elsevier. Cited by: §II-A.
- [2] (2016) A direct imaging method for electromagnetic scattering data without phase information. SIAM Journal on Imaging Sciences 9 (3), pp. 1273–1297. External Links: Document Cited by: §I.
- [3] (2022) Phase-noise compensation for OFDM systems exploiting coherence bandwidth: modeling, algorithms, and analysis. IEEE Transactions on Wireless Communications 21 (5), pp. 3040–3056. External Links: Document Cited by: §I.
- [4] (2005) Digital processing of synthetic aperture radar data: algorithms and implementation. Artech House, Norwood, MA, USA. Cited by: §I.
- [5] (2023) Sensing as a service in 6G perceptive networks: a unified framework for ISAC resource allocation. IEEE Transactions on Wireless Communications 22 (5), pp. 3522–3536. External Links: Document Cited by: §I.
- [6] (2019) Phaseless inverse source scattering problem: phase retrieval, uniqueness and direct sampling methods. Journal of Computational Physics: X 1, pp. 100003. External Links: Document Cited by: §I.
- [7] (2025) Electromagnetic property sensing in ISAC with multiple base stations: algorithm, pilot design, and performance analysis. IEEE Transactions on Wireless Communications 24 (4), pp. 3400–3416. External Links: Document Cited by: §I.
- [8] (2024) Electromagnetic property sensing: a new paradigm of integrated sensing and communication. IEEE Transactions on Wireless Communications 23 (10), pp. 13471–13483. External Links: Document Cited by: §I.
- [9] (1993) Fundamentals of statistical signal processing, volume I: estimation theory. Prentice Hall. Cited by: §III-A2.
- [10] (2019) A belief propagation algorithm for multipath-based SLAM. IEEE Transactions on Wireless Communications 18 (12), pp. 5613–5629. External Links: Document Cited by: §I.
- [11] (2024) Environment reconstruction based on multi-user selection and multi-modal fusion in ISAC. IEEE Transactions on Wireless Communications 23 (10), pp. 15083–15095. External Links: Document Cited by: §I.
- [12] (2022) Integrated sensing and communications: toward dual-functional wireless networks for 6G and beyond. IEEE Journal on Selected Areas in Communications 40 (6), pp. 1728–1767. External Links: Document Cited by: §I.
- [13] (2024) Cooperative sensing for 6G mobile cellular networks: feasibility, performance, and field trial. IEEE Journal on Selected Areas in Communications 42 (10), pp. 2863–2876. External Links: Document Cited by: §I.
- [14] (2026) A robust CSI-based scatterer geometric reconstruction method for 6G ISAC system. IEEE Wireless Communications Letters 15, pp. 2129–2133. External Links: Document Cited by: §I.
- [15] (2019) Message passing-based joint CFO and channel estimation in mmWave systems with one-bit ADCs. IEEE Transactions on Wireless Communications 18 (6), pp. 3064–3077. External Links: Document Cited by: §I.
- [16] (2024) A direct sampling-based deep learning approach for inverse medium scattering problems. Inverse Problems 40 (1), pp. 015005. Cited by: §I.
- [17] (2010) Microwave imaging. John Wiley & Sons. Cited by: §I.
- [18] (1981) An overview on the time delay estimate in active and passive systems for target localization. IEEE Transactions on Acoustics, Speech, and Signal Processing 29 (3), pp. 527–533. Cited by: §III-A2.
- [19] (2020) Photonic-assisted high-resolution incoherent back projection synthetic aperture radar imaging. Optics Communications 466, pp. 125633. External Links: Document Cited by: §I.
- [20] (2001) Detection, estimation, and modulation theory, part I: detection, estimation, and linear modulation theory. John Wiley & Sons. Cited by: §III-A2.
- [21] (2025) Digital twin channel for 6G: concepts, architectures and potential applications. IEEE Communications Magazine 63 (3), pp. 24–30. External Links: Document Cited by: §I.
- [22] (2025) Radio environment knowledge pool for 6G digital twin channel. IEEE Communications Magazine 63 (5), pp. 158–164. External Links: Document Cited by: §I.
- [23] (2024) Cramér-Rao bound analysis and beamforming design for integrated sensing and communication with extended targets. IEEE Transactions on Wireless Communications 23 (11), pp. 15987–16000. External Links: Document Cited by: §I.
- [24] (2026) Wireless environmental information theory: a new paradigm toward 6G online and proactive environment intelligence communication. Engineering 56, pp. 186–200. External Links: Document Cited by: §I.
- [25] (2026) A unified RCS modeling of typical targets for 3GPP ISAC channel standardization and experimental analysis. IEEE Journal on Selected Areas in Communications 44, pp. 702–716. External Links: Document Cited by: §I, §IV-C1.
- [26] (2012) A fast back-projection algorithm based on cross correlation for GPR imaging. IEEE Geoscience and Remote Sensing Letters 9 (2), pp. 228–232. External Links: Document Cited by: §I.