Calibrating the Digital Twin Channel: Statistics-Consistent Sim-to-Lab Adaptation for W-Band Industrial OFDM Links
Abstract
Digital twins (DTs) can reduce over-the-air validation cost in industrial wireless networks, but their utility depends on the fidelity of the underlying channel twin (CT). At W-band, site-specific ray tracing captures deterministic propagation geometry, yet its channel frequency responses (CFRs) do not reproduce the small-scale impairments and capture-to-capture variability observed in laboratory orthogonal frequency-division multiplexing (OFDM) measurements above 90 GHz. This paper proposes Statistics-Consistent Sim-to-Lab Adaptation (SC-SLA), a calibration framework that improves the fidelity of a 95 GHz Sionna ray-traced CT toward that of the testbed by aligning the mean power delay profile (PDP), the distribution of root-mean-square delay spread (), and per-subcarrier statistics at sampling bandwidth. SC-SLA uses a generative adversarial network (GAN)-inspired, cycle-consistent architecture with ResNet generators and batch-level channel-statistics losses on the PDP, sub-band PDP, moments and quantiles, and normalized mean-square error (NMSE) of the mean CFR-magnitude profile. The framework is non-adversarial and requires neither paired simulated/measured samples nor discriminators. On held-out 95 GHz data, SC-SLA reduces the Kolmogorov–Smirnov (KS) statistic from 0.86 to 0.050 relative to the impairment-augmented ray-traced input, and by () relative to the strongest of four supervised baselines (FCNN, CNN1D, BiLSTM, and UNet1D). Without retraining, the same checkpoint also generalizes to 92–94 GHz carriers, where it achieves the lowest PDP and CFR-magnitude errors among all baselines while reducing the -distribution mismatch relative to the uncalibrated twin.
I Introduction
Industry 4.0 wireless deployments increasingly rely on dense collections of sensors, controllers, and mobile robots that require high-rate, low-latency, and reliable private connectivity [3, 26]. Validating every physical-layer configuration on hardware is costly, particularly in fixed industrial cells where machinery, racks, reflectors, and access points occupy known locations. This setting has motivated the use of digital twins (DTs), measurement-synchronized virtual representations that can be queried before candidate configurations are transferred to the physical system [10, 44, 23]. Industrial radio environments are well suited to this paradigm because their propagation geometry is largely static over the time scale of link-level design.
In a wireless DT, the channel determines link-level quantities such as throughput, pilot overhead, cyclic-prefix margin, and beam coherence. The digital twin channel, referred to in this paper as the Channel Twin (CT), is therefore a foundation layer for wireless network-level twins [38, 43, 36]. Under the high-frequency approximation of Maxwell’s equations, a CT can model electromagnetic propagation through geometric rays and support geometry-aware analysis in complex three-dimensional environments [46]. Site-specific ray tracing (RT) is a practical CT engine for industrial Internet of Things (IIoT) deployments. Once the cell geometry and material properties are specified, solvers such as NVIDIA Sionna RT [15, 14] and Wireless InSite [29] can produce deterministic channel frequency responses (CFRs) and synthesize large datasets at low marginal cost. Recent advances further bring such solvers close to real-time operation [48, 5].
The usefulness of this closed loop depends on CT fidelity, which is difficult to guarantee in the frequency ranges targeted by 6G IIoT. The ITU-R IMT-2030 framework anticipates IMT operation in bands beyond those used for IMT-2020 [18], and 3GPP has extended new radio (NR) operation beyond the original frequency range 2 (FR2) ceiling to 71 GHz [2]. The W-band (75–110 GHz) follows this spectrum trajectory. In this paper, W-band refers to the 92–95 GHz carriers under study. These carriers lie near the upper limit of TR 38.901 [1], which is specified only up to 100 GHz and was parameterized primarily from sub-6 GHz and lower millimeter-wave (mmWave) measurements. Its small-scale channel assumptions therefore require measurement-based recalibration before they can support W-band physical-layer studies [42, 12].
Calibration is precisely where ray-traced twins begin to fall short. In our comparison between 95 GHz Sionna CFRs and in-lab orthogonal frequency-division multiplexing (OFDM) measurements, two systematic fidelity gaps emerged. (i) Early-tap energy underestimation. The solver reproduces specular and refractive interactions but omits diffuse near-field scattering, antenna-to-cable mismatch ripple, and analog-front-end group-delay distortion, all of which contribute strongly to the first few channel-impulse-response (CIR) taps in a real W-band setup. (ii) Narrow delay-spread distribution. Because the geometry is static, the twin produces an almost deterministic root-mean-square (RMS) delay spread, , whereas the physical testbed exhibits a much broader capture-to-capture spread, attributable to effects such as oscillator phase noise, sampling-clock jitter, and signal-to-noise ratio (SNR) drift in the W-band frequency converter. A twin with these discrepancies provides only limited fidelity, and any IIoT link-level study built on it, including pilot-density selection, cyclic-prefix design, and beam-coherence analysis, inherits the resulting bias.
Learning the missing propagation and hardware effects from data is a natural calibration strategy. However, industrial measurements impose an important constraint. Supervised calibration would require paired simulated and measured CFRs for the same channel realization, which is rarely available in field or laboratory captures. Simulated and measured acquisitions are typically generated independently, leaving no sample-wise correspondence between domains. The calibration problem must therefore be treated as an unpaired domain-alignment task. The learned correction should also preserve channel statistics that affect OFDM receiver design, including the power delay profile (PDP), , and per-subcarrier magnitude, rather than merely fitting individual waveform samples.
Motivated by these requirements, we propose Statistics-Consistent Sim-to-Lab Adaptation (SC-SLA), a calibration layer between a ray-traced CT and a physical W-band testbed. SC-SLA learns an unpaired, non-adversarial mapping from Sionna-generated CFRs to testbed-like CFRs by directly aligning receiver-relevant channel statistics. The calibrated twin retains the scalability of RT while reproducing the delay and frequency-domain statistics observed in the measured industrial link.
I-A Related Work
Outside explicit wireless-DT calibration, prior efforts to reduce the gap between modeled and measured wireless channels have followed three main directions. Measurement-based and site-specific models extend geometry-based stochastic modeling beyond mmWave by incorporating empirical path-loss, clustering, and delay statistics [28, 22]. Recent studies also infer channel statistics from environment geometry or co-located sensor data [35, 27]. These models capture propagation behavior. However, they do not calibrate the capture-specific impairments introduced by a particular measurement testbed.
Supervised networks trained on measured channels have been used for channel prediction and estimation, including vehicular channel state information (CSI) prediction [21] and massive multiple-input multiple-output (MIMO) channel prediction validated on field measurements [34]. In sim-to-real calibration, however, these methods require synchronized simulated and measured channel pairs, which are rarely available in field or industrial captures.
Generative models relax the need for paired samples by learning the underlying channel distribution. Examples include generative adversarial networks (GANs) trained on measured channel-sounder data [9], conditional GANs for air-to-ground channels [39], and physics-informed generative variants [7]. Nevertheless, these approaches typically optimize adversarial or parameter-domain objectives without explicitly preserving OFDM-relevant statistics such as PDP shape, , and frequency-domain magnitude structure.
A parallel line of work focuses explicitly on reducing the sim-to-real gap in wireless DTs. Following the taxonomy in [31], this gap can be addressed through (i) direct calibration of the DT using real measurements, (ii) uncertainty-aware modeling of the residual environment mismatch, and (iii) correction of the task-level AI training objective. SC-SLA belongs to the first category and, more specifically, performs post-RT channel-output calibration: the RT solver, reconstructed geometry, and nominal material properties remain fixed while a learned mapping aligns the simulated and measured CFR distributions.
Direct DT calibration can operate at different stages of the channel-generation process. Model-space approaches learn object-level electromagnetic properties and radio-wave interaction functions [19], or estimate RT material parameters while compensating for path-phase errors caused by geometric mismatch [32]. Geometry-space calibration instead corrects physical scene variables, such as transmitter and receiver locations, using measured and simulated PDPs [45]. Propagation-consistent environment twins can also embed a learnable scene-level electromagnetic field within differentiable RT and calibrate it using sparse position-labeled CSI [4].
Channel-output calibration avoids reconstructing the complete internal environment model. Haider et al. [11] first increase the fidelity of a site-specific RT model by selecting suitable propagation settings and material assignments, and then train a supervised U-Net to refine the resulting RT CIRs using measured over-the-air CIRs as labels. The refined channels are subsequently used for precoding and evaluated through end-to-end bit error rate (BER). Similarly, Luo et al. [25] use supervised learning to refine discrete Fourier transform (DFT)-domain channel weights generated by a low-complexity DT for CSI compression and feedback. Table I summarizes these calibration approaches and positions SC-SLA relative to them.
| Work | Calibration space | Calibrated component | Measurement or supervision | Calibration objective | Relation to SC-SLA |
|---|---|---|---|---|---|
| Jiang et al. [19] | Model | Object-level electromagnetic representation and interaction model | Scene-associated wireless observations | Learn electromagnetic properties and radio-wave interactions | Re-learns the environment model. SC-SLA keeps the RT model fixed and calibrates its CFR outputs. |
| Ruah et al. [32] | Model | RT material parameters and path-phase errors | Position-associated channel observations | Estimate material parameters while accounting for geometric phase errors | Calibrates inside a differentiable RT solver. SC-SLA does not backpropagate through the solver. |
| Ying et al. [45] | Geometry | Transmitter and receiver locations | Measured and simulated PDPs at corresponding sites | Correct location uncertainty and multipath mismatch | Corrects TX/RX placement. SC-SLA assumes fixed nominal node positions. |
| Ai et al. [4] | Environment | Scene-level electromagnetic property field | Sparse position-labeled CSI | Construct a propagation-consistent environment representation | Modifies the environment model and requires differentiable RT. SC-SLA applies a post-RT CFR mapping. |
| Haider et al. [11] | Output | RT configuration, material assignments, and generated CIR | Paired RT-generated and measured CIRs | Refine CIRs using a supervised U-Net for precoding and BER evaluation | Uses paired CIR supervision and task-level validation. SC-SLA uses unpaired data and ensemble statistics. |
| Luo et al. [25] | Output | DFT-domain channel output | Aligned low-/high-fidelity channel information or historical CSI | Refine DFT-domain channel weights by supervised learning | Uses an aligned, task-oriented representation. SC-SLA calibrates unpaired complex OFDM CFRs. |
| SC-SLA | Output | Post-RT complex CFR distribution | Independently acquired Sionna RT and W-band testbed captures | Align PDP, sub-band PDP, , and per-subcarrier magnitude statistics | Unpaired, non-adversarial calibration that leaves the RT scene, solver, and communication model unchanged. |
Adjacent sim-to-real studies adapt channel generators or communication models using limited measurements. Hu et al. [16] propose a transfer-learning transformer GAN pretrained on THz channels from a geometry-based stochastic channel model and fine-tuned using a smaller vector network analyzer measurement set. The model remains adversarial and operates in the multipath-parameter domain. Baytekin et al. [6] instead pretrain a neural receiver on a randomized 3GPP urban-microcell channel model and fine-tune it with real 5G NR physical uplink shared channel (PUSCH) measurements. Their objective is end-to-end block-error-rate performance rather than calibration of the underlying channel distribution.
Taken together, prior approaches leave three requirements unresolved for the present setting: (i) calibrating the generated channel without modifying or differentiating through the RT model, (ii) learning from independently acquired simulated and measured CFRs without sample-wise correspondence, and (iii) correcting the channel distribution itself rather than optimizing a task-specific communication objective. Fidelity calibration of a ray-traced CT for a single-antenna W-band OFDM link therefore remains largely unaddressed, particularly under the unpaired acquisition conditions considered in this work.
I-B Contributions
The proposed SC-SLA addresses this gap by providing an unpaired, GAN-inspired but non-adversarial calibration layer for the CT of a static industrial communication network. Instead of requiring sample-wise simulated/measured correspondence, SC-SLA directly aligns measured channel statistics that govern OFDM link behavior. The main contributions are summarized as follows.
-
•
An unpaired distributional CT-calibration formulation: We formulate sim-to-lab adaptation as the problem of bringing the statistical distribution of DT-generated CSI toward that of independently acquired real-world industrial CSI captures. The formulation explicitly aligns complementary delay and frequency-domain attributes, including the mean PDP, local sub-band delay structure, the distribution of , and the per-subcarrier magnitude profile, without imposing an arbitrary one-to-one correspondence between channel realizations.
-
•
A W-band OFDM testbed for CT calibration: We develop a commercial off-the-shelf (COTS) W-band measurement platform that combines USRP-B200 radios, WR-10 frequency-conversion modules, Faraday isolators, and pyramidal horn antennas. The testbed supports complex OFDM CFR acquisition over 92–95 GHz on the same subcarrier grid used by the ray-traced CT, thereby providing the real-world CSI required for sim-to-lab calibration and cross-frequency evaluation.
-
•
A CycleGAN-style non-adversarial channel translator: We develop a two-generator, cycle-consistent translation architecture inspired by GAN-based unpaired domain adaptation, but replace adversarial discriminators with receiver-relevant channel-statistics losses. The training objective combines losses on the truncated PDP and ten sub-band PDPs, moment and quantile losses, normalized mean-square error (NMSE) of the mean CFR-magnitude profile, and cycle/identity regularization.
-
•
An empirical study at 92–95 GHz: We evaluate SC-SLA using Sionna RT CFRs from the reconstructed laboratory CT and OFDM CFR measurements from the W-band testbed. At 95 GHz, SC-SLA reduces the Kolmogorov–Smirnov (KS) statistic by relative to the strongest of four supervised baselines while also achieving the lowest PDP and CFR-magnitude NMSE. Without retraining or carrier-specific tuning, the same checkpoint transfers to 92, 93, and 94 GHz and retains the lowest PDP and CFR-magnitude errors among the evaluated baselines, demonstrating calibration robustness across a 3 GHz carrier shift. To support reproducibility and further research on CT calibration, we publicly release the Sionna datasets, testbed datasets, trained checkpoints, and source code111Available at https://github.com/puloktarafder/thz-pcsla upon publication..
The remainder of this paper is organized as follows. Section II presents the signal model and the simulated and measured datasets. Section III details the SC-SLA generator and its channel-statistics losses. Section IV reports in-domain and cross-frequency results against the supervised baselines, together with inference scalability, sensitivity to the amount of measured training data, and an ablation of the loss terms and generator head that isolates the mechanisms behind the observed gains. Section V concludes the paper.
II System Model and Experimental Setup
This section specifies the OFDM signal model, the simulated and measured CFR datasets, and the common normalization protocol used by SC-SLA. The simulated ray-traced CT dataset is denoted by , and its measured testbed counterpart is denoted by . Both datasets use the same OFDM grid and normalization procedure before calibration.
II-A OFDM Signal and Channel Model
We consider a static indoor W-band link over 92–95 GHz with a dominant line-of-sight (LOS) component. To make simulation and measurement directly comparable, both domains use the same OFDM reference grid. Each OFDM symbol has subcarriers, of which are active. The active subcarriers occupy
which excludes the direct-current (DC) bin . This bin is left unused because practical radio-frequency (RF) front ends can exhibit DC offset, local-oscillator (LO) leakage, mixer feedthrough, and residual in-phase/quadrature (I/Q) imbalance around the carrier. Excluding the DC tone therefore avoids a hardware-sensitive subcarrier and yields a cleaner active CFR. The sampling bandwidth is , yielding a subcarrier spacing of and a useful OFDM symbol duration of . The active subcarriers therefore span approximately of the sampling bandwidth, with the remainder serving as guard band. A cyclic prefix of samples is prepended to each symbol. The active CFR on subcarrier is modeled as
| (1) |
where and denote the complex gain and excess delay of path , respectively, is the number of resolvable paths, and represents the residual receiver noise after least-squares (LS) channel estimation.
Let denote the full -point CFR obtained by placing the active samples on their corresponding fast Fourier transform (FFT) bins and assigning zeros to the DC and guard-band bins. The truncated CIR is obtained from the -point inverse FFT (IFFT) of by retaining the first taps,
| (2) |
Because the IFFT is defined over the full sampling bandwidth , adjacent CIR taps are spaced by . Retaining taps gives a uniformly sampled delay window. From this truncated CIR, the normalized PDP is computed as . The resulting discrete delay-energy distribution defines as
| (3) |
The finite delay support reduces the influence of late-delay noise-floor samples on the second-order delay moment. This choice is consistent with prior observations that noise and spurious PDP components at large delays can bias measured estimates if they are not excluded before computation [33]. The CFR, truncated CIR, PDP, and definitions above provide the common channel representation used to construct and .
II-B Physical Testbed (W-Band USRP Measurements)
The measured dataset is captured with the W-band OFDM testbed in Fig. 1, which operates over 92–95 GHz and uses COTS components. Separate USRP-B200 devices serve as transmitter (TX) and receiver (RX). The transmitter USRP generates the baseband OFDM waveform, while the receiver USRP records the received waveform with a master clock and a receiver-gain setting. The waveform is generated at an intermediate frequency (IF), , upconverted to W-band at the transmitter, and downconverted back to IF at the receiver using matched COTS Eravant [8] WR-10 up/down-conversion modules. The chain uses a -multiplied phase-locked oscillator and yields , where GHz. Retuning produces carriers from to GHz. A W-band Faraday isolator is inserted at each antenna port using COTS WR-10 waveguide components operating over 75–110 GHz, with isolation, insertion loss, and a twist. Both ends use co-polarized, boresight-aligned Eravant WR-10 pyramidal horn antennas, with a TX–RX separation of .
The transmitted reference is a flat-power OFDM symbol whose active subcarriers carry unit-magnitude symbols with uniformly random phases. For each carrier, the receiver records a burst. Offline synchronization cross-correlates the received burst with the known reference symbol. Candidate frame peaks are detected with a minimum peak distance of and a threshold equal to three times the mean correlation magnitude. Each valid detected reference symbol yields one channel capture. For every detected capture, the receiver computes the FFT of the synchronized OFDM symbol and estimates the active CFR by the LS method, , where is the received pilot on active subcarrier and is the transmitted reference symbol. The estimated active CFR is stored over the same active tones and in the same ordering used by the simulated dataset. A full -point CFR is also constructed by inserting zeros on the DC and guard-band bins, and its IFFT gives the corresponding full-length CIR.
An instantaneous SNR is estimated for each capture from the full CIR. The signal power is taken as the maximum tap power within the first taps, and the noise power is estimated as the mean tap power over the remaining taps. Captures with SNR below are rejected. After SNR filtering, we retain valid CFRs per carrier. At 95 GHz, the retained captures have a mean estimated SNR of and a standard deviation of . For 95 GHz training, the measured dataset is split into training samples and held-out samples, and held-out captures are subsampled for each reported metric.
The retained measurements exhibit substantial delay-domain variation even though the antenna positions and link geometry remain fixed. Across the held-out 95 GHz captures in Fig. 3, the mean is , with a standard deviation of . Its 10th and 90th percentiles are and , respectively. The normalized CIR and PDP statistics indicate that the variation redistributes energy among the retained taps rather than producing only a common gain change. These measurements define the reference distribution used to assess RT fidelity.
II-C Ray-Traced CT (Sionna RT)
The simulated CT dataset is generated with the Sionna RT path solver [14] on a calibrated indoor scene. The geometry corresponds to a single-room laboratory reconstructed in Blender, with walls and floor assigned ITU-R indoor material properties at W-band. The scene is exported as a Mitsuba XML file. Fig. 2 depicts the Blender reconstruction, the imported Sionna scene, and the traced paths. The TX and RX are placed at the nominal horn locations with a LOS distance of , matching the testbed. At both ends, we use a custom vertically polarized antenna element with a horn-like directive pattern, since Sionna RT does not provide a native horn model. The antenna power pattern is modeled as , where is the angular offset from boresight. The exponent solves the half-power condition at a half-angle, giving a full 3-dB beamwidth. The cosine-power model captures the dominant boresight directivity while omitting detailed sidelobe structure. We configure the path solver to support LOS, specular reflection, and refraction with a maximum interaction depth of and disable diffuse reflection. The resulting RT model represents the deterministic geometric component of the link, while PC-SLA learns the residual, comprising both stochastic capture variation and the deterministic response of the measurement chain. For each carrier, the static scene is solved once to obtain a deterministic base CFR over the same active OFDM tones used by the measurement waveform. The active-frequency vector follows the ordering , which matches the stored active-bin convention of the measured CFRs. The path delays are normalized so that the first arrival falls at tap 0, matching the synchronization convention used in the receiver processing.
A single RT solution yields one deterministic base CFR per carrier, whereas the physical testbed produces a distribution of captures because each acquisition is affected by noise, timing, and phase fluctuations. To emulate this behavior and match the captures retained per carrier in Section II-B, we generate simulated samples per carrier by passing the deterministic base CFR through a measurement-inspired impairment layer. Let denote the deterministic base CFR on active subcarrier , and let index the simulated capture. For each capture, a timing offset is drawn independently from a zero-mean Gaussian distribution with a standard deviation of . This offset is applied as the subcarrier-dependent phase rotation , which introduces a random linear phase slope across frequency. A common phase offset is independently drawn from a zero-mean Gaussian distribution with a standard deviation of and applied identically to all active subcarriers. The resulting impairment-augmented ray-traced CFR is
| (4) |
where is zero-mean complex Gaussian noise scaled from the mean base-CFR power to realize a per-capture SNR drawn from a Gaussian distribution with mean and standard deviation , limited to 8–60 dB. These nominal levels provide conservative capture variation without tuning to the held-out distribution. The timing spread equals one eighth of the sampling interval, and the common-phase spread is about . These phase perturbations leave the path geometry unchanged. The simulated SNR mean and standard deviation, and , approximate the corresponding measured values of and for the retained 95 GHz captures. The lower limit matches the measurement rejection threshold, while the upper limit suppresses extreme Gaussian draws. Further details on the typical Sionna RT workflow and site-specific dataset construction are available in our prior works [37, 38].
Fig. 3 summarizes the residual delay-domain gap after capture-side impairment augmentation. Across the held-out samples from each domain, the Sionna RT channels have a mean of and a standard deviation of , compared with and for the measured channels. The measured standard deviation is larger. Let and denote the th ordered samples from the simulated and measured domains, respectively. For equally weighted samples per domain, the empirical one-dimensional 1-Wasserstein distance is [41]
| (5) |
The two distributions yield a two-sample KS statistic of , defined in Eq. (14), and . These distributional discrepancies demonstrate that nominal capture-side impairments do not reproduce the measured delay-domain variability and motivate the learned SC-SLA calibration.
II-D Joint Normalization
The simulated and measured datasets pass through a common normalization procedure before training and evaluation. The goal is to remove absolute-power and receiver-gain differences while preserving the relative phase, frequency selectivity, and delay-domain structure needed for channel-statistics alignment.
First, each complex CFR sample is divided by its own RMS magnitude. This per-sample normalization removes capture-to-capture gain variation without forcing the two datasets to share a global power scale. Second, the normalized complex CFR is converted into a real-valued tensor by interleaving the real and imaginary components along the channel dimension, yielding samples of size and dataset tensors of shape and for the simulated and measured sets, respectively. Third, let denote the joint quantile level. The threshold is the -quantile of the pooled simulated and measured magnitudes. Both datasets are clipped at and scaled to the normalized support expected by the generator.
The normalization parameters are fixed after training and reused during evaluation. This keeps the input normalization identical across carriers, allowing the GHz checkpoint to be applied directly to the – GHz carriers without retuning. All SC-SLA losses and reported metrics are computed in this shared normalized representation.
III SC-SLA Framework
SC-SLA calibrates the ray-traced CT dataset toward the measured dataset . The proposed SC-SLA adopts a two-generator, cycle-consistent translation structure inspired by CycleGAN [47], but removes the adversarial discriminators and replaces them with batch-level channel-statistics losses. The result is a GAN-inspired, non-adversarial, unpaired channel translator. For a normalized simulated CFR sample , the forward generator maps the simulated sample to a real-world-like sample, . Applying this map to the full simulated dataset gives the calibrated dataset . For a normalized measured CFR sample , a reverse generator maps measured samples back to the ray-traced domain, . Applying the reverse map to all measured samples gives , which is used for cycle-consistency regularization. Both generators share the same architecture.
The calibration objective is distributional. SC-SLA does not force a simulated CFR to match a particular measured CFR, because the datasets are not synchronized and no sample-wise correspondence exists. Instead, the forward map is constrained to reproduce the channel statistics that govern OFDM receiver design, including the delay-domain energy distribution, , and frequency-domain magnitude response. This matches the data collection process, where simulated and measured datasets represent the same nominal static link and OFDM grid. However, their sample-to-sample variations arise independently. Fig. 4 depicts the data flow and the complete SC-SLA framework.
In an adversarial translator, discriminators learn implicit criteria for separating translated outputs from samples in the corresponding target domains, and the generators are updated through a minimax objective. SC-SLA instead minimizes explicit, differentiable discrepancies in the PDP, distribution, and CFR-magnitude profile because these receiver-relevant target statistics are available from the measured captures. Our non-adversarial formulation removes discriminator design and minimax tuning while making the physical role of each alignment term explicit. The reverse generator does not replace a discriminator. We retain to provide the inverse path required by cycle-consistency regularization. We use only to calibrate simulated CFRs during inference.
III-A Generator Architecture
The SC-SLA generator is a 1D ResNet [13] with an encoder–residual-trunk–decoder structure, illustrated in the bottom panel of Fig. 4. The topology adapts the translation generator of Johnson et al. [20], also used by CycleGAN [47], from 2D images to the 1D active CFR. The real and imaginary components of each normalized sample form two input channels over the active subcarriers. Since SC-SLA retains the CycleGAN translation structure and removes only its adversarial discriminators, the generator family is kept fixed and the layer dimensions are set by the OFDM numerology in Section II-A. A -tap input convolution lifts the two-channel input to feature channels over a spectral aperture of . By Eq. (1), a path at excess delay varies across frequency with period . Consequently, the longest delay retained in Eq. (2), , induces the fastest CFR ripple with a period. The input aperture covers at most of this cycle, so it extracts local CFR level, slope, and curvature while leaving multipath-induced ripples to deeper layers. Two stride-2 convolutions with -tap kernels then reduce the active-subcarrier length from to and , while increasing the channel dimension from to and . Two strided stages provide the deepest exact compression of the active grid. A third halving would require a fractional length of . The channel doubling keeps the activation volume constant ( values per CFR), trading spectral resolution for feature richness without reducing representational capacity [13].
After the strided stages, the compressed spectral grid contains one feature vector per four subcarriers, or approximately of spectrum. A residual trunk of six instance-normalized blocks refines the encoded CFR features at this resolution, matching the CycleGAN depth for inputs of comparable size [47]. Across the input convolution, strided stages, twelve trunk convolutions, mirrored decoder, and output head, the end-to-end receptive field reaches subcarriers, or approximately . This span exceeds the period of the fastest ripple, so each output subcarrier observes at least one complete cycle of the finest spectral structure generated by the truncated delay window. Without downsampling, the same depth would cover only subcarriers, or approximately . The receptive field is also close to the -subcarrier slices used by the sub-band PDP loss in Section III-B, so learned corrections act at the spectral granularity penalized by the loss. Instance normalization [40] standardizes each capture individually rather than across the mini-batch, which is appropriate for unpaired translation with capture-dependent receive power.
The decoder mirrors the encoder and restores the original -subcarrier length through two stride-2 transposed convolutions that reduce the channel dimension from back to , with explicit length matching for the odd upsampling. A final un-normalized -tap convolutional head predicts the residual correction on the amplitude scale of the input CFR. For a generic normalized input sample from either domain, the output is formed as . The residual skip biases the map toward the identity, so the network refines the ray-traced CFR rather than reconstructing it from scratch. The head confines the output to the normalized support shared by both datasets. Since the generator is fully feed-forward, each calibrated CFR requires only one forward pass, allowing SC-SLA to be inserted into a W-band link-level simulator without iterative calibration.
III-B Channel-Statistics Losses
The generator architecture determines the form of the CFR correction, while the loss function determines which channel properties it preserves or reproduces. We use two types of constraints. The first group consists of channel-statistics losses, which compare ensemble-level quantities computed over independently sampled simulated and measured mini-batches. These losses are not applied sample by sample, because the captures are unpaired and no measured realization corresponds to a particular simulated one. The quantities most relevant to OFDM receiver design, including PDP shape, , and the average subcarrier-magnitude profile, are distributional channel properties rather than deterministic labels for individual samples.
The second group consists of instantaneous regularizers, namely the cycle and identity losses. These losses act directly on each sample and prevent the learned mapping from becoming an unconstrained distribution-matching transform. A purely statistical objective is insensitive to permutations within a mini-batch and cannot alone ensure that each translated CFR remains tied to the geometry and structure of its input. The cycle and identity losses therefore complement the residual generator parameterization in Section III-A. We retain multiple statistical losses because each constrains a different receiver-relevant projection of the channel. The ablation results in Section IV-F demonstrate that no single term subsumes the others.
Let denote the mini-batch size and the number of retained CIR taps within the common delay window. We denote a simulated mini-batch by and an independently drawn measured mini-batch of the same size by . For a CFR mini-batch , let be the collection of per-sample normalized PDPs obtained using the CIR and PDP definitions in Section II. The losses below compare mini-batch statistics rather than enforcing arbitrary one-to-one matching between simulated and measured samples.
Truncated PDP loss: The first term matches the average truncated PDP of the measured batch to that of the translated simulated batch as
| (6) |
where denotes averaging over the mini-batch.
Sub-band PDP loss: To constrain local frequency-dependent delay behavior, the active subcarriers are divided into contiguous slices, each containing subcarriers, or approximately . For each slice , we compute a 150-point IFFT and form a sub-band PDP from the first taps. Since this sub-band transform has a tap spacing of , the sub-band transform provides a coarser delay grid than the global CIR. The resulting loss captures slower frequency-selective structure over a wider effective delay window and is given by
| (7) |
The factor averages the absolute PDP discrepancy over the sub-bands and retained taps.
moment loss: Let and denote the mini-batch mean and standard deviation of the per-sample values defined in Eq. (3). Since these are scalar batch statistics, we penalize their mismatch using absolute differences as
| (8) |
Here, the standard-deviation term receives twice the weight of the mean term to emphasize the dispersion error identified in Section I, where the simulated twin exhibits a narrower capture-to-capture distribution than the measured link.
quantile loss: The moment loss controls only the first two summary statistics. To better align the full empirical distribution, especially the upper tail, we also match the sorted values. Let denote the vector of per-sample values from in ascending order. We define
| (9) |
For equal-sized, equally weighted mini-batches, equals the empirical 1-Wasserstein distance between the measured and generated samples in Eq. (5). Fig. 3 reports this distance for the held-out pre-calibration data, whereas training evaluates it on independently drawn mini-batches.
CFR-magnitude NMSE loss: The delay-domain terms constrain the PDP and statistics. However, they do not fully determine the average spectral profile across active subcarriers. We therefore include a frequency-domain magnitude loss. For a mini-batch represented by real and imaginary channels, let denote the corresponding per-sample, per-subcarrier magnitude. The loss is
| (10) |
Cycle and identity losses: Cycle and identity regularization keep the unpaired translator close to a physically meaningful correction rather than an unconstrained remapping. We define
| (11) | ||||
| (12) |
The cycle-consistency loss requires a sample translated to the other domain and mapped back to reconstruct the original input. Cycle consistency therefore preserves sample-specific CFR structure and discourages many-to-one mappings. However, this loss constrains the compositions and , while the individual generators may still learn compensating transformations that cancel under composition. The identity loss reduces this ambiguity by regularizing each generator toward the identity mapping for inputs that already belong to its output domain. Identity regularization limits unnecessary amplitude and phase changes to CFRs that already exhibit the intended domain statistics. In the present static single-link setting, the bounded residual head provides much of the same input–output coupling even when the cycle and identity losses are removed, as quantified in Section IV-F. We nevertheless retain both losses as conservative sample-level regularizers.
The complete generator objective is
| (13) |
where is the vector of scalar loss weights listed in Table II. All SC-SLA loss weights are selected once using the 95 GHz training/validation split and kept fixed for the in-domain test, cross-frequency transfer, and ablation experiments. The weighting follows a coarse validation calibration in which the cycle and identity terms provide conservative sample-level regularization, while the delay-domain terms receive larger weights to address the PDP and mismatches discussed in Section I.
III-C Training Procedure
We implement the mappings as separate ResNet generators with independent parameter sets and . Consequently, we learn directly through the overall objective rather than obtain it by inverting . During each mini-batch, we evaluate both networks, and a common Adam optimizer [24] updates both parameter sets by minimizing . The reverse generator receives gradients through the cycle and identity terms, while the forward generator is additionally constrained by the channel-statistics losses. Training is therefore bidirectional, whereas sim-to-lab inference uses only .
At 95 GHz, simulated and measured captures are used for optimization, and captures are held-out from training. Both generators are trained for epochs with a learning rate of , linear decay after epochs, BF16 automatic mixed precision, and an gradient-norm clip of . Checkpoint selection, final evaluation, and the inference benchmark use FP32. Table II lists the system, RT, training, and model-selection parameters.
| Parameter | Value |
| System Parameters | |
| Carrier frequencies | – GHz |
| IF / LO frequency (GHz) | / () |
| FFT size / active subcarriers | / |
| Cyclic-prefix length | samples |
| Sampling bandwidth / | MHz / kHz |
| Occupied bandwidth | MHz |
| CIR taps | at ns ( ns window) |
| TX–RX separation | m (LOS) |
| Antennas | WR-10 pyramidal horns, co-polarized |
| Software-defined radio | USRP-B200, MHz master clock |
| Receiver gain / SNR gate | dB / dB |
| Capture burst per carrier | s |
| Captures per carrier (, ) | each |
| Training frequency | GHz |
| Testing frequency | – GHz |
| Ray-Traced CT Parameters | |
| Path solver | Sionna RT |
| Enabled interactions | LOS, specular, refraction |
| Max interaction depth | (diffuse reflection disabled) |
| Antenna power pattern | Cosine-power, , 3-dB beamwidth |
| Timing jitter (std) | ns |
| Common phase jitter (std) | rad |
| Per-capture SNR | dB mean, dB std |
| SC-SLA Hyperparameters | |
| Clipping quantile | |
| Sub-band count | |
| Residual blocks / channel widths | / –– |
| Optimizer | Adam (, ) |
| Learning rate / decay start | / epoch |
| Training epochs / batch size | / |
| Precision / gradient-norm clip | BF16 / |
| 95 GHz training samples | |
| Reserved samples | |
| Model-selection interval | every epochs |
| coefficients | |
| Evaluation subsample | captures per carrier |
Checkpoint selection uses three lower-is-better metrics evaluated on reserved captures. Let denote the output CFR set of a candidate model and the measured reference set. Both are evaluation sets rather than the training mini-batches and of Section III-B. Let and denote the empirical cumulative distribution functions (CDFs) of the per-sample values for and , respectively. Their two-sample KS statistic is
| (14) |
Here denotes the supremum, that is, the least upper bound of over all delay-spread values . Both empirical CDFs are right-continuous step functions with finitely many jumps. The difference therefore takes finitely many values, and the supremum is attained as a maximum over the pooled values of and . Accordingly, is the largest vertical gap between the two empirical CDFs and is bounded to . The two NMSE metrics are
| (15) | ||||
| (16) |
The metric has the same mathematical form as the training loss , but it is evaluated on the reserved or held-out sets and rather than on training mini-batches. In contrast, uses normalized squared error between ensemble-mean PDPs, whereas uses mean absolute error. The PDP metric therefore provides an in-objective consistency check without duplicating the training loss.
The three metrics form the scalar model-selection criterion
| (17) |
The fixed coefficients and weight the two NMSE terms relative to during checkpoint selection. They act only on the selection criterion and are distinct from the training-loss weights in Eq. (13). The coefficient compensates for the small numerical range of , whereas limits the contribution of the numerically larger . Both values are fixed once on the 95 GHz validation split and reused for every reported experiment. The statistic enters as a sup-norm on the CDF. No training term takes this form, so it scores the distribution targeted by and under a different functional and penalizes checkpoints that match the moments and quantiles while leaving a localized gap in the CDF.
Every epochs, the checkpoint is evaluated on a fixed subset of reserved captures using the criterion . The checkpoint with the lowest is retained. The reported 95 GHz results are therefore in-domain, model-selected results on non-training captures. The 92–94 GHz carriers are used only for cross-frequency evaluation, not for optimization, model selection, or carrier-specific tuning. Algorithm 1 summarizes the joint generator optimization and checkpoint selection.
We conduct all simulations on an Ubuntu 24.04 LTS workstation equipped with an AMD Ryzen 9 9950X (16 cores, 32 threads, 5.7 GHz) processor, 128 GB RAM, and an NVIDIA RTX 4000 Ada Generation GPU with 20 GB memory. We implement SC-SLA in Python 3.12 with PyTorch 2.8.
IV Experiments and Result Analysis
IV-A Baselines and Evaluation Metrics
We compare SC-SLA with four supervised neural generators that use the same normalization method and the same zero-initialized residual-delta form, . The baselines include a fully connected neural network (FCNN) with four -unit layers, LayerNorm, and dropout. The one-dimensional convolutional neural network (CNN1D) has six residual convolutional blocks and filters. The three-layer bidirectional long short-term memory network (BiLSTM) has hidden units. The depth- one-dimensional U-Net (UNet1D) [30] uses base channels.
The supervised baselines are constrained by the data collection process. Simulated and measured captures are triggered independently, so no physical one-to-one correspondence exists between a simulated CFR and a measured CFR. We therefore train the baselines using Huber regression [17] on index-aligned pairs, together with weak batch regularizers on CFR magnitude, sub-band magnitude, and statistics. We use this Huber-based configuration as the strongest paired-supervision setting supported by the data and as a direct comparison with the unpaired SC-SLA objective. All models are selected using the same composite validation criterion, with family-specific weights chosen to place the constituent terms on comparable numerical scales. The three lower-is-better metrics are defined in Section III-C. They comprise , , and . The KS statistic is not directly optimized by any training loss. The PDP metric compares the same ensemble-mean PDPs as but uses normalized squared error and therefore serves mainly as an in-objective consistency check. The CFR-magnitude metric measures preservation of the average frequency-domain magnitude profile.
IV-B In-Domain Results at 95 GHz
We first evaluate SC-SLA and the four supervised baselines on held-out captures at GHz. For context, the uncalibrated twin, denoted as “Sionna (identity)”, applies the impairment-augmented ray-traced output directly through for an input CFR sample . Fig. 3 shows that this input has a large -distribution mismatch, with . Thus, the impairment-augmented twin alone does not reproduce the measured delay-domain statistics.
The supervised baselines reduce the KS statistic to –, as shown in Fig. 5. However, CNN1D and BiLSTM yield CFR-magnitude NMSE values of and , respectively. This spectral distortion is consistent with the limitation of index-aligned supervision because the independently triggered simulated and measured captures are not physically paired. SC-SLA achieves the lowest error across all three metrics in Fig. 5. It reduces the KS statistic to , a reduction from the obtained by UNet1D, the best supervised baseline on this metric. Furthermore, SC-SLA attains a CFR-magnitude NMSE of and a PDP NMSE of . The latter is more than two orders of magnitude below every supervised baseline. Because PDP NMSE compares the same ensemble-mean PDPs as using a different error norm, it serves as an in-objective consistency check rather than a fully independent metric.
IV-C Cross-Frequency Generalization at 92–94 GHz
We next examine whether the calibration learned at GHz transfers to nearby W-band carriers. For this experiment, we apply the same GHz checkpoint directly to Sionna inputs at , , and GHz, without retraining, fine-tuning, or carrier-specific model selection. We then compare the calibrated outputs with measured captures at each carrier.
The PDP NMSE increases when the model is evaluated away from the training carrier, rising from the in-domain value of to the range. This degradation is expected because no – GHz data are used during training. Even under this shift, SC-SLA achieves the lowest PDP NMSE at every held-out carrier. Compared with the strongest supervised baseline at each frequency, SC-SLA improves the PDP NMSE by , , and at , , and GHz, respectively (Fig. 6). SC-SLA also achieves the lowest CFR-magnitude NMSE across all three held-out carriers (Fig. 7).
The results follow the same overall trend, although the margin varies by carrier. SC-SLA keeps the KS statistic below that of the uncalibrated Sionna input, with values of , , and at , , and GHz, respectively. The identity twin yields larger KS values of , , and at the same carriers. The only exception among the baselines occurs at GHz, where FCNN gives a slightly lower KS value of . However, FCNN’s lower 93 GHz KS is accompanied by worse PDP and CFR-magnitude errors and therefore does not represent the same overall calibration quality. Because the – GHz datasets are not used for training, model selection, or carrier-specific tuning, these results provide direct evidence of zero-shot cross-frequency generalization to three adjacent W-band carriers. They also constitute an unseen-carrier stress test. The test remains limited to the measured – GHz span and the present static LOS scene. Denser multipath, non-line-of-sight (NLOS) propagation, and geometry changes require corresponding RT and measured datasets.
Three complementary mechanisms support the observed stability of SC-SLA under carrier shift, and the ablation in Section IV-F isolates each of them. The quantile loss aligns the upper tail of the distribution, rather than only its mean and variance. The -bounded residual head keeps the translated CFR within the shared normalized channel range, limiting unstable extrapolation at unseen carriers. The sub-band PDP loss further encourages preservation of delay structure over local frequency regions, making the correction less dependent on the exact GHz training carrier.
IV-D Inference Scalability
Cross-frequency transfer avoids carrier-specific retraining. However, the translated CFRs must also be generated at the scale of an RT dataset. Table III therefore reports post-RT calibration time for subsets of the -sample 95 GHz Sionna dataset. The deployed contains million parameters and sustains approximately CFRs/s across the evaluated dataset sizes. It processes the complete set in s, corresponding to an amortized inference time of ms per CFR. Joint training incorporates two generators with million parameters in total, whereas inference retains only . The timing includes host-to-device transfer, one generator pass, and output return. The reported timing excludes RT generation, model loading, normalization, and metric computation.
| CFRs | Input (MB) | Parameters () | Runtime (s) | Throughput (CFRs/s) |
|---|---|---|---|---|
IV-E Sensitivity to Measured Training Data
Practical calibration also depends on how much measured CSI is available for training. We therefore retrain SC-SLA with , , and measured 95 GHz CFRs, corresponding to nested , , and subsets of the measured training set. All simulated training CFRs and the fixed -sample reserved split remain unchanged. Each variant is initialized independently and trained once with seed and the original 300-epoch schedule, and every epoch contains 150 optimizer updates, so all fractions receive the same number of gradient steps. The loss weights and checkpoint-selection criterion are unchanged, the normalization parameters remain those estimated from the complete 95 GHz dataset, and the row reuses the full-data checkpoint. This experiment therefore isolates the number of measured CFRs entering optimization rather than the total measurement-acquisition requirement.
Table IV reports the results. At the 95 GHz training carrier, the reduced-data models attain KS values of –, below the of the full-data model, and CFR-magnitude NMSE of –, within of the full-data . PDP NMSE is the most sensitive in-domain metric at – times the full-data value of , although even its worst case remains more than times below the uncalibrated error.
| (a) In-domain evaluation at 95 GHz | ||||
| Measured CFRs | Fraction (%) | KS | PDP NMSE | CFR magnitude NMSE |
| (b) Zero-shot mean at 92–94 GHz | ||||
| Measured CFRs | Fraction (%) | KS | PDP NMSE | CFR magnitude NMSE |
Cross-frequency behavior is less uniform. The mean KS at 92–94 GHz is , , and for the , , and models, against for the full-data model and for the uncalibrated twin. Mean PDP NMSE spans –, compared with , and mean CFR-magnitude NMSE stays within of the full-data . The model yields the largest mean KS but the lowest mean CFR-magnitude NMSE, so sensitivity to the measured-data fraction depends on which channel statistic is evaluated. Under the fourfold reduction, the model still holds in-domain KS below the full-data value and differs from it by in unseen-carrier mean KS. Each fraction is nevertheless represented by a single run, so these results do not establish a monotonic data-scaling relationship or across-seed uncertainty.
IV-F Loss Function Ablation Study
We retrain six SC-SLA variants to isolate the role of the main loss and architecture components. Each variant changes one ingredient by removing one channel-statistics loss, removing all four delay-domain losses jointly (only mag), removing cycle and identity regularization, or replacing the -bounded head with a linear residual head, . All variants use the same seed, training schedule, and checkpoint-selection rule as the full model. Because the study uses a single seed, we focus on clear multi-fold differences and avoid overinterpreting small KS changes below roughly .
The ablation results demonstrate that the delay and frequency-domain objectives are complementary. Removing the quantile loss increases the in-domain KS statistic by () and worsens the GHz KS by , while PDP and magnitude errors remain nearly unchanged. These multi-fold KS increases indicate that the quantile loss primarily controls the upper tail of the distribution. In contrast, the only mag variant gives the lowest CFR-magnitude NMSE in Fig. 8 (), but leaves the delay domain largely uncorrected, with KS . Removing the magnitude loss produces the opposite failure mode. The delay-domain terms remain active, but the spectral profile drifts and the CFR-magnitude NMSE rises to , worse than the uncalibrated input.
The architecture ablations clarify the source of sample-level input–output coupling. A linear residual head fits the training carrier reasonably well, with in-domain KS , but generalizes poorly. The GHz KS more than doubles from to , and the in-domain PDP NMSE increases by . Thus, the head mainly stabilizes the correction under carrier shift rather than improving the 95 GHz fit alone. Removing the cycle and identity losses does not cause collapse. On held-out inputs, the mean input–output correlation remains , compared with for the full model. These correlation values indicate that, in the present static single-link setting, the bounded residual form provides most of the input–output coupling. We retain cycle and identity regularization because it is conservative and may become more important when geometry varies across samples.
The sub-band PDP loss contributes mainly to cross-frequency robustness. Removing it slightly improves some in-domain metrics, but worsens the GHz KS statistic by () and also degrades KS at and GHz. Across Figs. 8 and 9, no ablated model is uniformly superior. Each variant improves one metric at the cost of another. The full configuration is therefore not the per-metric optimum, but it provides the most balanced operating point across , PDP shape, spectral magnitude, and carrier transfer.
V Conclusion
In this paper, we presented SC-SLA, a non-adversarial, cycle-consistent framework that calibrates a ray-traced W-band CT against unpaired testbed measurements by aligning receiver-relevant OFDM channel statistics. These statistics include the PDP, sub-band PDP, distribution, and per-subcarrier CFR magnitude. On held-out captures at GHz, SC-SLA reduced the KS statistic of the uncalibrated twin from to , improved on the strongest of four supervised baselines by , and achieved the lowest PDP and CFR-magnitude NMSE among all evaluated models. The same checkpoint transferred zero-shot to – GHz, providing evidence of cross-frequency calibration over the measured GHz span using training measurements from 95 GHz only. A fourfold reduction of the measured training set to CFRs preserved calibration quality under a fixed seed, yielding KS at 95 GHz and mean KS across the three unseen carriers. The ablation study demonstrated that the quantile loss governs the upper tail of the distribution, the delay- and frequency-domain objectives are jointly necessary, and the -bounded residual head stabilizes the correction under carrier shift. The deployed generator calibrates CFRs in s, corresponding to approximately CFRs/s. Future work could repeat the sensitivity and ablation studies across seeds, quantify sensitivity to the impairment parameters, extend the captures beyond the present sampling bandwidth, and address NLOS, geometry-varying, and multi-antenna deployments.
References
- [1] (2022-03) Study on channel model for frequencies from 0.5 to 100 GHz. Tech. Rep. Technical Report TR 38.901, V17.0.0, 3GPP. Cited by: §I.
- [2] (2021-03) Study on supporting NR from 52.6 GHz to 71 GHz. Tech. Rep. Technical Report TR 38.808, V17.0.0, 3GPP. Cited by: §I.
- [3] (2019) A survey on information and communication technologies for industry 4.0: state-of-the-art, taxonomies, perspectives, and challenges. IEEE Communications Surveys & Tutorials 21 (4), pp. 3467–3501. Cited by: §I.
- [4] (2026) Propagation-consistent wireless environment digital twin construction under sparse measurements. arXiv preprint arXiv:2605.22361. Cited by: §I-A, TABLE I.
- [5] (2023) Real-time digital twins: vision and research directions for 6G and beyond. IEEE Commun. Mag. 61 (11), pp. 128–134. Cited by: §I.
- [6] (2026) Site-specific finetuning of neural receivers with real-world 5G NR measurements. Note: arXiv:2603.09644 Cited by: §I-A.
- [7] (2025) Physics-informed generative modeling of wireless channels. In Proc. Int. Conf. Mach. Learn. (ICML), PMLR, Vol. 267, pp. 4602–4626. Cited by: §I-A.
- [8] Eravant Millimeter Wave Products & Components. Note: https://www.eravant.com/Accessed: 2026-07-02 Cited by: §II-B.
- [9] (2024) GAN-based massive MIMO channel model trained on measured data. In Proc. 27th Int. ITG Workshop Smart Antennas (WSA), pp. 109–116. External Links: Document Cited by: §I-A.
- [10] (2016) Digital twin: mitigating unpredictable, undesirable emergent behavior in complex systems. In Transdisciplinary Perspectives on Complex Systems, pp. 85–113. Cited by: §I.
- [11] (2025-07) Digital twin enabled site specific channel precoding: over the air CIR inference. IEEE Commun. Lett. 29 (7), pp. 1559–1563. External Links: Document Cited by: §I-A, TABLE I.
- [12] (2022) Terahertz wireless channels: a holistic survey on measurement, modeling, and analysis. IEEE Commun. Surveys Tuts. 24 (3), pp. 1670–1707. Cited by: §I.
- [13] (2016) Deep residual learning for image recognition. In Proc. IEEE Conf. Comput. Vis. Pattern Recognit. (CVPR), pp. 770–778. Cited by: §III-A.
- [14] (2023) Sionna RT: differentiable ray tracing for radio propagation modeling. Note: arXiv:2303.11103 Cited by: §I, §II-C.
- [15] (2022) Sionna: an open-source library for next-generation physical-layer research. Note: arXiv:2203.11854 Cited by: §I.
- [16] (2024) Transfer learning enabled transformer-based generative adversarial networks for modeling and generating terahertz channels. Commun. Eng. 3, pp. 153. Cited by: §I-A.
- [17] (1964) Robust estimation of a location parameter. The Annals of Mathematical Statistics 35 (1), pp. 73–101. External Links: Document Cited by: §IV-A.
- [18] (2023-11) Framework and overall objectives of the future development of IMT for 2030 and beyond. Recommendation Technical Report M.2160-0, ITU-R. Cited by: §I.
- [19] (2025) Learnable wireless digital twins: reconstructing electromagnetic field with neural representations. IEEE Open Journal of the Communications Society 6, pp. 1568–1590. External Links: Document Cited by: §I-A, TABLE I.
- [20] (2016) Perceptual losses for real-time style transfer and super-resolution. In Proc. Eur. Conf. Comput. Vis. (ECCV), pp. 694–711. Cited by: §III-A.
- [21] (2019) Deep learning-based channel prediction in realistic vehicular communications. IEEE Access 7, pp. 27846–27858. Cited by: §I-A.
- [22] (2021-06) Millimeter wave and sub-terahertz spatial statistical channel model for an indoor office building. IEEE J. Sel. Areas Commun. 39 (6), pp. 1561–1575. Cited by: §I-A.
- [23] (2022) Digital-twin-enabled 6G: vision, architectural trends, and future directions. IEEE Commun. Mag. 60 (1), pp. 74–80. Cited by: §I.
- [24] (2015) Adam: a method for stochastic optimization. In Proc. Int. Conf. Learn. Represent. (ICLR), Cited by: §III-C.
- [25] (2026) Wireless digital twin calibration: refining DFT-domain channel information. arXiv preprint arXiv:2603.16126. Cited by: §I-A, TABLE I.
- [26] (2023) A functional architecture for 6G special-purpose industrial iot networks. IEEE Transactions on Industrial Informatics 19 (3), pp. 2530–2540. Cited by: §I.
- [27] (2024) Measurement-based prediction of mmWave channel parameters using deep learning and point cloud. IEEE Open J. Veh. Technol. 5, pp. 1059–1072. External Links: Document Cited by: §I-A.
- [28] (2019-06) Wireless communications and applications above 100 GHz: opportunities and challenges for 6G and beyond. IEEE Access 7, pp. 78729–78757. Cited by: §I-A.
- [29] (2024) Wireless InSite: 3D wireless propagation software. Note: https://www.remcom.com/wireless-insite-em-propagation-softwareState College, PA Cited by: §I.
- [30] (2015) U-Net: convolutional networks for biomedical image segmentation. In Proc. Med. Image Comput. Comput.-Assist. Intervent. (MICCAI), pp. 234–241. Cited by: §IV-A.
- [31] (2025) How to bridge the sim-to-real gap in digital twin-aided telecommunication networks. arXiv preprint arXiv:2507.07067. Cited by: §I-A.
- [32] (2024) Calibrating wireless ray tracing for digital twinning using local phase error estimates. IEEE Transactions on Machine Learning in Communications and Networking 2, pp. 1193–1215. External Links: Document Cited by: §I-A, TABLE I.
- [33] (2023-03) Ambiguity in RMS delay spread of millimeter-wave channel measurements. In Proc. Eur. Conf. Antennas Propag. (EuCAP), Cited by: §II-A.
- [34] (2022-04) ML-based massive MIMO channel prediction: does it work on real-world data?. IEEE Wireless Commun. Lett. 11 (4), pp. 811–815. Cited by: §I-A.
- [35] (2025) A novel site-specific inference model for urban canyon channels: from measurements to modeling. Note: arXiv:2509.19275 Cited by: §I-A.
- [36] (2024) Wireless network digital twin for 6G: generative AI as a key enabler. IEEE Wireless Commun. 31 (4), pp. 24–31. Cited by: §I.
- [37] (2025) Digital-twin empowered site-specific radio resource management in 5G aerial corridor. In MILCOM 2025-2025 IEEE Military Communications Conference (MILCOM), pp. 1–6. Cited by: §II-C.
- [38] (2026) Digital-twin empowered deep reinforcement learning for site-specific radio resource management in NextG wireless aerial corridor. Note: arXiv:2602.03801 Cited by: §I, §II-C.
- [39] (2024-04) Generative network-based channel modeling and generation for air-to-ground communication scenarios. IEEE Commun. Lett. 28 (4), pp. 892–896. Cited by: §I-A.
- [40] (2016) Instance normalization: the missing ingredient for fast stylization. arXiv preprint arXiv:1607.08022. Cited by: §III-A.
- [41] (2009) Optimal transport: old and new. Springer, Berlin, Germany. Cited by: §II-C.
- [42] (2020-12) 6G wireless channel measurements and models: trends and challenges. IEEE Veh. Technol. Mag. 15 (4), pp. 22–32. Cited by: §I.
- [43] (2025) Electromagnetic wave property inspired radio environment knowledge construction and artificial intelligence based verification for 6G digital twin channel. Frontiers Inf. Technol. Electron. Eng. 26 (2), pp. 260–277. Cited by: §I.
- [44] (2021) Digital twin networks: a survey. IEEE Internet Things J. 8 (18), pp. 13789–13804. Cited by: §I.
- [45] (2026) Site-specific location calibration and validation of ray-tracing simulator NYURay at upper mid-band frequencies. npj Wireless Technology 2, pp. 8. External Links: Document Cited by: §I-A, TABLE I.
- [46] (2015) Ray tracing for radio propagation modeling: principles and applications. IEEE Access 3, pp. 1089–1100. Cited by: §I.
- [47] (2017) Unpaired image-to-image translation using cycle-consistent adversarial networks. In Proc. IEEE Int. Conf. Comput. Vis. (ICCV), pp. 2242–2251. Cited by: §III-A, §III-A, §III.
- [48] (2024) Toward real-time digital twins of EM environments: computational benchmark for ray launching software. IEEE Open J. Commun. Soc. 5, pp. 6291–6302. Cited by: §I.