arXiv is now an independent nonprofit! Learn more
License: CC BY 4.0
arXiv:2607.26501v1 [eess.SP] 29 Jul 2026

Calibrating the Digital Twin Channel: Statistics-Consistent Sim-to-Lab Adaptation for W-Band Industrial OFDM Links

Pulok Tarafder, , Abigail O. Oyekola, Jasni Areepatta Mannil, Imtiaz Ahmed, , Zoheb Hassan, , Danda B. Rawat, , IEEE, Wenjie Che P. Tarafder, A.O. Oyekola, J.A. Mannil, I. Ahmed, D. B. Rawat, and W. Che are with the Department of Electrical Engineering and Computer Science, Howard University, Washington, DC 20059, USA (emails: {pulok.tarafder, abigail.oyekola, jasni.mannil}@bison.howard.edu; imtiaz.ahmed@howard.edu; danda.rawat@Howard.edu; wenjie.che@howard.edu).Z. Hassan is with Dept. of Electrical and Computer Engineering, Université Laval, Québec, G1V 0A6, Canada (email: md-zoheb.hassan@gel.ulaval.ca).This work was supported in part by the NSF Grant # 2200640, in part by DoD/US Army Contract W911NF-22-1-022, and in part by the US DoD Center of Excellence in AI/ML at Howard University under Contract W911NF-20-2-0277 with the US ARL. However, the views and conclusions expressed herein are those of the authors and do not necessarily represent the official policies of the funding agencies.
Abstract

Digital twins (DTs) can reduce over-the-air validation cost in industrial wireless networks, but their utility depends on the fidelity of the underlying channel twin (CT). At W-band, site-specific ray tracing captures deterministic propagation geometry, yet its channel frequency responses (CFRs) do not reproduce the small-scale impairments and capture-to-capture variability observed in laboratory orthogonal frequency-division multiplexing (OFDM) measurements above 90 GHz. This paper proposes Statistics-Consistent Sim-to-Lab Adaptation (SC-SLA), a calibration framework that improves the fidelity of a 95 GHz Sionna ray-traced CT toward that of the testbed by aligning the mean power delay profile (PDP), the distribution of root-mean-square delay spread (τrms\tau_{\text{rms}}), and per-subcarrier statistics at 50 MHz50\text{\,}\mathrm{MHz} sampling bandwidth. SC-SLA uses a generative adversarial network (GAN)-inspired, cycle-consistent architecture with ResNet generators and batch-level channel-statistics losses on the PDP, sub-band PDP, τrms\tau_{\text{rms}} moments and quantiles, and normalized mean-square error (NMSE) of the mean CFR-magnitude profile. The framework is non-adversarial and requires neither paired simulated/measured samples nor discriminators. On held-out 95 GHz data, SC-SLA reduces the τrms\tau_{\text{rms}} Kolmogorov–Smirnov (KS) statistic from 0.86 to 0.050 relative to the impairment-augmented ray-traced input, and by 44%44\% (0.0890.0500.089\!\to\!0.050) relative to the strongest of four supervised baselines (FCNN, CNN1D, BiLSTM, and UNet1D). Without retraining, the same checkpoint also generalizes to 92–94 GHz carriers, where it achieves the lowest PDP and CFR-magnitude errors among all baselines while reducing the τrms\tau_{\text{rms}}-distribution mismatch relative to the uncalibrated twin.

I Introduction

Industry 4.0 wireless deployments increasingly rely on dense collections of sensors, controllers, and mobile robots that require high-rate, low-latency, and reliable private connectivity [3, 26]. Validating every physical-layer configuration on hardware is costly, particularly in fixed industrial cells where machinery, racks, reflectors, and access points occupy known locations. This setting has motivated the use of digital twins (DTs), measurement-synchronized virtual representations that can be queried before candidate configurations are transferred to the physical system [10, 44, 23]. Industrial radio environments are well suited to this paradigm because their propagation geometry is largely static over the time scale of link-level design.

In a wireless DT, the channel determines link-level quantities such as throughput, pilot overhead, cyclic-prefix margin, and beam coherence. The digital twin channel, referred to in this paper as the Channel Twin (CT), is therefore a foundation layer for wireless network-level twins [38, 43, 36]. Under the high-frequency approximation of Maxwell’s equations, a CT can model electromagnetic propagation through geometric rays and support geometry-aware analysis in complex three-dimensional environments [46]. Site-specific ray tracing (RT) is a practical CT engine for industrial Internet of Things (IIoT) deployments. Once the cell geometry and material properties are specified, solvers such as NVIDIA Sionna RT [15, 14] and Wireless InSite [29] can produce deterministic channel frequency responses (CFRs) and synthesize large datasets at low marginal cost. Recent advances further bring such solvers close to real-time operation [48, 5].

The usefulness of this closed loop depends on CT fidelity, which is difficult to guarantee in the frequency ranges targeted by 6G IIoT. The ITU-R IMT-2030 framework anticipates IMT operation in bands beyond those used for IMT-2020 [18], and 3GPP has extended new radio (NR) operation beyond the original frequency range 2 (FR2) ceiling to 71 GHz [2]. The W-band (75–110 GHz) follows this spectrum trajectory. In this paper, W-band refers to the 92–95 GHz carriers under study. These carriers lie near the upper limit of TR 38.901 [1], which is specified only up to 100 GHz and was parameterized primarily from sub-6 GHz and lower millimeter-wave (mmWave) measurements. Its small-scale channel assumptions therefore require measurement-based recalibration before they can support W-band physical-layer studies [42, 12].

Calibration is precisely where ray-traced twins begin to fall short. In our comparison between 95 GHz Sionna CFRs and in-lab orthogonal frequency-division multiplexing (OFDM) measurements, two systematic fidelity gaps emerged. (i) Early-tap energy underestimation. The solver reproduces specular and refractive interactions but omits diffuse near-field scattering, antenna-to-cable mismatch ripple, and analog-front-end group-delay distortion, all of which contribute strongly to the first few channel-impulse-response (CIR) taps in a real W-band setup. (ii) Narrow delay-spread distribution. Because the geometry is static, the twin produces an almost deterministic root-mean-square (RMS) delay spread, τrms\tau_{\text{rms}}, whereas the physical testbed exhibits a much broader capture-to-capture spread, attributable to effects such as oscillator phase noise, sampling-clock jitter, and signal-to-noise ratio (SNR) drift in the W-band frequency converter. A twin with these discrepancies provides only limited fidelity, and any IIoT link-level study built on it, including pilot-density selection, cyclic-prefix design, and beam-coherence analysis, inherits the resulting bias.

Learning the missing propagation and hardware effects from data is a natural calibration strategy. However, industrial measurements impose an important constraint. Supervised calibration would require paired simulated and measured CFRs for the same channel realization, which is rarely available in field or laboratory captures. Simulated and measured acquisitions are typically generated independently, leaving no sample-wise correspondence between domains. The calibration problem must therefore be treated as an unpaired domain-alignment task. The learned correction should also preserve channel statistics that affect OFDM receiver design, including the power delay profile (PDP), τrms\tau_{\text{rms}}, and per-subcarrier magnitude, rather than merely fitting individual waveform samples.

Motivated by these requirements, we propose Statistics-Consistent Sim-to-Lab Adaptation (SC-SLA), a calibration layer between a ray-traced CT and a physical W-band testbed. SC-SLA learns an unpaired, non-adversarial mapping from Sionna-generated CFRs to testbed-like CFRs by directly aligning receiver-relevant channel statistics. The calibrated twin retains the scalability of RT while reproducing the delay and frequency-domain statistics observed in the measured industrial link.

I-A Related Work

Outside explicit wireless-DT calibration, prior efforts to reduce the gap between modeled and measured wireless channels have followed three main directions. Measurement-based and site-specific models extend geometry-based stochastic modeling beyond mmWave by incorporating empirical path-loss, clustering, and delay statistics [28, 22]. Recent studies also infer channel statistics from environment geometry or co-located sensor data [35, 27]. These models capture propagation behavior. However, they do not calibrate the capture-specific impairments introduced by a particular measurement testbed.

Supervised networks trained on measured channels have been used for channel prediction and estimation, including vehicular channel state information (CSI) prediction [21] and massive multiple-input multiple-output (MIMO) channel prediction validated on field measurements [34]. In sim-to-real calibration, however, these methods require synchronized simulated and measured channel pairs, which are rarely available in field or industrial captures.

Generative models relax the need for paired samples by learning the underlying channel distribution. Examples include generative adversarial networks (GANs) trained on measured channel-sounder data [9], conditional GANs for air-to-ground channels [39], and physics-informed generative variants [7]. Nevertheless, these approaches typically optimize adversarial or parameter-domain objectives without explicitly preserving OFDM-relevant statistics such as PDP shape, τrms\tau_{\text{rms}}, and frequency-domain magnitude structure.

A parallel line of work focuses explicitly on reducing the sim-to-real gap in wireless DTs. Following the taxonomy in [31], this gap can be addressed through (i) direct calibration of the DT using real measurements, (ii) uncertainty-aware modeling of the residual environment mismatch, and (iii) correction of the task-level AI training objective. SC-SLA belongs to the first category and, more specifically, performs post-RT channel-output calibration: the RT solver, reconstructed geometry, and nominal material properties remain fixed while a learned mapping aligns the simulated and measured CFR distributions.

Direct DT calibration can operate at different stages of the channel-generation process. Model-space approaches learn object-level electromagnetic properties and radio-wave interaction functions [19], or estimate RT material parameters while compensating for path-phase errors caused by geometric mismatch [32]. Geometry-space calibration instead corrects physical scene variables, such as transmitter and receiver locations, using measured and simulated PDPs [45]. Propagation-consistent environment twins can also embed a learnable scene-level electromagnetic field within differentiable RT and calibrate it using sparse position-labeled CSI [4].

Channel-output calibration avoids reconstructing the complete internal environment model. Haider et al. [11] first increase the fidelity of a site-specific RT model by selecting suitable propagation settings and material assignments, and then train a supervised U-Net to refine the resulting RT CIRs using measured over-the-air CIRs as labels. The refined channels are subsequently used for precoding and evaluated through end-to-end bit error rate (BER). Similarly, Luo et al. [25] use supervised learning to refine discrete Fourier transform (DFT)-domain channel weights generated by a low-complexity DT for CSI compression and feedback. Table I summarizes these calibration approaches and positions SC-SLA relative to them.

TABLE I: Representative wireless-DT calibration approaches
Work Calibration space Calibrated component Measurement or supervision Calibration objective Relation to SC-SLA
Jiang et al. [19] Model Object-level electromagnetic representation and interaction model Scene-associated wireless observations Learn electromagnetic properties and radio-wave interactions Re-learns the environment model. SC-SLA keeps the RT model fixed and calibrates its CFR outputs.
Ruah et al. [32] Model RT material parameters and path-phase errors Position-associated channel observations Estimate material parameters while accounting for geometric phase errors Calibrates inside a differentiable RT solver. SC-SLA does not backpropagate through the solver.
Ying et al. [45] Geometry Transmitter and receiver locations Measured and simulated PDPs at corresponding sites Correct location uncertainty and multipath mismatch Corrects TX/RX placement. SC-SLA assumes fixed nominal node positions.
Ai et al. [4] Environment Scene-level electromagnetic property field Sparse position-labeled CSI Construct a propagation-consistent environment representation Modifies the environment model and requires differentiable RT. SC-SLA applies a post-RT CFR mapping.
Haider et al. [11] Output RT configuration, material assignments, and generated CIR Paired RT-generated and measured CIRs Refine CIRs using a supervised U-Net for precoding and BER evaluation Uses paired CIR supervision and task-level validation. SC-SLA uses unpaired data and ensemble statistics.
Luo et al. [25] Output DFT-domain channel output Aligned low-/high-fidelity channel information or historical CSI Refine DFT-domain channel weights by supervised learning Uses an aligned, task-oriented representation. SC-SLA calibrates unpaired complex OFDM CFRs.
SC-SLA Output Post-RT complex CFR distribution Independently acquired Sionna RT and W-band testbed captures Align PDP, sub-band PDP, τrms\tau_{\text{rms}}, and per-subcarrier magnitude statistics Unpaired, non-adversarial calibration that leaves the RT scene, solver, and communication model unchanged.

Adjacent sim-to-real studies adapt channel generators or communication models using limited measurements. Hu et al. [16] propose a transfer-learning transformer GAN pretrained on THz channels from a geometry-based stochastic channel model and fine-tuned using a smaller vector network analyzer measurement set. The model remains adversarial and operates in the multipath-parameter domain. Baytekin et al. [6] instead pretrain a neural receiver on a randomized 3GPP urban-microcell channel model and fine-tune it with real 5G NR physical uplink shared channel (PUSCH) measurements. Their objective is end-to-end block-error-rate performance rather than calibration of the underlying channel distribution.

Taken together, prior approaches leave three requirements unresolved for the present setting: (i) calibrating the generated channel without modifying or differentiating through the RT model, (ii) learning from independently acquired simulated and measured CFRs without sample-wise correspondence, and (iii) correcting the channel distribution itself rather than optimizing a task-specific communication objective. Fidelity calibration of a ray-traced CT for a single-antenna W-band OFDM link therefore remains largely unaddressed, particularly under the unpaired acquisition conditions considered in this work.

I-B Contributions

The proposed SC-SLA addresses this gap by providing an unpaired, GAN-inspired but non-adversarial calibration layer for the CT of a static industrial communication network. Instead of requiring sample-wise simulated/measured correspondence, SC-SLA directly aligns measured channel statistics that govern OFDM link behavior. The main contributions are summarized as follows.

  • An unpaired distributional CT-calibration formulation: We formulate sim-to-lab adaptation as the problem of bringing the statistical distribution of DT-generated CSI toward that of independently acquired real-world industrial CSI captures. The formulation explicitly aligns complementary delay and frequency-domain attributes, including the mean PDP, local sub-band delay structure, the distribution of τrms\tau_{\text{rms}}, and the per-subcarrier magnitude profile, without imposing an arbitrary one-to-one correspondence between channel realizations.

  • A W-band OFDM testbed for CT calibration: We develop a commercial off-the-shelf (COTS) W-band measurement platform that combines USRP-B200 radios, WR-10 frequency-conversion modules, Faraday isolators, and pyramidal horn antennas. The testbed supports complex OFDM CFR acquisition over 92–95 GHz on the same subcarrier grid used by the ray-traced CT, thereby providing the real-world CSI required for sim-to-lab calibration and cross-frequency evaluation.

  • A CycleGAN-style non-adversarial channel translator: We develop a two-generator, cycle-consistent translation architecture inspired by GAN-based unpaired domain adaptation, but replace adversarial discriminators with receiver-relevant channel-statistics losses. The training objective combines 1\ell_{1} losses on the truncated PDP and ten sub-band PDPs, τrms\tau_{\text{rms}} moment and quantile losses, normalized mean-square error (NMSE) of the mean CFR-magnitude profile, and cycle/identity regularization.

  • An empirical study at 92–95 GHz: We evaluate SC-SLA using Sionna RT CFRs from the reconstructed laboratory CT and OFDM CFR measurements from the W-band testbed. At 95 GHz, SC-SLA reduces the τrms\tau_{\text{rms}} Kolmogorov–Smirnov (KS) statistic by 44%44\% relative to the strongest of four supervised baselines while also achieving the lowest PDP and CFR-magnitude NMSE. Without retraining or carrier-specific tuning, the same checkpoint transfers to 92, 93, and 94 GHz and retains the lowest PDP and CFR-magnitude errors among the evaluated baselines, demonstrating calibration robustness across a 3 GHz carrier shift. To support reproducibility and further research on CT calibration, we publicly release the Sionna datasets, testbed datasets, trained checkpoints, and source code111Available at https://github.com/puloktarafder/thz-pcsla upon publication..

The remainder of this paper is organized as follows. Section II presents the signal model and the simulated and measured datasets. Section III details the SC-SLA generator and its channel-statistics losses. Section IV reports in-domain and cross-frequency results against the supervised baselines, together with inference scalability, sensitivity to the amount of measured training data, and an ablation of the loss terms and generator head that isolates the mechanisms behind the observed gains. Section V concludes the paper.

II System Model and Experimental Setup

This section specifies the OFDM signal model, the simulated and measured CFR datasets, and the common normalization protocol used by SC-SLA. The simulated ray-traced CT dataset is denoted by XsX_{s}, and its measured testbed counterpart is denoted by XrX_{r}. Both datasets use the same OFDM grid and normalization procedure before calibration.

Refer to caption
(a) Transmitter side.
Refer to caption
(b) Receiver side.
Figure 1: W-band OFDM testbed setup. The transmitter and receiver sides use Eravant WR-10 pyramidal horn antennas that are co-polarized and boresight-aligned to form the LOS measurement link.
Refer to caption
(a) Blender scene.
Refer to caption
(b) Sionna scene.
Refer to caption
(c) Sionna RT scene.
Figure 2: Ray-traced CT construction. (a) Blender reconstruction of the indoor W-band testbed. (b) Imported Sionna RT scene with the transmitter and receiver placed at the nominal horn antenna locations. (c) Ray-traced Sionna RT scene with LOS, specular reflection, and refraction paths overlaid.

II-A OFDM Signal and Channel Model

We consider a static indoor W-band link over 92–95 GHz with a dominant line-of-sight (LOS) component. To make simulation and measurement directly comparable, both domains use the same OFDM reference grid. Each OFDM symbol has NFFT=2048N_{\text{FFT}}=2048 subcarriers, of which Na=1500N_{\text{a}}=1500 are active. The active subcarriers occupy

𝒦a={+1,,+750}{750,,1},\mathcal{K}_{\text{a}}=\{+1,\dots,+750\}\cup\{-750,\dots,-1\},

which excludes the direct-current (DC) bin k=0k=0. This bin is left unused because practical radio-frequency (RF) front ends can exhibit DC offset, local-oscillator (LO) leakage, mixer feedthrough, and residual in-phase/quadrature (I/Q) imbalance around the carrier. Excluding the DC tone therefore avoids a hardware-sensitive subcarrier and yields a cleaner active CFR. The sampling bandwidth is Bs=50 MHzB_{\text{s}}=$50\text{\,}\mathrm{MHz}$, yielding a subcarrier spacing of Δf=Bs/NFFT24.41 kHz\Delta f=B_{\text{s}}/N_{\text{FFT}}\approx$24.41\text{\,}\mathrm{kHz}$ and a useful OFDM symbol duration of Tu=1/Δf40.96 µsT_{u}=1/\Delta f\approx$40.96\text{\,}\mathrm{\SIUnitSymbolMicro s}$. The active subcarriers therefore span approximately NaΔf=1500×24.41 kHz36.6 MHzN_{\text{a}}\Delta f=1500\times$24.41\text{\,}\mathrm{kHz}$\approx$36.6\text{\,}\mathrm{MHz}$ of the 50 MHz50\text{\,}\mathrm{MHz} sampling bandwidth, with the remainder serving as guard band. A cyclic prefix of Ncp=512N_{\text{cp}}=512 samples is prepended to each symbol. The active CFR on subcarrier k𝒦ak\in\mathcal{K}_{\text{a}} is modeled as

H[k]==0L1αej2πkΔfτ+Z[k],H[k]=\sum_{\ell=0}^{L-1}\alpha_{\ell}\,e^{-j2\pi k\Delta f\tau_{\ell}}+Z[k], (1)

where α\alpha_{\ell} and τ\tau_{\ell} denote the complex gain and excess delay of path \ell, respectively, LL is the number of resolvable paths, and Z[k]Z[k] represents the residual receiver noise after least-squares (LS) channel estimation.

Let H~[m]\tilde{H}[m] denote the full NFFTN_{\text{FFT}}-point CFR obtained by placing the active samples H[k]H[k] on their corresponding fast Fourier transform (FFT) bins and assigning zeros to the DC and guard-band bins. The truncated CIR is obtained from the NFFTN_{\text{FFT}}-point inverse FFT (IFFT) of H~[m]\tilde{H}[m] by retaining the first T=20T=20 taps,

h~[n]=1NFFTm=0NFFT1H~[m]ej2πmn/NFFT,n=0,,T1.\tilde{h}[n]=\frac{1}{N_{\text{FFT}}}\sum_{m=0}^{N_{\text{FFT}}-1}\tilde{H}[m]\,e^{j2\pi mn/N_{\text{FFT}}},\quad n=0,\dots,T-1. (2)

Because the IFFT is defined over the full sampling bandwidth BsB_{\text{s}}, adjacent CIR taps are spaced by Δt=1/Bs=20 ns\Delta t=1/B_{\text{s}}=$20\text{\,}\mathrm{ns}$. Retaining T=20T=20 taps gives a uniformly sampled 400 ns400\text{\,}\mathrm{ns} delay window. From this truncated CIR, the normalized PDP is computed as p[n]=|h~[n]|2/m=0T1|h~[m]|2p[n]=|\tilde{h}[n]|^{2}/\sum_{m=0}^{T-1}|\tilde{h}[m]|^{2}. The resulting discrete delay-energy distribution defines τrms\tau_{\text{rms}} as

τrms=n=0T1(nΔt)2p[n](n=0T1nΔtp[n])2.\tau_{\text{rms}}=\sqrt{\sum_{n=0}^{T-1}\!(n\Delta t)^{2}p[n]-\!\Big(\sum_{n=0}^{T-1}\!n\Delta t\,p[n]\Big)^{2}}. (3)

The finite delay support reduces the influence of late-delay noise-floor samples on the second-order delay moment. This choice is consistent with prior observations that noise and spurious PDP components at large delays can bias measured τrms\tau_{\text{rms}} estimates if they are not excluded before computation [33]. The CFR, truncated CIR, PDP, and τrms\tau_{\text{rms}} definitions above provide the common channel representation used to construct XsX_{s} and XrX_{r}.

II-B Physical Testbed (W-Band USRP Measurements)

The measured dataset is captured with the W-band OFDM testbed in Fig. 1, which operates over 92–95 GHz and uses COTS components. Separate USRP-B200 devices serve as transmitter (TX) and receiver (RX). The transmitter USRP generates the baseband OFDM waveform, while the receiver USRP records the received waveform with a 50 MHz50\text{\,}\mathrm{MHz} master clock and a 70 dB70\text{\,}\mathrm{dB} receiver-gain setting. The waveform is generated at an intermediate frequency (IF), fIF{2,3,4,5}GHzf_{\text{IF}}\in\{2,3,4,5\}~$\mathrm{GHz}$, upconverted to W-band at the transmitter, and downconverted back to IF at the receiver using matched COTS Eravant [8] WR-10 up/down-conversion modules. The chain uses a ×8\times 8-multiplied phase-locked oscillator and yields fRF=8fLO+fIFf_{\text{RF}}=8f_{\text{LO}}+f_{\text{IF}}, where fLO=11.25f_{\text{LO}}=11.25 GHz. Retuning fIFf_{\text{IF}} produces carriers from 9292 to 9595 GHz. A W-band Faraday isolator is inserted at each antenna port using COTS WR-10 waveguide components operating over 75–110 GHz, with 28 dB28\text{\,}\mathrm{dB} isolation, 4.5 dB4.5\text{\,}\mathrm{dB} insertion loss, and a 90 °90\text{\,}\mathrm{\SIUnitSymbolDegree} twist. Both ends use co-polarized, boresight-aligned Eravant WR-10 pyramidal horn antennas, with a TX–RX separation of 1.75 m1.75\text{\,}\mathrm{m}.

The transmitted reference is a flat-power OFDM symbol whose active subcarriers carry unit-magnitude symbols with uniformly random phases. For each carrier, the receiver records a 15 s15\text{\,}\mathrm{s} burst. Offline synchronization cross-correlates the received burst with the known reference symbol. Candidate frame peaks are detected with a minimum peak distance of NFFT+NcpN_{\text{FFT}}+N_{\text{cp}} and a threshold equal to three times the mean correlation magnitude. Each valid detected reference symbol yields one channel capture. For every detected capture, the receiver computes the FFT of the synchronized OFDM symbol and estimates the active CFR by the LS method, H^[k]=Y[k]Xref[k]\hat{H}[k]=\frac{Y[k]}{X_{\text{ref}}[k]}, where Y[k]Y[k] is the received pilot on active subcarrier kk and Xref[k]X_{\text{ref}}[k] is the transmitted reference symbol. The estimated active CFR is stored over the same 15001500 active tones and in the same [+1,,+750,750,,1][+1,\dots,+750,-750,\dots,-1] ordering used by the simulated dataset. A full NFFTN_{\text{FFT}}-point CFR is also constructed by inserting zeros on the DC and guard-band bins, and its IFFT gives the corresponding full-length CIR.

An instantaneous SNR is estimated for each capture from the full CIR. The signal power is taken as the maximum tap power within the first 100100 taps, and the noise power is estimated as the mean tap power over the remaining taps. Captures with SNR below 8 dB8\text{\,}\mathrm{dB} are rejected. After SNR filtering, we retain Nr=50,000N_{r}=50{,}000 valid CFRs per carrier. At 95 GHz, the retained captures have a mean estimated SNR of 28.98 dB28.98\text{\,}\mathrm{dB} and a standard deviation of 2.87 dB2.87\text{\,}\mathrm{dB}. For 95 GHz training, the measured dataset is split into 45,00045{,}000 training samples and 5,0005{,}000 held-out samples, and 2,0002{,}000 held-out captures are subsampled for each reported metric.

The retained measurements exhibit substantial delay-domain variation even though the antenna positions and link geometry remain fixed. Across the 2,0002{,}000 held-out 95 GHz captures in Fig. 3, the mean τrms\tau_{\text{rms}} is 34.05 ns34.05\text{\,}\mathrm{ns}, with a standard deviation of 10.88 ns10.88\text{\,}\mathrm{ns}. Its 10th and 90th percentiles are 24.75 ns24.75\text{\,}\mathrm{ns} and 44.40 ns44.40\text{\,}\mathrm{ns}, respectively. The normalized CIR and PDP statistics indicate that the variation redistributes energy among the retained taps rather than producing only a common gain change. These measurements define the reference distribution used to assess RT fidelity.

Refer to caption
Figure 3: Delay-domain comparison of the impairment-augmented Sionna RT and measured testbed channels at 95 GHz using 2,0002{,}000 unpaired held-out samples per domain. Panels (a) and (b) depict the first 20 CIR taps, normalized to each sample’s peak power and ordered by τrms\tau_{\text{rms}}. Panel (c) presents the mean unit-energy PDP and its 10th–90th percentile range. Panel (d) compares the empirical cumulative distribution functions and reports the mean and standard deviation of τrms\tau_{\text{rms}} as (μ,σ)=(23.8,0.5)(\mu,\sigma)=(23.8,0.5) ns\mathrm{ns} for Sionna RT and (34.1,10.9)(34.1,10.9) ns\mathrm{ns} for the measured testbed, together with the KS statistic DKS=0.861D_{\mathrm{KS}}=0.861 and the empirical 1-Wasserstein distance W1=10.3 nsW_{1}=$10.3\text{\,}\mathrm{ns}$.

II-C Ray-Traced CT (Sionna RT)

The simulated CT dataset is generated with the Sionna RT path solver [14] on a calibrated indoor scene. The geometry corresponds to a single-room laboratory reconstructed in Blender, with walls and floor assigned ITU-R indoor material properties at W-band. The scene is exported as a Mitsuba XML file. Fig. 2 depicts the Blender reconstruction, the imported Sionna scene, and the traced paths. The TX and RX are placed at the nominal horn locations with a LOS distance of 1.75 m1.75\text{\,}\mathrm{m}, matching the testbed. At both ends, we use a custom vertically polarized antenna element with a horn-like directive pattern, since Sionna RT does not provide a native horn model. The antenna power pattern is modeled as Ph(ϑ)max(cosϑ,0)nhP_{\mathrm{h}}(\vartheta)\propto\max(\cos\vartheta,0)^{n_{\mathrm{h}}}, where ϑ\vartheta is the angular offset from boresight. The exponent nh=ln(1/2)/ln[cos(6)]126.18n_{\mathrm{h}}=\ln(1/2)/\ln[\cos(6^{\circ})]\approx 126.18 solves the half-power condition at a 66^{\circ} half-angle, giving a 12 °12\text{\,}\mathrm{\SIUnitSymbolDegree} full 3-dB beamwidth. The cosine-power model captures the dominant boresight directivity while omitting detailed sidelobe structure. We configure the path solver to support LOS, specular reflection, and refraction with a maximum interaction depth of 55 and disable diffuse reflection. The resulting RT model represents the deterministic geometric component of the link, while PC-SLA learns the residual, comprising both stochastic capture variation and the deterministic response of the measurement chain. For each carrier, the static scene is solved once to obtain a deterministic base CFR over the same active OFDM tones used by the measurement waveform. The active-frequency vector follows the ordering [+1,,+750,750,,1][+1,\dots,+750,-750,\dots,-1], which matches the stored active-bin convention of the measured CFRs. The path delays are normalized so that the first arrival falls at tap 0, matching the synchronization convention used in the receiver processing.

A single RT solution yields one deterministic base CFR per carrier, whereas the physical testbed produces a distribution of captures because each acquisition is affected by noise, timing, and phase fluctuations. To emulate this behavior and match the NrN_{r} captures retained per carrier in Section II-B, we generate Ns=50,000N_{s}=50{,}000 simulated samples per carrier by passing the deterministic base CFR through a measurement-inspired impairment layer. Let HRT[k]H_{\mathrm{RT}}[k] denote the deterministic base CFR on active subcarrier kk, and let ii index the simulated capture. For each capture, a timing offset Δτi\Delta\tau_{i} is drawn independently from a zero-mean Gaussian distribution with a standard deviation of 2.5 ns2.5\text{\,}\mathrm{ns}. This offset is applied as the subcarrier-dependent phase rotation ej2πkΔfΔτie^{-j2\pi k\Delta f\,\Delta\tau_{i}}, which introduces a random linear phase slope across frequency. A common phase offset ϕi\phi_{i} is independently drawn from a zero-mean Gaussian distribution with a standard deviation of 0.20 rad0.20\text{\,}\mathrm{rad} and applied identically to all active subcarriers. The resulting impairment-augmented ray-traced CFR is

Hi(s)[k]\displaystyle H_{i}^{(s)}[k] =HRT[k]ejϕij2πkΔfΔτi+Vi[k],\displaystyle=H_{\mathrm{RT}}[k]\,e^{\,j\phi_{i}-j2\pi k\Delta f\,\Delta\tau_{i}}+V_{i}[k], (4)

where Vi[k]V_{i}[k] is zero-mean complex Gaussian noise scaled from the mean base-CFR power to realize a per-capture SNR drawn from a Gaussian distribution with mean 28 dB28\text{\,}\mathrm{dB} and standard deviation 2 dB2\text{\,}\mathrm{dB}, limited to 8–60 dB. These nominal levels provide conservative capture variation without tuning to the held-out τrms\tau_{\text{rms}} distribution. The 2.5 ns2.5\text{\,}\mathrm{ns} timing spread equals one eighth of the 20 ns20\text{\,}\mathrm{ns} sampling interval, and the 0.20 rad0.20\text{\,}\mathrm{rad} common-phase spread is about 11.5 °11.5\text{\,}\mathrm{\SIUnitSymbolDegree}. These phase perturbations leave the path geometry unchanged. The simulated SNR mean and standard deviation, 28 dB28\text{\,}\mathrm{dB} and 2 dB2\text{\,}\mathrm{dB}, approximate the corresponding measured values of 28.98 dB28.98\text{\,}\mathrm{dB} and 2.87 dB2.87\text{\,}\mathrm{dB} for the retained 95 GHz captures. The 8 dB8\text{\,}\mathrm{dB} lower limit matches the measurement rejection threshold, while the 60 dB60\text{\,}\mathrm{dB} upper limit suppresses extreme Gaussian draws. Further details on the typical Sionna RT workflow and site-specific dataset construction are available in our prior works [37, 38].

Fig. 3 summarizes the residual delay-domain gap after capture-side impairment augmentation. Across the 2,0002{,}000 held-out samples from each domain, the Sionna RT channels have a mean τrms\tau_{\text{rms}} of 23.82 ns23.82\text{\,}\mathrm{ns} and a standard deviation of 0.46 ns0.46\text{\,}\mathrm{ns}, compared with 34.05 ns34.05\text{\,}\mathrm{ns} and 10.88 ns10.88\text{\,}\mathrm{ns} for the measured channels. The measured standard deviation is 23.6×23.6\times larger. Let τ(i)(s)\tau^{(s)}_{(i)} and τ(i)(r)\tau^{(r)}_{(i)} denote the iith ordered τrms\tau_{\text{rms}} samples from the simulated and measured domains, respectively. For N=2,000N=2{,}000 equally weighted samples per domain, the empirical one-dimensional 1-Wasserstein distance is [41]

W1=1Ni=1N|τ(i)(s)τ(i)(r)|.W_{1}=\frac{1}{N}\sum_{i=1}^{N}\left|\tau^{(s)}_{(i)}-\tau^{(r)}_{(i)}\right|. (5)

The two distributions yield a two-sample KS statistic of DKS=0.861D_{\mathrm{KS}}=0.861, defined in Eq. (14), and W1=10.3 nsW_{1}=$10.3\text{\,}\mathrm{ns}$. These distributional discrepancies demonstrate that nominal capture-side impairments do not reproduce the measured delay-domain variability and motivate the learned SC-SLA calibration.

II-D Joint Normalization

The simulated and measured datasets pass through a common normalization procedure before training and evaluation. The goal is to remove absolute-power and receiver-gain differences while preserving the relative phase, frequency selectivity, and delay-domain structure needed for channel-statistics alignment.

First, each complex CFR sample is divided by its own RMS magnitude. This per-sample normalization removes capture-to-capture gain variation without forcing the two datasets to share a global power scale. Second, the normalized complex CFR is converted into a real-valued tensor by interleaving the real and imaginary components along the channel dimension, yielding samples of size (Na,2)(N_{\text{a}},2) and dataset tensors of shape (Ns,Na,2)(N_{s},N_{\text{a}},2) and (Nr,Na,2)(N_{r},N_{\text{a}},2) for the simulated and measured sets, respectively. Third, let ρclip=0.995\rho_{\text{clip}}=0.995 denote the joint quantile level. The threshold qclipq_{\text{clip}} is the ρclip\rho_{\text{clip}}-quantile of the pooled simulated and measured magnitudes. Both datasets are clipped at qclipq_{\text{clip}} and scaled to the normalized support expected by the generator.

The normalization parameters are fixed after training and reused during evaluation. This keeps the input normalization identical across carriers, allowing the 9595 GHz checkpoint to be applied directly to the 92929494 GHz carriers without retuning. All SC-SLA losses and reported metrics are computed in this shared normalized representation.

III SC-SLA Framework

Refer to caption
Figure 4: SC-SLA framework. The simulation branch (top left) runs the Sionna RT path solver on an indoor Mitsuba scene, adds capture-side impairments, and yields the CT dataset XsNs×Na×2X_{s}\in\mathbb{R}^{N_{s}\times N_{\text{a}}\times 2}. The testbed branch (top right) feeds the W-band USRP-B200 chain through OFDM synchronization, LS CFR estimation H^[k]=Y[k]/Xref[k]\hat{H}[k]=Y[k]/X_{\text{ref}}[k], and an SNR gate, producing the measured dataset. Both datasets share a per-sample RMS normalization with Re/Im interleaving (middle). The 1D ResNet generator (bottom) is an encoder–residual-trunk–decoder with a tanh\tanh-bounded residual delta head 𝐱^=tanh(𝐱+Δ𝐱)\hat{\mathbf{x}}=\tanh(\mathbf{x}+\Delta\mathbf{x}) for an input CFR sample 𝐱\mathbf{x}, forming the forward map GSR:XsX^rG_{S\to R}\!:X_{s}\!\to\!\hat{X}_{r} and the reverse map GRS:XrX^sG_{R\to S}\!:X_{r}\!\to\!\hat{X}_{s} in the cycle-consistent translator.

SC-SLA calibrates the ray-traced CT dataset XsX_{s} toward the measured dataset XrX_{r}. The proposed SC-SLA adopts a two-generator, cycle-consistent translation structure inspired by CycleGAN [47], but removes the adversarial discriminators and replaces them with batch-level channel-statistics losses. The result is a GAN-inspired, non-adversarial, unpaired channel translator. For a normalized simulated CFR sample 𝐱sNa×2\mathbf{x}_{s}\in\mathbb{R}^{N_{\text{a}}\times 2}, the forward generator GSR:Na×2Na×2G_{S\to R}:\mathbb{R}^{N_{\text{a}}\times 2}\to\mathbb{R}^{N_{\text{a}}\times 2} maps the simulated sample to a real-world-like sample, 𝐱^r=GSR(𝐱s)\hat{\mathbf{x}}_{r}=G_{S\to R}(\mathbf{x}_{s}). Applying this map to the full simulated dataset gives the calibrated dataset X^r=GSR(Xs)\hat{X}_{r}=G_{S\to R}(X_{s}). For a normalized measured CFR sample 𝐱rNa×2\mathbf{x}_{r}\in\mathbb{R}^{N_{\text{a}}\times 2}, a reverse generator GRSG_{R\to S} maps measured samples back to the ray-traced domain, 𝐱^s=GRS(𝐱r)\hat{\mathbf{x}}_{s}=G_{R\to S}(\mathbf{x}_{r}). Applying the reverse map to all measured samples gives X^s=GRS(Xr)\hat{X}_{s}=G_{R\to S}(X_{r}), which is used for cycle-consistency regularization. Both generators share the same architecture.

The calibration objective is distributional. SC-SLA does not force a simulated CFR to match a particular measured CFR, because the datasets are not synchronized and no sample-wise correspondence exists. Instead, the forward map is constrained to reproduce the channel statistics that govern OFDM receiver design, including the delay-domain energy distribution, τrms\tau_{\text{rms}}, and frequency-domain magnitude response. This matches the data collection process, where simulated and measured datasets represent the same nominal static link and OFDM grid. However, their sample-to-sample variations arise independently. Fig. 4 depicts the data flow and the complete SC-SLA framework.

In an adversarial translator, discriminators learn implicit criteria for separating translated outputs from samples in the corresponding target domains, and the generators are updated through a minimax objective. SC-SLA instead minimizes explicit, differentiable discrepancies in the PDP, τrms\tau_{\text{rms}} distribution, and CFR-magnitude profile because these receiver-relevant target statistics are available from the measured captures. Our non-adversarial formulation removes discriminator design and minimax tuning while making the physical role of each alignment term explicit. The reverse generator GRSG_{R\to S} does not replace a discriminator. We retain GRSG_{R\to S} to provide the inverse path required by cycle-consistency regularization. We use only GSRG_{S\to R} to calibrate simulated CFRs during inference.

III-A Generator Architecture

The SC-SLA generator is a 1D ResNet [13] with an encoder–residual-trunk–decoder structure, illustrated in the bottom panel of Fig. 4. The topology adapts the translation generator of Johnson et al. [20], also used by CycleGAN [47], from 2D images to the 1D active CFR. The real and imaginary components of each normalized sample form two input channels over the Na=1500N_{\text{a}}=1500 active subcarriers. Since SC-SLA retains the CycleGAN translation structure and removes only its adversarial discriminators, the generator family is kept fixed and the layer dimensions are set by the OFDM numerology in Section II-A. A 77-tap input convolution lifts the two-channel input to 6464 feature channels over a spectral aperture of 7Δf171 kHz7\Delta f\approx$171\text{\,}\mathrm{kHz}$. By Eq. (1), a path at excess delay τ\tau varies across frequency with period 1/τ1/\tau. Consequently, the longest delay retained in Eq. (2), 400 ns400\text{\,}\mathrm{ns}, induces the fastest CFR ripple with a 2.5 MHz2.5\text{\,}\mathrm{MHz} period. The input aperture covers at most 7%7\% of this cycle, so it extracts local CFR level, slope, and curvature while leaving multipath-induced ripples to deeper layers. Two stride-2 convolutions with 33-tap kernels then reduce the active-subcarrier length from 15001500 to 750750 and 375375, while increasing the channel dimension from 6464 to 128128 and 256256. Two strided stages provide the deepest exact compression of the active grid. A third halving would require a fractional length of 187.5187.5. The channel doubling keeps the activation volume constant (1500×64=750×128=375×256=96,0001500\times 64=750\times 128=375\times 256=96{,}000 values per CFR), trading spectral resolution for feature richness without reducing representational capacity [13].

After the strided stages, the compressed spectral grid contains one feature vector per four subcarriers, or approximately 97.7 kHz97.7\text{\,}\mathrm{kHz} of spectrum. A residual trunk of six instance-normalized blocks refines the encoded CFR features at this resolution, matching the CycleGAN depth for inputs of comparable size [47]. Across the input convolution, strided stages, twelve trunk convolutions, mirrored decoder, and output head, the end-to-end receptive field reaches 127127 subcarriers, or approximately 3.1 MHz3.1\text{\,}\mathrm{MHz}. This span exceeds the 2.5 MHz2.5\text{\,}\mathrm{MHz} period of the fastest ripple, so each output subcarrier observes at least one complete cycle of the finest spectral structure generated by the truncated delay window. Without downsampling, the same depth would cover only 4545 subcarriers, or approximately 1.1 MHz1.1\text{\,}\mathrm{MHz}. The receptive field is also close to the 150150-subcarrier slices used by the sub-band PDP loss in Section III-B, so learned corrections act at the spectral granularity penalized by the loss. Instance normalization [40] standardizes each capture individually rather than across the mini-batch, which is appropriate for unpaired translation with capture-dependent receive power.

The decoder mirrors the encoder and restores the original 15001500-subcarrier length through two stride-2 transposed convolutions that reduce the channel dimension from 256256 back to 6464, with explicit length matching for the odd 375750375\rightarrow 750 upsampling. A final un-normalized 77-tap convolutional head predicts the residual correction Δ𝐱\Delta\mathbf{x} on the amplitude scale of the input CFR. For a generic normalized input sample 𝐱Na×2\mathbf{x}\in\mathbb{R}^{N_{\text{a}}\times 2} from either domain, the output is formed as 𝐱^=tanh(𝐱+Δ𝐱)\hat{\mathbf{x}}=\tanh(\mathbf{x}+\Delta\mathbf{x}). The residual skip biases the map toward the identity, so the network refines the ray-traced CFR rather than reconstructing it from scratch. The tanh\tanh head confines the output to the normalized [1,1][-1,1] support shared by both datasets. Since the generator is fully feed-forward, each calibrated CFR requires only one forward pass, allowing SC-SLA to be inserted into a W-band link-level simulator without iterative calibration.

III-B Channel-Statistics Losses

The generator architecture determines the form of the CFR correction, while the loss function determines which channel properties it preserves or reproduces. We use two types of constraints. The first group consists of channel-statistics losses, which compare ensemble-level quantities computed over independently sampled simulated and measured mini-batches. These losses are not applied sample by sample, because the captures are unpaired and no measured realization corresponds to a particular simulated one. The quantities most relevant to OFDM receiver design, including PDP shape, τrms\tau_{\text{rms}}, and the average subcarrier-magnitude profile, are distributional channel properties rather than deterministic labels for individual samples.

The second group consists of instantaneous regularizers, namely the cycle and identity losses. These losses act directly on each sample and prevent the learned mapping from becoming an unconstrained distribution-matching transform. A purely statistical objective is insensitive to permutations within a mini-batch and cannot alone ensure that each translated CFR remains tied to the geometry and structure of its input. The cycle and identity losses therefore complement the residual generator parameterization in Section III-A. We retain multiple statistical losses because each constrains a different receiver-relevant projection of the channel. The ablation results in Section IV-F demonstrate that no single term subsumes the others.

Let MM denote the mini-batch size and T=20T=20 the number of retained CIR taps within the common 400 ns400\text{\,}\mathrm{ns} delay window. We denote a simulated mini-batch by sM×Na×2\mathcal{B}_{s}\in\mathbb{R}^{M\times N_{\text{a}}\times 2} and an independently drawn measured mini-batch of the same size by r\mathcal{B}_{r}. For a CFR mini-batch \mathcal{B}, let P()M×TP(\mathcal{B})\in\mathbb{R}^{M\times T} be the collection of per-sample normalized PDPs obtained using the CIR and PDP definitions in Section II. The losses below compare mini-batch statistics rather than enforcing arbitrary one-to-one matching between simulated and measured samples.

Truncated PDP loss: The first term matches the average truncated PDP of the measured batch to that of the translated simulated batch as

PDP=1TP(r)¯P(GSR(s))¯1,\mathcal{L}_{\text{PDP}}=\frac{1}{T}\big\|\,\overline{P(\mathcal{B}_{r})}-\overline{P(G_{S\to R}(\mathcal{B}_{s}))}\,\big\|_{1}, (6)

where ()¯\overline{(\cdot)} denotes averaging over the mini-batch.

Sub-band PDP loss: To constrain local frequency-dependent delay behavior, the NaN_{\text{a}} active subcarriers are divided into K=10K=10 contiguous slices, each containing Na/K=150N_{\text{a}}/K=150 subcarriers, or approximately 3.66 MHz3.66\text{\,}\mathrm{MHz}. For each slice cc, we compute a 150-point IFFT and form a sub-band PDP Pc()P_{c}(\cdot) from the first TT taps. Since this sub-band transform has a tap spacing of 1/(3.66 MHz)273 ns1/($3.66\text{\,}\mathrm{MHz}$)\approx$273\text{\,}\mathrm{ns}$, the sub-band transform provides a coarser delay grid than the global 20 ns20\text{\,}\mathrm{ns} CIR. The resulting loss captures slower frequency-selective structure over a wider effective delay window and is given by

sPDP=1KTc=1KPc(r)¯Pc(GSR(s))¯1.\mathcal{L}_{\text{sPDP}}=\frac{1}{KT}\sum_{c=1}^{K}\big\|\,\overline{P_{c}(\mathcal{B}_{r})}-\overline{P_{c}(G_{S\to R}(\mathcal{B}_{s}))}\,\big\|_{1}. (7)

The factor 1/(KT)1/(KT) averages the absolute PDP discrepancy over the KK sub-bands and TT retained taps.

τrms\tau_{\text{rms}} moment loss: Let μτ()\mu_{\tau}(\cdot) and στ()\sigma_{\tau}(\cdot) denote the mini-batch mean and standard deviation of the per-sample τrms\tau_{\text{rms}} values defined in Eq. (3). Since these are scalar batch statistics, we penalize their mismatch using absolute differences as

τ\displaystyle\mathcal{L}_{\tau} =|μτ(r)μτ(GSR(s))|\displaystyle=\big|\mu_{\tau}(\mathcal{B}_{r})-\mu_{\tau}(G_{S\to R}(\mathcal{B}_{s}))\big|
+2|στ(r)στ(GSR(s))|.\displaystyle\quad+2\big|\sigma_{\tau}(\mathcal{B}_{r})-\sigma_{\tau}(G_{S\to R}(\mathcal{B}_{s}))\big|. (8)

Here, the standard-deviation term receives twice the weight of the mean term to emphasize the dispersion error identified in Section I, where the simulated twin exhibits a narrower capture-to-capture τrms\tau_{\text{rms}} distribution than the measured link.

τrms\tau_{\text{rms}} quantile loss: The moment loss controls only the first two summary statistics. To better align the full empirical distribution, especially the upper tail, we also match the sorted τrms\tau_{\text{rms}} values. Let 𝝉~()M\tilde{\boldsymbol{\tau}}(\mathcal{B})\in\mathbb{R}^{M} denote the vector of per-sample τrms\tau_{\text{rms}} values from \mathcal{B} in ascending order. We define

τq=1M𝝉~(r)𝝉~(GSR(s))1.\mathcal{L}_{\tau q}=\frac{1}{M}\big\|\tilde{\boldsymbol{\tau}}(\mathcal{B}_{r})-\tilde{\boldsymbol{\tau}}(G_{S\to R}(\mathcal{B}_{s}))\big\|_{1}. (9)

For equal-sized, equally weighted mini-batches, τq\mathcal{L}_{\tau q} equals the empirical 1-Wasserstein distance between the measured and generated τrms\tau_{\text{rms}} samples in Eq. (5). Fig. 3 reports this distance for the held-out pre-calibration data, whereas training evaluates it on independently drawn mini-batches.

CFR-magnitude NMSE loss: The delay-domain terms constrain the PDP and τrms\tau_{\text{rms}} statistics. However, they do not fully determine the average spectral profile across active subcarriers. We therefore include a frequency-domain magnitude loss. For a mini-batch \mathcal{B} represented by real and imaginary channels, let |||\mathcal{B}| denote the corresponding per-sample, per-subcarrier magnitude. The loss is

mag=|r|¯|GSR(s)|¯22|r|¯22.\mathcal{L}_{\text{mag}}=\frac{\big\|\overline{|\mathcal{B}_{r}|}-\overline{|G_{S\to R}(\mathcal{B}_{s})|}\big\|_{2}^{2}}{\big\|\overline{|\mathcal{B}_{r}|}\big\|_{2}^{2}}. (10)

Cycle and identity losses: Cycle and identity regularization keep the unpaired translator close to a physically meaningful correction rather than an unconstrained remapping. We define

cyc\displaystyle\mathcal{L}_{\text{cyc}} =GRS(GSR(s))s1\displaystyle=\|G_{R\to S}(G_{S\to R}(\mathcal{B}_{s}))-\mathcal{B}_{s}\|_{1}
+GSR(GRS(r))r1,\displaystyle\quad+\|G_{S\to R}(G_{R\to S}(\mathcal{B}_{r}))-\mathcal{B}_{r}\|_{1}, (11)
id\displaystyle\mathcal{L}_{\text{id}} =GSR(r)r1+GRS(s)s1.\displaystyle=\|G_{S\to R}(\mathcal{B}_{r})-\mathcal{B}_{r}\|_{1}+\|G_{R\to S}(\mathcal{B}_{s})-\mathcal{B}_{s}\|_{1}. (12)

The cycle-consistency loss requires a sample translated to the other domain and mapped back to reconstruct the original input. Cycle consistency therefore preserves sample-specific CFR structure and discourages many-to-one mappings. However, this loss constrains the compositions GRSGSRG_{R\to S}\circ G_{S\to R} and GSRGRSG_{S\to R}\circ G_{R\to S}, while the individual generators may still learn compensating transformations that cancel under composition. The identity loss reduces this ambiguity by regularizing each generator toward the identity mapping for inputs that already belong to its output domain. Identity regularization limits unnecessary amplitude and phase changes to CFRs that already exhibit the intended domain statistics. In the present static single-link setting, the bounded residual head provides much of the same input–output coupling even when the cycle and identity losses are removed, as quantified in Section IV-F. We nevertheless retain both losses as conservative sample-level regularizers.

The complete generator objective is

G\displaystyle\mathcal{L}_{G} =λcyccyc+λidid+λττ+λτqτq\displaystyle=\lambda_{\text{cyc}}\mathcal{L}_{\text{cyc}}+\lambda_{\text{id}}\mathcal{L}_{\text{id}}+\lambda_{\tau}\mathcal{L}_{\tau}+\lambda_{\tau q}\mathcal{L}_{\tau q}
+λPDPPDP+λsPDPsPDP+λmagmag,\displaystyle\quad+\lambda_{\text{PDP}}\mathcal{L}_{\text{PDP}}+\lambda_{\text{sPDP}}\mathcal{L}_{\text{sPDP}}+\lambda_{\text{mag}}\mathcal{L}_{\text{mag}}, (13)

where 𝝀\boldsymbol{\lambda} is the vector of scalar loss weights listed in Table II. All SC-SLA loss weights are selected once using the 95 GHz training/validation split and kept fixed for the in-domain test, cross-frequency transfer, and ablation experiments. The weighting follows a coarse validation calibration in which the cycle and identity terms provide conservative sample-level regularization, while the delay-domain terms receive larger weights to address the PDP and τrms\tau_{\text{rms}} mismatches discussed in Section I.

III-C Training Procedure

We implement the mappings as separate ResNet generators with independent parameter sets θSR\theta_{S\to R} and θRS\theta_{R\to S}. Consequently, we learn GRSG_{R\to S} directly through the overall objective rather than obtain it by inverting GSRG_{S\to R}. During each mini-batch, we evaluate both networks, and a common Adam optimizer [24] updates both parameter sets by minimizing G\mathcal{L}_{G}. The reverse generator receives gradients through the cycle and identity terms, while the forward generator is additionally constrained by the channel-statistics losses. Training is therefore bidirectional, whereas sim-to-lab inference uses only GSRG_{S\to R}.

At 95 GHz, 45,00045{,}000 simulated and measured captures are used for optimization, and 5,0005{,}000 captures are held-out from training. Both generators are trained for Nep=300N_{\mathrm{ep}}=300 epochs with a learning rate of 10410^{-4}, linear decay after Ndec=150N_{\mathrm{dec}}=150 epochs, BF16 automatic mixed precision, and an 2\ell_{2} gradient-norm clip of 1.01.0. Checkpoint selection, final evaluation, and the inference benchmark use FP32. Table II lists the system, RT, training, and model-selection parameters.

TABLE II: System, RT, training, and model-selection parameters.
Parameter Value
System Parameters
Carrier frequencies 92929595 GHz
IF / LO frequency (GHz) {2,3,4,5}\{2,3,4,5\} / 11.2511.25 (×8\times 8)
FFT size / active subcarriers 20482048 / 15001500
Cyclic-prefix length 512512 samples
Sampling bandwidth / Δf\Delta f 5050 MHz / 24.4124.41 kHz
Occupied bandwidth 36.636.6 MHz
CIR taps T=20T=20 at 2020 ns (400400 ns window)
TX–RX separation 1.751.75 m (LOS)
Antennas WR-10 pyramidal horns, co-polarized
Software-defined radio USRP-B200, 5050 MHz master clock
Receiver gain / SNR gate 7070 dB / 88 dB
Capture burst per carrier 1515 s
Captures per carrier (NsN_{s}, NrN_{r}) 50,00050{,}000 each
Training frequency 9595 GHz
Testing frequency 92929595 GHz
Ray-Traced CT Parameters
Path solver Sionna RT
Enabled interactions LOS, specular, refraction
Max interaction depth 55 (diffuse reflection disabled)
Antenna power pattern Cosine-power, nh=126.18n_{\mathrm{h}}=126.18, 1212^{\circ} 3-dB beamwidth
Timing jitter (std) 2.52.5 ns
Common phase jitter (std) 0.200.20 rad
Per-capture SNR 2828 dB mean, 22 dB std
SC-SLA Hyperparameters
Clipping quantile ρclip\rho_{\text{clip}} 0.9950.995
Sub-band count KK 1010
Residual blocks / channel widths 66 / 6464128128256256
Optimizer Adam (β1=0.5\beta_{1}=0.5, β2=0.999\beta_{2}=0.999)
Learning rate / decay start NdecN_{\mathrm{dec}} 10410^{-4} / epoch 150150
Training epochs NepN_{\mathrm{ep}} / batch size MM 300300 / 6464
Precision / gradient-norm clip BF16 / 1.01.0
λcyc,λid\lambda_{\text{cyc}},\ \lambda_{\text{id}} 10, 210,\ 2
λτ,λτq\lambda_{\tau},\ \lambda_{\tau q} 5, 35,\ 3
λPDP,λsPDP,λmag\lambda_{\text{PDP}},\ \lambda_{\text{sPDP}},\ \lambda_{\text{mag}} 15, 20, 115,\ 20,\ 1
95 GHz training samples 45,00045{,}000
Reserved samples 5,0005{,}000
Model-selection interval NselN_{\mathrm{sel}} every 55 epochs
JJ coefficients ωPDP,ωmag\omega_{\mathrm{PDP}},\ \omega_{\mathrm{mag}} 50, 0.2550,\ 0.25
Evaluation subsample 2,0002{,}000 captures per carrier
Algorithm 1 SC-SLA training
1:Input: Xs,Xr,𝝀,Nep,Nsel,ρclipX_{s},X_{r},\boldsymbol{\lambda},N_{\mathrm{ep}},N_{\mathrm{sel}},\rho_{\text{clip}}
2:Split XsX_{s} and XrX_{r} into optimization and reserved subsets
3:Initialize GSRG_{S\to R} and GRSG_{R\to S} with parameters θSR\theta_{S\to R} and θRS\theta_{R\to S}
4:Set qclipq_{\text{clip}} to the ρclip\rho_{\text{clip}}-quantile of the pooled 95 GHz magnitudes
5:Set JJ^{\star}\leftarrow\infty
6:for e=1,,Nepe=1,\dots,N_{\mathrm{ep}} do
7:  Form independently shuffled mini-batch index sets {πs(i)}\{\pi_{s}^{(i)}\} and {πr(i)}\{\pi_{r}^{(i)}\}
8:  for each mini-batch ii do
9:   sXs[πs(i)]\mathcal{B}_{s}\!\leftarrow\!X_{s}[\pi_{s}^{(i)}], rXr[πr(i)]\mathcal{B}_{r}\!\leftarrow\!X_{r}[\pi_{r}^{(i)}]
10:   Evaluate both mappings and compute G\mathcal{L}_{G} from Eqs. (6)–(13)
11:   Compute θSR,θRSG\nabla_{\theta_{S\to R},\theta_{R\to S}}\mathcal{L}_{G}
12:   Clip G21\|\nabla\mathcal{L}_{G}\|_{2}\!\leq\!1 and jointly update both generators with Adam
13:  end for
14:  if ee is divisible by NselN_{\mathrm{sel}} then
15:   Evaluate DKSD_{\mathrm{KS}}, NMSEPDP\operatorname{NMSE}_{\mathrm{PDP}}, and NMSEmag\operatorname{NMSE}_{\mathrm{mag}} on reserved captures
16:   Compute J=DKS+ωPDPNMSEPDPJ=D_{\mathrm{KS}}+\omega_{\mathrm{PDP}}\,\operatorname{NMSE}_{\mathrm{PDP}}
17:  +ωmagNMSEmag+\omega_{\mathrm{mag}}\,\operatorname{NMSE}_{\mathrm{mag}}
18:   if J<JJ<J^{\star} then
19:     JJJ^{\star}\leftarrow J and save the current GSRG_{S\to R} checkpoint
20:   end if
21:  end if
22:end for
23:return GSRG_{S\to R} checkpoint with minimum JJ

Checkpoint selection uses three lower-is-better metrics evaluated on reserved captures. Let g\mathcal{E}_{g} denote the output CFR set of a candidate model and r\mathcal{E}_{r} the measured reference set. Both are evaluation sets rather than the training mini-batches s\mathcal{B}_{s} and r\mathcal{B}_{r} of Section III-B. Let F^g(t)\widehat{F}_{g}(t) and F^r(t)\widehat{F}_{r}(t) denote the empirical cumulative distribution functions (CDFs) of the per-sample τrms\tau_{\text{rms}} values for g\mathcal{E}_{g} and r\mathcal{E}_{r}, respectively. Their two-sample KS statistic is

DKS=supt0|F^g(t)F^r(t)|.D_{\mathrm{KS}}=\sup_{t\geq 0}\left|\widehat{F}_{g}(t)-\widehat{F}_{r}(t)\right|. (14)

Here sup\sup denotes the supremum, that is, the least upper bound of |F^g(t)F^r(t)||\widehat{F}_{g}(t)-\widehat{F}_{r}(t)| over all delay-spread values t0t\geq 0. Both empirical CDFs are right-continuous step functions with finitely many jumps. The difference therefore takes finitely many values, and the supremum is attained as a maximum over the pooled τrms\tau_{\text{rms}} values of g\mathcal{E}_{g} and r\mathcal{E}_{r}. Accordingly, DKSD_{\mathrm{KS}} is the largest vertical gap between the two empirical CDFs and is bounded to [0,1][0,1]. The two NMSE metrics are

NMSEPDP\displaystyle\operatorname{NMSE}_{\mathrm{PDP}} =P(g)¯P(r)¯22P(r)¯22,\displaystyle=\frac{\big\|\overline{P(\mathcal{E}_{g})}-\overline{P(\mathcal{E}_{r})}\big\|_{2}^{2}}{\big\|\overline{P(\mathcal{E}_{r})}\big\|_{2}^{2}}, (15)
NMSEmag\displaystyle\operatorname{NMSE}_{\mathrm{mag}} =|g|¯|r|¯22|r|¯22.\displaystyle=\frac{\big\|\overline{|\mathcal{E}_{g}|}-\overline{|\mathcal{E}_{r}|}\big\|_{2}^{2}}{\big\|\overline{|\mathcal{E}_{r}|}\big\|_{2}^{2}}. (16)

The metric NMSEmag\operatorname{NMSE}_{\mathrm{mag}} has the same mathematical form as the training loss mag\mathcal{L}_{\text{mag}}, but it is evaluated on the reserved or held-out sets g\mathcal{E}_{g} and r\mathcal{E}_{r} rather than on training mini-batches. In contrast, NMSEPDP\operatorname{NMSE}_{\mathrm{PDP}} uses normalized squared 2\ell_{2} error between ensemble-mean PDPs, whereas PDP\mathcal{L}_{\text{PDP}} uses mean absolute error. The PDP metric therefore provides an in-objective consistency check without duplicating the training loss.

The three metrics form the scalar model-selection criterion

J=DKS+ωPDPNMSEPDP+ωmagNMSEmag.J=D_{\mathrm{KS}}+\omega_{\mathrm{PDP}}\,\operatorname{NMSE}_{\mathrm{PDP}}+\omega_{\mathrm{mag}}\,\operatorname{NMSE}_{\mathrm{mag}}. (17)

The fixed coefficients ωPDP=50\omega_{\mathrm{PDP}}=50 and ωmag=0.25\omega_{\mathrm{mag}}=0.25 weight the two NMSE terms relative to DKSD_{\mathrm{KS}} during checkpoint selection. They act only on the selection criterion and are distinct from the training-loss weights 𝝀\boldsymbol{\lambda} in Eq. (13). The coefficient ωPDP\omega_{\mathrm{PDP}} compensates for the small numerical range of NMSEPDP\operatorname{NMSE}_{\mathrm{PDP}}, whereas ωmag\omega_{\mathrm{mag}} limits the contribution of the numerically larger NMSEmag\operatorname{NMSE}_{\mathrm{mag}}. Both values are fixed once on the 95 GHz validation split and reused for every reported experiment. The statistic DKSD_{\mathrm{KS}} enters JJ as a sup-norm on the τrms\tau_{\text{rms}} CDF. No training term takes this form, so it scores the distribution targeted by τ\mathcal{L}_{\tau} and τq\mathcal{L}_{\tau q} under a different functional and penalizes checkpoints that match the moments and quantiles while leaving a localized gap in the CDF.

Every Nsel=5N_{\mathrm{sel}}=5 epochs, the checkpoint is evaluated on a fixed subset of reserved captures using the criterion JJ. The checkpoint with the lowest JJ is retained. The reported 95 GHz results are therefore in-domain, model-selected results on non-training captures. The 92–94 GHz carriers are used only for cross-frequency evaluation, not for optimization, model selection, or carrier-specific tuning. Algorithm 1 summarizes the joint generator optimization and checkpoint selection.

We conduct all simulations on an Ubuntu 24.04 LTS workstation equipped with an AMD Ryzen 9 9950X (16 cores, 32 threads, 5.7 GHz) processor, 128 GB RAM, and an NVIDIA RTX 4000 Ada Generation GPU with 20 GB memory. We implement SC-SLA in Python 3.12 with PyTorch 2.8.

IV Experiments and Result Analysis

IV-A Baselines and Evaluation Metrics

We compare SC-SLA with four supervised neural generators that use the same normalization method and the same zero-initialized residual-delta form, 𝐱^=𝐱+Δ𝐱\hat{\mathbf{x}}=\mathbf{x}+\Delta\mathbf{x}. The baselines include a fully connected neural network (FCNN) with four 512512-unit layers, LayerNorm, and dropout. The one-dimensional convolutional neural network (CNN1D) has six residual convolutional blocks and 128128 filters. The three-layer bidirectional long short-term memory network (BiLSTM) has 128128 hidden units. The depth-22 one-dimensional U-Net (UNet1D) [30] uses 6464 base channels.

The supervised baselines are constrained by the data collection process. Simulated and measured captures are triggered independently, so no physical one-to-one correspondence exists between a simulated CFR and a measured CFR. We therefore train the baselines using Huber regression [17] on index-aligned pairs, together with weak batch regularizers on CFR magnitude, sub-band magnitude, and τrms\tau_{\text{rms}} statistics. We use this Huber-based configuration as the strongest paired-supervision setting supported by the data and as a direct comparison with the unpaired SC-SLA objective. All models are selected using the same composite validation criterion, with family-specific weights chosen to place the constituent terms on comparable numerical scales. The three lower-is-better metrics are defined in Section III-C. They comprise DKSD_{\mathrm{KS}}, NMSEPDP\operatorname{NMSE}_{\mathrm{PDP}}, and NMSEmag\operatorname{NMSE}_{\mathrm{mag}}. The KS statistic is not directly optimized by any training loss. The PDP metric compares the same ensemble-mean PDPs as PDP\mathcal{L}_{\text{PDP}} but uses normalized squared 2\ell_{2} error and therefore serves mainly as an in-objective consistency check. The CFR-magnitude metric measures preservation of the average frequency-domain magnitude profile.

Refer to caption
Figure 5: In-domain calibration at 95 GHz on 2,0002{,}000 held-out captures. The panels report (a) the τrms\tau_{\text{rms}} KS statistic, (b) truncated PDP NMSE on a logarithmic axis, and (c) per-subcarrier CFR-magnitude NMSE across SC-SLA and the four supervised baselines. SC-SLA achieves the lowest error on all three metrics, with KS 0.04950.0495 and PDP NMSE 3.0×1053.0\times 10^{-5}.
Refer to caption
Figure 6: Cross-frequency truncated PDP NMSE at 92–94 GHz using the 95 GHz checkpoint without retraining. The 95 GHz point is the in-domain carrier. SC-SLA achieves the lowest PDP error at every held-out carrier, with 2.1×2.1\times, 1.7×1.7\times, and 5.0×5.0\times gains over the strongest supervised baseline at 92, 93, and 94 GHz, respectively.
Refer to caption
Figure 7: Cross-frequency per-subcarrier CFR-magnitude NMSE at 92–94 GHz using the 95 GHz checkpoint without retraining. The 95 GHz point is the in-domain carrier. SC-SLA preserves the magnitude profile under carrier shift and remains below every supervised baseline at each held-out carrier.

IV-B In-Domain Results at 95 GHz

Refer to caption
Figure 8: In-domain ablation at 95 GHz. The panels report (a) τrms\tau_{\text{rms}} KS statistic, (b) truncated PDP NMSE, and (c) per-subcarrier CFR-magnitude NMSE across the full model and six single-component variants. Bars are sorted within each panel. Panels (a) and (b) use logarithmic axes. No ablated variant dominates all three metrics, while the full model avoids severe degradation in any metric.
Refer to caption
Figure 9: Cross-frequency ablation using the 95 GHz checkpoint without retraining. The 95 GHz point is the in-domain held-out split. Panel (a) plots the τrms\tau_{\text{rms}} KS statistic and panel (b) plots truncated PDP NMSE versus carrier, both on logarithmic axes. Removing the τrms\tau_{\text{rms}} quantile loss or all delay-domain terms consistently worsens the τrms\tau_{\text{rms}} match, with the clearest separation at 94 GHz.

We first evaluate SC-SLA and the four supervised baselines on 2,0002{,}000 held-out captures at 9595 GHz. For context, the uncalibrated twin, denoted as “Sionna (identity)”, applies the impairment-augmented ray-traced output directly through Gid(𝐱)=𝐱G_{\mathrm{id}}(\mathbf{x})=\mathbf{x} for an input CFR sample 𝐱\mathbf{x}. Fig. 3 shows that this input has a large τrms\tau_{\text{rms}}-distribution mismatch, with DKS=0.861D_{\mathrm{KS}}=0.861. Thus, the impairment-augmented twin alone does not reproduce the measured delay-domain statistics.

The supervised baselines reduce the τrms\tau_{\text{rms}} KS statistic to 0.08900.08900.11800.1180, as shown in Fig. 5. However, CNN1D and BiLSTM yield CFR-magnitude NMSE values of 0.57080.5708 and 0.52990.5299, respectively. This spectral distortion is consistent with the limitation of index-aligned supervision because the independently triggered simulated and measured captures are not physically paired. SC-SLA achieves the lowest error across all three metrics in Fig. 5. It reduces the KS statistic to 0.04950.0495, a 44%44\% reduction from the 0.08900.0890 obtained by UNet1D, the best supervised baseline on this metric. Furthermore, SC-SLA attains a CFR-magnitude NMSE of 0.10770.1077 and a PDP NMSE of 3.0×1053.0\times 10^{-5}. The latter is more than two orders of magnitude below every supervised baseline. Because PDP NMSE compares the same ensemble-mean PDPs as PDP\mathcal{L}_{\text{PDP}} using a different error norm, it serves as an in-objective consistency check rather than a fully independent metric.

IV-C Cross-Frequency Generalization at 92–94 GHz

We next examine whether the calibration learned at 9595 GHz transfers to nearby W-band carriers. For this experiment, we apply the same 9595 GHz checkpoint directly to Sionna inputs at 9292, 9393, and 9494 GHz, without retraining, fine-tuning, or carrier-specific model selection. We then compare the calibrated outputs with 2,0002{,}000 measured captures at each carrier.

The PDP NMSE increases when the model is evaluated away from the training carrier, rising from the in-domain value of 3.0×1053.0\times 10^{-5} to the 10310^{-3} range. This degradation is expected because no 92929494 GHz data are used during training. Even under this shift, SC-SLA achieves the lowest PDP NMSE at every held-out carrier. Compared with the strongest supervised baseline at each frequency, SC-SLA improves the PDP NMSE by 2.1×2.1\times, 1.7×1.7\times, and 5.0×5.0\times at 9292, 9393, and 9494 GHz, respectively (Fig. 6). SC-SLA also achieves the lowest CFR-magnitude NMSE across all three held-out carriers (Fig. 7).

The τrms\tau_{\text{rms}} results follow the same overall trend, although the margin varies by carrier. SC-SLA keeps the τrms\tau_{\text{rms}} KS statistic below that of the uncalibrated Sionna input, with values of 0.2100.210, 0.2240.224, and 0.08450.0845 at 9292, 9393, and 9494 GHz, respectively. The identity twin yields larger KS values of 0.6680.668, 0.7520.752, and 0.9110.911 at the same carriers. The only exception among the baselines occurs at 9393 GHz, where FCNN gives a slightly lower KS value of 0.2010.201. However, FCNN’s lower 93 GHz KS is accompanied by worse PDP and CFR-magnitude errors and therefore does not represent the same overall calibration quality. Because the 92929494 GHz datasets are not used for training, model selection, or carrier-specific tuning, these results provide direct evidence of zero-shot cross-frequency generalization to three adjacent W-band carriers. They also constitute an unseen-carrier stress test. The test remains limited to the measured 92929595 GHz span and the present static LOS scene. Denser multipath, non-line-of-sight (NLOS) propagation, and geometry changes require corresponding RT and measured datasets.

Three complementary mechanisms support the observed stability of SC-SLA under carrier shift, and the ablation in Section IV-F isolates each of them. The τrms\tau_{\text{rms}} quantile loss aligns the upper tail of the τrms\tau_{\text{rms}} distribution, rather than only its mean and variance. The tanh\tanh-bounded residual head keeps the translated CFR within the shared normalized channel range, limiting unstable extrapolation at unseen carriers. The sub-band PDP loss further encourages preservation of delay structure over local frequency regions, making the correction less dependent on the exact 9595 GHz training carrier.

IV-D Inference Scalability

Cross-frequency transfer avoids carrier-specific retraining. However, the translated CFRs must also be generated at the scale of an RT dataset. Table III therefore reports post-RT calibration time for subsets of the 50,00050{,}000-sample 95 GHz Sionna dataset. The deployed GSRG_{S\to R} contains 2.622.62 million parameters and sustains approximately 5.1×1035.1\times 10^{3} CFRs/s across the evaluated dataset sizes. It processes the complete set in 9.75±0.099.75\pm 0.09 s, corresponding to an amortized inference time of 0.1950.195 ms per CFR. Joint training incorporates two generators with 5.245.24 million parameters in total, whereas inference retains only GSRG_{S\to R}. The timing includes host-to-device transfer, one generator pass, and output return. The reported timing excludes RT generation, model loading, normalization, and metric computation.

TABLE III: FP32 inference scalability of GSRG_{S\to R} on an RTX 4000 Ada GPU with batch size 64 and 1,500 complex64 subcarriers per CFR. Runtime is the mean ±\pm standard deviation over five runs.
CFRs Input (MB) Parameters (10610^{6}) Runtime (s) Throughput (CFRs/s)
2,0002{,}000 2424 2.6182.618 0.380±0.0010.380\pm 0.001 5,2575{,}257
10,00010{,}000 120120 2.6182.618 1.910±0.0051.910\pm 0.005 5,2355{,}235
50,00050{,}000 600600 2.6182.618 9.751±0.0859.751\pm 0.085 5,1275{,}127

IV-E Sensitivity to Measured Training Data

Practical calibration also depends on how much measured CSI is available for training. We therefore retrain SC-SLA with 11,25011{,}250, 22,50022{,}500, and 33,75033{,}750 measured 95 GHz CFRs, corresponding to nested 25%25\%, 50%50\%, and 75%75\% subsets of the measured training set. All 45,00045{,}000 simulated training CFRs and the fixed 5,0005{,}000-sample reserved split remain unchanged. Each variant is initialized independently and trained once with seed 12341234 and the original 300-epoch schedule, and every epoch contains 150 optimizer updates, so all fractions receive the same number of gradient steps. The loss weights and checkpoint-selection criterion are unchanged, the normalization parameters remain those estimated from the complete 95 GHz dataset, and the 100%100\% row reuses the full-data checkpoint. This experiment therefore isolates the number of measured CFRs entering optimization rather than the total measurement-acquisition requirement.

Table IV reports the results. At the 95 GHz training carrier, the reduced-data models attain τrms\tau_{\text{rms}} KS values of 0.02500.02500.03500.0350, below the 0.04950.0495 of the full-data model, and CFR-magnitude NMSE of 0.10080.10080.10670.1067, within 0.00690.0069 of the full-data 0.10770.1077. PDP NMSE is the most sensitive in-domain metric at 1.341.343.493.49 times the full-data value of 3.02×1053.02\times 10^{-5}, although even its worst case remains more than 5050 times below the uncalibrated error.

TABLE IV: Measured-data sensitivity at 95 GHz and under zero-shot transfer to 92–94 GHz. Lower values indicate better agreement. The 100%100\% rows reproduce the full-data SC-SLA results.
(a) In-domain evaluation at 95 GHz
Measured CFRs Fraction (%) τrms\tau_{\text{rms}} KS PDP NMSE CFR magnitude NMSE
11,25011{,}250 2525 0.03500.0350 1.05×1041.05\times 10^{-4} 0.10550.1055
22,50022{,}500 5050 0.02950.0295 4.05×1054.05\times 10^{-5} 0.10080.1008
33,75033{,}750 7575 0.02500.0250 6.73×1056.73\times 10^{-5} 0.10670.1067
45,00045{,}000 100100 0.04950.0495 3.02×1053.02\times 10^{-5} 0.10770.1077
(b) Zero-shot mean at 92–94 GHz
Measured CFRs Fraction (%) τrms\tau_{\text{rms}} KS PDP NMSE CFR magnitude NMSE
11,25011{,}250 2525 0.19350.1935 2.13×1032.13\times 10^{-3} 0.07260.0726
22,50022{,}500 5050 0.21120.2112 5.36×1035.36\times 10^{-3} 0.07020.0702
33,75033{,}750 7575 0.29020.2902 4.46×1034.46\times 10^{-3} 0.06830.0683
45,00045{,}000 100100 0.17280.1728 5.29×1035.29\times 10^{-3} 0.07080.0708

Cross-frequency behavior is less uniform. The mean τrms\tau_{\text{rms}} KS at 92–94 GHz is 0.19350.1935, 0.21120.2112, and 0.29020.2902 for the 25%25\%, 50%50\%, and 75%75\% models, against 0.17280.1728 for the full-data model and 0.77650.7765 for the uncalibrated twin. Mean PDP NMSE spans 2.13×1032.13\times 10^{-3}5.36×1035.36\times 10^{-3}, compared with 5.29×1035.29\times 10^{-3}, and mean CFR-magnitude NMSE stays within 0.00250.0025 of the full-data 0.07080.0708. The 75%75\% model yields the largest mean KS but the lowest mean CFR-magnitude NMSE, so sensitivity to the measured-data fraction depends on which channel statistic is evaluated. Under the fourfold reduction, the 25%25\% model still holds in-domain KS below the full-data value and differs from it by 0.02070.0207 in unseen-carrier mean KS. Each fraction is nevertheless represented by a single run, so these results do not establish a monotonic data-scaling relationship or across-seed uncertainty.

IV-F Loss Function Ablation Study

We retrain six SC-SLA variants to isolate the role of the main loss and architecture components. Each variant changes one ingredient by removing one channel-statistics loss, removing all four delay-domain losses jointly (only mag), removing cycle and identity regularization, or replacing the tanh\tanh-bounded head with a linear residual head, 𝐱^=𝐱+Δ𝐱\hat{\mathbf{x}}=\mathbf{x}+\Delta\mathbf{x}. All variants use the same seed, training schedule, and checkpoint-selection rule as the full model. Because the study uses a single seed, we focus on clear multi-fold differences and avoid overinterpreting small KS changes below roughly 0.020.02.

The ablation results demonstrate that the delay and frequency-domain objectives are complementary. Removing the τrms\tau_{\text{rms}} quantile loss increases the in-domain KS statistic by 3.6×3.6\times (0.04950.17750.0495\rightarrow 0.1775) and worsens the 9494 GHz KS by 2.6×2.6\times, while PDP and magnitude errors remain nearly unchanged. These multi-fold KS increases indicate that the quantile loss primarily controls the upper tail of the τrms\tau_{\text{rms}} distribution. In contrast, the only mag variant gives the lowest CFR-magnitude NMSE in Fig. 8 (0.09990.0999), but leaves the delay domain largely uncorrected, with KS 0.6940.694. Removing the magnitude loss produces the opposite failure mode. The delay-domain terms remain active, but the spectral profile drifts and the CFR-magnitude NMSE rises to 0.16080.1608, worse than the uncalibrated input.

The architecture ablations clarify the source of sample-level input–output coupling. A linear residual head fits the training carrier reasonably well, with in-domain KS 0.04250.0425, but generalizes poorly. The 9494 GHz KS more than doubles from 0.08450.0845 to 0.18750.1875, and the in-domain PDP NMSE increases by 5.7×5.7\times. Thus, the tanh\tanh head mainly stabilizes the correction under carrier shift rather than improving the 95 GHz fit alone. Removing the cycle and identity losses does not cause collapse. On 512512 held-out inputs, the mean input–output correlation remains 0.820.82, compared with 0.750.75 for the full model. These correlation values indicate that, in the present static single-link setting, the bounded residual form provides most of the input–output coupling. We retain cycle and identity regularization because it is conservative and may become more important when geometry varies across samples.

The sub-band PDP loss contributes mainly to cross-frequency robustness. Removing it slightly improves some in-domain metrics, but worsens the 9494 GHz KS statistic by 43%43\% (0.08450.12050.0845\rightarrow 0.1205) and also degrades KS at 9292 and 9393 GHz. Across Figs. 8 and 9, no ablated model is uniformly superior. Each variant improves one metric at the cost of another. The full configuration is therefore not the per-metric optimum, but it provides the most balanced operating point across τrms\tau_{\text{rms}}, PDP shape, spectral magnitude, and carrier transfer.

V Conclusion

In this paper, we presented SC-SLA, a non-adversarial, cycle-consistent framework that calibrates a ray-traced W-band CT against unpaired testbed measurements by aligning receiver-relevant OFDM channel statistics. These statistics include the PDP, sub-band PDP, τrms\tau_{\text{rms}} distribution, and per-subcarrier CFR magnitude. On 2,0002{,}000 held-out captures at 9595 GHz, SC-SLA reduced the τrms\tau_{\text{rms}} KS statistic of the uncalibrated twin from 0.860.86 to 0.04950.0495, improved on the strongest of four supervised baselines by 44%44\%, and achieved the lowest PDP and CFR-magnitude NMSE among all evaluated models. The same checkpoint transferred zero-shot to 92929494 GHz, providing evidence of cross-frequency calibration over the measured 33 GHz span using training measurements from 95 GHz only. A fourfold reduction of the measured training set to 11,25011{,}250 CFRs preserved calibration quality under a fixed seed, yielding KS 0.03500.0350 at 95 GHz and mean KS 0.19350.1935 across the three unseen carriers. The ablation study demonstrated that the τrms\tau_{\text{rms}} quantile loss governs the upper tail of the τrms\tau_{\text{rms}} distribution, the delay- and frequency-domain objectives are jointly necessary, and the tanh\tanh-bounded residual head stabilizes the correction under carrier shift. The deployed generator calibrates 50,00050{,}000 CFRs in 9.75±0.099.75\pm 0.09 s, corresponding to approximately 5.1×1035.1\times 10^{3} CFRs/s. Future work could repeat the sensitivity and ablation studies across seeds, quantify sensitivity to the impairment parameters, extend the captures beyond the present 50 MHz50\text{\,}\mathrm{MHz} sampling bandwidth, and address NLOS, geometry-varying, and multi-antenna deployments.

References

  • [1] 3GPP (2022-03) Study on channel model for frequencies from 0.5 to 100 GHz. Tech. Rep. Technical Report TR 38.901, V17.0.0, 3GPP. Cited by: §I.
  • [2] 3GPP (2021-03) Study on supporting NR from 52.6 GHz to 71 GHz. Tech. Rep. Technical Report TR 38.808, V17.0.0, 3GPP. Cited by: §I.
  • [3] G. Aceto, V. Persico, and A. Pescapé (2019) A survey on information and communication technologies for industry 4.0: state-of-the-art, taxonomies, perspectives, and challenges. IEEE Communications Surveys & Tutorials 21 (4), pp. 3467–3501. Cited by: §I.
  • [4] J. Ai, S. Xu, Y. Ren, Z. Liu, W. Chen, W. Tang, X. Li, C. Wen, and S. Jin (2026) Propagation-consistent wireless environment digital twin construction under sparse measurements. arXiv preprint arXiv:2605.22361. Cited by: §I-A, TABLE I.
  • [5] A. Alkhateeb, S. Jiang, and G. Charan (2023) Real-time digital twins: vision and research directions for 6G and beyond. IEEE Commun. Mag. 61 (11), pp. 128–134. Cited by: §I.
  • [6] N. B. Baytekin, R. Wiesmayr, S. Cammerer, C. Dick, and C. Studer (2026) Site-specific finetuning of neural receivers with real-world 5G NR measurements. Note: arXiv:2603.09644 Cited by: §I-A.
  • [7] B. Böck, A. Oeldemann, T. Mayer, F. Rossetto, and W. Utschick (2025) Physics-informed generative modeling of wireless channels. In Proc. Int. Conf. Mach. Learn. (ICML), PMLR, Vol. 267, pp. 4602–4626. Cited by: §I-A.
  • [8] Eravant Eravant Millimeter Wave Products & Components. Note: https://www.eravant.com/Accessed: 2026-07-02 Cited by: §II-B.
  • [9] F. Euchner, J. Sanzi, M. Henninger, and S. ten Brink (2024) GAN-based massive MIMO channel model trained on measured data. In Proc. 27th Int. ITG Workshop Smart Antennas (WSA), pp. 109–116. External Links: Document Cited by: §I-A.
  • [10] M. Grieves and J. Vickers (2016) Digital twin: mitigating unpredictable, undesirable emergent behavior in complex systems. In Transdisciplinary Perspectives on Complex Systems, pp. 85–113. Cited by: §I.
  • [11] M. Haider, I. Ahmed, Z. Hassan, T. J. O’Shea, L. Liu, and D. B. Rawat (2025-07) Digital twin enabled site specific channel precoding: over the air CIR inference. IEEE Commun. Lett. 29 (7), pp. 1559–1563. External Links: Document Cited by: §I-A, TABLE I.
  • [12] C. Han, Y. Wu, Y. Li, et al. (2022) Terahertz wireless channels: a holistic survey on measurement, modeling, and analysis. IEEE Commun. Surveys Tuts. 24 (3), pp. 1670–1707. Cited by: §I.
  • [13] K. He, X. Zhang, S. Ren, and J. Sun (2016) Deep residual learning for image recognition. In Proc. IEEE Conf. Comput. Vis. Pattern Recognit. (CVPR), pp. 770–778. Cited by: §III-A.
  • [14] J. Hoydis, F. Ait Aoudia, S. Cammerer, et al. (2023) Sionna RT: differentiable ray tracing for radio propagation modeling. Note: arXiv:2303.11103 Cited by: §I, §II-C.
  • [15] J. Hoydis, S. Cammerer, F. Ait Aoudia, et al. (2022) Sionna: an open-source library for next-generation physical-layer research. Note: arXiv:2203.11854 Cited by: §I.
  • [16] Z. Hu, Y. Li, and C. Han (2024) Transfer learning enabled transformer-based generative adversarial networks for modeling and generating terahertz channels. Commun. Eng. 3, pp. 153. Cited by: §I-A.
  • [17] P. J. Huber (1964) Robust estimation of a location parameter. The Annals of Mathematical Statistics 35 (1), pp. 73–101. External Links: Document Cited by: §IV-A.
  • [18] ITU-R (2023-11) Framework and overall objectives of the future development of IMT for 2030 and beyond. Recommendation Technical Report M.2160-0, ITU-R. Cited by: §I.
  • [19] S. Jiang, Q. Qu, X. Pan, A. K. Agrawal, R. Newcombe, and A. Alkhateeb (2025) Learnable wireless digital twins: reconstructing electromagnetic field with neural representations. IEEE Open Journal of the Communications Society 6, pp. 1568–1590. External Links: Document Cited by: §I-A, TABLE I.
  • [20] J. Johnson, A. Alahi, and L. Fei-Fei (2016) Perceptual losses for real-time style transfer and super-resolution. In Proc. Eur. Conf. Comput. Vis. (ECCV), pp. 694–711. Cited by: §III-A.
  • [21] J. Joo, M. Park, D. S. Han, and V. Pejović (2019) Deep learning-based channel prediction in realistic vehicular communications. IEEE Access 7, pp. 27846–27858. Cited by: §I-A.
  • [22] S. Ju, Y. Xing, O. Kanhere, and T. S. Rappaport (2021-06) Millimeter wave and sub-terahertz spatial statistical channel model for an indoor office building. IEEE J. Sel. Areas Commun. 39 (6), pp. 1561–1575. Cited by: §I-A.
  • [23] L. U. Khan, W. Saad, D. Niyato, Z. Han, and C. S. Hong (2022) Digital-twin-enabled 6G: vision, architectural trends, and future directions. IEEE Commun. Mag. 60 (1), pp. 74–80. Cited by: §I.
  • [24] D. P. Kingma and J. Ba (2015) Adam: a method for stochastic optimization. In Proc. Int. Conf. Learn. Represent. (ICLR), Cited by: §III-C.
  • [25] H. Luo, S. R. Khosravirad, and A. Alkhateeb (2026) Wireless digital twin calibration: refining DFT-domain channel information. arXiv preprint arXiv:2603.16126. Cited by: §I-A, TABLE I.
  • [26] N. H. Mahmood, G. Berardinelli, E. J. Khatib, R. Hashemi, C. De Lima, and M. Latva-aho (2023) A functional architecture for 6G special-purpose industrial iot networks. IEEE Transactions on Industrial Informatics 19 (3), pp. 2530–2540. Cited by: §I.
  • [27] H. Mi, B. Ai, R. He, et al. (2024) Measurement-based prediction of mmWave channel parameters using deep learning and point cloud. IEEE Open J. Veh. Technol. 5, pp. 1059–1072. External Links: Document Cited by: §I-A.
  • [28] T. S. Rappaport, Y. Xing, O. Kanhere, et al. (2019-06) Wireless communications and applications above 100 GHz: opportunities and challenges for 6G and beyond. IEEE Access 7, pp. 78729–78757. Cited by: §I-A.
  • [29] Remcom, Inc. (2024) Wireless InSite: 3D wireless propagation software. Note: https://www.remcom.com/wireless-insite-em-propagation-softwareState College, PA Cited by: §I.
  • [30] O. Ronneberger, P. Fischer, and T. Brox (2015) U-Net: convolutional networks for biomedical image segmentation. In Proc. Med. Image Comput. Comput.-Assist. Intervent. (MICCAI), pp. 234–241. Cited by: §IV-A.
  • [31] C. Ruah, H. Sifaou, O. Simeone, and B. M. Al-Hashimi (2025) How to bridge the sim-to-real gap in digital twin-aided telecommunication networks. arXiv preprint arXiv:2507.07067. Cited by: §I-A.
  • [32] C. Ruah, O. Simeone, J. Hoydis, and B. M. Al-Hashimi (2024) Calibrating wireless ray tracing for digital twinning using local phase error estimates. IEEE Transactions on Machine Learning in Communications and Networking 2, pp. 1193–1215. External Links: Document Cited by: §I-A, TABLE I.
  • [33] R. Schulpen, U. Johannsen, A. Smolders, and L. A. Bronckers (2023-03) Ambiguity in RMS delay spread of millimeter-wave channel measurements. In Proc. Eur. Conf. Antennas Propag. (EuCAP), Cited by: §II-A.
  • [34] M. K. Shehzad, L. Rose, S. Wesemann, and M. Assaad (2022-04) ML-based massive MIMO channel prediction: does it work on real-world data?. IEEE Wireless Commun. Lett. 11 (4), pp. 811–815. Cited by: §I-A.
  • [35] J. Song, R. He, M. Yang, Z. Zhang, X. Chen, X. Zhang, and B. Ai (2025) A novel site-specific inference model for urban canyon channels: from measurements to modeling. Note: arXiv:2509.19275 Cited by: §I-A.
  • [36] Z. Tao, W. Xu, Y. Huang, X. Wang, and X. You (2024) Wireless network digital twin for 6G: generative AI as a key enabler. IEEE Wireless Commun. 31 (4), pp. 24–31. Cited by: §I.
  • [37] P. Tarafder, I. Ahmed, D. B. Rawat, Z. Hassan, and K. Hasan (2025) Digital-twin empowered site-specific radio resource management in 5G aerial corridor. In MILCOM 2025-2025 IEEE Military Communications Conference (MILCOM), pp. 1–6. Cited by: §II-C.
  • [38] P. Tarafder, Z. Hassan, I. Ahmed, D. B. Rawat, K. Hasan, and C. Pu (2026) Digital-twin empowered deep reinforcement learning for site-specific radio resource management in NextG wireless aerial corridor. Note: arXiv:2602.03801 Cited by: §I, §II-C.
  • [39] Y. Tian, H. Li, Q. Zhu, K. Mao, F. Ali, X. Chen, and W. Zhong (2024-04) Generative network-based channel modeling and generation for air-to-ground communication scenarios. IEEE Commun. Lett. 28 (4), pp. 892–896. Cited by: §I-A.
  • [40] D. Ulyanov, A. Vedaldi, and V. Lempitsky (2016) Instance normalization: the missing ingredient for fast stylization. arXiv preprint arXiv:1607.08022. Cited by: §III-A.
  • [41] C. Villani (2009) Optimal transport: old and new. Springer, Berlin, Germany. Cited by: §II-C.
  • [42] C. Wang, J. Huang, H. Wang, et al. (2020-12) 6G wireless channel measurements and models: trends and challenges. IEEE Veh. Technol. Mag. 15 (4), pp. 22–32. Cited by: §I.
  • [43] J. Wang, J. Zhang, Y. Sun, Y. Zhang, T. Jiang, and L. Xia (2025) Electromagnetic wave property inspired radio environment knowledge construction and artificial intelligence based verification for 6G digital twin channel. Frontiers Inf. Technol. Electron. Eng. 26 (2), pp. 260–277. Cited by: §I.
  • [44] Y. Wu, K. Zhang, and Y. Zhang (2021) Digital twin networks: a survey. IEEE Internet Things J. 8 (18), pp. 13789–13804. Cited by: §I.
  • [45] M. Ying, D. Shakya, P. Ma, G. Qian, and T. S. Rappaport (2026) Site-specific location calibration and validation of ray-tracing simulator NYURay at upper mid-band frequencies. npj Wireless Technology 2, pp. 8. External Links: Document Cited by: §I-A, TABLE I.
  • [46] Z. Yun and M. F. Iskander (2015) Ray tracing for radio propagation modeling: principles and applications. IEEE Access 3, pp. 1089–1100. Cited by: §I.
  • [47] J. Zhu, T. Park, P. Isola, and A. A. Efros (2017) Unpaired image-to-image translation using cycle-consistent adversarial networks. In Proc. IEEE Int. Conf. Comput. Vis. (ICCV), pp. 2242–2251. Cited by: §III-A, §III-A, §III.
  • [48] M. Zhu, L. Cazzella, F. Linsalata, M. Magarini, M. Matteucci, and U. Spagnolini (2024) Toward real-time digital twins of EM environments: computational benchmark for ray launching software. IEEE Open J. Commun. Soc. 5, pp. 6291–6302. Cited by: §I.