arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:2605.23682v2 [eess.SP] 29 Jul 2026

Tri-Domain Multiuser MIMO Precoding Optimization and Channel Estimation with Spatial-EM Reconfigurable Antenna

Yining Li, Ziwei Wan, Zhen Gao, , Keke Ying, Lipeng Zhu, , and Rui Zhang
Abstract

In this paper, we propose a tri-domain reconfigurable multiuser multiple-input multiple-output (MIMO) communication system that integrates the electromagnetic (EM) reconfigurable antenna (EMRA) with the spatially movable antenna (SMA), termed the spatial-EM reconfigurable antenna (SEMRA). The proposed system offers EM, spatial, and digital domain degrees of freedom (DoFs) for joint channel reconfiguration, yet introduces new challenges in channel estimation (CE) and precoding optimization. Specifically, for a multiuser orthogonal frequency division multiplexing (OFDM) downlink, the precoding design is formulated as a tri-domain optimization problem over antenna positions, EM-domain radiation-pattern weights, and digital precoders. We first develop a zero-forcing (ZF)-based baseline algorithm to decouple the design of spatial reconfiguration, and then propose a weighted minimum mean square error (WMMSE)-based tri-domain joint optimization algorithm to further improve the spectral efficiency (SE). Furthermore, we propose a low-overhead movement-aided channel estimation scheme in which coordinated antenna repositioning across pilot slots synthesizes a denser virtual array, enabling more accurate angle-of-departure (AoD) estimation and EM-domain channel state information (eCSI) reconstruction under the same per-user pilot overhead as the EMRA baseline. The resulting parametric representation enables eCSI assembly at desired antenna positions without additional pilots. Simulation results show that the proposed CE scheme improves eCSI estimation accuracy and the proposed SEMRA achieves higher SE than the EMRA baseline under the same pilot overhead.

I Introduction

wbsphack @@writeaux“newlabelS1wcurrentlabel1wesphack

Future wideband multiuser wireless systems require higher spectral efficiency (SE) under increasingly tight hardware constraints [Wang2024COMSTXLMIMO, Gong2024COMSTHMIMO]. In traditional fixed-antenna (TFA) settings, performance is improved mainly through digital precoding and array scaling, whereas antenna locations and radiation patterns remain largely fixed after fabrication. This limitation has motivated growing interest in flexible hardware architectures that introduce new physical-layer degrees of freedom (DoFs). Prior work on reconfigurable intelligent surfaces (RIS) has demonstrated the promise of programmable propagation environments [DiRenzo2020JSACRIS], yet its passive nature also brings cascaded-channel acquisition and double-fading penalties, complicating practical deployment [Wu2021TCOM]. On the other hand, active transceiver architectures [Krikidis2024TCOMRAMIMO, Wong2022TWCFAMA], such as multiple-input multiple-output (MIMO) systems with electromagnetic reconfigurable antenna (EMRA), have shown substantial SE improvement without extra antennas or power-hungry radio-frequency (RF) chains. More recently, spatially movable antenna (SMA) architectures, including movable antenna (MA) [Zhu2024TWC, Zhu2025COMSTTutorialMA] and fluid antenna systems (FAS) [Wong2022TWCFAMA, New2025COMST], have emerged to unlock additional spatial DoFs for MIMO communications. Although EMRA and SMA have each attracted much attention, their integration remains largely underexplored [Ma2026COMSTSurveyRMA].

I-A Prior Work

For EMRA architectures, one appealing benefit is the ability to adapt the radiation field [Hasan2018TWC, Bahceci2017TWC, Wang2023TWC]. By electronically tuning loads or switch states, the current distribution of each antenna element can be changed, thereby reshaping its radiation pattern. At the system level, this effect can be modeled through antenna-mode selection or basis-function expansions whose coefficients serve as electromagnetic (EM)-domain precoders [Zhao2021TWCMode, Wang2023TWC, Ying2025TCOM]. In this setting, transmitter design depends not only on conventional spatial-domain channel state information (sCSI), but also on EM-domain channel state information (eCSI), which links the propagation geometry to the radiation-pattern basis and enables joint EM-domain and digital-domain precoding for substantial SE improvement over TFA systems [Wang2023TWC, Ying2025TCOM].

One core challenge for EMRA systems is channel estimation (CE). Recent studies have also examined channel state information (CSI) extrapolation across radiation states [Liang2024TVTExtrap]. In wideband parametric CE pipelines, the channel is first characterized by multipath delays, angle-of-departure (AoD), angle-of-arrival (AoA), and equivalent path gains, and the eCSI required for EM-domain design is then reconstructed from these geometric parameters [Bahceci2017TWC, Gao2014LCOMM, Liao2019TCOM, Ying2025TCOM]. Although this representation reduces the estimation dimension, its accuracy still hinges on spatial sampling. In orthogonal frequency division multiplexing (OFDM) systems, frequency diversity mainly assists delay estimation, whereas reliable AoD recovery still requires sufficiently rich spatial observations and is further affected by spatial-wideband and frequency-wideband effects [Wang2018TSPWideband]. For fixed-geometry arrays, simply enlarging the inter-element spacing is not always effective. Although a larger aperture can improve angular resolution, excessive spacing may also introduce spatial aliasing and grating-lobe ambiguity [Gao2014LCOMM, Liao2019TCOM]. Since AoD errors directly perturb the steering and basis matrices used to assemble eCSI, they propagate to EM-domain precoding and degrade the subsequent transmission design [Costa2010TSPSpherical, Nadeem2015TSP3D, Ying2025TCOM].

Work on SMA architectures now spans channel modeling, antenna-position optimization, and multiuser precoding for MA and FAS systems [Zhu2024TWC, Yang2025TSPFWMMSE, Xiao2024TWCJointPositioning, Zhu2023LCOMMNullSteering], and recent surveys provide broader accounts of these developments [Zhu2025COMSTTutorialMA, New2025COMST]. Antenna translation can also be combined with rotation to provide position and orientation DoFs [Ma2026COMSTSurveyRMA]. Linear minimum mean square error and Bayesian estimators have been studied for CE [Skouroumounis2023TCOM, Zhang2025TWC]. Compressed sensing (CS) methods recover AoDs, AoAs, and path gains from measurements at selected MA positions [Ma2023LCOMMCSMA, Xiao2024TWCCSMA]; with an EM radiation basis, these parameters can be used to assemble eCSI. For these methods, the angular estimates depend on the grid resolution and sensing geometry. In the proposed design, coordinated motion forms a uniform virtual array for two-dimensional estimation of signal parameters via rotational invariance techniques (ESPRIT), and the recovered parameters are used to assemble eCSI at the transmission positions.

Recent studies combine spatial and EM reconfiguration through electronically synthesized radiation locations [Ma2026COMSTSurveyRMA]. Pixel-based reconfigurable antennas emulate FAS ports on a fixed aperture [Zhang2025OJAPPRAFAS], and the reconfigurable pixel antenna (RPA)-based electronic movable-antenna array (REMAA) selects from a finite set of candidate radiation positions for multiuser transmission [Chen2025TCOMREMAA]. Radiation-center reconfigurable arrays similarly combine discrete radiation-center selection with analog and digital beamforming [Li2026TWCRCRAA]. These architectures provide electronic spatial flexibility through effective positions or radiation centers on a fixed aperture.

Physical repositioning directly controls the sampling coordinates and supplies a regular observation geometry for virtual-array CE. The EM coefficients remain available for radiation-pattern adaptation during data transmission, so sensing and transmission retain separate control variables [Zhang2025OJAPPRAFAS, Chen2025TCOMREMAA, Li2026TWCRCRAA].

The spatial-EM reconfigurable antenna (SEMRA) architecture considered here coordinates antenna repositioning across pilot slots to form a denser virtual array under the same per-user pilot overhead as EMRA. The resulting parametric eCSI model supports tri-domain precoding over antenna positions, EM-domain pattern weights, and digital precoders.

I-B Contributions

This paper presents a SEMRA framework that combines movement-aided parametric CE with tri-domain precoding. Antenna repositioning supplies the spatial samples for virtual-array eCSI reconstruction and the position variables for transmission optimization, while EM-domain pattern weights and digital precoders are optimized for data transmission. The contributions are summarized as follows.

  • Within the proposed SEMRA framework, we develop a concrete low-overhead movement-aided parametric CE procedure. By coordinating local repositioning of the My×MzM_{y}\times M_{z} antennas over a prescribed multi-block pilot schedule, a denser virtual array is synthesized while preserving the same per-user pilot overhead as the EMRA baseline. The resulting virtual array supplies additional spatial samples for the parametric estimation pipeline, improves AoD estimation, and yields a position-independent parametric channel model from which the eCSI can be assembled at desired antenna positions without additional pilots.

  • We formulate the SEMRA SE-maximization problem and develop a tri-domain alternating optimization framework based on block coordinate descent (BCD) over antenna positions, EM-domain pattern weights, and digital precoders. A zero-forcing (ZF)-based joint design is first constructed as a low-complexity baseline to isolate the gain of spatial reconfiguration beyond existing EM-digital designs for EMRA systems. We then develop a weighted minimum mean square error (WMMSE)-based design to further improve the system SE.

  • Numerical results show that the proposed CE scheme improves eCSI accuracy over the EMRA baseline under the same per-user pilot overhead. Within the tested range, both SEMRA-ZF and SEMRA-WMMSE achieve higher SE than the EMRA baseline under both estimated and perfect eCSI scenarios. For the tested multiuser loads, SEMRA-WMMSE further improves over SEMRA-ZF. The SEMRA-over-EMRA advantage is especially evident at larger inter-element spacings, where the EMRA baseline is more sensitive to spatial aliasing.

The remainder of this paper is organized as follows. Section LABEL:S2 presents the system model and formulates the tri-domain SE-maximization problem. Sections LABEL:S3 and LABEL:S4 develop the SEMRA-ZF and SEMRA-WMMSE designs, respectively. Section LABEL:S5 introduces the movement-aided parametric CE procedure. Section LABEL:S6 reports the simulation results, and Section LABEL:S7 concludes the paper.

Notation: Matrices and column vectors are denoted by uppercase and lowercase boldface letters, respectively. ()T(\cdot)^{T}, ()(\cdot)^{*}, ()H(\cdot)^{H}, ()1(\cdot)^{-1}, ()(\cdot)^{\dagger}, and 𝔼{}\mathbb{E}\{\cdot\} denote the transpose, conjugate, Hermitian transpose, inversion, Moore-Penrose pseudoinverse, and expectation. [𝒙]k[\bm{x}]_{k} and [𝑨]i,j[\bm{A}]_{i,j} denote the kk-th entry of vector 𝒙\bm{x} and the (i,j)(i,j)-th entry of matrix 𝑨\bm{A}. diag(𝒂)\mathrm{diag}(\bm{a}) forms a diagonal matrix from vector 𝒂\bm{a}, and Blkdiag{𝒂1,,𝒂M}\mathrm{Blkdiag}\{\bm{a}_{1},\ldots,\bm{a}_{M}\} constructs a block diagonal matrix. \odot denotes the Hadamard product. 2\|\cdot\|_{2} and F\|\cdot\|_{\mathrm{F}} denote the 2\ell_{2} norm and Frobenius norm, respectively. {}\Re\{\cdot\} and {}\Im\{\cdot\} denote the real and imaginary parts, respectively. j\mathrm{j} is the imaginary unit. 𝑰N\bm{I}_{N} denotes the N×NN\times N identity matrix. 𝒞𝒩(μ,σ2)\mathcal{CN}(\mu,\sigma^{2}) denotes a complex Gaussian distribution.

II System Model and Problem Formulation

wbsphack @@writeaux“newlabelS2wcurrentlabel1wesphack

II-A Downlink Signal Model

wbsphack @@writeaux“newlabelS2.1wcurrentlabel1wesphack

Refer to caption
Figure 1: SEMRA-enabled multiuser downlink with EM-domain radiation-pattern reconfigurability and spatial-domain position reconfigurability.wbsphack @@writeaux“newlabelfig:system˙modelwcurrentlabel1wesphack

We consider a time-division duplex (TDD) multiuser MIMO-OFDM downlink system implemented with the SEMRA architecture illustrated in Fig. LABEL:fig:system_model. The base station (BS) employs an M=MyMzM=M_{y}M_{z} uniform planar array (UPA) of SEMRA elements to serve UU single-antenna user equipments (UEs), each with a fixed position and radiation pattern. The inter-element spacings of the UPA along the y- and z-axes are dyd_{y} and dzd_{z}, respectively. The number of OFDM subcarriers is GG with subcarrier spacing ΔfBw/G\Delta f\triangleq B_{w}/G, and fg(g1)Δff_{g}\triangleq(g-1)\Delta f (1gG1\leq g\leq G) denotes the frequency of the gg-th subcarrier. The received signal at the uu-th UE (1uU1\leq u\leq U) and the gg-th subcarrier (1gG1\leq g\leq G) is given by

wbsphack@@writeaux\newlabelequ.rxwcurrentlabel1wesphackyu,g=𝒉u,gH𝑾g𝒔g+nu,g,wbsphack@@writeaux{}{\newlabel{equ.rx}{{wcurrentlabel}{1}}}wesphacky_{u,g}=\bm{h}_{u,g}^{H}\bm{W}_{g}\bm{s}_{g}+n_{u,g}, (1)

where 𝒉u,g=[hu,1,g,,hu,M,g]TM\bm{h}_{u,g}=[h_{u,1,g},\ldots,h_{u,M,g}]^{T}\in\mathbb{C}^{M} is the downlink channel vector between the BS and the uu-th UE, 𝑾g=[𝒘1,g,,𝒘U,g]M×U\bm{W}_{g}=[\bm{w}_{1,g},\ldots,\bm{w}_{U,g}]\in\mathbb{C}^{M\times U} is the digital precoding matrix with 𝒘u,gM×1\bm{w}_{u,g}\in\mathbb{C}^{M\times 1} denoting the digital precoder vector for the uu-th UE on the gg-th subcarrier, 𝒔gU\bm{s}_{g}\in\mathbb{C}^{U} is the data symbol vector satisfying 𝔼{𝒔g𝒔gH}=𝑰U\mathbb{E}\{\bm{s}_{g}\bm{s}_{g}^{H}\}=\bm{I}_{U}, and nu,g𝒞𝒩(0,σn2)n_{u,g}\sim\mathcal{CN}(0,\sigma_{n}^{2}) is additive white Gaussian noise (AWGN).

II-B Channel Model

wbsphack @@writeaux“newlabelS2.2wcurrentlabel1wesphack The channel between the mm-th BS antenna and the uu-th UE on the gg-th subcarrier is modeled as [38.901]

wbsphack@@writeaux\newlabelequ.chwcurrentlabel1wesphackhu,m,g=i=1Lux~i,ufrx,u(ϑi,u,φi,u)ftx,m(θi,u,ϕi,u)×ej2πλ(𝒌tx,i,uT𝒑m+𝒌rx,i,uT𝒓u)ej2πτi,ufg,\displaystyle\begin{split}wbsphack@@writeaux{}{\newlabel{equ.ch}{{wcurrentlabel}{1}}}wesphackh_{u,m,g}&=\sum_{i=1}^{L_{u}}\tilde{x}_{i,u}f_{\mathrm{rx},u}(\vartheta_{i,u},\varphi_{i,u})f_{\mathrm{tx},m}(\theta_{i,u},\phi_{i,u})\\ &\times e^{-\mathrm{j}\frac{2\pi}{\lambda}(\bm{k}_{\mathrm{tx},i,u}^{T}\bm{p}_{m}+\bm{k}_{\mathrm{rx},i,u}^{T}\bm{r}_{u})}\cdot e^{-\mathrm{j}2\pi\tau_{i,u}f_{g}},\end{split} (2)

where LuL_{u} is the number of channel paths, x~i,u\tilde{x}_{i,u} and τi,u\tau_{i,u} denote the complex gain and delay of the ii-th path, respectively, (θi,u,ϕi,u)(\theta_{i,u},\phi_{i,u}) and (ϑi,u,φi,u)(\vartheta_{i,u},\varphi_{i,u}) are respectively the AoD and AoA, ftx,m()f_{\mathrm{tx},m}(\cdot) and frx,u()f_{\mathrm{rx},u}(\cdot) denote the real-valued scalar radiation-pattern gains of the BS and UE antennas, respectively, and 𝒑m\bm{p}_{m}, 𝒓u\bm{r}_{u} are the position vectors of BS antenna mm and UE uu, respectively. Throughout this paper, the BS-side departure azimuth is restricted to the front half-space, i.e., ϕi,u[π/2,π/2]\phi_{i,u}\in[-\pi/2,\pi/2]. The unit direction vectors 𝒌tx,i,u,𝒌rx,i,u3\bm{k}_{\mathrm{tx},i,u},\bm{k}_{\mathrm{rx},i,u}\in\mathbb{R}^{3} are defined as 𝒌tx,i,u=[sinθi,ucosϕi,u,sinθi,usinϕi,u,cosθi,u]T\bm{k}_{\mathrm{tx},i,u}=[\sin\theta_{i,u}\cos\phi_{i,u},\sin\theta_{i,u}\sin\phi_{i,u},\cos\theta_{i,u}]^{T} and 𝒌rx,i,u=[sinϑi,ucosφi,u,sinϑi,usinφi,u,cosϑi,u]T\bm{k}_{\mathrm{rx},i,u}=[\sin\vartheta_{i,u}\cos\varphi_{i,u},\sin\vartheta_{i,u}\sin\varphi_{i,u},\cos\vartheta_{i,u}]^{T}.

Incorporating the delay into the path gain as xi,u,g=x~i,uej2πτi,ufgx_{i,u,g}=\tilde{x}_{i,u}e^{-\mathrm{j}2\pi\tau_{i,u}f_{g}}, the channel can be written in the compact matrix form as

wbsphack@@writeaux\newlabelequ.chcompactwcurrentlabel1wesphackhu,m,g=𝒇rx,uT𝑨u𝚺u,g𝑩u,m𝒇tx,u,m,wbsphack@@writeaux{}{\newlabel{equ.ch_{c}ompact}{{wcurrentlabel}{1}}}wesphackh_{u,m,g}=\bm{f}_{\mathrm{rx},u}^{T}\bm{A}_{u}\bm{\Sigma}_{u,g}\bm{B}_{u,m}\bm{f}_{\mathrm{tx},u,m}, (3)

where 𝒇rx,uLu\bm{f}_{\mathrm{rx},u}\in\mathbb{R}^{L_{u}} and 𝒇tx,u,mLu\bm{f}_{\mathrm{tx},u,m}\in\mathbb{R}^{L_{u}} collect the receive and transmit pattern gains along all LuL_{u} paths, with the ii-th elements being frx,u(ϑi,u,φi,u)f_{\mathrm{rx},u}(\vartheta_{i,u},\varphi_{i,u}) and ftx,m(θi,u,ϕi,u)f_{\mathrm{tx},m}(\theta_{i,u},\phi_{i,u}), respectively, 𝑨u=diag(𝒂u)Lu×Lu\bm{A}_{u}=\mathrm{diag}(\bm{a}_{u})\in\mathbb{C}^{L_{u}\times L_{u}} and 𝑩u,m=diag(𝒃u,m)Lu×Lu\bm{B}_{u,m}=\mathrm{diag}(\bm{b}_{u,m})\in\mathbb{C}^{L_{u}\times L_{u}} capture the array steering phases, with [𝒂u]i=ej2πλ𝒌rx,i,uT𝒓u[\bm{a}_{u}]_{i}=e^{-\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{rx},i,u}^{T}\bm{r}_{u}} and [𝒃u,m]i=ej2πλ𝒌tx,i,uT𝒑m[\bm{b}_{u,m}]_{i}=e^{-\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{tx},i,u}^{T}\bm{p}_{m}}, and 𝚺u,g=diag([x1,u,g,,xLu,u,g])Lu×Lu\bm{\Sigma}_{u,g}=\mathrm{diag}([x_{1,u,g},\ldots,x_{L_{u},u,g}])\in\mathbb{C}^{L_{u}\times L_{u}} collects the per-path gains.

II-C EM-Domain CSI Representation

wbsphack @@writeaux“newlabelS2.3wcurrentlabel1wesphack To enable EM-domain radiation pattern optimization, we represent the transmit pattern as a weighted combination of KK orthonormal basis functions

wbsphack@@writeaux\newlabelequ.shwcurrentlabel1wesphackftx,m(θ,ϕ)=k=1Kαk,mωk(θ,ϕ),wbsphack@@writeaux{}{\newlabel{equ.sh}{{wcurrentlabel}{1}}}wesphackf_{\mathrm{tx},m}(\theta,\phi)=\sum_{k=1}^{K}\alpha_{k,m}\omega_{k}(\theta,\phi), (4)

where {ωk(θ,ϕ)}k=1K\{\omega_{k}(\theta,\phi)\}_{k=1}^{K} are real-valued orthonormal basis functions over the unit sphere, constructed from real spherical harmonics as detailed in [Ying2025TCOM]. The orthonormality condition reads 02π0πωk(θ,ϕ)ωk(θ,ϕ)sinθdθdϕ=δk,k\int_{0}^{2\pi}\!\int_{0}^{\pi}\omega_{k}(\theta,\phi)\omega_{k^{\prime}}(\theta,\phi)\sin\theta\,\mathrm{d}\theta\,\mathrm{d}\phi=\delta_{k,k^{\prime}}, and 𝜶m=[α1,m,,αK,m]TK\bm{\alpha}_{m}=[\alpha_{1,m},\ldots,\alpha_{K,m}]^{T}\in\mathbb{R}^{K} is the weight coefficient vector for the mm-th antenna. Under the adopted decomposition, ftx,m(θ,ϕ)f_{\mathrm{tx},m}(\theta,\phi) models the real-valued directional gain, and we further assume that 𝜶m2=1\|\bm{\alpha}_{m}\|_{2}=1. Applying (LABEL:equ.sh) to the transmit pattern gain vector yields

wbsphack@@writeaux\newlabelequ.ftxwcurrentlabel1wesphack𝒇tx,u,m=𝛀u𝜶mLu,u,m,wbsphack@@writeaux{}{\newlabel{equ.ftx}{{wcurrentlabel}{1}}}wesphack\bm{f}_{\mathrm{tx},u,m}=\bm{\Omega}_{u}\bm{\alpha}_{m}\in{\mathbb{R}^{L_{u}}},\quad\forall u,m, (5)

where 𝜶m\bm{\alpha}_{m} parameterizes real gain-pattern coefficients, and the (i,k)(i,k)-th element of 𝛀uLu×K\bm{\Omega}_{u}\in\mathbb{R}^{L_{u}\times K} is [𝛀u]i,k=ωk(θi,u,ϕi,u)[\bm{\Omega}_{u}]_{i,k}=\omega_{k}(\theta_{i,u},\phi_{i,u}). Substituting (LABEL:equ.ftx) into (LABEL:equ.ch_compact), the channel can be expressed as

wbsphack@@writeaux\newlabelequ.ecsiwcurrentlabel1wesphackhu,m,g=𝒇rx,uT𝑨u𝚺u,g𝑩u,m𝛀u𝜶m=𝒒u,m,gH𝜶m.wbsphack@@writeaux{}{\newlabel{equ.ecsi}{{wcurrentlabel}{1}}}wesphackh_{u,m,g}=\bm{f}_{\mathrm{rx},u}^{T}\bm{A}_{u}\bm{\Sigma}_{u,g}\bm{B}_{u,m}\bm{\Omega}_{u}\bm{\alpha}_{m}=\bm{q}_{u,m,g}^{H}\bm{\alpha}_{m}. (6)

We define the eCSI vector as

wbsphack@@writeaux\newlabelequ.qdefwcurrentlabel1wesphack𝒒u,m,gH𝒇rx,uT𝑨u𝚺u,g𝑩u,m𝛀u1×K,wbsphack@@writeaux{}{\newlabel{equ.q_{d}ef}{{wcurrentlabel}{1}}}wesphack\bm{q}_{u,m,g}^{H}\triangleq\bm{f}_{\mathrm{rx},u}^{T}\bm{A}_{u}\bm{\Sigma}_{u,g}\bm{B}_{u,m}\bm{\Omega}_{u}\in\mathbb{C}^{1\times K}, (7)

and define the corresponding position-independent equivalent channel gain vector as

wbsphack@@writeaux\newlabelequ.vdefwcurrentlabel1wesphack𝝌u,g(𝒇rx,uT𝑨u𝚺u,g)HLu.wbsphack@@writeaux{}{\newlabel{equ.v_{d}ef}{{wcurrentlabel}{1}}}wesphack\bm{\chi}_{u,g}\triangleq\left(\bm{f}_{\mathrm{rx},u}^{T}\bm{A}_{u}\bm{\Sigma}_{u,g}\right)^{H}\in\mathbb{C}^{L_{u}}. (8)

It collects the position-independent UE-side and path-gain terms. The corresponding column-vector form of the eCSI is 𝒒u,m,g=𝛀uT𝑩u,mH𝝌u,gK×1.\bm{q}_{u,m,g}=\bm{\Omega}_{u}^{T}\bm{B}_{u,m}^{H}\bm{\chi}_{u,g}\in\mathbb{C}^{K\times 1}.

By aggregating the eCSI across all MM antennas, we define the stacked eCSI vector 𝒒u,g[𝒒u,1,gT,,𝒒u,M,gT]TMK×1\bm{q}_{u,g}\triangleq[\bm{q}_{u,1,g}^{T},\ldots,\bm{q}_{u,M,g}^{T}]^{T}\in\mathbb{C}^{MK\times 1} and its conjugated form 𝒒¯u,g𝒒u,g\bar{\bm{q}}_{u,g}\triangleq\bm{q}_{u,g}^{*}. We also define the block-diagonal EM precoding matrix 𝚲Blkdiag{𝜶1,,𝜶M}MK×M\bm{\Lambda}\triangleq\mathrm{Blkdiag}\{\bm{\alpha}_{1},\ldots,\bm{\alpha}_{M}\}\in\mathbb{R}^{MK\times M}. The effective channel then satisfies

wbsphack@@writeaux\newlabelequ.hdefwcurrentlabel1wesphack𝒉u,gH=𝒒¯u,gH𝚲.wbsphack@@writeaux{}{\newlabel{equ.h_{d}ef}{{wcurrentlabel}{1}}}wesphack\bm{h}_{u,g}^{H}=\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}. (9)

Substituting (LABEL:equ.h_def) into (LABEL:equ.rx) yields yu,g=𝒒¯u,gH𝚲𝑾g𝒔g+nu,gy_{u,g}=\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{W}_{g}\bm{s}_{g}+n_{u,g}, so the signal-to-interference-plus-noise ratio (SINR) for the uu-th UE on the gg-th subcarrier is

wbsphack@@writeaux\newlabelequ.sinrwcurrentlabel1wesphackSINRu,g=|𝒒¯u,gH𝚲𝒘u,g|2=1uU|𝒒¯u,gH𝚲𝒘,g|2+σn2.wbsphack@@writeaux{}{\newlabel{equ.sinr}{{wcurrentlabel}{1}}}wesphack\mathrm{SINR}_{u,g}=\frac{|\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{w}_{u,g}|^{2}}{\sum_{\begin{subarray}{c}\ell=1\\ \ell\neq u\end{subarray}}^{U}|\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{w}_{\ell,g}|^{2}+\sigma_{n}^{2}}. (10)

Accordingly, the sum SE is given by

wbsphack@@writeaux\newlabelequ.sewcurrentlabel1wesphackR=g=1Gu=1Ulog2(1+SINRu,g).wbsphack@@writeaux{}{\newlabel{equ.se}{{wcurrentlabel}{1}}}wesphackR=\sum_{g=1}^{G}\sum_{u=1}^{U}\log_{2}(1+\mathrm{SINR}_{u,g}). (11)

II-D SEMRA Model and Problem Formulation

wbsphack @@writeaux“newlabelS2.4wcurrentlabel1wesphack Unlike the EMRA with TFA architecture, the proposed SEMRA architecture jointly optimizes the per-antenna EM-domain pattern weights 𝜶m\bm{\alpha}_{m} and antenna positions 𝒑m\bm{p}_{m}. Let 𝒑¯m\bar{\bm{p}}_{m} denote the reference position of the mm-th antenna in the reference UPA. In SEMRA, each transmit antenna is permitted to move within a local feasible region 𝒮m\mathcal{S}_{m} centered at 𝒑¯m\bar{\bm{p}}_{m} (depicted as the local movement region in Fig. LABEL:fig:system_model), which is defined as

wbsphack@@writeaux\newlabelequ.constraintwcurrentlabel1wesphack𝒑m𝒮m{𝒙3𝒙𝒑¯mDmax,[𝒙]1=0},wbsphack@@writeaux{}{\newlabel{equ.constraint}{{wcurrentlabel}{1}}}wesphack\bm{p}_{m}\in\mathcal{S}_{m}\triangleq\left\{\bm{x}\in\mathbb{R}^{3}\mid\|\bm{x}-\bar{\bm{p}}_{m}\|_{\infty}\leq D_{\max},\ [\bm{x}]_{1}=0\right\}, (12)

where DmaxD_{\max} is the maximum allowable displacement for each antenna. To avoid antenna collisions, we require a minimum inter-antenna separation Dsep>0D_{\mathrm{sep}}>0 over all feasible positions, which is ensured for the reference UPA if dy2Dmax+Dsepd_{y}\geq 2D_{\max}+D_{\mathrm{sep}} and dz2Dmax+Dsepd_{z}\geq 2D_{\max}+D_{\mathrm{sep}}. The position 𝒑m\bm{p}_{m} enters the channel through the BS-side steering matrix 𝑩u,m\bm{B}_{u,m} in (LABEL:equ.ch_compact), introducing an additional spatial degree of freedom for performance optimization.

According to (LABEL:equ.h_def), the effective channels are jointly determined by the antenna positions 𝒑[𝒑1T,,𝒑MT]T3M\bm{p}\triangleq[\bm{p}_{1}^{T},\ldots,\bm{p}_{M}^{T}]^{T}\in\mathbb{R}^{3M} and the per-antenna EM-domain weights 𝜶m\bm{\alpha}_{m}, whose effects are captured by 𝒒¯u,g\bar{\bm{q}}_{u,g} and 𝚲\bm{\Lambda}, respectively. Let 𝑾{𝑾g}g=1G\bm{W}\triangleq\{\bm{W}_{g}\}_{g=1}^{G} denote the collection of digital precoders across all subcarriers. We aim to maximize the sum SE of all UEs in (LABEL:equ.se) by jointly optimizing digital-domain precoder 𝑾\bm{W}, EM-domain radiation-pattern weights 𝜶m\bm{\alpha}_{m}, and spatial-domain antenna position 𝒑\bm{p}, leading to

wbsphack@@writeaux\newlabelequ.problemwcurrentlabel1wesphack𝒫SE:max{𝜶m},𝑾,𝒑Rs.t.g=1G𝑾gF2PT,𝜶m2=1,𝒑m𝒮m,m.\displaystyle wbsphack@@writeaux{}{\newlabel{equ.problem}{{wcurrentlabel}{1}}}wesphack\begin{aligned} \mathcal{P}_{\mathrm{SE}}:\quad\max_{\{\bm{\alpha}_{m}\},\bm{W},\bm{p}}\quad&R\\ \text{s.t.}\quad&\sum_{g=1}^{G}\|\bm{W}_{g}\|_{\mathrm{F}}^{2}\leq P_{T},\\ &\|\bm{\alpha}_{m}\|_{2}=1,\ \bm{p}_{m}\in\mathcal{S}_{m},\ \forall m.\end{aligned} (13)

where RR is defined in (LABEL:equ.se) and depends on 𝑾\bm{W}, 𝜶m\bm{\alpha}_{m}, and 𝒑\bm{p} through the effective channel vectors 𝒉u,g\bm{h}_{u,g} in (LABEL:equ.h_def). The resulting problem is highly nonconvex due to the coupled variables and the unit-norm constraints, making it challenging to solve optimally.

Refer to caption
Figure 2: Concrete workflow of the proposed SEMRA method, including movement-aided CE, position-dependent eCSI reconstruction, and ZF/WMMSE tri-domain precoding.wbsphack @@writeaux“newlabelfig:semra˙workflowwcurrentlabel1wesphack

III Tri-Domain Alternating Optimization: ZF-Based Baseline

wbsphack @@writeaux“newlabelS3wcurrentlabel1wesphack

Building upon the EM-digital alternating ZF design in [Ying2025TCOM], we extend it to SEMRA by adding a spatial-domain position-reconfigurable block. The detailed CE-to-transmission workflow is summarized in Fig. LABEL:fig:semra_workflow, with the ZF, WMMSE, and CE components developed in Sections LABEL:S3-LABEL:S5, respectively.

SEMRA-ZF is used as the lower complexity reference design. Its normalized ZF update suppresses multiuser interference and makes the gain of the added spatial block directly comparable with the EM and digital baseline, but it does not optimize user power allocation under the global power constraint.

III-A Digital and EM-Domain Optimization

wbsphack @@writeaux“newlabelS3.1zfwcurrentlabel1wesphack

Given the current effective channels 𝒉u,g\bm{h}_{u,g} in (LABEL:equ.h_def), we adopt the ZF precoder. Define the normalized channel direction and the stacked matrix as in (LABEL:equ.O_g_def).

wbsphack@@writeaux\newlabelequ.Ogedfwcurrentlabel1wesphack𝒉~u,g𝒉u,g𝒉u,g2,𝑯~g[𝒉~1,g,,𝒉~U,g]M×U.wbsphack@@writeaux{}{\newlabel{equ.O_{g}{}_{d}ef}{{wcurrentlabel}{1}}}wesphack\tilde{\bm{h}}_{u,g}\triangleq\frac{\bm{h}_{u,g}}{\|\bm{h}_{u,g}\|_{2}},\qquad\tilde{\bm{H}}_{g}\triangleq[\tilde{\bm{h}}_{1,g},\ldots,\tilde{\bm{h}}_{U,g}]\in\mathbb{C}^{M\times U}. (14)

For MUM\geq U, the normalized-channel ZF precoder is computed as

wbsphack@@writeaux\newlabelequ.zfnormalizewcurrentlabel1wesphack𝑾g=PTG𝑾~g𝑾~gF,wbsphack@@writeaux{}{\newlabel{equ.zf_{n}ormalize}{{wcurrentlabel}{1}}}wesphack\bm{W}_{g}=\sqrt{\frac{P_{T}}{G}}\frac{\tilde{\bm{W}}_{g}}{\|\tilde{\bm{W}}_{g}\|_{\mathrm{F}}}, (15)

where 𝑾~g𝑯~g(𝑯~gH𝑯~g)1M×U\tilde{\bm{W}}_{g}\triangleq\tilde{\bm{H}}_{g}\big(\tilde{\bm{H}}_{g}^{H}\tilde{\bm{H}}_{g}\big)^{-1}\in\mathbb{C}^{M\times U} is the unnormalized ZF precoder built from the normalized channel directions in (LABEL:equ.O_g_def).

With 𝒑\bm{p} and the digital precoders fixed, we update the EM-domain precoder by optimizing the block-diagonal matrix 𝚲=Blkdiag{𝜶1,,𝜶M}\bm{\Lambda}=\mathrm{Blkdiag}\{\bm{\alpha}_{1},\ldots,\bm{\alpha}_{M}\} under per-antenna unit-norm constraints. This is formulated on a masked oblique manifold. We define the oblique manifold

wbsphack@@writeaux\newlabelequ.obliquedefwcurrentlabel1wesphack𝒪{𝚲KM×M:[𝚲T𝚲]m,m=1,m}.wbsphack@@writeaux{}{\newlabel{equ.oblique_{d}ef}{{wcurrentlabel}{1}}}wesphack\mathcal{OB}\triangleq\left\{\bm{\Lambda}\in\mathbb{R}^{KM\times M}:\left[\bm{\Lambda}^{T}\bm{\Lambda}\right]_{m,m}=1,\;\forall m\right\}. (16)

Let 𝑴0Blkdiag{𝟏K,,𝟏K}KM×M\bm{M}_{0}\triangleq\mathrm{Blkdiag}\{\bm{1}_{K},\ldots,\bm{1}_{K}\}\in\mathbb{R}^{KM\times M} and impose the structural constraint 𝚲=𝚲𝑴0\bm{\Lambda}=\bm{\Lambda}\odot\bm{M}_{0}. Equivalently, the feasible EM set is {𝚲𝒪:𝚲=𝚲𝑴0}\{\bm{\Lambda}\in\mathcal{OB}:\bm{\Lambda}=\bm{\Lambda}\odot\bm{M}_{0}\}. We define the cost function 𝔣(𝚲)R\mathfrak{f}(\bm{\Lambda})\triangleq-R. The Riemannian gradient at 𝚲\bm{\Lambda} is the orthogonal projection of the Euclidean gradient 𝔣(𝚲)\nabla\mathfrak{f}(\bm{\Lambda}) onto the tangent space, given by

wbsphack@@writeaux\newlabelequ.obliquegradwcurrentlabel1wesphackgrad𝔣(𝚲)=𝔣(𝚲)𝚲ddiag(𝚲T𝔣(𝚲)),wbsphack@@writeaux{}{\newlabel{equ.oblique_{g}rad}{{wcurrentlabel}{1}}}wesphack\mathrm{grad}\mathfrak{f}(\bm{\Lambda})=\nabla\mathfrak{f}(\bm{\Lambda})-\bm{\Lambda}\,\mathrm{ddiag}\!\left(\bm{\Lambda}^{T}\nabla\mathfrak{f}(\bm{\Lambda})\right), (17)

where ddiag()\mathrm{ddiag}(\cdot) sets all off-diagonal entries of its matrix argument to zero. The retraction is column-wise normalization Retr𝚲(γ𝑫)normalize(𝚲+γ𝑫)\mathrm{Retr}_{\bm{\Lambda}}(\gamma\bm{D})\triangleq\mathrm{normalize}(\bm{\Lambda}+\gamma\bm{D}).

We adopt the Riemannian steepest-descent direction on this masked-oblique set, namely

wbsphack@@writeaux\newlabelequ.obliquecgdirwcurrentlabel1wesphack𝑫(n)=grad𝔣(𝚲(n)),wbsphack@@writeaux{}{\newlabel{equ.oblique_{c}g_{d}ir}{{wcurrentlabel}{1}}}wesphack\bm{D}^{(n)}=-\mathrm{grad}\mathfrak{f}\!\left(\bm{\Lambda}^{(n)}\right), (18)

and choose a suitable step size γ(n)\gamma^{(n)} for the current EM subproblem with the digital precoders fixed. The EM precoder update is then

wbsphack@@writeaux\newlabelequ.lambdaupdatezfwcurrentlabel1wesphack𝚲(n+1)=Retr𝚲(n)(γ(n)𝑫(n))=normalize(𝚲(n)+γ(n)𝑫(n)).wbsphack@@writeaux{}{\newlabel{equ.lambda_{u}pdate_{z}f}{{wcurrentlabel}{1}}}wesphack\begin{split}\bm{\Lambda}^{(n+1)}&=\mathrm{Retr}_{\bm{\Lambda}^{(n)}}\!\left(\gamma^{(n)}\bm{D}^{(n)}\right)\\ &=\mathrm{normalize}\!\left(\bm{\Lambda}^{(n)}+\gamma^{(n)}\bm{D}^{(n)}\right).\end{split} (19)

To keep the per-outer-iteration cost moderate, the EM block in both tri-domain algorithms below uses a single manifold-gradient step in each outer iteration rather than an inner loop to full convergence. Since the digital and spatial blocks are refreshed between successive EM updates, the EM descent direction is recomputed from the current gradient in every outer iteration. For compactness, variables without outer-iteration superscripts in the algorithm below denote the latest available block values within the current outer iteration. To express the Euclidean gradient compactly, we use (LABEL:equ.h_def), i.e., 𝒉u,gH=𝒒¯u,gH𝚲\bm{h}_{u,g}^{H}=\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}. Let ζ1/σn2\zeta\triangleq 1/\sigma_{n}^{2}, and 𝑾u¯,g\bm{W}_{\bar{u},g} denote the digital precoder matrix obtained from 𝑾g\bm{W}_{g} by removing its uu-th column. The Euclidean gradient 𝔣(𝚲)\nabla\mathfrak{f}(\bm{\Lambda}) is given by [Ying2025TCOM]

wbsphack@@writeaux\newlabelequ.gradeuczflambdawcurrentlabel1wesphack𝔣(𝚲)=(g=1Gu=1U𝚪u,g(1)𝚪u,g(2)ln2)𝑴0.wbsphack@@writeaux{}{\newlabel{equ.grad_{e}uc_{z}f_{l}ambda}{{wcurrentlabel}{1}}}wesphack\nabla\mathfrak{f}\!\left(\bm{\Lambda}\right)=\left(-\sum_{g=1}^{G}\sum_{u=1}^{U}\frac{\bm{\Gamma}_{u,g}^{(1)}-\bm{\Gamma}_{u,g}^{(2)}}{\ln 2}\right)\odot\bm{M}_{0}. (20)

The Hadamard masking enforces the block-diagonal sparsity of 𝚲\bm{\Lambda}, such that each column update only acts on its corresponding KK-dimensional pattern-weight block. Starting from a feasible block-diagonal initialization, the masked gradient in (LABEL:equ.grad_euc_zf_lambda) and the retraction in (LABEL:equ.lambda_update_zf) preserve this structure throughout the ZF-EM updates. Moreover,

wbsphack@@writeaux\newlabelequ.Gammazfdefwcurrentlabel1wesphack𝚪u,g(1)\displaystyle wbsphack@@writeaux{}{\newlabel{equ.Gamma_{z}f_{d}ef}{{wcurrentlabel}{1}}}wesphack\bm{\Gamma}_{u,g}^{(1)} 2ζ{𝒒¯u,g𝒒¯u,gH𝚲𝑾g𝑾gH}1+ζ𝒒¯u,gH𝚲𝑾g22,\displaystyle\triangleq\frac{2\zeta\Re\!\left\{\bar{\bm{q}}_{u,g}\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{W}_{g}\bm{W}_{g}^{H}\right\}}{1+\zeta\|\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{W}_{g}\|_{2}^{2}}, (21)
𝚪u,g(2)\displaystyle\bm{\Gamma}_{u,g}^{(2)} 2ζ{𝒒¯u,g𝒒¯u,gH𝚲𝑾u¯,g𝑾u¯,gH}1+ζ𝒒¯u,gH𝚲𝑾u¯,g22.\displaystyle\triangleq\frac{2\zeta\Re\!\left\{\bar{\bm{q}}_{u,g}\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{W}_{\bar{u},g}\bm{W}_{\bar{u},g}^{H}\right\}}{1+\zeta\|\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}\bm{W}_{\bar{u},g}\|_{2}^{2}}. (22)

III-B Spatial-Domain Antenna Position Optimization

wbsphack @@writeaux“newlabelS3.2zfwcurrentlabel1wesphack

Using the available parametric eCSI model, we derive the spatial gradient of the SE objective RR with respect to the conjugate channel coefficient hu,m,gh_{u,m,g}^{*}.

III-B1 Multipath Channel Decomposition

Based on (LABEL:equ.v_def) and hu,m,g=𝜶mT𝒒u,m,gh_{u,m,g}^{*}=\bm{\alpha}_{m}^{T}\bm{q}_{u,m,g}, the conjugate channel admits the pathwise decomposition

wbsphack@@writeaux\newlabelequ.hdecompzfwcurrentlabel1wesphackhu,m,g=𝜶mT𝛀uT𝑩u,mH𝝌u,g=i=1Luhu,m,g,i,wbsphack@@writeaux{}{\newlabel{equ.h_{d}ecomp_{z}f}{{wcurrentlabel}{1}}}wesphackh_{u,m,g}^{*}=\bm{\alpha}_{m}^{T}\bm{\Omega}_{u}^{T}\bm{B}_{u,m}^{H}\bm{\chi}_{u,g}=\sum_{i=1}^{L_{u}}h_{u,m,g,i}^{*}, (23)

where the ii-th path component is

wbsphack@@writeaux\newlabelequ.vdefzfwcurrentlabel1wesphackhu,m,g,i[𝛀u𝜶m]i[𝝌u,g]iej2πλ𝒌tx,i,uT𝒑m.wbsphack@@writeaux{}{\newlabel{equ.v_{d}ef_{z}f}{{wcurrentlabel}{1}}}wesphackh_{u,m,g,i}^{*}\triangleq[\bm{\Omega}_{u}\bm{\alpha}_{m}]_{i}\,[\bm{\chi}_{u,g}]_{i}\,e^{\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{tx},i,u}^{T}\bm{p}_{m}}. (24)

Here 𝒌tx,i,u\bm{k}_{\mathrm{tx},i,u} is determined by the AoD pair (θi,u,ϕi,u)(\theta_{i,u},\phi_{i,u}), while 𝝌u,g\bm{\chi}_{u,g} absorbs the remaining position-independent factors. Therefore, the spatial update depends only on the position-dependent eCSI model, not on explicit AoA or UE-side path parameters.

III-B2 Sum-SE Gradient Kernel

Define the effective coupling coefficient from the \ell-th column of 𝑾g\bm{W}_{g} to the uu-th UE on the g-th subcarrier as

wbsphack@@writeaux\newlabelequ.cdefwcurrentlabel1wesphackcu,,g𝒉u,gH𝒘,g,{1,,U}.wbsphack@@writeaux{}{\newlabel{equ.c_{d}ef}{{wcurrentlabel}{1}}}wesphackc_{u,\ell,g}\triangleq\bm{h}_{u,g}^{H}\bm{w}_{\ell,g},\quad\ell\in\{1,\ldots,U\}. (25)

In particular, cu,u,gc_{u,u,g} corresponds to the desired signal term of the uu-th UE, whereas cu,,gc_{u,\ell,g} for u\ell\neq u represent inter-user interference. Let Pu,g=1U|cu,,g|2+σn2P_{u,g}\triangleq\sum_{\ell=1}^{U}|c_{u,\ell,g}|^{2}+\sigma_{n}^{2} and u,g=1uU|cu,,g|2+σn2\mathcal{I}_{u,g}\triangleq\sum_{\begin{subarray}{c}\ell=1\\ \ell\neq u\end{subarray}}^{U}|c_{u,\ell,g}|^{2}+\sigma_{n}^{2}, so that SINRu,g=|cu,u,g|2/u,g\mathrm{SINR}_{u,g}=|c_{u,u,g}|^{2}/\mathcal{I}_{u,g}, 1+SINRu,g=Pu,g/u,g1+\mathrm{SINR}_{u,g}=P_{u,g}/\mathcal{I}_{u,g} and ru,g=log2(Pu,g)log2(u,g)r_{u,g}=\log_{2}(P_{u,g})-\log_{2}(\mathcal{I}_{u,g}). Moreover, |cu,,g|2hu,m,g=cu,,gw,m,g\frac{\partial|c_{u,\ell,g}|^{2}}{\partial h_{u,m,g}^{*}}=c_{u,\ell,g}^{*}w_{\ell,m,g} with w,m,g=[𝒘,g]mw_{\ell,m,g}=[\bm{w}_{\ell,g}]_{m}. Applying the chain rule yields the sensitivity of RR with respect to hu,m,gh_{u,m,g}^{*} as

wbsphack@@writeaux\newlabelequ.xizfdefwcurrentlabel1wesphackξu,m,g(R)\displaystyle wbsphack@@writeaux{}{\newlabel{equ.xi_{z}f_{d}ef}{{wcurrentlabel}{1}}}wesphack\xi_{u,m,g}^{(R)} Rhu,m,g=1ln2Pu,g(cu,u,gwu,m,g\displaystyle\triangleq\frac{\partial R}{\partial h_{u,m,g}^{*}}=\frac{1}{\ln 2\cdot P_{u,g}}\bigg(c_{u,u,g}^{*}w_{u,m,g}
SINRu,g=1uUcu,,gw,m,g).\displaystyle\quad-\mathrm{SINR}_{u,g}\textstyle\sum_{\begin{subarray}{c}\ell=1\\ \ell\neq u\end{subarray}}^{U}c_{u,\ell,g}^{*}w_{\ell,m,g}\bigg). (26)

III-B3 Spatial Gradient Assembly and Projection

From (LABEL:equ.v_def_zf), differentiating the conjugate path component hu,m,g,ih_{u,m,g,i}^{*} with respect to 𝒑m\bm{p}_{m} yields hu,m,g,i𝒑m=j2πλ𝒌tx,i,uhu,m,g,i\frac{\partial h_{u,m,g,i}^{*}}{\partial\bm{p}_{m}}=\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{tx},i,u}h_{u,m,g,i}^{*}. Using Wirtinger calculus and noting that 𝒑m3\bm{p}_{m}\in\mathbb{R}^{3} and RR is real-valued, the real gradient satisfies 𝒑mR=2{u,gRhh𝒑m}\nabla_{\bm{p}_{m}}R=2\Re\{\sum_{u,g}\frac{\partial R}{\partial h^{*}}\frac{\partial h^{*}}{\partial\bm{p}_{m}}\}. Exploiting 2{jz}=2{z}2\Re\{\mathrm{j}z\}=-2\Im\{z\}, we obtain

wbsphack@@writeaux\newlabelequ.gradposzfwcurrentlabel1wesphack𝒑mR=4πλ{g=1Gu=1Uξu,m,g(R)i=1Luhu,m,g,i𝒌tx,i,u}.wbsphack@@writeaux{}{\newlabel{equ.grad_{p}os_{z}f}{{wcurrentlabel}{1}}}wesphack\nabla_{\bm{p}_{m}}R=-\frac{4\pi}{\lambda}\Im\left\{\sum_{g=1}^{G}\sum_{u=1}^{U}\xi_{u,m,g}^{(R)}\sum_{i=1}^{L_{u}}h_{u,m,g,i}^{*}\bm{k}_{\mathrm{tx},i,u}\right\}. (27)

The ZF spatial block updates the antenna positions by the projected-gradient step

wbsphack@@writeaux\newlabelequ.posupdatezfwcurrentlabel1wesphack𝒑m(n+1)=Π𝒮m(𝒑m(n)+ηp𝒑mR),wbsphack@@writeaux{}{\newlabel{equ.pos_{u}pdate_{z}f}{{wcurrentlabel}{1}}}wesphack\bm{p}_{m}^{(n+1)}=\Pi_{\mathcal{S}_{m}}\left(\bm{p}_{m}^{(n)}+\eta_{p}\nabla_{\bm{p}_{m}}R\right), (28)

where ηp>0\eta_{p}>0 denotes a suitable position-update step size with 𝚲\bm{\Lambda} and the digital precoders fixed. The projection Π𝒮m()\Pi_{\mathcal{S}_{m}}(\cdot) maps the updated point onto the feasible region (LABEL:equ.constraint) via component-wise clipping. Specifically, [Π𝒮m(𝒙)]1=0[\Pi_{\mathcal{S}_{m}}(\bm{x})]_{1}=0 enforces the planar constraint, while [Π𝒮m(𝒙)]k=min(max([𝒙]k,[𝒑¯m]kDmax),[𝒑¯m]k+Dmax)[\Pi_{\mathcal{S}_{m}}(\bm{x})]_{k}=\min\!\big(\max([\bm{x}]_{k},\,[\bar{\bm{p}}_{m}]_{k}-D_{\max}),\,[\bar{\bm{p}}_{m}]_{k}+D_{\max}\big) for k{2,3}k\in\{2,3\} confines the displacement to the box region. The gradients of all antenna positions are evaluated at the current iterate, and one projected-gradient position trial is formed in each outer iteration. The feasible trial is retained if it does not decrease the current sum SE; otherwise, the current antenna positions are preserved. The overall ZF-based tri-domain procedure is summarized in Algorithm LABEL:alg:mm_zf_tridomain.

Input: eCSI model; feasible sets {𝒮m}\{\mathcal{S}_{m}\}; power budget PTP_{T}; (Nmax,ϵ)(N_{\max},\epsilon); step sizes.
Output: Design {𝑾g}\{\bm{W}_{g}\}, {𝜶m}\{\bm{\alpha}_{m}\}, and {𝒑m}\{\bm{p}_{m}\}.
1 Initialization: set 𝒑m(0)=𝒑¯m\bm{p}_{m}^{(0)}=\bar{\bm{p}}_{m} and choose feasible unit-norm 𝜶m(0)\bm{\alpha}_{m}^{(0)}, m\forall m;
2 Build 𝚲(0)\bm{\Lambda}^{(0)} and {𝒉u,g}\{\bm{h}_{u,g}\} via (LABEL:equ.h_def);
3 Digital initialization: obtain {𝑾g(0)}\{\bm{W}_{g}^{(0)}\} via (LABEL:equ.O_g_def)-(LABEL:equ.zf_normalize) and R(0)R^{(0)} via (LABEL:equ.se);
4 for n=0n=0 to Nmax1N_{\max}-1 do
5  Channel update: evaluate {𝒒u,m,g}\{\bm{q}_{u,m,g}\} at {𝒑m(n)}\{\bm{p}_{m}^{(n)}\} and form {𝒉u,g}\{\bm{h}_{u,g}\} via (LABEL:equ.h_def);
6  Digital update: update {𝑾g}\{\bm{W}_{g}\} by normalized ZF via (LABEL:equ.O_g_def)-(LABEL:equ.zf_normalize);
7  EM update: update 𝚲\bm{\Lambda} via (LABEL:equ.oblique_cg_dir) and (LABEL:equ.lambda_update_zf), and extract {𝜶m(n+1)}\{\bm{\alpha}_{m}^{(n+1)}\};
8  EM refresh: rebuild {𝒉u,g}\{\bm{h}_{u,g}\} and update {𝑾g}\{\bm{W}_{g}\};
9  Position update: form a trial via (LABEL:equ.grad_pos_zf)-(LABEL:equ.pos_update_zf); accept it if the sum SE does not decrease and otherwise retain {𝒑m(n)}\{\bm{p}_{m}^{(n)}\};
10  Position refresh: update {𝒒u,m,g}\{\bm{q}_{u,m,g}\}, {𝒉u,g}\{\bm{h}_{u,g}\}, and {𝑾g}\{\bm{W}_{g}\} at {𝒑m(n+1)}\{\bm{p}_{m}^{(n+1)}\};
11  Objective update: evaluate R(n+1)R^{(n+1)} via (LABEL:equ.se);
12  if |R(n+1)R(n)|max{1,|R(n)|}ϵ\frac{|R^{(n+1)}-R^{(n)}|}{\max\{1,|R^{(n)}|\}}\leq\epsilon then
13     break;
14    
15  end if
16 
17 end for
18return {𝑾g}\{\bm{W}_{g}\}, {𝜶m}\{\bm{\alpha}_{m}\}, and {𝒑m}\{\bm{p}_{m}\};
Algorithm 1 ZF-based tri-domain alternating optimizationwbsphack @@writeaux“newlabelalg:mm˙zf˙tridomainwcurrentlabel1wesphack

IV Tri-Domain Alternating Optimization: WMMSE-Based Design

wbsphack @@writeaux“newlabelS4wcurrentlabel1wesphack

Beyond the ZF-based baseline algorithm introduced in Section LABEL:S3, we develop a tri-domain WMMSE-based algorithm to solve (LABEL:equ.problem) under a global total-power constraint in this section. Specifically, the WMMSE reformulation yields closed-form auxiliary and digital updates, while the EM and spatial blocks retain the update structure of Section LABEL:S3.

SEMRA-WMMSE retains the same EM and position variables and the same outer alternating structure. It replaces normalized ZF with auxiliary receiver, mean square error weight, and digital precoder updates that jointly account for interference and the total power budget. It is therefore the higher complexity extension of SEMRA-ZF rather than an independent design.

IV-A WMMSE Problem Transformation

wbsphack @@writeaux“newlabelS4.1wcurrentlabel1wesphack

IV-A1 Signal Estimation and Mean Square Error

For each user-subcarrier pair, we introduce a scalar receive equalizer βu,g\beta_{u,g}\in\mathbb{C} and the estimated symbol s^u,gβu,gyu,g\hat{s}_{u,g}\triangleq\beta_{u,g}^{*}y_{u,g}. The mean square error (MSE) εu,g𝔼{|βu,gyu,gsu,g|2}\varepsilon_{u,g}\triangleq\mathbb{E}\{|\beta_{u,g}^{*}y_{u,g}-s_{u,g}|^{2}\} then expands via (LABEL:equ.rx) as

wbsphack@@writeaux\newlabelequ.mseexplicitwcurrentlabel1wesphackεu,g=|βu,g𝒉u,gH𝒘u,g1|2signal distortion+=1uU|βu,g𝒉u,gH𝒘,g|2multiuser interference+σn2|βu,g|2noise amplification,\displaystyle wbsphack@@writeaux{}{\newlabel{equ.mse_{e}xplicit}{{wcurrentlabel}{1}}}wesphack\begin{split}\varepsilon_{u,g}&=\underbrace{\left|\beta_{u,g}^{*}\bm{h}_{u,g}^{H}\bm{w}_{u,g}-1\right|^{2}}_{\text{signal distortion}}\\ &\quad+\underbrace{\sum_{\begin{subarray}{c}\ell=1\\ \ell\neq u\end{subarray}}^{U}\left|\beta_{u,g}^{*}\bm{h}_{u,g}^{H}\bm{w}_{\ell,g}\right|^{2}}_{\text{multiuser interference}}+\underbrace{\sigma_{n}^{2}|\beta_{u,g}|^{2}}_{\text{noise amplification}},\end{split} (29)

where 𝒉u,g\bm{h}_{u,g} is defined in (LABEL:equ.h_def) and its dependence on 𝒑\bm{p} and 𝜶m\bm{\alpha}_{m} is omitted for notational brevity.

IV-A2 WMMSE Equivalence

By the WMMSE equivalence [shi2011iteratively], the sum-SE maximization problem (LABEL:equ.problem) can be equivalently reformulated by introducing auxiliary minimum mean square error (MMSE) receivers βu,g\beta_{u,g} and positive MSE weights ρu,g\rho_{u,g} for each user-subcarrier pair, yielding

wbsphack@@writeaux\newlabelequ.wmmseproblemwcurrentlabel1wesphack𝒫WMMSE:min𝒱FWMMSEs.t.g=1Gu=1U𝒘u,g22PT,𝜶m2=1,𝜶mK,m,𝒑m𝒮m,m,ρu,g>0,u,g,wbsphack@@writeaux{}{\newlabel{equ.wmmse_{p}roblem}{{wcurrentlabel}{1}}}wesphack\begin{aligned} \mathcal{P}_{\text{WMMSE}}:\quad\min_{\mathcal{V}}\quad&F_{\mathrm{WMMSE}}\\ \text{s.t.}\quad&\sum_{g=1}^{G}\sum_{u=1}^{U}\|\bm{w}_{u,g}\|_{2}^{2}\leq P_{T},\\ &\|\bm{\alpha}_{m}\|_{2}=1,\;\bm{\alpha}_{m}\in\mathbb{R}^{K},\;\forall m,\\ &\bm{p}_{m}\in\mathcal{S}_{m},\;\forall m,\;\rho_{u,g}>0,\;\forall u,g,\end{aligned} (30)

where 𝒱{{𝒘u,g},{𝜶m},{𝒑m},{βu,g},{ρu,g}}\mathcal{V}\triangleq\{\{\bm{w}_{u,g}\},\{\bm{\alpha}_{m}\},\{\bm{p}_{m}\},\{\beta_{u,g}\},\{\rho_{u,g}\}\} denotes the collection of all optimization variables, and the WMMSE objective function is defined as

wbsphack@@writeaux\newlabelequ.wmmseobjwcurrentlabel1wesphackFWMMSEg=1Gu=1U(ρu,gεu,glnρu,g).wbsphack@@writeaux{}{\newlabel{equ.wmmse_{o}bj}{{wcurrentlabel}{1}}}wesphackF_{\mathrm{WMMSE}}\triangleq\sum_{g=1}^{G}\sum_{u=1}^{U}\left(\rho_{u,g}\varepsilon_{u,g}-\ln\rho_{u,g}\right). (31)

For fixed 𝑾\bm{W}, 𝜶m\bm{\alpha}_{m}, and 𝒑\bm{p}, optimizing the auxiliary variables βu,g\beta_{u,g} and ρu,g\rho_{u,g} yields min{βu,g},{ρu,g>0}FWMMSE=UG(ln2)R\min_{\{\beta_{u,g}\},\,\{\rho_{u,g}>0\}}F_{\mathrm{WMMSE}}=UG-(\ln 2)\,R. Hence, after eliminating the auxiliary variables, minimizing FWMMSEF_{\mathrm{WMMSE}} over the remaining design variables is equivalent to maximizing RR in (LABEL:equ.se). The reformulation also yields closed-form updates for βu,g\beta_{u,g}, ρu,g\rho_{u,g}, and 𝒘u,g\bm{w}_{u,g}, whereas the EM- and spatial-domain blocks remain nonconvex and are handled in Sections LABEL:S4.3 and LABEL:S4.4.

Substituting (LABEL:equ.mse_explicit) into (LABEL:equ.wmmse_obj) and expanding, the WMMSE objective becomes

wbsphack@@writeaux\newlabelequ.wmmseobjfullwcurrentlabel1wesphackFWMMSE=g=1Gu=1U[ρu,g(|βu,g𝒉u,gH𝒘u,g1|2+=1uU|βu,g𝒉u,gH𝒘,g|2+σn2|βu,g|2)lnρu,g].wbsphack@@writeaux{}{\newlabel{equ.wmmse_{o}bj_{f}ull}{{wcurrentlabel}{1}}}wesphack\begin{split}F_{\mathrm{WMMSE}}=\sum_{g=1}^{G}\sum_{u=1}^{U}\Bigg[\rho_{u,g}\Bigg(\left|\beta_{u,g}^{*}\bm{h}_{u,g}^{H}\bm{w}_{u,g}-1\right|^{2}\\ +\sum_{\begin{subarray}{c}\ell=1\\ \ell\neq u\end{subarray}}^{U}\left|\beta_{u,g}^{*}\bm{h}_{u,g}^{H}\bm{w}_{\ell,g}\right|^{2}+\sigma_{n}^{2}|\beta_{u,g}|^{2}\Bigg)-\ln\rho_{u,g}\Bigg].\end{split} (32)

IV-B Digital Precoding and Auxiliary Variable Optimization

wbsphack @@writeaux“newlabelS4.2wcurrentlabel1wesphack

With antenna positions 𝒑\bm{p} and EM-domain precoders 𝜶m\bm{\alpha}_{m} fixed, the first block minimizes FWMMSEF_{\mathrm{WMMSE}} over the auxiliary variables βu,g\beta_{u,g}, ρu,g\rho_{u,g}, and the digital precoders 𝒘u,g\bm{w}_{u,g} via BCD. Each subproblem is convex in its own variable block and therefore admits a closed-form update.

IV-B1 Optimal Receive Equalizer Update

We first optimize the scalar receive equalizer βu,g\beta_{u,g} for each user-subcarrier pair while keeping 𝒘u,g\bm{w}_{u,g} and ρu,g\rho_{u,g} fixed. Setting εu,g/βu,g=0\partial\varepsilon_{u,g}/\partial\beta_{u,g}^{*}=0 in (LABEL:equ.mse_explicit) yields the MMSE receiver given by

wbsphack@@writeaux\newlabelequ.betaoptwcurrentlabel1wesphackβu,g=𝒉u,gH𝒘u,g=1U|𝒉u,gH𝒘,g|2+σn2.wbsphack@@writeaux{}{\newlabel{equ.beta_{o}pt}{{wcurrentlabel}{1}}}wesphack\beta_{u,g}^{\star}=\frac{\bm{h}_{u,g}^{H}\bm{w}_{u,g}}{\sum_{\ell=1}^{U}|\bm{h}_{u,g}^{H}\bm{w}_{\ell,g}|^{2}+\sigma_{n}^{2}}. (33)

IV-B2 Optimal MSE Weight Update

Setting (ρu,gεu,glnρu,g)/ρu,g=0\partial(\rho_{u,g}\varepsilon_{u,g}-\ln\rho_{u,g})/\partial\rho_{u,g}=0 gives ρu,g=1/εu,g\rho_{u,g}^{\star}=1/\varepsilon_{u,g}. When the optimal equalizer (LABEL:equ.beta_opt) is substituted, the minimum MSE reduces to εu,gmin=(1+SINRu,g)1\varepsilon_{u,g}^{\min}=(1+\mathrm{SINR}_{u,g})^{-1} by the orthogonality principle, yielding

wbsphack@@writeaux\newlabelequ.rhooptwcurrentlabel1wesphackρu,g=1+SINRu,g.wbsphack@@writeaux{}{\newlabel{equ.rho_{o}pt}{{wcurrentlabel}{1}}}wesphack\rho_{u,g}^{\star}=1+\mathrm{SINR}_{u,g}. (34)

IV-B3 Optimal Digital Precoder Update

Finally, we optimize the digital precoders 𝒘u,g\bm{w}_{u,g} subject to the total power constraint. Removing the terms independent of 𝒘u,g\bm{w}_{u,g} from (LABEL:equ.wmmse_obj_full), the subproblem becomes the following quadratically constrained quadratic program (QCQP)

wbsphack@@writeaux\newlabelequ.wsubproblemwcurrentlabel1wesphackmin{𝒘u,g}g=1Gu=1Uρu,gεu,gs.t.g=1Gu=1U𝒘u,g22PT.wbsphack@@writeaux{}{\newlabel{equ.w_{s}ubproblem}{{wcurrentlabel}{1}}}wesphack\begin{aligned} \min_{\{\bm{w}_{u,g}\}}\quad&\sum_{g=1}^{G}\sum_{u=1}^{U}\rho_{u,g}\varepsilon_{u,g}\\ \text{s.t.}\quad&\sum_{g=1}^{G}\sum_{u=1}^{U}\|\bm{w}_{u,g}\|_{2}^{2}\leq P_{T}.\end{aligned} (35)

To decouple the precoders across users, we first expand the weighted MSE. Interchanging the summation order over users uu and interferers \ell, the objective function can be rewritten as (with constants independent of 𝒘u,g\bm{w}_{u,g} omitted)

wbsphack@@writeaux\newlabelequ.Jwwcurrentlabel1wesphackJ({𝒘u,g})=g=1Gu=1U(𝒘u,gH𝚿g𝒘u,g2{𝒘u,gH𝒛u,g}),wbsphack@@writeaux{}{\newlabel{equ.J_{w}}{{wcurrentlabel}{1}}}wesphackJ(\{\bm{w}_{u,g}\})=\sum_{g=1}^{G}\sum_{u=1}^{U}\left(\bm{w}_{u,g}^{H}\bm{\Psi}_{g}\bm{w}_{u,g}-2\Re\left\{\bm{w}_{u,g}^{H}\bm{z}_{u,g}\right\}\right), (36)

where we define the weighted channel covariance matrix as 𝚿g=1Uρ,g|β,g|2𝒉,g𝒉,gHM×M\bm{\Psi}_{g}\triangleq\sum_{\ell=1}^{U}\rho_{\ell,g}|\beta_{\ell,g}|^{2}\bm{h}_{\ell,g}\bm{h}_{\ell,g}^{H}\in\mathbb{C}^{M\times M} and the equivalent target vector as 𝒛u,gρu,gβu,g𝒉u,gM\bm{z}_{u,g}\triangleq\rho_{u,g}\beta_{u,g}\bm{h}_{u,g}\in\mathbb{C}^{M}.

We solve (LABEL:equ.w_subproblem) via the Lagrangian method. Let ν0\nu\geq 0 denote the dual variable associated with the power constraint. The Karush-Kuhn-Tucker (KKT) stationarity condition /𝒘u,g=𝟎\partial\mathcal{L}/\partial\bm{w}_{u,g}^{*}=\bm{0} yields (𝚿g+ν𝑰M)𝒘u,g=𝒛u,g(\bm{\Psi}_{g}+\nu\bm{I}_{M})\bm{w}_{u,g}=\bm{z}_{u,g}. Since 𝚿g𝟎\bm{\Psi}_{g}\succeq\bm{0}, the matrix 𝚿g+ν𝑰M\bm{\Psi}_{g}+\nu\bm{I}_{M} is positive definite for any ν>0\nu>0, and the optimal precoder admits the closed-form solution

wbsphack@@writeaux\newlabelequ.woptwcurrentlabel1wesphack𝒘u,g(ν)=(𝚿g+ν𝑰M)1𝒛u,g.wbsphack@@writeaux{}{\newlabel{equ.w_{o}pt}{{wcurrentlabel}{1}}}wesphack\bm{w}_{u,g}^{\star}(\nu)=(\bm{\Psi}_{g}+\nu\bm{I}_{M})^{-1}\bm{z}_{u,g}. (37)

Since 𝚿g\bm{\Psi}_{g} has rank at most UU, it is singular whenever U<MU<M. However, 𝒛u,g=ρu,gβu,g𝒉u,g\bm{z}_{u,g}=\rho_{u,g}\beta_{u,g}\bm{h}_{u,g} is collinear with 𝒉u,g\bm{h}_{u,g}, and the uu-th term of 𝚿g\bm{\Psi}_{g} is scaled by ρu,g|βu,g|2\rho_{u,g}|\beta_{u,g}|^{2}, so 𝒛u,g\bm{z}_{u,g} lies in the column space range(𝚿g)\mathrm{range}(\bm{\Psi}_{g}) of 𝚿g\bm{\Psi}_{g} whenever βu,g0\beta_{u,g}\neq 0, while the case βu,g=0\beta_{u,g}=0 gives 𝒛u,g=𝟎\bm{z}_{u,g}=\bm{0} trivially. Therefore, the limit limν0+𝒘u,g(ν)=𝚿g𝒛u,g\lim_{\nu\to 0^{+}}\bm{w}_{u,g}^{\star}(\nu)=\bm{\Psi}_{g}^{\dagger}\bm{z}_{u,g} is well-defined, and we define 𝒘u,g(0)𝚿g𝒛u,g\bm{w}_{u,g}^{\star}(0)\triangleq\bm{\Psi}_{g}^{\dagger}\bm{z}_{u,g}. The optimal Lagrange multiplier ν\nu^{\star} is determined by complementary slackness. If g,u𝒘u,g(0)22PT\sum_{g,u}\|\bm{w}_{u,g}^{\star}(0)\|_{2}^{2}\leq P_{T}, then ν=0\nu^{\star}=0 and the power constraint is inactive. Otherwise, ν>0\nu^{\star}>0 is the unique root of g,u𝒘u,g(ν)22=PT\sum_{g,u}\|\bm{w}_{u,g}^{\star}(\nu)\|_{2}^{2}=P_{T}, which can be found by bisection since the left-hand side is monotonically decreasing in ν\nu. Unlike the normalized ZF baseline in Section LABEL:S3, this digital update enforces only the global total-power constraint and can therefore redistribute digital power across users and subcarriers.

IV-C EM-Domain Radiation Pattern Optimization

wbsphack @@writeaux“newlabelS4.3wcurrentlabel1wesphack

With 𝑾\bm{W}, 𝒑\bm{p}, βu,g\beta_{u,g}, and ρu,g\rho_{u,g} fixed, we optimize the EM-domain weights 𝜶m\bm{\alpha}_{m}, or equivalently the block-diagonal matrix 𝚲=Blkdiag{𝜶1,,𝜶M}KM×M\bm{\Lambda}=\mathrm{Blkdiag}\{\bm{\alpha}_{1},\ldots,\bm{\alpha}_{M}\}\in\mathbb{R}^{KM\times M}. As in Section LABEL:S3.1zf, we reuse the masked-oblique manifold framework while adapting the objective and its Euclidean gradient to the WMMSE formulation.

IV-C1 Subproblem Formulation

With 𝑾\bm{W}, 𝒑\bm{p}, βu,g\beta_{u,g}, and ρu,g\rho_{u,g} fixed, the EM-domain subproblem is

wbsphack@@writeaux\newlabelequ.alphasubproblemlambdawcurrentlabel1wesphackmin𝚲𝒪𝔣EM(𝚲)g=1Gu=1Uρu,gεu,gs.t.𝚲=𝚲𝑴0,\displaystyle wbsphack@@writeaux{}{\newlabel{equ.alpha_{s}ubproblem_{l}ambda}{{wcurrentlabel}{1}}}wesphack\begin{split}\min_{\bm{\Lambda}\in\mathcal{OB}}\quad&\mathfrak{f}_{\mathrm{EM}}(\bm{\Lambda})\triangleq\sum_{g=1}^{G}\sum_{u=1}^{U}\rho_{u,g}\varepsilon_{u,g}\\ \text{s.t.}\quad&\bm{\Lambda}=\bm{\Lambda}\odot\bm{M}_{0},\end{split} (38)

where εu,g\varepsilon_{u,g} is given in (LABEL:equ.mse_explicit) and 𝑴0\bm{M}_{0} is the block-diagonal mask from Section LABEL:S3.1zf. Since the weights ρu,g\rho_{u,g} are fixed, the term g,ulnρu,g-\sum_{g,u}\ln\rho_{u,g} in (LABEL:equ.wmmse_obj) is an additive constant and is omitted.

IV-C2 Euclidean Gradient of the EM Objective

From (LABEL:equ.h_def), the effective channel satisfies 𝒉u,gH=𝒒¯u,gH𝚲\bm{h}_{u,g}^{H}=\bar{\bm{q}}_{u,g}^{H}\bm{\Lambda}. Let cu,,g𝒉u,gH𝒘,gc_{u,\ell,g}\triangleq\bm{h}_{u,g}^{H}\bm{w}_{\ell,g} be defined as in (LABEL:equ.c_def). We define the WMMSE MSE-residual kernel as

wbsphack@@writeaux\newlabelequ.xidefwcurrentlabel1wesphackξu,m,g(ε)\displaystyle wbsphack@@writeaux{}{\newlabel{equ.xi_{d}ef}{{wcurrentlabel}{1}}}wesphack\xi_{u,m,g}^{(\varepsilon)} (ρu,gεu,g)hu,m,g\displaystyle\triangleq\frac{\partial(\rho_{u,g}\varepsilon_{u,g})}{\partial h_{u,m,g}^{*}}
=ρu,g(|βu,g|2=1Ucu,,gw,m,gβu,gwu,m,g),\displaystyle=\rho_{u,g}\left(|\beta_{u,g}|^{2}\sum_{\ell=1}^{U}c_{u,\ell,g}^{*}w_{\ell,m,g}-\beta_{u,g}^{*}w_{u,m,g}\right), (39)

where w,m,g=[𝒘,g]mw_{\ell,m,g}=[\bm{w}_{\ell,g}]_{m}. Stacking the kernel into 𝝃u,g(ε)[ξu,1,g(ε),,ξu,M,g(ε)]TM\bm{\xi}_{u,g}^{(\varepsilon)}\triangleq[\xi_{u,1,g}^{(\varepsilon)},\ldots,\xi_{u,M,g}^{(\varepsilon)}]^{T}\in\mathbb{C}^{M} and noting that 𝚲\bm{\Lambda} is real-valued, the chain rule yields

wbsphack@@writeaux\newlabelequ.gradeucwmmselambdawcurrentlabel1wesphack𝔣EM(𝚲)=(2{g=1Gu=1U𝒒u,g(𝝃u,g(ε))T})𝑴0.wbsphack@@writeaux{}{\newlabel{equ.grad_{e}uc_{w}mmse_{l}ambda}{{wcurrentlabel}{1}}}wesphack\nabla\mathfrak{f}_{\mathrm{EM}}(\bm{\Lambda})=\left(2\Re\left\{\sum_{g=1}^{G}\sum_{u=1}^{U}\bm{q}_{u,g}\big(\bm{\xi}_{u,g}^{(\varepsilon)}\big)^{T}\right\}\right)\odot\bm{M}_{0}. (40)

IV-C3 Masked-Oblique Manifold Update

With the Euclidean gradient in (LABEL:equ.grad_euc_wmmse_lambda), we reuse the masked-oblique update (LABEL:equ.oblique_cg_dir)-(LABEL:equ.lambda_update_zf) with 𝔣()\mathfrak{f}(\cdot) replaced by 𝔣EM()\mathfrak{f}_{\mathrm{EM}}(\cdot) in (LABEL:equ.alpha_subproblem_lambda). The updated per-antenna weights 𝜶m\bm{\alpha}_{m} are extracted from the diagonal blocks of 𝚲\bm{\Lambda} before the digital refresh in Algorithm LABEL:alg:wmmse_tridomain. The initialization uses a maximum-ratio transmission (MRT) digital precoder followed by one global normalization to satisfy the total-power constraint at n=0n=0.

Input: eCSI model; feasible sets {𝒮m}\{\mathcal{S}_{m}\}; power budget PTP_{T}; (Nmax,ϵ,Iinmax=15)(N_{\max},\epsilon,I_{\mathrm{in}}^{\max}=15); step sizes.
Output: Design {𝑾g}\{\bm{W}_{g}\}, {𝜶m}\{\bm{\alpha}_{m}\}, and {𝒑m}\{\bm{p}_{m}\}.
1 Initialization: set 𝒑m(0)=𝒑¯m\bm{p}_{m}^{(0)}=\bar{\bm{p}}_{m} and choose feasible unit-norm 𝜶m(0)\bm{\alpha}_{m}^{(0)}, m\forall m;
2 Build 𝚲(0)\bm{\Lambda}^{(0)} and {𝒉u,g}\{\bm{h}_{u,g}\} via (LABEL:equ.h_def);
3 Digital initialization: obtain {𝑾g(0)}\{\bm{W}_{g}^{(0)}\} by globally normalized MRT, update βu,g(0)\beta_{u,g}^{(0)} and ρu,g(0)\rho_{u,g}^{(0)} via (LABEL:equ.beta_opt)-(LABEL:equ.rho_opt), and evaluate FWMMSE(0)F_{\mathrm{WMMSE}}^{(0)} via (LABEL:equ.wmmse_obj);
4 for n=0n=0 to Nmax1N_{\max}-1 do
5  Channel update: evaluate {𝒒u,m,g}\{\bm{q}_{u,m,g}\} at {𝒑m(n)}\{\bm{p}_{m}^{(n)}\} and form {𝒉u,g}\{\bm{h}_{u,g}\} via (LABEL:equ.h_def);
6  Digital-WMMSE update: update βu,g\beta_{u,g}, ρu,g\rho_{u,g}, and {𝑾g}\{\bm{W}_{g}\} via (LABEL:equ.beta_opt)-(LABEL:equ.w_opt) until the relative objective change is at most ϵ\epsilon or IinmaxI_{\mathrm{in}}^{\max} sweeps;
7  EM update: update 𝚲\bm{\Lambda} via (LABEL:equ.oblique_cg_dir) and (LABEL:equ.lambda_update_zf), and extract {𝜶m(n+1)}\{\bm{\alpha}_{m}^{(n+1)}\};
8  Digital refresh: rebuild {𝒉u,g}\{\bm{h}_{u,g}\} and rerun the inner BCD under the same stopping rule;
9  Position update: form a trial via (LABEL:equ.grad_pos)-(LABEL:equ.pos_update_wmmse); accept it if the sum SE does not decrease and otherwise retain {𝒑m(n)}\{\bm{p}_{m}^{(n)}\};
10  Objective update: update {𝒒u,m,g}\{\bm{q}_{u,m,g}\} and {𝒉u,g}\{\bm{h}_{u,g}\} at {𝒑m(n+1)}\{\bm{p}_{m}^{(n+1)}\}, rerun the inner BCD under the same stopping rule, and evaluate FWMMSE(n+1)F_{\mathrm{WMMSE}}^{(n+1)};
11  if |FWMMSE(n+1)FWMMSE(n)|max{1,|FWMMSE(n)|}ϵ\frac{|F_{\mathrm{WMMSE}}^{(n+1)}-F_{\mathrm{WMMSE}}^{(n)}|}{\max\{1,|F_{\mathrm{WMMSE}}^{(n)}|\}}\leq\epsilon then
12     break;
13    
14  end if
15 
16 end for
17return {𝑾g}\{\bm{W}_{g}\}, {𝜶m}\{\bm{\alpha}_{m}\}, and {𝒑m}\{\bm{p}_{m}\};
Algorithm 2 WMMSE-based tri-domain alternating optimizationwbsphack @@writeaux“newlabelalg:wmmse˙tridomainwcurrentlabel1wesphack

IV-D Spatial-Domain Antenna Position Optimization

wbsphack @@writeaux“newlabelS4.4wcurrentlabel1wesphack

With 𝑾\bm{W}, 𝜶m\bm{\alpha}_{m}, βu,g\beta_{u,g}, and ρu,g\rho_{u,g} fixed, the final block updates 𝒑[𝒑1T,,𝒑MT]T\bm{p}\triangleq[\bm{p}_{1}^{T},\ldots,\bm{p}_{M}^{T}]^{T} by minimizing 𝔣pos(𝒑)g,uρu,gεu,g\mathfrak{f}_{\mathrm{pos}}(\bm{p})\triangleq\sum_{g,u}\rho_{u,g}\varepsilon_{u,g} subject to 𝒑m𝒮m\bm{p}_{m}\in\mathcal{S}_{m}, where the constant g,ulnρu,g-\sum_{g,u}\ln\rho_{u,g} is omitted as in (LABEL:equ.alpha_subproblem_lambda).

IV-D1 Spatial Gradient

Repeating the derivation of Section LABEL:S3.2zf, i.e., using the same decomposition in (LABEL:equ.h_decomp_zf)-(LABEL:equ.v_def_zf) together with hu,m,g,i/𝒑m=j2πλ𝒌tx,i,uhu,m,g,i\partial h_{u,m,g,i}^{*}/\partial\bm{p}_{m}=\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{tx},i,u}h_{u,m,g,i}^{*}, but replacing ξu,m,g(R)\xi_{u,m,g}^{(R)} by ξu,m,g(ε)\xi_{u,m,g}^{(\varepsilon)} from (LABEL:equ.xi_def), gives

wbsphack@@writeaux\newlabelequ.gradposwcurrentlabel1wesphack𝒈mpos\displaystyle wbsphack@@writeaux{}{\newlabel{equ.grad_{p}os}{{wcurrentlabel}{1}}}wesphack\bm{g}_{m}^{\text{pos}} 𝒑m𝔣pos(𝒑)\displaystyle\triangleq\nabla_{\bm{p}_{m}}\mathfrak{f}_{\mathrm{pos}}(\bm{p})
=4πλ{g=1Gu=1Uξu,m,g(ε)i=1Luhu,m,g,i𝒌tx,i,u}.\displaystyle=-\frac{4\pi}{\lambda}\Im\left\{\sum_{g=1}^{G}\sum_{u=1}^{U}\xi_{u,m,g}^{(\varepsilon)}\sum_{i=1}^{L_{u}}h_{u,m,g,i}^{*}\bm{k}_{\mathrm{tx},i,u}\right\}. (41)

IV-D2 Position Update with Projection

The spatial block forms one projected-gradient position trial in each outer iteration. The trial is computed over all antenna positions according to

wbsphack@@writeaux\newlabelequ.posupdatewmmsewcurrentlabel1wesphack𝒑m(n+1)=Π𝒮m(𝒑m(n)ηp𝒈mpos),wbsphack@@writeaux{}{\newlabel{equ.pos_{u}pdate_{w}mmse}{{wcurrentlabel}{1}}}wesphack\bm{p}_{m}^{(n+1)}=\Pi_{\mathcal{S}_{m}}\left(\bm{p}_{m}^{(n)}-\eta_{p}\bm{g}_{m}^{\text{pos}}\right), (42)

which is the WMMSE counterpart of (LABEL:equ.pos_update_zf). The feasible trial is retained if it does not decrease the current sum SE; otherwise, the current positions are preserved before the next outer tri-domain iteration.

V Low-Overhead Movement-Aided CE Scheme

wbsphack @@writeaux“newlabelS5wcurrentlabel1wesphack

The tri-domain optimization developed in Sections LABEL:S3-LABEL:S4 requires a position-dependent eCSI model so that 𝒒u,m,g\bm{q}_{u,m,g} can be evaluated for all users, antennas, and subcarriers at the current antenna positions. In the fixed-position EMRA baseline following [Ying2025TCOM], the same parametric pipeline is applied on the reference array. One-dimensional (1D) ESPRIT is used for delay estimation, followed by two-dimensional (2D) ESPRIT for AoD estimation. Because the array geometry remains fixed, the angular resolution is fundamentally limited by the MM physical spatial samples. For a moderate array size, AoD estimation accuracy degrades and the resulting eCSI quality deteriorates.

We overcome this limitation by exploiting the spatial reconfigurability of SEMRA. At the uplink CE stage, a known training pattern is assigned to each of NpatN_{\mathrm{pat}} training blocks. Within every block, all MM antennas are simultaneously repositioned over NsN_{s} successive position slots, with each antenna visiting the same NsN_{s} prescribed observation positions. The resulting NsMN_{s}M distinct spatial samples form a densified virtual array while the NpatN_{\mathrm{pat}} blocks provide repeated observations under different known patterns. The proposed scheme adopts a parametric estimation framework in which the multipath delays, AoD angles, and an equivalent channel gain vector, all independent of the BS antenna positions, are extracted as intermediate parameters. This position-invariant parameterization decouples CE from the subsequent data-transmission design. During CE, the antennas move to observation positions for high-resolution parameter extraction, and the eCSI at the optimized positions is then assembled directly from the estimated parameters without additional pilot overhead. Since uplink-downlink channel reciprocity holds in TDD, the BS-side angles estimated during the uplink CE phase coincide with the downlink AoD parameters (θi,u,ϕi,u)(\theta_{i,u},\phi_{i,u}) defined in Section LABEL:S2.2, and we retain the downlink notation throughout this section.

V-A Virtual Array Construction

wbsphack @@writeaux“newlabelS5.1wcurrentlabel1wesphack

V-A1 Observation Position Design

For this specific CE procedure, we assume equal inter-element spacings ddy=dzd\triangleq d_{y}=d_{z} and define dmind/4d_{\min}\triangleq d/4. For each BS antenna mm, the reference position 𝒑¯m\bar{\bm{p}}_{m} serves as the center around which Ns=4N_{s}=4 observation positions are defined as

wbsphack@@writeaux\newlabelequ.posschedulewcurrentlabel1wesphack𝒑m(s)=𝒑¯m+𝚫s,s=1,,Ns,wbsphack@@writeaux{}{\newlabel{equ.pos_{s}chedule}{{wcurrentlabel}{1}}}wesphack\bm{p}_{m}^{(s)}=\bar{\bm{p}}_{m}+\bm{\Delta}_{s},\quad s=1,\ldots,N_{s}, (43)

with the displacement vectors

wbsphack@@writeaux\newlabelequ.deltadefwcurrentlabel1wesphack𝚫s=[0ψy,sdminψz,sdmin],(ψy,s,ψz,s){(+1,+1),(+1,1),(1,+1),(1,1)},wbsphack@@writeaux{}{\newlabel{equ.delta_{d}ef}{{wcurrentlabel}{1}}}wesphack\bm{\Delta}_{s}=\begin{bmatrix}0\\ \psi_{y,s}\,d_{\min}\\ \psi_{z,s}\,d_{\min}\end{bmatrix},\;\;(\psi_{y,s},\psi_{z,s})\in\left\{\!\begin{aligned} &(+1,+1),\;(+1,-1),\\ &(-1,+1),\;(-1,-1)\end{aligned}\right\}\!, (44)

for s=1,,4s=1,\ldots,4, respectively. Thus, each observation position is offset from 𝒑¯m\bar{\bm{p}}_{m} by ±dmin\pm d_{\min} along both the yy- and zz-axes. The design requires Dmaxdmin=d/4D_{\max}\geq d_{\min}=d/4 so that all observation positions lie within the feasible region (LABEL:equ.constraint). This is compatible with the non-overlapping condition d2Dmax+Dsepd\geq 2D_{\max}+D_{\mathrm{sep}} in Section LABEL:S2.4. Both constraints are jointly satisfiable whenever Dsep<d/2D_{\mathrm{sep}}<d/2.

The CE stage uses a known set {𝜶ce(t)}t=1Npat\{\bm{\alpha}_{\mathrm{ce}}^{(t)}\}_{t=1}^{N_{\mathrm{pat}}} of real-valued pattern coefficient vectors in K\mathbb{R}^{K}, where 𝜶ce(t)2=1\|\bm{\alpha}_{\mathrm{ce}}^{(t)}\|_{2}=1. Pattern 𝜶ce(t)\bm{\alpha}_{\mathrm{ce}}^{(t)} is held fixed across the NsN_{s} position slots and all antennas within training block tt, so every block retains the coherent virtual-array steering used next. The set is used directly without CE-pattern optimization, while the different patterns provide additional observations for eCSI reconstruction. We consider block-level operation in which all NpatNsN_{\mathrm{pat}}N_{s} pilot slots are completed within one quasi-static channel interval.

Refer to caption
Figure 3: Virtual-array construction for movement-aided CE. Within each training block, the observation positions visited over Ns=4N_{s}=4 position slots interleave an My×MzM_{y}\times M_{z} reference UPA with spacing dd into a uniform 2My×2Mz2M_{y}\times 2M_{z} virtual UPA with spacing dv=d/2d_{v}=d/2. The same position schedule is repeated for NpatN_{\mathrm{pat}} known training patterns. The illustration is shown for My=Mz=4M_{y}=M_{z}=4.wbsphack @@writeaux“newlabelfig:virtual˙upa˙constructionwcurrentlabel1wesphack

V-A2 Virtual UPA Formation

The resulting geometry is illustrated in Fig. LABEL:fig:virtual_upa_construction. Within each training block, the four observation positions of antenna mm form a 2×22\times 2 arrangement centered at 𝒑¯m\bar{\bm{p}}_{m} with spacing 2dmin2d_{\min}, while the closest positions of adjacent antennas are separated by d2dmind-2d_{\min}. Setting dmin=d/4d_{\min}=d/4 makes both spacings equal to d/2d/2, so the NsM=4MN_{s}M=4M samples form a uniform 2My×2Mz2M_{y}\times 2M_{z} virtual UPA with spacing dv=d/2d_{v}=d/2 and aperture (My12)d×(Mz12)d(M_{y}-\tfrac{1}{2})d\times(M_{z}-\tfrac{1}{2})d. This array is used for 2D-ESPRIT AoD estimation, whereas delay estimation continues to rely mainly on frequency-domain shift invariance.

V-B Parametric Channel Estimation

wbsphack @@writeaux“newlabelS5.2wcurrentlabel1wesphack

The discussion assumes the standard subspace-identifiability conditions for the minimum description length (MDL)/ESPRIT pipeline, under which the delay and spatial steering structures are resolvable for the chosen pilot and virtual-array dimensions, and we focus on a generic UE uu.

V-B1 Delay Estimation and sCSI Recovery

In position slot ss of training block tt, UE uu transmits pilots on a dedicated subset 𝒢u={gu,1,,gu,J}\mathcal{G}_{u}=\{g_{u,1},\ldots,g_{u,J}\} of JJ uniformly spaced subcarriers (JG/UJ\leq G/U), where the subsets of different users are disjoint to avoid mutual interference. The signal received at BS antenna mm on pilot subcarrier gjg_{j} is

wbsphack@@writeaux\newlabelequ.ulobswcurrentlabel1wesphackyu,m,gj(t,s)=(hu,m,gj(t,s))su,gj+nu,m,gj(t,s),wbsphack@@writeaux{}{\newlabel{equ.ul_{o}bs}{{wcurrentlabel}{1}}}wesphacky_{u,m,g_{j}}^{(t,s)}=\big(h_{u,m,g_{j}}^{(t,s)}\big)^{*}\,s_{u,g_{j}}+n_{u,m,g_{j}}^{(t,s)}, (45)

where su,gjs_{u,g_{j}} is the known pilot symbol and nu,m,gj(t,s)n_{u,m,g_{j}}^{(t,s)} denotes receiver noise. Dividing by the pilot gives the effective observation y~u,m,gj(t,s)=(hu,m,gj(t,s))+n~u,m,gj(t,s)\tilde{y}_{u,m,g_{j}}^{(t,s)}=\big(h_{u,m,g_{j}}^{(t,s)}\big)^{*}+\tilde{n}_{u,m,g_{j}}^{(t,s)}. Under the downlink convention used here, reciprocity gives the uplink scalar coefficient as (hu,m,gj(t,s))\big(h_{u,m,g_{j}}^{(t,s)}\big)^{*}. The frequency-domain observations from all NpatNsMN_{\mathrm{pat}}N_{s}M position-pattern pairs are divided by the known pilot symbols and stacked column-wise into a data matrix 𝒀dJ×NpatNsM\bm{Y}_{d}\in\mathbb{C}^{J\times N_{\mathrm{pat}}N_{s}M}. Because the multipath channel (LABEL:equ.ch) is a sum of complex exponentials across frequency, 𝒀d\bm{Y}_{d} can be written as 𝒀d=𝑬d𝑿d+𝑵d\bm{Y}_{d}=\bm{E}_{d}\bm{X}_{d}+\bm{N}_{d}. Here, 𝑬dJ×Lu\bm{E}_{d}\in\mathbb{C}^{J\times L_{u}} is a Vandermonde-structured delay steering matrix, 𝑿dLu×NpatNsM\bm{X}_{d}\in\mathbb{C}^{L_{u}\times N_{\mathrm{pat}}N_{s}M} collects the equivalent path gains across the position-pattern pairs, and 𝑵d\bm{N}_{d} is the corresponding noise matrix.

The number of multipath components LuL_{u} is estimated, rather than assumed known a priori, by applying the MDL criterion to the eigenvalues of the sample covariance matrix 𝒀d𝒀dHJ×J\bm{Y}_{d}\bm{Y}_{d}^{H}\in\mathbb{C}^{J\times J} [Wax1985Detection], yielding the model-order estimate L^u\hat{L}_{u}. The estimated order L^u\hat{L}_{u} is then used in the subsequent 1D-ESPRIT stage.

With L^u\hat{L}_{u} determined, the 1D-ESPRIT algorithm extracts the multipath delays {τ^i,u}i=1L^u\{\hat{\tau}_{i,u}\}_{i=1}^{\hat{L}_{u}} from the shift-invariance structure of 𝑬d\bm{E}_{d}. The full-band sCSI h^u,m,g(t,s)\hat{h}_{u,m,g}^{(t,s)} is then recovered by delay-domain least-squares (LS) fitting for every training block at the CE observation positions, i.e., the My×MzM_{y}\times M_{z} reference-array positions for EMRA and the NsMN_{s}M positions of the 2My×2Mz2M_{y}\times 2M_{z} virtual UPA for SEMRA.

V-B2 AoD Estimation on the Virtual Array

At each observation position 𝒑m(s)\bm{p}_{m}^{(s)} in training block tt, the spatial-domain channel on subcarrier gg can be expressed as

wbsphack@@writeaux\newlabelequ.chspatialwcurrentlabel1wesphackhu,m,g(t,s)=i=1Lux˙i,u,g(t)ej2πλ𝒌tx,i,uT𝒑m(s),wbsphack@@writeaux{}{\newlabel{equ.ch_{s}patial}{{wcurrentlabel}{1}}}wesphackh_{u,m,g}^{(t,s)}=\sum_{i=1}^{L_{u}}\dot{x}_{i,u,g}^{(t)}\,e^{-\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{tx},i,u}^{T}\bm{p}_{m}^{(s)}}, (46)

where 𝝎i,u[ωk(θi,u,ϕi,u)]k=1K\bm{\omega}_{i,u}\triangleq[\omega_{k}(\theta_{i,u},\phi_{i,u})]_{k=1}^{K}. The block-dependent composite gain is

wbsphack@@writeaux\newlabelequ.xdotblockwcurrentlabel1wesphackx˙i,u,g(t)xi,u,gfrx,u(ϑi,u,φi,u)((𝜶ce(t))T𝝎i,u)×ej2πλ𝒌rx,i,uT𝒓u.wbsphack@@writeaux{}{\newlabel{equ.xdot_{b}lock}{{wcurrentlabel}{1}}}wesphack\begin{aligned} \dot{x}_{i,u,g}^{(t)}&\triangleq x_{i,u,g}\,f_{\mathrm{rx},u}(\vartheta_{i,u},\varphi_{i,u})\big((\bm{\alpha}_{\mathrm{ce}}^{(t)})^{T}\bm{\omega}_{i,u}\big)\\[-2.84526pt] &\quad\times e^{-\mathrm{j}\frac{2\pi}{\lambda}\bm{k}_{\mathrm{rx},i,u}^{T}\bm{r}_{u}}.\end{aligned} (47)

Within training block tt, the radiation pattern 𝜶ce(t)\bm{\alpha}_{\mathrm{ce}}^{(t)} is fixed across all position slots and antennas. Hence, the composite gain x˙i,u,g(t)\dot{x}_{i,u,g}^{(t)} does not depend on the antenna index mm or the position-slot index ss, although it can change with tt. For every training block, the 4M4M spatial observations therefore exhibit the same canonical steering-vector structure of a 2My×2Mz2M_{y}\times 2M_{z} UPA with inter-element spacing dvd_{v}. The standard 2D-ESPRIT procedure [Liao2019TCOM] can consequently combine the training blocks after the re-indexing and spatial-smoothing steps described next. Concretely, the virtual-array pilot observations are arranged into a data matrix of size NsM×NpatJN_{s}M\times N_{\mathrm{pat}}J, where rows index the NsMN_{s}M virtual-array positions and columns jointly index the training pattern and pilot subcarrier. The columns share the same spatial steering structure while their composite gains depend on the training pattern and subcarrier. The 2D-ESPRIT algorithm estimates the spatial frequencies {μi,u,νi,u}\{\mu_{i,u},\nu_{i,u}\} from the sample covariance of this matrix after spatial smoothing.

Let {μ^i,u,ν^i,u}i=1L^u\{\hat{\mu}_{i,u},\hat{\nu}_{i,u}\}_{i=1}^{\hat{L}_{u}} denote the spatial frequencies returned by 2D-ESPRIT, and define the virtual-array spatial frequency constant κv2πdv/λ\kappa_{v}\triangleq 2\pi d_{v}/\lambda. Under the adopted virtual-array indexing, the spatial frequencies are related to the AoD by μi,u=κvsinθi,usinϕi,u\mu_{i,u}=\kappa_{v}\sin\theta_{i,u}\sin\phi_{i,u} and νi,u=κvcosθi,u\nu_{i,u}=-\kappa_{v}\cos\theta_{i,u}. The minus sign in νi,u\nu_{i,u} follows from the phase convention in (LABEL:equ.ch_spatial) together with the chosen indexing along the zz-direction of the virtual UPA. We focus on BS-side departures in the front half-space, i.e., ϕi,u[π/2,π/2]\phi_{i,u}\in[-\pi/2,\pi/2]. The AoD angles are then recovered on the principal branches as

wbsphack@@writeaux\newlabelequ.anglerecoverywcurrentlabel1wesphackθ^i,u=arccos(ν^i,uκv),ϕ^i,u=arcsin(μ^i,uκv2ν^i,u2),wbsphack@@writeaux{}{\newlabel{equ.angle_{r}ecovery}{{wcurrentlabel}{1}}}wesphack\hat{\theta}_{i,u}=\arccos\!\left(-\frac{\hat{\nu}_{i,u}}{\kappa_{v}}\right),\;\;\hat{\phi}_{i,u}=\arcsin\!\left(\frac{\hat{\mu}_{i,u}}{\sqrt{\kappa_{v}^{2}-\hat{\nu}_{i,u}^{2}}}\right), (48)

where θ^i,u[0,π]\hat{\theta}_{i,u}\in[0,\pi] and ϕ^i,u[π/2,π/2]\hat{\phi}_{i,u}\in[-\pi/2,\pi/2] by construction. Numerically, the arguments of arccos()\arccos(\cdot) and arcsin()\arcsin(\cdot) are clipped to their feasible intervals before inversion. This mapping is one-to-one over the front half-space when dvλ/2d_{v}\leq\lambda/2. For dv>λ/2d_{v}>\lambda/2, we retain the same principal-branch inversion, while global uniqueness is not claimed.

With the estimated AoD pairs {(θ^i,u,ϕ^i,u)}\{(\hat{\theta}_{i,u},\hat{\phi}_{i,u})\}, the estimated direction vector is 𝒌^tx,i,u=[sinθ^i,ucosϕ^i,u,sinθ^i,usinϕ^i,u,cosθ^i,u]T\hat{\bm{k}}_{\mathrm{tx},i,u}=[\sin\hat{\theta}_{i,u}\cos\hat{\phi}_{i,u},\,\sin\hat{\theta}_{i,u}\sin\hat{\phi}_{i,u},\,\cos\hat{\theta}_{i,u}]^{T}. The steering matrix 𝑩^u,mL^u×L^u\hat{\bm{B}}_{u,m}\in\mathbb{C}^{\hat{L}_{u}\times\hat{L}_{u}} is then assembled at any desired position 𝒑m\bm{p}_{m} via [𝒃^u,m]i=ej2πλ𝒌^tx,i,uT𝒑m[\hat{\bm{b}}_{u,m}]_{i}=e^{-\mathrm{j}\frac{2\pi}{\lambda}\hat{\bm{k}}_{\mathrm{tx},i,u}^{T}\bm{p}_{m}} (cf. Section LABEL:S2.2), and the basis-function matrix 𝛀^uL^u×K\hat{\bm{\Omega}}_{u}\in\mathbb{R}^{\hat{L}_{u}\times K} is constructed as [𝛀^u]i,k=ωk(θ^i,u,ϕ^i,u)[\hat{\bm{\Omega}}_{u}]_{i,k}=\omega_{k}(\hat{\theta}_{i,u},\hat{\phi}_{i,u}) (cf. Section LABEL:S2.3).

V-B3 Equivalent Channel Gain Estimation and eCSI Reconstruction

Mirroring (LABEL:equ.v_def), define 𝝌~u,gL^u\tilde{\bm{\chi}}_{u,g}\in\mathbb{C}^{\hat{L}_{u}} as the equivalent path-gain vector under the estimated parametric model. The reconstructed eCSI then satisfies 𝒒^u,m,gH=𝝌~u,gH𝑩^u,m𝛀^u\hat{\bm{q}}_{u,m,g}^{H}=\tilde{\bm{\chi}}_{u,g}^{H}\hat{\bm{B}}_{u,m}\hat{\bm{\Omega}}_{u}, where only 𝑩^u,m\hat{\bm{B}}_{u,m} depends on 𝒑m\bm{p}_{m}.

Accordingly, each virtual observation satisfies

wbsphack@@writeaux\newlabelequ.obsmodelwcurrentlabel1wesphackh^u,m,g(t,s)=(𝜶ce(t))T𝛀^uT𝑩^u,m(s)H𝝌~u,g,m,s,t,wbsphack@@writeaux{}{\newlabel{equ.obs_{m}odel}{{wcurrentlabel}{1}}}wesphack\hat{h}_{u,m,g}^{(t,s)*}=(\bm{\alpha}_{\mathrm{ce}}^{(t)})^{T}\hat{\bm{\Omega}}_{u}^{T}\hat{\bm{B}}_{u,m}^{(s)H}\tilde{\bm{\chi}}_{u,g},\quad\forall m,s,t, (49)

where 𝑩^u,m(s)\hat{\bm{B}}_{u,m}^{(s)} is evaluated at the observation position 𝒑m(s)\bm{p}_{m}^{(s)}. Concatenating all NpatNsMN_{\mathrm{pat}}N_{s}M observations into a single linear system gives 𝒉^u,g=𝚼u𝝌~u,g\hat{\bm{h}}_{u,g}^{*}=\bm{\Upsilon}_{u}\tilde{\bm{\chi}}_{u,g}, with the observation vector 𝒉^u,gNpatNsM\hat{\bm{h}}_{u,g}^{*}\in\mathbb{C}^{N_{\mathrm{pat}}N_{s}M} and the measurement matrix

wbsphack@@writeaux\newlabelequ.Upsilondefwcurrentlabel1wesphack𝚼u[𝚼u(1)𝚼u(Npat)]NpatNsM×L^u,𝚼u(t)[𝑩^u,1(1)𝛀^u𝜶ce(t),,𝑩^u,M(Ns)𝛀^u𝜶ce(t)]H.wbsphack@@writeaux{}{\newlabel{equ.Upsilon_{d}ef}{{wcurrentlabel}{1}}}wesphack\begin{aligned} \bm{\Upsilon}_{u}&\triangleq\begin{bmatrix}\bm{\Upsilon}_{u}^{(1)}\\[-2.84526pt] \vdots\\[-2.84526pt] \bm{\Upsilon}_{u}^{(N_{\mathrm{pat}})}\end{bmatrix}\in\mathbb{C}^{N_{\mathrm{pat}}N_{s}M\times\hat{L}_{u}},\\ \bm{\Upsilon}_{u}^{(t)}&\triangleq\left[\hat{\bm{B}}_{u,1}^{(1)}\hat{\bm{\Omega}}_{u}\bm{\alpha}_{\mathrm{ce}}^{(t)},\,\ldots,\,\hat{\bm{B}}_{u,M}^{(N_{s})}\hat{\bm{\Omega}}_{u}\bm{\alpha}_{\mathrm{ce}}^{(t)}\right]^{H}.\end{aligned} (50)

Provided that the columns of 𝚼u\bm{\Upsilon}_{u} induced by the recovered paths are linearly independent, the matrix 𝚼u\bm{\Upsilon}_{u} has full column rank. This requires nondegenerate steering responses over the NsMN_{s}M distinct observation positions, collectively nonzero responses from the training pattern set on every recovered path, and NpatNsML^uN_{\mathrm{pat}}N_{s}M\geq\hat{L}_{u}. Under this condition, the LS solution is

wbsphack@@writeaux\newlabelequ.vlswcurrentlabel1wesphack𝝌^u,g=(𝚼uH𝚼u)1𝚼uH𝒉^u,g,g.wbsphack@@writeaux{}{\newlabel{equ.v_{l}s}{{wcurrentlabel}{1}}}wesphack\hat{\bm{\chi}}_{u,g}=\left(\bm{\Upsilon}_{u}^{H}\bm{\Upsilon}_{u}\right)^{-1}\bm{\Upsilon}_{u}^{H}\hat{\bm{h}}_{u,g}^{*},\quad\forall g. (51)

The eCSI at any target position 𝒑m\bm{p}_{m} then follows from

wbsphack@@writeaux\newlabelequ.qreconwcurrentlabel1wesphack𝒒^u,m,gH=𝝌^u,gH𝑩^u,m𝛀^u,m,g,wbsphack@@writeaux{}{\newlabel{equ.q_{r}econ}{{wcurrentlabel}{1}}}wesphack\hat{\bm{q}}_{u,m,g}^{H}=\hat{\bm{\chi}}_{u,g}^{H}\hat{\bm{B}}_{u,m}\hat{\bm{\Omega}}_{u},\quad\forall m,g, (52)

where 𝑩^u,m\hat{\bm{B}}_{u,m} is evaluated at the target position, either the reference position 𝒑¯m\bar{\bm{p}}_{m} or the optimized position from Sections LABEL:S3-LABEL:S4. Because only 𝑩^u,m\hat{\bm{B}}_{u,m} depends on 𝒑m\bm{p}_{m}, the CE-phase observation positions and the data-transmission positions need not coincide.

This CE procedure uses Npat=3N_{\mathrm{pat}}=3 training blocks and Ns=4N_{s}=4 position slots per block, with JJ pilot subcarriers per slot per UE. The total per-user pilot overhead is therefore NpatNsJ=360N_{\mathrm{pat}}N_{s}J=360 symbols. For fair comparison, the EMRA baseline retains its fixed physical geometry over the same NpatNs=12N_{\mathrm{pat}}N_{s}=12 slots and uses the same training pattern set and the same JJ pilot subcarriers per slot, yielding the same total per-user pilot overhead.

Algorithm LABEL:alg:sera_ce summarizes the virtual-array parameter estimation and target-position eCSI assembly.

Input: Pilot observations {yu,m,gj(t,s)}\{y_{u,m,g_{j}}^{(t,s)}\}; training patterns; pilot allocation; position schedule.
Output: Position-dependent eCSI model.
1 Generate {𝒑m(s)}\{\bm{p}_{m}^{(s)}\} via (LABEL:equ.pos_schedule);
2 for u=1u=1 to UU do
3  Pilot stacking: form 𝒀d\bm{Y}_{d} from the normalized observations;
4  Model order: estimate L^u\hat{L}_{u} via MDL on 𝒀d𝒀dH\bm{Y}_{d}\bm{Y}_{d}^{H} [Wax1985Detection];
5  Delays: estimate {τ^i,u}\{\hat{\tau}_{i,u}\} via 1D-ESPRIT;
6  sCSI: recover {h^u,m,g(t,s)}\{\hat{h}_{u,m,g}^{(t,s)}\} on the virtual array by delay-domain LS;
7  Spatial frequencies: form the virtual-array snapshots and estimate {(μ^i,u,ν^i,u)}\{(\hat{\mu}_{i,u},\hat{\nu}_{i,u})\} via 2D-ESPRIT;
8  AoDs: obtain {(θ^i,u,ϕ^i,u)}\{(\hat{\theta}_{i,u},\hat{\phi}_{i,u})\} via (LABEL:equ.angle_recovery);
9  Basis construction: build 𝛀^u\hat{\bm{\Omega}}_{u} and the observation-position steering matrices;
10  Equivalent gains: assemble 𝚼u\bm{\Upsilon}_{u} and estimate 𝝌^u,g\hat{\bm{\chi}}_{u,g} via (LABEL:equ.v_ls), g\forall g;
11  Optional target assembly: obtain 𝒒^u,m,g\hat{\bm{q}}_{u,m,g} via (LABEL:equ.q_recon);
12 
13 end for
14return position-dependent eCSI model;
Algorithm 3 Movement-aided parametric CE for SEMRAwbsphack @@writeaux“newlabelalg:sera˙cewcurrentlabel1wesphack

VI Simulation Results

wbsphack @@writeaux“newlabelS6wcurrentlabel1wesphack

VI-A Simulation Setup

wbsphack @@writeaux“newlabelS6.1wcurrentlabel1wesphack

Table LABEL:tab:simulation_parameters consolidates the default simulation configuration. Unless otherwise stated, each parameter sweep changes only the quantity identified on its horizontal axis while retaining the remaining settings in the table.

TABLE I: Default simulation parameters.wbsphack @@writeaux“newlabeltab:simulation˙parameterswcurrentlabel1wesphack
Parameter Default value
Carrier and waveform fc=2.4f_{c}=2.4 GHz, Δf=30\Delta f=30 kHz, G=128G=128
Reference array and users My=Mz=4M_{y}=M_{z}=4 (M=16)(M=16), U=3U=3
Radiation-basis size K=102K=10^{2}
Multipath channel Lu=6L_{u}=6, στ=100\sigma_{\tau}=100 ns
Movement geometry d=λ/2d=\lambda/2, Dsep=0.01λD_{\mathrm{sep}}=0.01\lambda, Dmax=0.245λD_{\max}=0.245\lambda
Channel-estimation geometry Ns=4N_{s}=4, dv=λ/4d_{v}=\lambda/4, 8×88\times 8 virtual array
CE training schedule Npat=3N_{\mathrm{pat}}=3, NpatNs=12N_{\mathrm{pat}}N_{s}=12 total slots
Pilot allocation J=30J=30 subcarriers per user per slot
Signal-to-noise ratio (SNR) SNRC=SNRP=20\text{SNR}_{\text{C}}=\text{SNR}_{\text{P}}=20 dB, with 03030 dB sweeps
Outer iteration budget Nmax=10N_{\max}=10
Monte Carlo averaging 10001000 independent channel realizations

We compare SEMRA-ZF and SEMRA-WMMSE with EMRA and with TFA/SMA baselines using downtilt, 3GPP 38.901, and isotropic patterns. EMRA and SEMRA are reported under both estimated and perfect eCSI, while TFA/SMA are reported under perfect sCSI.

The evaluation focuses on communication-rate gains under a fixed transmit-power constraint. Actuator energy, movement latency, and calibration overhead are outside the present SE metric.

VI-B Channel Estimation Accuracy

wbsphack @@writeaux“newlabelS6.2wcurrentlabel1wesphack

We evaluate CE with NMSE-Su,m,g,t,s|h^u,m,g(t,s)hu,m,g(t,s)|2u,m,g,t,s|hu,m,g(t,s)|2\text{NMSE-S}\triangleq\frac{\sum_{u,m,g,t,s}|\hat{h}_{u,m,g}^{(t,s)}-h_{u,m,g}^{(t,s)}|^{2}}{\sum_{u,m,g,t,s}|h_{u,m,g}^{(t,s)}|^{2}} at each scheme’s CE positions and NMSE-Eu,m,g𝒒^u,m,g𝒒u,m,g22u,m,g𝒒u,m,g22\text{NMSE-E}\triangleq\frac{\sum_{u,m,g}\|\hat{\bm{q}}_{u,m,g}-\bm{q}_{u,m,g}\|_{2}^{2}}{\sum_{u,m,g}\|\bm{q}_{u,m,g}\|_{2}^{2}} on a common reference array. Both ratios are computed per channel realization, converted to decibels, and then averaged. EMRA and SEMRA use the same per-user pilot overhead NpatNsJ=360N_{\mathrm{pat}}N_{s}J=360; orthogonal matching pursuit (OMP) benchmarks supplement the ESPRIT-based pipeline following [Ying2025TCOM]. NMSE-S reflects the delay/sCSI stage, whereas NMSE-E with estimated or oracle sCSI reflects eCSI reconstruction.

Refer to caption
Figure 4: NMSE-S and NMSE-E versus SNRC\text{SNR}_{\text{C}} at d=λ/2d=\lambda/2 for the EMRA baseline and the proposed SEMRA CE scheme.wbsphack @@writeaux“newlabelfig:nmse˙vs˙snr˙cwcurrentlabel1wesphack

Fig. LABEL:fig:nmse_vs_snr_c shows both metrics versus SNRC\text{SNR}_{\text{C}} at d=λ/2d=\lambda/2. For NMSE-S, the EMRA and SEMRA curves almost overlap for each algorithm because this metric is dominated by frequency-domain delay estimation and LS-based sCSI recovery rather than by spatial sampling density. ESPRIT decreases steadily to about 34.1-34.1 dB at 3030 dB, whereas OMP saturates near 17.5-17.5 dB due to its finite dictionary resolution. For NMSE-E with estimated sCSI, SEMRA+ESPRIT reaches about 35.5-35.5 dB at 3030 dB, compared with about 22.9-22.9 dB for EMRA+ESPRIT. Their perfect-sCSI counterparts reach about 39.8-39.8 dB and 23.4-23.4 dB, respectively. In contrast, the OMP-based NMSE-E curves remain around 4-4 to 5-5 dB at moderate and high SNR and show little separation between EMRA and SEMRA, indicating a dictionary-limited error floor. Thus, movement leaves sCSI recovery nearly unchanged, while the denser virtual array improves ESPRIT-based AoD estimation and eCSI reconstruction.

Refer to caption
Figure 5: NMSE-S and NMSE-E versus inter-element spacing d/λd/\lambda at SNRC=20\text{SNR}_{\text{C}}=20 dB for the EMRA baseline and the proposed SEMRA CE scheme.wbsphack @@writeaux“newlabelfig:nmse˙vs˙spacingwcurrentlabel1wesphack

Fig. LABEL:fig:nmse_vs_spacing examines sensitivity to antenna spacing at SNRC=20\text{SNR}_{\text{C}}=20 dB. The NMSE-S curves are nearly flat, with ESPRIT near 25.1-25.1 dB and OMP near 17.1-17.1 dB, because spacing mainly changes the spatial geometry used for AoD estimation and eCSI reconstruction rather than delay estimation and sCSI recovery. For NMSE-E, EMRA+ESPRIT initially benefits from the larger aperture and reaches a minimum of about 21.1-21.1 dB at d=0.8λd=0.8\lambda with estimated sCSI. Beyond this point, sparse physical sampling causes principal-branch ambiguity and grating lobes, and its NMSE-E rises sharply to about 1.01.0 dB at d=1.5λd=1.5\lambda. SEMRA+ESPRIT shows substantially greater robustness. It reaches about 27.4-27.4 dB at d=0.5λd=0.5\lambda and remains near 32-32 dB from 0.9λ0.9\lambda to 1.5λ1.5\lambda. Its perfect-sCSI curve further improves from about 40-40 dB at d=0.9λd=0.9\lambda to about 43-43 dB at d=1.5λd=1.5\lambda. This advantage arises because the prescribed repositioning forms NsMN_{s}M distinct virtual spatial samples and observes them under NpatN_{\mathrm{pat}} known training patterns while retaining the same total pilot overhead as EMRA. The resulting denser observation geometry mitigates ambiguity at large spacing. SEMRA+OMP also outperforms EMRA+OMP from about 0.6λ0.6\lambda onward, but degrades beyond 1.2λ1.2\lambda because the finite dictionary remains sensitive to off-grid mismatch. Thus, movement-aided CE mainly improves eCSI reconstruction at large antenna spacings while leaving sCSI recovery nearly unchanged.

VI-C Spectral Efficiency

wbsphack @@writeaux“newlabelS6.3wcurrentlabel1wesphack

Unless otherwise stated, the SE plots compare EMRA, SEMRA-ZF, and SEMRA-WMMSE under perfect and estimated eCSI, with TFA/SMA baselines under perfect sCSI. EMRA versus SEMRA-ZF isolates the spatial-reconfiguration gain under normalized ZF, while SEMRA-WMMSE shows the additional digital-domain gain; estimated-eCSI curves use the ESPRIT-based CE pipeline and SNRP=20\text{SNR}_{\text{P}}=20 dB by default.

VI-C1 SE versus SNRC\text{SNR}_{\text{C}}

Refer to caption
Figure 6: Sum SE versus SNRC\text{SNR}_{\text{C}} when SNRP=20\text{SNR}_{\text{P}}=20 dB and d=λ/2d=\lambda/2.wbsphack @@writeaux“newlabelfig:se˙vs˙snr˙cwcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_snr_c plots SE versus SNRC\text{SNR}_{\text{C}}. All estimated-eCSI curves increase with SNRC\text{SNR}_{\text{C}} and approach their perfect-eCSI benchmarks, while the ordering SEMRA-WMMSE >> SEMRA-ZF >> EMRA is maintained over the whole range. The improvement is especially visible for EMRA, whose estimated-eCSI SE rises from the low-SNR regime but still remains below the SEMRA curves at high SNRC\text{SNR}_{\text{C}}. At SNRC=20\text{SNR}_{\text{C}}=20 dB, the SEMRA perfect-versus-estimated gaps remain within about 0.500.50 bps/Hz, much smaller than the EMRA gap, which is consistent with the lower NMSE-E in Fig. LABEL:fig:nmse_vs_snr_c. The remaining SEMRA-ZF gain under perfect eCSI confirms that spatial reconfiguration improves SE beyond CE accuracy alone, while SEMRA-WMMSE further benefits from stronger digital precoding.

VI-C2 SE versus SNRP\text{SNR}_{\text{P}}

Refer to caption
Figure 7: Sum SE versus SNRP\text{SNR}_{\text{P}} when SNRC=20\text{SNR}_{\text{C}}=20 dB and d=λ/2d=\lambda/2.wbsphack @@writeaux“newlabelfig:se˙vs˙snr˙pwcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_snr_p shows SE versus SNRP\text{SNR}_{\text{P}} at SNRC=20\text{SNR}_{\text{C}}=20 dB. All curves increase approximately linearly with SNRP\text{SNR}_{\text{P}}, and the SEMRA advantage over EMRA persists across the entire range. At the default SNRP=20\text{SNR}_{\text{P}}=20 dB, SEMRA-WMMSE remains above EMRA and the fixed-pattern TFA/SMA baselines under estimated eCSI. The smaller perfect-versus-estimated gaps of the two SEMRA schemes again reflect their more accurate eCSI, while the persistent WMMSE-over-ZF gap of about 1.701.70 bps/Hz at high SNRP\text{SNR}_{\text{P}} shows the benefit of the stronger digital design.

VI-C3 SE versus Antenna Spacing

Refer to caption
Figure 8: Sum SE versus inter-element spacing d/λd/\lambda when SNRC=20\text{SNR}_{\text{C}}=20 dB and SNRP=20\text{SNR}_{\text{P}}=20 dB.wbsphack @@writeaux“newlabelfig:se˙vs˙spacingwcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_spacing examines the impact of spacing when both SNRs are fixed at 2020 dB. Varying dd jointly changes the physical aperture, CE-side aliasing behavior, and, for SEMRA, the feasible motion region. The perfect-eCSI curves increase with dd, reflecting the aperture gain available without estimation errors. Under estimated eCSI, EMRA first benefits from aperture enlargement and peaks at about 15.4015.40 bps/Hz around d0.8λd\approx 0.8\lambda, but then collapses to about 6.006.00 bps/Hz at d=1.5λd=1.5\lambda, leaving the marked gap of about 10.7010.70 bps/Hz to its perfect-eCSI counterpart. This SE turning point coincides with the NMSE-E minimum at d=0.8λd=0.8\lambda in Fig. LABEL:fig:nmse_vs_spacing, showing that the aperture benefit remains effective up to the point of the most accurate eCSI reconstruction. Once dd grows further, principal-branch ambiguity and grating lobes dominate, so the EMRA estimated-eCSI SE drops sharply despite the larger aperture. By contrast, SEMRA-ZF and SEMRA-WMMSE increase monotonically with dd and remain close to their perfect-eCSI benchmarks because SEMRA maintains low eCSI NMSE over the large-spacing region. At d=1.5λd=1.5\lambda, their estimated-eCSI SEs are about 17.0017.00 and 18.6018.60 bps/Hz, respectively, confirming SEMRA’s stronger robustness to enlarged inter-element spacings.

VI-C4 SE versus Number of Users

Refer to caption
Figure 9: Sum SE and paired gain over EMRA versus UU when SNRP=20\text{SNR}_{\text{P}}=20 dB and d=λ/2d=\lambda/2 under perfect eCSI.wbsphack @@writeaux“newlabelfig:se˙vs˙userswcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_users reports the sum SE and paired SEMRA gains over EMRA for U=1,,7U=1,\ldots,7 under the default physical and channel settings. All three designs improve from the single-user case through moderate user loads, while both SEMRA variants retain positive paired gains over EMRA throughout the sweep. At U=4U=4 and U=5U=5, the SEMRA-ZF gains over EMRA are 0.690.69 and 0.790.79 bps/Hz, while the SEMRA-WMMSE gains are 3.603.60 and 4.644.64 bps/Hz. The lower panel removes the common variation in absolute SE and shows that the SEMRA-WMMSE gain becomes more pronounced as UU increases, which is consistent with the greater role of joint spatial, EM, and digital optimization under heavier multiuser interference. SEMRA-ZF also remains above EMRA for every tested UU, confirming a benefit from position optimization under the same normalized ZF design.

VI-C5 SE versus Maximum Antenna Displacement

Refer to caption
Figure 10: Sum SE versus Dmax/λD_{\max}/\lambda when d=λ/2d=\lambda/2, SNRC=20\text{SNR}_{\text{C}}=20 dB, and SNRP=20\text{SNR}_{\text{P}}=20 dB.wbsphack @@writeaux“newlabelfig:se˙vs˙dmaxwcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_dmax reports monotonic SE increases across the feasible displacement range. Since dd is held fixed, the changes isolate the value of additional position freedom from any enlargement of the reference aperture. At Dmax=0D_{\max}=0, the antennas retain the reference geometry while the EM coefficients and digital precoders remain optimized. Relative to this reference, the endpoint gains across the two designs and CSI conditions range from 0.130.13 to 0.620.62 bps/Hz, and the perfect-eCSI curves remain above their estimated-eCSI counterparts. Most of the endpoint improvement is already obtained at Dmax=0.20λD_{\max}=0.20\lambda, so the curves begin to level before the full displacement allowance is reached. With d=λ/2d=\lambda/2 and Dsep=0.01λD_{\mathrm{sep}}=0.01\lambda, the collision-limited displacement bound is 0.245λ0.245\lambda. Both designs therefore benefit from the available local position freedom, and SEMRA-WMMSE yields the highest SE for every tested DmaxD_{\max}.

VI-C6 Convergence versus Outer Iteration Index

Refer to caption
Figure 11: Sum SE versus the outer iteration index under perfect eCSI when SNRP=20\text{SNR}_{\text{P}}=20 dB and d=λ/2d=\lambda/2. The default iteration budget is Nmax=10N_{\max}=10.wbsphack @@writeaux“newlabelfig:se˙vs˙iterationwcurrentlabel1wesphack

Fig. LABEL:fig:se_vs_iteration shows that the two algorithms obtain most of their observed improvement within the default iteration budget. The first 1010 iterations capture 99.3%99.3\% and 96.4%96.4\% of the improvements observed by iteration 2020, leaving only 0.040.04 and 0.250.25 bps/Hz of additional gain. SEMRA-ZF plateaus earlier, while SEMRA-WMMSE continues refining the solution and approaches a higher SE more gradually. These trends support Nmax=10N_{\max}=10 as a practical budget for both methods, while SEMRA-WMMSE remains above SEMRA-ZF throughout the common iteration horizon.

VII Conclusions

wbsphack @@writeaux“newlabelS7wcurrentlabel1wesphack

The proposed SEMRA framework combines movement-aided parametric CE with spatial, EM, and digital precoding for multiuser MIMO. During training, coordinated repositioning synthesizes a denser virtual array under the same per-user pilot overhead as EMRA, enabling AoD estimation and eCSI assembly at the transmission positions. SEMRA-ZF isolates the gain of spatial reconfiguration, while SEMRA-WMMSE further coordinates interference and power allocation. The simulations show improved eCSI accuracy and higher SE than EMRA, with larger gains at wider antenna spacings, positive paired gains across the tested user loads, and continued improvement throughout the tested displacement range. The convergence study further shows that the default ten iterations capture most of the improvement observed over twenty iterations and support the adopted ten-iteration budget. The study considers quasi-static geometric channels and block-level movement, and future work will examine measurement-based validation, online position control, and scalable optimization for larger systems.

winput@maintmp.bbl

uxtagasecondoftwo