Multi-User Localization via Active Sensing with Electromagnetically Reconfigurable Antennas
Abstract
This paper investigates multi-user localization in uplink wireless systems assisted by electromagnetically reconfigurable antennas (ERAs). Unlike traditional localization schemes, we formulate an active sensing problem where a base station (BS) exploits historical pilot observations accumulated over previous sensing stages to adapt the shared ERA configuration and progressively refine position estimates. To capture both theoretical flexibility and practical hardware constraints, we establish a unified wideband geometric signal model accommodating two complementary ERA paradigms: a synthesis-based model utilizing spherical-harmonic basis functions, and a finite-state model based on measured radiation codebooks. Because analytically solving the resulting joint design problem is highly intractable due to the high-dimensional observation and the shared-aperture coupling among multiple users, we develop a learning-based active sensing framework. Specifically, pilot-matched wideband observations are compressed into compact user-wise features and sequentially accumulated by a long short-term memory (LSTM) module. These temporal features are then processed by a graph neural network (GNN) to capture multi-user shared-aperture coupling. Model-specific output heads generate either continuous synthesis coefficients or finite-state ERA selections, while a localization head produces stage-wise position estimates. Numerical results under a specific channel distribution show that the proposed ERA-assisted active sensing framework achieves progressive localization refinement across sensing stages and obtains better performance than conventional non-reconfigurable arrays and representative ablation baselines.
I Introduction
Accurate and reliable wireless localization has emerged as a key ability for next-generation 6G networks [1, 2]. It enables a multitude of location-critical applications, including autonomous driving and robotic navigation to immersive extended reality (XR) [3, 4]. To support these services, future localization systems must achieve ultra-high precision, even in complex propagation environments characterized by dense multi-user deployments and multipath fading [5].
To achieve high-precision localization, existing base stations (BSs) typically rely on massive multiple-input multiple-output (MIMO) technology [6]. By deploying large antenna arrays, MIMO systems provide high spatial resolution and abundant spatial degrees of freedom (DoFs), enabling the separation of closely spaced users and multipath components [7, 8]. However, conventional MIMO systems typically employ static and fixed-pattern antenna elements. The spatial flexibility of the MIMO is purely derived from adjusting the baseband or radio frequency (RF) phase shifts across the array [9]. Consequently, the sensing capability of the MIMO is fundamentally constrained by the static electromagnetic (EM) characteristics.
To overcome the above limitations, electromagnetically reconfigurable antennas (ERAs) have recently emerged as a promising technology for next-generation wireless systems [10, 11, 12, 13, 14, 15]. Unlike conventional antenna arrays with fixed electromagnetic responses, ERAs retain a fixed physical location while reshaping their metallic patterns or dielectric substrates to provide on-demand control over operating frequency, polarization, and radiation pattern [16]. The capability introduces an additional DoF in EM domain, enabling adaptive directional response control. As a result, ERAs can potentially improve beam adaptability and sensing performance in complex wireless systems [17]. Compared with spatially reconfigurable architectures such as movable antennas or fluid antennas [18, 19, 20], ERAs can be dynamically adjusted through EM reconfiguration without altering antennas’ positions.
Motivated by the above advantages, ERAs have been investigated for wireless sensing and localization. Early studies in [21, 22, 23] mainly focused on one-dimensional (1-D) or two-dimensional (2-D) angle-of-arrival (AoA) estimation with reconfigurable radiation patterns, demonstrating the angular sensing potential of ERAs under relatively simple narrowband assumption. The works in [24] further studied single-anchor indoor localization using ERAs, showing the feasibility of low-cost ERAs-assisted localization. More recently, [17] investigated wideband localization with radiation-pattern-based ERAs and proposed hybrid baseband (BB)/EM-domain designs, verifying the benefit of ERA reconfigurability over conventional fixed-pattern arrays. Nevertheless, the existing works mainly focus on non-adaptive or single-shot sensing designs, and do not study how a shared ERA aperture should be sequentially configured for multi-user localization based on the observations accumulated over multiple sensing stages.
Active sensing provides a natural way to address the above limitation. Instead of using a fixed sensing configuration, active sensing designs the next-stage probing strategy according to the current information state obtained from previous measurements [26, 27, 28]. From this aspect, active sensing has been applied to beam alignment, channel acquisition, and reconfigurable intelligent surfaces (RIS)-aided sensing. In particular, RIS-aided active localization has been investigated in [30, 31, 32, 33], where the phase configuration is sequentially adapted according to received pilots to improve localization performance. In addition, active sensing was further combined with ERA-assisted localization in [29], showing that the additional DoFs offered by ERAs can provide richer measurement diversity. However, these studies are mainly developed for RIS-assisted systems or single-user localization, and therefore do not address the shared-aperture sensing design.
Nevertheless, directly extending active sensing from single-user to multi-user localization is non-trivial. Although pilot-domain inter-user interference can be largely mitigated by orthogonal pilot design and pilot matching [32], multiple users remain coupled through the shared ERA aperture. In addition, the sensing strategy must extract useful information from high-dimensional historical observations and convert it into ERA configurations. Therefore, learning-based active sensing has been shown to be effective for parameterizing such sequential sensing policies [26, 30, 31, 32, 33, 29]. In particular, recurrent architectures have been used to summarize historical observations [26, 30], which motivates the use of a long short-term memory (LSTM) module to accumulate information across sensing stages. Meanwhile, graph-based neural architectures have shown advantages in multi-user wireless system and active sensing [34, 35], since they can model interactions among users while preserving permutation-aware processing. This motivates the use of a graph neural network (GNN) to learn the shared-aperture coupling among multiple users. These considerations motivate the learning-based active sensing framework developed in this paper.
| Work | Main scope | Difference from this work |
| [17] | ERA-assisted localization with hybrid/codebook design | Does not consider shared ERA configuration for multi-user scenario. |
| [29] | ERA-assisted active sensing for localization | Focuses on single-user localization. |
| [30, 31, 32, 33] | RIS-aided active localization with sequential configuration | Does not consider ERA-based localization design under shared aperture. |
| [34] | GNN-based multi-user active sensing for beam tracking | Targets multi-user beam tracking rather than localization. |
Although ERA-assisted localization and active sensing have been separately investigated, existing studies have not considered ERA-assisted multi-user localization with stage-wise adaptation of a shared ERA configuration. In particular, the joint design of multi-user localization, shared-aperture active sensing, and practical finite-state ERA radiation patterns remains unexplored. Motivated by this gap, this paper investigates ERA-assisted multi-user localization via learning-based active sensing. To clarify the relationship with related studies, Table I summarizes representative ERA localization and active sensing works. The main contributions of this work are summarized as follows:
-
•
We formulate an active sensing framework for ERA-assisted multi-user localization. An ERA-driven two-timescale sensing protocol is developed, where the shared ERA configuration is generated at the stage level based on accumulated pilot observations, while multiple block-wise ERA configurations are applied within each stage to provide measurement diversity. To capture both theoretical flexibility and practical hardware constraints within a unified framework, we develop a signal model that incorporates two complementary ERA paradigms: a synthesis-based model built on spherical-harmonic basis functions, under which the idealized performance reference can be investigated, and a measured finite-state selection model based on a physically realizable radiation pattern library obtained in [10] to evaluate practical performance 111Some fabrication technologies have been developed to realize ERAs, including pixel-based parasitic layouts [11], electronically steerable parasitic array radiator structures [13], and liquid-metal fluidic implementations [10]. Without loss of generality, we adopt the implementation and modeling paradigm in [10], where each element’s radiation pattern is reconfigured via software-controlled fluidic actuation. It is worth noting that the proposed framework is agnostic to the specific fabrication technology. .
-
•
We propose a learning-based active sensing framework that jointly performs stage-wise localization and ERA configuration design. The framework combines an observation encoder, an LSTM-based recurrent state update, and a GNN-based multi-user interaction module. The LSTM-based module accumulates historical sensing information across stages, while the GNN-based module produces user representations and a shared global sensing context for shared ERA configuration. Moreover, model-specific output heads generate either continuous synthesis coefficients or finite-state ERA selections, enabling a unified design for both ERA paradigms.
-
•
We provide numerical comparisons under the considered channel distribution. The results quantify the gains brought by ERA reconfigurability and active sensing. We also compare the synthesis-based model with the measured-based model, showing that the practically implementable finite-state ERA model can achieve performance close to the idealized synthesis-based model. The results also demonstrate the generality of the proposed design under moderate variations from the considered channel distribution, such as different numbers of users.
The remainder of this paper is organized as follows. Section II introduces the system model, and presents two ERA modeling paradigms. Section III presents the proposed active sensing scheme for ERA-assisted multi-user localization and formulates the corresponding joint design problem. Section IV develops a learning-based active sensing framework and describes its training procedure. Numerical results are provided in Section V. Finally, Section VI concludes the paper.
In this paper, the real and the imaginary parts of a vector/matrix are denoted as and , respectively. and denote the transpose and Hermitian transpose, respectively. The notation denotes the Euclidean norm. The operator stacks the columns of a matrix into a vector, and denotes element-wise multiplication. In addition, we use and to denote the spaces of real and complex matrices, respectively. Finally, we use to denote the complex Gaussian distribution.
II System and Signal Model
II-A System Model
As shown in Fig. 1, we consider a multi-user uplink localization system, where single-antenna users (UEs) with fixed positions simultaneously transmit pilot signals to a BS. The BS is equipped with ERAs to enable flexible EM reconfiguration, which provides additional spatial DoFs for the subsequent active sensing-based localization. Let and denote the positions of the -th UE and the BS, respectively. We assume that the BS and all UEs are perfectly synchronized, and synchronization errors are not considered in this work. The localization procedure is performed over active sensing stages. At each stage , the ERA configurations are applied over pilot blocks, each consisting of pilot symbols transmitted on active subcarriers.
To enhance sensing adaptability while avoiding excessive ERA optimization and reconfiguration overhead, we adopt an ERA-driven two-timescale active sensing protocol. Specifically, the ERA configurations are designed at the stage level based on the observations accumulated from previous stages, while the generated ERA configurations are applied at the block level for the next stage. Through local EM reconfiguration, different blocks correspond to different spatial probing patterns, which provide block level measurement diversity with relatively lightweight reconfiguration overhead [10]. In this way, the ERA-enabled BS can progressively explore the propagation environment and collect informative measurements for multi-user localization. The proposed pilot transmission protocol is illustrated in Fig. 2. The received pilots at the -th stage are used for localization and for designing the block-wise ERA configurations for the -th stage.
Denote the fixed pilot matrix used in each pilot block by
| (1) |
where the -th column vector contains the pilot symbols transmitted by the -th UE within each pilot block. The pilot sequences are assumed to be mutually orthogonal, i.e., , where denotes the uplink transmit power of each UE. The same pilot matrix is repeatedly used over the pilot blocks in each stage.
During the -th pilot block and on the -th active subcarrier of the -th stage, the BS first collects the antenna-domain received signal matrix
| (2) |
where and denote the received signal and additive white Gaussian noise (AWGN), respectively. The entries of are independently distributed as with , where is the subcarrier spacing. Moreover, is the effective channel from the -th UE to the BS on the -th subcarrier at the -th pilot block of the -th stage. Since the pilot sequences are orthogonal among UEs, the BS can extract the user-specific observation for the -th UE through pilot matching as
| (3) |
where is the decorrelated antenna-domain observation of the -th UE, and is the corresponding effective noise vector after pilot matching. Since the entries of are independently distributed as and , we have . With orthogonal pilots, direct pilot-domain inter-user interference is removed after pilot matching.
Collecting the user-specific antenna-domain observations over all pilot blocks, antennas, and active subcarriers within the -th sensing stage, we define the stage-wise observation matrix for the -th UE as
| (4) |
where denotes the block-wise observation matrix of the -th UE over all active subcarriers at the -th active sensing stage. organizes the block level observations along its row dimension, where different rows correspond to different ERA configurations applied over different pilot blocks. Its column dimension jointly contains the antenna-domain observations and frequency-domain observations across all subcarriers.
Following the geometric model in [10], we represent the channel between the -th UE and the BS by resolvable paths, then the channel vector in (2) can be given by
| (5) |
where corresponds to line-of-sight (LoS) component, while are non-line-of-sight (NLoS) paths generated by scatterers. represents the complex gain of the -th path between the -th UE and the BS, with and denoting its modulus and phase components, respectively. denotes the azimuth-elevation direction of the -th path observed at the BS. denote the subcarrier index. Since the carrier frequency phase can be absorbed into the complex path gain , only the baseband offset appears in the frequency-dependent phase term. is the BS array response toward , and is the ERA-induced directional gain vector toward , which will be elaborated in II-B in detail. It is noted that is assumed to be frequency-flat over the considered bandwidth, since the ERA is modeled as an antenna-domain reconfiguration mechanism that changes the effective directional radiation gain of the array. Therefore, the angle-dependent gain is evaluated at the carrier frequency and shared by all subcarriers, provided that the ERA hardware does not exhibit strong frequency selectivity over the active band.
In (5), denotes the delay of the -th path between the -th UE and the BS. For the LoS path, we have
| (6) |
where is the speed of light. And for an NLoS path associated with a scatterer at , we have
| (7) |
Under the assumption of perfect synchronization, the delay terms in (6) and (7) only contain geometric propagation delays, without additional clock bias or timing offset 222A possible clock bias would enter as , which is beyond the scope of this work..
We assume that the BS employs a uniform planar array (UPA) located on the - plane, with its boresight oriented along the positive -axis. Let denote the coordinates of the -th antenna element with respect to the array reference point. The -th entry of the steering vector is
| (8) |
where denotes the wavelength. Here, the azimuth angle is measured in the - plane from the positive -axis toward the positive -axis, and the elevation angle is measured from the - plane toward the positive -axis.
II-B Modeling Reconfigurable Gains
Following [14] and [17], the reconfigurable complex amplitude response of the ERA in (5), can be described through two modeling paradigms. First, an idealized ERA is assumed to realize arbitrary beampatterns on demand by projecting onto a pre-defined set of orthonormal basis functions [8, 12], such as spherical-harmonic (SH) basis functions [17, 12, 36]. Second, each ERA switches among a finite set of discrete states, where each state corresponds to a distinct radiation response. A representative hardware realization of this type is reported in [10]. In the following, both the synthesis-based model and the finite-state model are incorporated into the unified signal framework in (5).
II-B1 Model I (Spherical-harmonic Synthesis Model)
In the first model, each antenna element synthesizes its directional response from a truncated SH basis. Let denote a vector of orthonormal basis functions. During the -th pilot block of -th stage, the radiation response of the -th element of the ERA is
| (9) |
where is the EM coefficients that assigns weights to the basis functions to generate the corresponding ERA gain of the -th element, and subject to the per-element normalization [12]. By stacking , the ERA configuration can be represented by .
In this paper, we adopt spherical harmonic orthogonal decomposition (SHOD) functions for basis functions [37] due to their simplicity. Under the SHOD paradigm, any radiation pattern admits an infinite-series expansion in spherical harmonics [8]. We can truncate these bases to the first basis functions for convenient simulations.
II-B2 Model II (Measured Finite-state Model)
In the second model, the ERA gain can be described by a calibrated finite library of realizable patterns, which provides the most faithful representation of practical hardware implementations [12, 17]. It is noted that the same state library is shared by all antenna elements, but the selected state may vary with the antenna index in a specific pilot block. Let denote a set of candidate beampatterns, where denotes the measured amplitude response of the -th ERA state. In practice, the library is available on a discrete azimuth-elevation grid. Accordingly, the radiation response of the -th element during the -th pilot block of -th stage can be expressed as
| (10) |
where is a one-hot state-selection vector satisfying . It should be noted that the measured library adopted in this work contains the directional amplitude responses of the finite ERA states. State-dependent phase variations are not included in the available pattern data, while the spatial phase progression across the array is represented by the array response . In addition, the total power of each state is assumed to be equal to to ensure the energy conservation, i.e., . Similarly, by stacking , the finite-state selection of all antenna elements in the -th stage can be represented by 333In our implementation, the measured library contains states sampled on a angular grid, with azimuth and elevation , both with spacing. For an off-grid direction, the response of each state is obtained by bilinear interpolation over the four nearest azimuth-elevation grid points.. Element-wise mutual coupling among elements is not explicitly modeled, and the measured state library is assumed to represent the effective response.
III Active Sensing For Localization
To estimate UEs’ positions, we proposes an active sensing strategy that configures the ERA in the EM domain for the blocks of each stage based on pilots received in previous stages. Within each stage, we also design the localization scheme based on the extracted features of the received pilots.
For the first stage, no previous observation is available. We therefore initialize using predefined configurations. For , the configurations are generated according to
| (11) |
Here, we refer to as the ERA configuration scheme in the -th stage. denotes a generic notation for the block-wise ERA configuration variable, whose specific form depends on the different ERA modeling paradigm. In particular,
| (12) |
Next, by performing the ERA configuration, we can obtain the observation at this -th stage. Then the UEs’ position can be estimated as follows
| (13) |
where denotes the set of position estimates, and is the corresponding localization scheme. The stage-wise active sensing workflow is also illustrated in part B of Fig. 2.
Based on the above analysis, the multi-user localization problem can be formulated as the joint design of the ERA configuration schemes and localization schemes to minimize the cumulative stage-wise localization error, i.e.,
| (14) | ||||
| s.t. |
The objective in (14) is designed for progressive localization, where a position estimate is available after each sensing stage. The stage weights determine the tradeoff between intermediate-stage availability and final-stage accuracy. Equal weights, i.e., , promote consistently reliable estimates and are suitable for online applications with possible early termination. Larger weights on later stages instead favor final-stage accuracy and allow earlier stages to emphasize exploration. Unless otherwise stated, we use equal weights to ensure usable intermediate estimates.
However, analytically solving is highly challenging, because it involves the joint optimization of the high-dimensional mappings and , which are coupled across sensing and localization stages. Furthermore, in the multi-user scenario, the users are coupled through the shared ERA aperture and the unified sensing configuration, thereby making both the observation process and the active sensing design more complicated. In addition, as the input dimensions of these mappings increase with the number of sensing stages, deriving an analytically tractable and scalable solution becomes nearly impossible.
To address the above challenges, we propose to utilize deep neural network as a powerful function approximator [38] to parameterize the above two mappings. In this way, the computational complexity of the optimization is transferred to the neural network training process [39]. The essence is to identify a neural network architecture capable of summarizing historical observation across different stages in designing an optimal sensing strategy for the considered system.
IV Learning-Based Active Sensing Framework for Multiuser Localization
This section develops a learning-based active sensing framework for solving problem , in which the stage-wise ERA configuration and localization policies are jointly parameterized by a neural network. At each stage, matched observations of each UE are compressed into low-dimensional user-wise features, which are then processed by an LSTM to capture cross-stage temporal information and update the corresponding hidden states. A GNN further models inter-user coupling caused by the shared ERA aperture. Based on the resulting user states, task-specific output heads generate the ERA configuration for the next stage and the position estimates for the current stage. This architecture follows the active sensing principle that current measurements should both improve immediate localization and enable more informative sensing configurations in subsequent stages.
IV-A Observation Encoder and Recurrent State Update
The recurrent module aims to summarize the information accumulated up to current stage into a fixed-dimensional state vector. Rather than directly vectorizing , we preserve its block, antenna, and frequency structures by reshaping its real and imaginary parts into a tensor over the block and joint antenna-frequency dimensions. The resulting tensor is then processed by a convolutional observation encoder
| (15) |
where denotes the observation encoder. The convolutional layers operate on both pilot-block and frequency domain, thereby capturing joint features across the block dimension and the delay dimension.
Next, we apply a shared-weight LSTM cell independently to summarize the evolution of the observations of each UE over stages. Equivalently, the architecture contains parallel LSTMs that share the same learnable parameters. For the -th UE, the hidden state and cell state are updated as
| (16) | ||||
where and are the activation vectors of the forget gate, input gate and output gate within the -th LSTM cell, respectively. The updating rules for different gates are given as , where , and is the element-wise sigmoid function. For each , all the and are fully connected layers. The parameters of the LSTM gates are shared across all UEs, while each UE maintains its own hidden state and cell state. In this way, the recurrent module can capture and extract important features of the received pilots, while the following modules need to handle the multi-user coupling at the current stage.
IV-B GNN-Based Multiuser Interaction and Joint Decision
In multi-user localization, the ERA configuration is shared by all UEs, which creates coupling among their sensing requirements. To model this interaction while preserving permutation symmetry with respect to UE ordering, we apply a GNN to the user-wise recurrent states. The shared ERA configuration is required to be permutation invariant, whereas the per-UE position estimates should be permutation equivariant. Although several set-based architectures such as Deep Sets and Set Transformers can provide these properties [40, 41], we adopt a complete-graph GNN as realization because it explicitly captures pairwise UE dependencies through edge features and message aggregation [42].
Based on the above analysis, after the recurrent update in (16), each UE is represented by a node feature (hidden state of LSTM cell) vector . We construct a complete interaction graph among the active UEs, where each node corresponds to one UE and each edge captures the pairwise dependency between two UEs. Let denote the input node embedding. At the -th message-passing layer, the message sent from -th node to -th node is computed as
| (17) |
where denotes the edge-update network, and is the pairwise edge feature constructed from the embeddings of -th and -th nodes. Specifically, is defined as
| (18) |
Here, the edge feature combines the two UEs embeddings with their difference and element-wise product. The individual embeddings retain UE-specific information, while the difference and product provide simple representations of relative and multiplicative interactions between the two UEs. Then the aggregated message for -th UE is obtained as
| (19) |
where denotes the neighbor set of -th node. The node state is updated by a node-update network as
| (20) |
After graph layers, we obtain the interaction-aware node embedding for each UE. For notational simplicity, we omit the superscript when no ambiguity arises. Based on these node embeddings, we employ two specific feed-forward heads to generate the current-stage position estimates and the ERA configurations for the next stage. Specifically,
| (21) | ||||
where and denote the localization and ERA configuration heads, respectively, while is the global context vector for generating the shared ERA configuration. The normalization layer enforces the corresponding constraints on the EM coefficients. The configuration head is model specific: it outputs continuous ERA synthesis coefficients for Model I and discrete codebook state-selection variables for Model II. These output designs are detailed in subsection IV-C.
IV-C Model-Specific ERA Output Heads
The above observation encoder, LSTM recurrent update module, and GNN interaction module can be shared by Model I and Model II under the same aperture architecture and propagation environment. The only ERA model-dependent module lies in the output parameterization of the ERA configuration head. Since the ERA configuration is shared across UEs at the BS side, we first aggregate the user-wise graph embeddings into a global context vector
| (22) |
where the attention weights are obtained by a shared scoring network , where is a lightweight network. Then the weight is obtained by .
IV-C1 Model I: Continuous SH-Coefficient Head
For the synthesis-based ERA model, the output head generates continuous SH coefficients for all antennas and pilot blocks as
| (23) |
where . To satisfy the per-element normalization constraint in (9), the output is normalized as
| (24) |
The resulting is then used to synthesize the directional response for the next stage.
IV-C2 Model II: Discrete State-Selection Head
For the measured finite-state ERA model, the output head instead generates the logits associated over candidate states
| (25) |
where contains the state-selection scores for all antenna elements and pilot blocks. Its -th slice, denoted by , represents the logits over the candidate states for -th antenna in -th pilot block of -th stage. Because discrete state selection is non-differentiable, we employ a straight-through estimator (STE) during training to enable gradient-based optimization [43]. Specifically, a soft state-selection vector is first constructed from using the Gumbel–Softmax relaxation [44]
| (26) |
where denotes the Gumbel–Softmax relaxation with temperature . A hard one-hot state is then generated as
| (27) |
Here, converts an index into a one-hot vector. To preserve end-to-end trainability, the effective training state vector is formed via the straight-through [43]
| (28) |
where denotes the stop-gradient manipulation. The forward process uses the hard-selection results, while the back-propagation reuses the gradient of the soft relaxed vector.
IV-D Extension to Variable Numbers of UEs
In practical wireless systems, the number of UEs may vary with the load and service demand. Nevertheless, the proposed framework is inherently compatible with such variations, providing the possibility to achieve such generalizability. The observation encoder and recurrent module operate independently on each UE, allowing their parameters to be shared across different UEs. Similarly, the GNN uses shared edge- and node-update networks to capture inter-user coupling induced by the shared ERA aperture and sensing configuration. The main scalability challenge therefore lies in the two output heads, as directly concatenating user features would result in input dimensions that depend on the number of UEs.
To distinguish the maximum network capacity from the actual number of active UEs, we denote as the maximum number of UEs supported by the network architecture and pilot design. An activity indicator is introduced, where indicates that the -th UE is active. Accordingly, the number of active UEs in the sample is .
As illustrated in Fig. 3, the proposed active sensing unit is designed to accommodate different numbers of UEs within a unified architecture. Specifically, the user-wise observation encoder and LSTM cell adopt a shared-weight design, such that the same feature extraction module can be reused for all UEs. In this way, pilot observations from different UEs are processed in a consistent manner, and the feature extraction stage does not depend on the number or ordering of active UEs. To further support variable-size multi-user scenarios, the extracted user features are organized as graph nodes and processed by a GNN, which is particularly well suited to this setting. Under a prescribed , an activity mask is introduced to indicate which UEs are active in the current sample. The mask is used to suppress inactive nodes and restrict graph interaction, feature aggregation, masked softmax for user weighting, and output evaluation to the active UEs only. As a result, the same GNN and output heads can be directly applied to scenarios with different numbers of UEs without architectural changes.
IV-E Training Procedure
The active sensing units are concatenated to form a deep recurrent architecture, as illustrated in Fig. 4, where each unit corresponds to one active sensing stage. The sensing unit is recursively unrolled over stages to form an end-to-end recurrent architecture. At each stage, the current pilot observations are encoded and combined with the user-wise recurrent states, after which the GNN produces the current position estimates and the ERA configuration for the next stage. Since the same sensing policy is repeatedly applied throughout this sequential procedure, the proposed architecture adopts a structured parameter-sharing design.
Specifically, the observation encoder and the LSTM parameters are shared across all UEs and all sensing stages. Each UE maintains its own recurrent hidden and cell states. At the -th graph layer, the edge-update network is shared across all directed edges, and the node-update network is shared across all UE nodes. These networks are also reused across sensing stages, whereas different graph layers employ distinct parameters. The attention and localization heads are also shared across UEs and stages, whereas the ERA configuration head is model specific. Consequently, all sensing stages reuse the same unit, and the stage dependence is represented through the recurrent states and the accumulated observations rather than through stage-specific network.
For variable-user training, we employ a mask-normalized implementation of the objective of . Specifically, the loss for each training sample is computed as
| (29) |
Therefore, inactive UEs do not contribute to the loss, and the localization error of each sample are normalized by its number of active UEs. The proposed architecture is trained offline in an end-to-end supervised manner using Adam to minimize the mask-normalized empirical objective of . By unrolling a sufficiently large number of stages during training, the network jointly learns the ERA configuration policy and the localization scheme to progressively refine sensing and position estimates from successive pilot observations. Since the recurrent sensing unit shares parameters across stages, the trained model can be recursively applied for any number of stages up to the training length and terminated early according to the available sensing budget.
V Numerical Results
In this section, we evaluate the localization performance of the proposed framework through numerical simulations.
V-A Simulation Scenario and Channel Generation
The considered scenario consists of a BS equipped with a square UPA, i.e., the UPA has the same number of rows and columns. The UEs are independently distributed within the front-side semicircular annulus of the BS. More specifically, the UE azimuth is uniformly sampled from the front-side sector with respect to the BS boresight, the horizontal location is sampled over the annulus with the distance range in Table II, and the UE height is uniformly sampled within the specified -coordinate range. Unless otherwise specified, the default system and scenario parameters are summarized in Table II, where the main settings are selected with reference to representative studies [17, 29, 30, 34].
| Default System Parameters | Value |
| BS position | m |
| UPA size | |
| Carrier frequency | 3.5 GHz |
| System bandwidth | 100 MHz |
| Number of active subcarriers | 256 |
| Noise PSD | dBm/Hz |
| Number of SHOD bases | 25 |
| UE horizontal distance range | m |
| UE -coordinate range | m |
| Scatterer horizontal distance range | m |
| Scatterer -coordinate range | m |
To make the statistical channel model explicit, all training and testing samples are generated independently from the same geometric wideband channel distribution described in Section II. The BS and all UEs are assumed to be perfectly synchronized, and thus the path delays only include geometric propagation delays. For each UE, the LoS component is always present, and the number of NLoS paths is fixed as . Accordingly, each UE-BS uplink channel contains resolvable paths in total. Each NLoS path is generated by an independent single-bounce scatterer , whose azimuth is uniformly sampled from and whose horizontal distance and height are independently sampled according to the scatterer ranges in Table II. Given the UE and scatterer locations, the angle and delay of each path are calculated according to the geometric relations in (6) and (7). The complex path coefficients are generated according to a free-space propagation model. Specifically, the LoS path amplitude is modeled as , while the amplitude of the -th NLoS path is modeled as , , following [17]. The path phase components are independently generated from a uniform distribution over . The ERA response is evaluated at the carrier frequency and is assumed to be frequency-flat.
V-B Network Implementation and Training Setup
During training, the number of active users in each sample is randomly drawn from to , and the proposed active sensing framework is trained with sensing stages. Each sensing stage contains pilot blocks, and each block contains mutually orthogonal pilot symbols. For the first sensing stage, the blocks use predefined structured ERA probing directions. The target azimuth angles are uniformly spaced over the front sector , with elevation fixed at . For Model I, the initial probing patterns are synthesized using normalized low-order SH coefficients, where only orders are retained and the coefficients of order are tapered by . For Model II, the state of each block is selected from the measured codebook as the state with the largest gain toward the corresponding target direction. Within each block, the same initial configuration is applied to all antenna elements.
After training, since the sensing unit is shared across stages, the learned model can be recursively reused and early-terminated to support any number of stages satisfying under the fixed blocks, which allows us to evaluate its early-termination capability. In addition, because each block contains mutually orthogonal pilot symbols, the maximum number of active UEs supported during testing is . Therefore, the generalization should be interpreted within the considered synchronized geometric channel distribution, rather than as robustness to arbitrary propagation environments.
For the network, the observation encoder is implemented as a lightweight 2D CNN with base channels, while the multi-user interaction module is realized by a two-layer message-passing GNN. Inter-stage information is propagated through an LSTM cell. All output heads are implemented using fully connected networks. In the implementation, the hidden feature dimension of all modules is set to . The model is trained with a batch size of 128 for 50,000 iterations. To enable the stable training, the network is optimized using Adam with an initial learning rate of and the weight decay coefficient is set to . Gradient clipping is further applied with a clipping threshold of 1.0 to improve stability [46]. In Model II experiments, the temperature is initialized as and annealed to . Denote as training iteration and the total number of iterations. We use an annealing schedule . During training, the forward pass uses hard selections through the STE, while gradients are propagated through the relaxed Gumbel-Softmax sample. During inference, the selected state is obtained by the deterministic argmax over the logits. We implement the proposed framework using PyTorch 2.0.0 on a RTX 3060 GPU, and follow the training procedure described in Section IV. The uplink transmit power is fixed within each training process, and a separate model is trained for different considered power setting. Specifically, we separately train the models at dBm and dBm for the corresponding comparisons, while all other evaluations reuse the models trained at dBm. The DL-based baselines are also trained under the same transmit power setting as the proposed models.
V-C Baselines
To evaluate the effectiveness of the proposed framework, we consider the following representative baselines.
1) DL-based localization with fixed ERA: This baseline uses the same measured finite-state codebook as Model II, but the ERA state is kept unchanged during the sensing process. Specifically, we choose a fixed state from the measured codebook whose radiation response provides a relatively high gain over the angular support of the considered UE distribution. In our implementation, the same selected state is applied to all antenna elements, pilot blocks, and sensing stages. Therefore, this baseline preserves the directional radiation property of a realizable ERA state, but does not perform active EM-domain reconfiguration. This baseline is used to distinguish the performance gain brought by adaptive ERA reconfiguration from the gain brought by using a fixed directional ERA pattern.
2) DL-based localization with omnidirectional antenna: This baseline replaces the ERA elements with conventional omnidirectional antenna elements. Therefore, the element-wise EM response is direction-independent and remains fixed over all pilot blocks and sensing stages, which can be written as . This baseline represents a conventional fixed-pattern antenna array and is used to evaluate the localization gain provided by directional ERA patterns and adaptive EM-domain sensing. Fig. 5 compares the azimuth-domain effective gains of the two fixed-pattern baselines (computing by ), whereas the fixed ERA provides a relatively stronger directional response over the considered angular region.
3) DL-based localization with random ERA configurations: This baseline uses the same measured finite-state codebook and DL-based localization network as Model II, but replaces the learnable ERA configuration with random policy. At each stage, the states for all ERAs and pilot blocks are randomly sampled from the measured codebook and resampled for the next stage. Therefore, this baseline provides stage-wise EM-domain measurement diversity but does not perform observation-adaptive sensing. This comparison isolates the gain of learned ERA adaptation from that provided by random ERA reconfiguration and the DL-based localization module.
4) Codebook-based active sensing with ERA: In each stage, UE-specific observations are first separated by pilot matching, as in the proposed method. The azimuth/elevation angle of each UE is then estimated using multiple signal classification (MUSIC) method, where the scan grid covers azimuth with 181 points and elevation with 111 points. The delay/range is then obtained from an oversampled inverse-DFT delay profile with an oversampling factor of 8. Since the UE labels are preserved after pilot matching, the angle and delay estimates are paired within the same pilot label. For next stage, the estimated UE directions are assigned to the pilot blocks in a round-robin manner. Each block selects the measured ERA state with the largest interpolated gain toward its assigned direction, and the selected state is applied to all antenna elements444For each pilot block, a common state is applied to all antenna elements to provide a low-complexity benchmark, since independent per-element selection would introduce a much larger combinatorial design problem.. If no valid direction estimate is available, the broadside direction is used as the fallback.
5) Model-based localization with omnidirectional antenna array: The BS equipped with conventional omnidirectional antenna collects all pilot observations and then applies a MUSIC-based angular estimator and DFT-based delay estimation to obtain the user angles and ranges for localization. This baseline serves as a non-adaptive model-based benchmark using an omnidirectional antenna array.
6) Ablations: We consider three variants of the proposed framework by removing the GNN module, the LSTM module, and learned attention pooling, respectively. Without the GNN, user representations are processed independently before global aggregation; without the LSTM, the current-stage observation features are directly fed to the GNN without recurrent memory; and without attention pooling, the shared ERA context is obtained by uniformly averaging the active-user embeddings.
To maintain the consistency of the comparison, baselines 1) – 3) follow the same transmission process as the proposed framework. At each sensing stage, the BS collects pilot-block observations under the fixed/random ERA response, and the observations are processed stage-by-stage by a DL-based localization network. In addition, ablations are also constructed from the Model II mechanism, where the ERA-related operations follow the discrete state-selection model. For each following evaluation experiment, we generate another independent samples and evaluate the fully trained proposed framework and all baselines using a batch size of . All compared schemes use the same geometric channel distribution, active-user masks, and total pilot-block budget.
V-D Training Convergence
We first show the convergence behavior of the proposed framework and the two fixed-pattern baselines in Fig. 6. As shown in Fig. 6(a), under the synthesis model, the proposed framework converges rapidly in the early stage and both the training and validation losses gradually approach a low stable level, demonstrating favorable optimization behavior under the idealized ERA configuration methodology. For the measured finite-state model in Fig. 6(b), the overall convergence trend remains largely consistent, although a moderate gap between training and validation appears in the later iteration. This indicates some loss in generalization compared with the synthesis model, which is expected since Model II is constrained by the measured finite-state radiation codebook and therefore represents a more practical but less flexible hardware configuration than Model I. Importantly, the validation loss still decreases and eventually stabilizes, confirming that Model II remains practically effective. By comparison, the fixed-pattern baselines in Fig. 6(c) and Fig. 6(d) exhibit substantially larger fluctuations during training. Although their losses still follow an overall decreasing trend, their validation losses fluctuate more significantly and do not converge as smoothly as those of the proposed ERA-assisted active sensing models. One possible explanation is that the non-adaptive responses provide less measurement diversity, which may make the localization mapping more difficult to optimize. Without adaptive ERA configuration, the network has to infer the positions under fixed spatial responses, making it more sensitive to varying number of UEs, difficult geometries, and multipath conditions.
V-E Localization Performance and Ablation Comparisons
We first examine whether the network should prioritize uniformly reliable position estimates across all stages or mainly optimize the final-stage estimate. To provide an intuitive measure of the localization error, we adopt the root mean square error (RMSE) as the performance metric. The stage-wise weights is controlled by a decay factor , where . When , all stages are equally weighted; when , larger weights are assigned to later stages. As shown in Fig. 7, compared with the equal-weight case, using a smaller decay factor can slightly improve the accuracy at the final stage, but it also leads to degraded early stage performance. Under limited initial observations, high EM-domain flexibility allows the network to allocate the early stage toward broad spatial exploration rather than immediate fine localization. After more observations are accumulated, however, the EM-domain flexibility enables to refine the localization more effectively, leading to improved RMSE in the later stages. These results confirm the expected tradeoff introduced by the stage-wise objective in (14). Since the proposed framework is intended to provide progressively refined estimates and may be terminated before the maximum number of stages, we use in all subsequent experiments.
Fig. 8 shows the localization RMSE versus the sensing stage under different uplink transmit powers. It can be observed that the ERA-assisted active sensing framework achieves a clear and progressive RMSE reduction as the stage increases. This improvement is especially significant in the first few stages, which indicates that the proposed active sensing framework can effectively exploit the accumulated pilot observations to refine the subsequent sensing strategy. It should be noted that the first-stage gain is attributed to the predefined structured probing and the supervised localization network, rather than ERA configuration. Adaptive refinement starts from the second stage and is reflected in the subsequent RMSE reduction and the advantage over the fixed and random ERA baselines.
Moreover, Model I consistently achieves the best performance, while Model II remains close to Model I over all stages. This is reasonable since Model I relies on the synthesis-based ERA model with continuous EM-domain control, whereas Model II is constrained by the measured finite-state radiation codebook. Nevertheless, Model II still outperforms the baselines, demonstrating its effectiveness as a practically deployable implementation. When the transmit power decreases from dBm to dBm, the RMSE of all methods increases, but the proposed Model I and Model II still preserve clear performance advantages. The results of the fixed ERA, random ERA, and omnidirectional antenna baselines clarify the role of EM-domain pattern design and observation-adaptive reconfiguration. The fixed-ERA baseline outperforms the omnidirectional baseline due to its selected directional gain, while the random ERA baseline remains close to the fixed-ERA baseline despite providing stage-wise pattern diversity. This indicates that random reconfiguration alone is insufficient, and that the main gain of Model II comes from adapting ERA configuration. The gradual RMSE reduction of the non-adaptive baselines is mainly attributed to LSTM-based observation accumulation rather than active ERA adaptation.
Fig. 9 shows the localization RMSE versus the number of UEs. The proposed Model I and Model II maintain nearly stable RMSE as the number of UEs increases, indicating that the learned framework can accommodate varying user counts within the considered range. Model I consistently achieves the lowest RMSE, while Model II remains comparable, showing that the measured finite-state ERA preserves most of the localization gain provided by the ERAs. In contrast, the fixed ERA, random ERA, and omnidirectional antenna baselines exhibit substantially larger RMSEs. The codebook-based and model-based baselines perform even worse, highlighting the limitations of hand-crafted angle/delay estimation. Overall, the results demonstrate that the proposed framework generalizes well to different numbers of UEs within the considered channel distribution and user count range.
Fig. 10 compares the localization RMSE and per-sample neural inference time under different stage/block allocations, with the total pilot-block budget fixed at . We train and evaluate separate models under each allocation, and per-sample inference time is obtained by averaging the forward-pass latency over all samples in one batch. As shown in the upper figure, increasing the number of sensing stages generally improves the RMSE for both models, since the sensing strategy can be updated frequently. The improvement is most significant from the allocation to the allocation, while the RMSE curves become stable when the number of stages further increases. The lower figure shows that the per-sample inference time increases with the number of stages, because more stages require more neural processing. Nevertheless, both Model I and Model II achieve millisecond-level neural inference latency under all considered allocations. Although the reported time only measures inference latency for one sample on the adopted computing platform and does not include signal transmission or hardware-control delays, the millisecond-level inference results still indicate that the proposed framework has relatively low computational overhead. Overall, Fig. 10 reveals a tradeoff between localization accuracy and computational complexity. Under the considered setting and channel distribution, the allocation provides a favorable balance between performance and inference cost.
To evaluate the robustness of the proposed framework to practical mismatches, we further test the trained Model II under LoS blockage, measured-pattern perturbations, and synchronization errors. For the LoS blockage, the LoS path of each UE is independently retained with probability . For the pattern mismatch, an independent Gaussian perturbation is applied to each angular gain sample of each measured finite-state pattern. Specifically, the perturbed amplitude response is generated as , where denotes the independent zero-mean Gaussian perturbation, where controls the perturbation strength. For the synchronization error, each UE is assigned an uncompensated clock bias during testing.
Fig. 11 evaluates the stage-wise localization performance of Model II under several mismatches. Measurement perturbations increase the RMSE as the perturbation strength grows, but both the -dB and -dB cases still exhibit clear improvement over the first few sensing stages. This indicates that the learned framework can continue to exploit multi-stage observations under moderate codebook mismatch. In addition, a -ns clock bias causes a moderate performance degradation, whereas a -ns bias leads to a larger localization error. In both cases, the RMSE remains nearly unchanged across stages, indicating that multi-stage observation accumulation cannot compensate for a systematic delay. LoS blockage leads to meter-level errors with limited stage-wise refinement, because removing the LoS component significantly changes the learned angle-delay relationship. Overall, the proposed framework exhibits partial robustness to small clock bias and moderate pattern perturbations, but remains sensitive to large synchronization errors and LoS blockage, motivating further mismatch modeling or training augmentation.
To provide a diagnostic view of the learned ERA configuration, we examine the normalized effective pattern gain toward the LoS and dominant NLoS directions under a two-UE scenario with 5,120 samples. For each block, the effective pattern is normalized by its maximum angular power response, and the resulting gains are averaged over all test samples and pilot blocks. The dominant NLoS path is defined as the NLoS component with the largest path amplitude for each UE. As shown in Fig. 12, the normalized gains toward both the LoS and dominant NLoS directions generally increase during the first few sensing stages and then become stable. In particular, the learned policy maintains enhanced relative responses toward multiple propagation directions rather than concentrating exclusively on a single UE or path. It is worth noting that the measured finite-state ERA patterns are predominantly broadside-oriented, with their main lobes generally formed around [17], as also illustrated in Fig. 5. Consequently, the selected ERA states cannot in general place their maximum response exactly along arbitrary LoS or NLoS direction. Therefore, the path gains in Fig. 12 remain below dB even after adaptation. Their progressive increase should be interpreted as stronger relative emphasis on these informative propagation directions. Overall, the results indicate that the adaptive ERA policy progressively reallocates the shared-aperture response toward LoS and multipath components.
| Variant | Degradation at dBm | Degradation at dBm |
| w/o GNN | 4.3% | 10.7% |
| w/o LSTM | 13.1% | 36.4% |
| w/o Attn. | 6.1% | 16.8% |
Table III summarizes the localization RMSE degradation of the architectural variants relative to the full Model II. Removing the LSTM leads to the largest degradation, particularly at dBm, indicating that temporal accumulation of multi-stage observations becomes more important under lower-SNR conditions. Replacing attention pooling with mean pooling causes the second-largest performance loss, while removing the GNN results in a smaller but consistent degradation. Overall, the ablation results show that all three modules contribute to the proposed framework, with the recurrent memory providing the most pronounced benefit. These architectural gains, however, remain secondary to the larger system-level improvement brought by observation-adaptive ERA reconfiguration.
VI Conclusion
In this paper, we investigated ERA-assisted multi-user localization in a wideband uplink system via learning-based active sensing. A unified framework was developed to accommodate both the synthesis-based and the finite-state model. To solve the resulting coupled problem, we proposed a deep active sensing architecture integrating LSTM-based state accumulation, GNN-based multi-user interaction, and adaptive weighting. Numerical results showed that the proposed method outperforms representative benchmarks and achieves refinement across sensing stages. Moreover, the measured-based modeling paradigm obtained performance close to the synthesis-based counterpart, demonstrating its potential practical applicability.
However, although the present results are promising, they should be interpreted within the considered channel distribution and under ideal training assumptions. Specificity, the main framework assumes synchronized clocks and calibrated finite-state radiation patterns. Moreover, the reported latency accounts only for neural inference and excludes pilot transmission and hardware delays. Although the robustness experiments provide an initial evaluation under LoS blockage, clock bias, and pattern perturbations, the model is still designed under nominal synchronized and calibrated conditions. Extending the framework to mismatch-aware training and experimental hardware validation constitutes important future work.
References
- [1] S. E. Trevlakis et al., “Localization as a Key Enabler of 6G Wireless Systems: A Comprehensive Survey and an Outlook,” IEEE Open J. Commun. Soc., vol. 4, pp. 2733-2801, 2023.
- [2] X. Cai, X. Cheng and F. Tufvesson, “Toward 6G with Terahertz Communications: Understanding the Propagation Channels,” IEEE Commun. Mag., vol. 62, no. 2, pp. 32-38, Feb. 2024.
- [3] K. Witrisal et al., “High-accuracy localization for assisted living: 5G systems will turn multipath channels from foe to friend,” IEEE Signal Process. Mag., vol. 33, no. 2, pp. 59-70, Mar. 2016.
- [4] X. Mu, Y. Liu, L. Guo, J. Lin, and R. Schober, “Intelligent reflecting surface enhanced indoor robot path planning: A radio map-based approach,” IEEE Trans. Wireless Commun., vol. 20, no. 7, pp. 4732-4747, Jul. 2021.
- [5] S. Aditya, A. F. Molisch and H. M. Behairy, "A Survey on the Impact of Multipath on Wideband Time-of-Arrival-Based Localization," Proc. IEEE, vol. 106, no. 7, pp. 1183-1203, Jul. 2018.
- [6] F. Rusek et al., “Scaling up MIMO: Opportunities and challenges with very large arrays,” IEEE Signal Process. Mag., vol. 30, no. 1, pp. 40-60, Jan. 2013.
- [7] Q. Xue et al., “A Survey of Beam Management for mmWave and THz Communications Towards 6G,” IEEE Commun. Surveys Tuts., vol. 26, no. 3, pp. 1520-1559, 2024.
- [8] K. Ying et al., “Reconfigurable Massive MIMO: Precoding Design and Channel Estimation in the Electromagnetic Domain,” IEEE Trans. Commun., vol. 73, no. 5, pp. 3423-3440, May 2025.
- [9] B. Zhou, A. Liu and V. Lau, “Successive Localization and Beamforming in 5G mmWave MIMO Communication Systems,” IEEE Trans. Signal Process., vol. 67, no. 6, pp. 1620-1635, 15 March, 2019.
- [10] R. Wang et al., “Electromagnetically Reconfigurable Fluid Antenna System for Wireless Communications: Design, Modeling, Algorithm, Fabrication, and Experiment,” IEEE J. Sel. Areas Commun., vol. 44, pp. 1464-1479, 2026.
- [11] J. Zhang et al., “A novel pixel-based reconfigurable antenna applied in fluid antenna systems with high switching speed,” IEEE Open J. Antennas Propag., vol. 6, no. 1, pp. 212-228, Feb. 2025.
- [12] M. Liu et al., “Tri-timescale beamforming design for tri-hybrid architectures with reconfigurable antennas,” arXiv preprint arXiv:2503.03620, 2025.
- [13] Z. Han et al., “Characteristic mode analysis of ESPAR for single-RF MIMO systems,” IEEE Trans. Wireless Commun., vol. 20, no. 4, pp. 2353-2367, Apr. 2021.
- [14] P. Zheng et al., “Tri-Hybrid Multi-User Precoding Using Pattern-Reconfigurable Antennas: Fundamental Models and Practical Algorithms,” arXiv preprint arXiv:2505.08938, 2025.
- [15] W. Ma et al., "A Survey on Reconfigurable and Movable Antennas for Wireless Communications and Sensing," IEEE Commun. Surveys Tuts., vol. 28, pp. 4842-4882, 2026.
- [16] P. Zheng et al., “Electromagnetically Reconfigurable Antennas for 6G: Enabling Technologies, Prototype Studies, and Research Outlook,” arXiv preprint arXiv:2506.00657, 2025.
- [17] A. Fadakar et al., "Hybrid Codebook Design for Localization Using Electromagnetically Reconfigurable Fluid Antenna System," IEEE J. Sel. Topics Signal Process., early access, pp. 1-16, 2026.
- [18] L. Zhu et al., “A Tutorial on Movable Antennas for Wireless Networks,” IEEE Commun. Surveys Tuts., vol. 28, pp. 3002-3054, 2026.
- [19] Y. Zhang et al., “6D Movable Antenna-Aided Hybrid Beamforming for Multi-User Communications,” 2024 IEEE Globecom Workshops (GC Wkshps), Cape Town, South Africa, 2024, pp. 1-6.
- [20] W. K. New et al., “A Tutorial on Fluid Antenna System for 6G Networks: Encompassing Communication Theory, Optimization Methods and Hardware Designs,” IEEE Commun. Surveys Tuts., vol. 27, no. 4, pp. 2325-2377, Aug. 2025.
- [21] E. Taillefer et al., “Direction-of-arrival estimation using radiation power pattern with an ESPAR antenna,” IEEE Trans. Antennas Propag., vol. 53, no. 2, pp. 678-684, 2005.
- [22] R. Qian et al., “Direction-of-arrival estimation with single-RF ESPAR antennas via sparse signal reconstruction,” in 2015 IEEE 16th International Workshop on Signal Processing Advances in Wireless Communications (SPAWC), 2015, pp. 485-489.
- [23] L. Kulas, “Simple 2-D direction-of-arrival estimation using an ESPAR antenna,” IEEE Antennas Wireless Propag. Lett. vol. 16, pp. 2513-2516, 2017.
- [24] M. Rzymowski et al., “Single-anchor indoor localization using ESPAR antenna,” IEEE Antennas Wireless Propag. Lett., vol. 15, pp. 1183-1186, 2016.
- [25] S.-E. Chiu et al., “Active learning and CSI acquisition for mmWave initial alignment,” IEEE J. Sel. Areas Commun., vol. 37, no. 11, pp. 2474-2489, Nov. 2019.
- [26] F. Sohrabi et al., “Active sensing for communications by learning,” IEEE J. Sel. Areas Commun., vol. 40, no. 6, pp. 1780-1794, Jun. 2022.
- [27] T. Jiang, F. Sohrabi, and W. Yu, “Active sensing for two-sided beam alignment using ping-pong pilots,” in Proc. Asilomar Conf. Signals Syst. Comput., Pacific Grove, California, USA, Nov. 2022, pp. 913-918.
- [28] T. Jiang, and W. Yu, “Active sensing for reciprocal MIMO channels,” IEEE Trans. Signal Process., vol. 72, pp. 2905-2920, 2024.
- [29] R. Zhang, Y. Zhang, and Y. Zhang, “User Localization via Active Sensing with Electromagnetically Reconfigurable Antennas,” arXiv preprint arXiv:2601.20501, 2026.
- [30] Y. Li and W. Yu, “Localization in multipath environments via active sensing with reconfigurable intelligent surfaces,” IEEE Commun. Lett., vol. 28, no. 9, pp. 2061-2065, Sep. 2024.
- [31] Z. Zhang, T. Jiang, and W. Yu, “Active sensing for localization with reconfigurable intelligent surface,” in Proc. IEEE Int. Conf. Commun. (ICC), Rome, Italy, May 2023, pp. 4261-4266.
- [32] Z. Zhang, T. Jiang, and W. Yu, “Localization with reconfigurable intelligent surface: An active sensing approach,” IEEE Trans. Wireless Commun., vol. 23, no. 7, pp. 7698-7711, Jul. 2024.
- [33] Z. Zhang and W. Yu, "Learning Beamforming Codebooks for Active Sensing With Reconfigurable Intelligent Surface," IEEE Trans. Wireless Commun., vol. 24, no. 8, pp. 6504-6517, Aug. 2025.
- [34] H. Han, T. Jiang, and W. Yu, “Active Sensing for Multiuser Beam Tracking With Reconfigurable Intelligent Surface,” IEEE Trans. Wireless Commun., vol. 24, no. 1, pp. 540-554, Jan. 2025.
- [35] Y. Shen, et al., “Graph neural networks for scalable radio resource management: Architecture design and theoretical analysis,” IEEE J. Sel. Areas Commun., vol. 39, no. 1, pp. 101–115, Jan. 2021.
- [36] J. Chen et al., “Integrated Sensing and Communication with Tri-Hybrid Beamforming Across Electromagnetically Reconfigurable Antennas,” arXiv preprint, arXiv:2510.14530, 2025.
- [37] M. Costa et al., “Unified array manifold decomposition based on spherical harmonics and 2-D fourier basis,” IEEE Trans. Signal Process., vol. 58, no. 9, pp. 4634-4645, 2010.
- [38] S. Liang and R. Srikant, “Why deep neural networks for function approximation?” in Proc. Int. Conf. Learn. Represent. (ICLR), Toulon, France, 2017.
- [39] W. Yu et al., “Role of deep learning in wireless communications,” IEEE BITS Inf. Theory Mag., vol. 2, no. 2, pp. 56-72, Nov. 2022.
- [40] M. Zaheer, et al., “Deep Sets,” Adv. Neural Inf. Process. Syst. (NeurIPS), vol. 30, pp. 3391–3401, 2017.
- [41] J. Lee, et al., “Set Transformer: A framework for attention-based permutation-invariant neural networks,” Proc. 36th Int. Conf. Mach. Learn. (ICML), vol. 97, pp. 3744–3753, Jun. 2019.
- [42] T. Jiang, H. V. Cheng, and W. Yu, “Learning to reflect and to beamform for intelligent reflecting surface with implicit channel estimation,” IEEE J. Sel. Areas Commun., vol. 39, no. 7, pp. 1931-1945, Jul. 2021.
- [43] R. Zhang et al., "A Deep Learning Framework for Joint Channel Acquisition and Communication Optimization in Movable Antenna Systems," IEEE Trans. Wireless Commun., vol. 25, pp. 14471-14485, 2026.
- [44] H. Xuan, B. Yang, and X. Li, “Exploring the impact of temperature scaling in softmax for classification and adversarial robustness,” arXiv preprint arXiv:2502.20604, 2025.
- [45] D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” Proc. Int. Conf. Learn. Represent. (ICLR), San Diego, CA, USA, 2015.
- [46] A. Sayal et al., “Neural networks and machine learning,” Proc. 2023 IEEE 5th Int. Conf. Cybernetics, Cognition Machine Learn. Appl. (ICCCMLA), Hamburg, Germany, pp. 58-63, 2023.