arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:2607.26547v1 [eess.SP] 29 Jul 2026

Stay or Switch: Online Conformal Bayesian Optimization Guided Fluid Antenna Configuration

Gangyong Zhu, Jia Yan, and Shijian Gao
Abstract

Fluid antenna systems (FAS) introduce additional spatial degrees of freedom to enable integrated sensing and communication (ISAC) in air-ground networks. However, conventional studies often overlook or simplify the physical overheads and switching costs of FAS. In practice, port switching incurs non-negligible time, during which communication and sensing may continue but with potentially degraded slot-level performance. This leads to two key challenges: (1) the characterization of a slot-level, cost-aware ISAC metric is difficult, and (2) the large port space and accompanying abrupt environmental variations demand more reliable online decision-making. To address these challenges, a cost-aware multi-objective FAS switching problem is formulated, jointly considering slot-level ISAC performance and switching energy. The online conformal Bayesian optimization (OCBO) algorithm is then proposed to learn the unknown gray-box ISAC objectives and calibrate surrogate uncertainty for robust stay-or-switch decisions. Simulation results demonstrate that the proposed cost-aware optimization framework achieves substantially improved long-term ISAC performance compared to existing baselines.

I Introduction

Future 6G networks demand the seamless integration of high-rate communication and high-precision environmental sensing. Fluid antenna systems (FAS), characterized by their ability to move radiation positions within a physical aperture, have emerged as a promising technology for achieving integrated sensing and communication (ISAC) in air-ground systems [6, 5, 9]. By dynamically optimizing antenna positions, FAS can reshape the spatial channel and provide additional spatial degrees of freedom to enhance ISAC integration gain [10]. Motivated by this, extensive research efforts have been devoted to port-selection for FAS-ISAC systems [13, 16].

Most existing FAS-ISAC studies optimize the active antenna positions based on the instantaneous performance of each selected configuration, while simplifying the physical switching process as an instantaneous or cost-free switch [18, 19]. This idealization may lead to overly greedy port switching, where the system frequently moves toward configurations with higher instantaneous gains but ignores the associated movement delay, and energy consumption. As a result, frequent switching may even compromise the performance gains brought by the spatial flexibility of FAS. To address this issue, recent studies start to consider cost-aware FAS and movable-antenna designs from different perspectives. The author in [8] incorporates a switching penalty into the long-term resource-allocation objective, discouraging frequent changes of the active FAS ports and improving the stability of scheduling decisions. Moreover, exhaustive port scanning is shown to become impractical in time-varying channels due to switching latency and measurement offsets in [4]. For liquid-metal FAS, movement delay and actuation energy are incorporated into energy-efficiency-oriented port selection, indicating that ignoring these costs can result in overly aggressive movement and degraded performance [15]. The movement-duration tradeoff of the movable-antenna systems further shows that the time spent on antenna movement may improve the subsequent channel condition but reduces the remaining time for data transmission [7]. However, these studies mainly model the switching interval as an energy penalty, a measurement delay, or a move-then-transmit overhead. For liquid-metal FAS-ISAC, a more fine-grained question remains: how should communication and sensing objectives during antenna movement be modeled and exploited for online switching?

To answer this question, this paper formulates a cost-aware FAS port-switching problem for liquid-metal FAS-ISAC, where the switching process is treated as an unstable but still usable transmission interval. During this interval, communication and sensing signals are not simply discarded; instead, they are accumulated as part of the slot-level cost-aware physical utility together with the dwell-stage performance and mechanical movement energy. Solving this problem faces two key challenges: first, the cost-aware ISAC metric is difficult to characterize, since the slot-level utility jointly depends on switching-stage performance, dwell-stage performance, switching time, and movement energy, while the real-time channel state information (CSI) and target responses required for explicit evaluation are difficult to obtain in dynamic air-ground environments; second, the multi-position FAS configuration space, moving users and targets, and potential abrupt environmental changes may cause model mismatch and overconfident surrogate predictions, requiring more stable online decision-making [14]. To address these challenges, we propose an online conformal Bayesian optimization (OCBO) algorithm. Specifically, Bayesian optimization (BO) is employed to learn the unknown mapping from discrete FAS port switching decisions to cost-aware ISAC objectives using limited evaluations, while online conformal calibration adjusts the surrogate uncertainty through a dynamic residual buffer to improve robustness under environmental variations. Finally, the port switching decision is made by maximizing the cost-aware acquisition function that balances the calibrated utility improvement against the switching energy.

II System Model

In this section, we first introduce the network architecture of the considered FAS-ISAC systems, and then establish the communication and sensing system models.

II-A Network Architecture

We consider the FAS-ISAC systems comprising BB base stations (BSs), KK single-antenna users 𝒦\mathcal{K}, and LL targets denoted by {𝐩,α}=1L\{\mathbf{p}_{\ell},\alpha_{\ell}\}_{\ell=1}^{L}, as illustrated in Fig. 1. Each BS is equipped with a rotatable FAS, where a set of candidate ports 𝒞b\mathcal{C}_{b} is deployed on a planar aperture. In each time slot, BS bb activates MbM_{b} transmit ports and NbN_{b} receive ports from 𝒞b\mathcal{C}_{b}, ensuring that the activated port sets are disjoint. Specifically, let btx\mathcal{M}_{b}^{\text{tx}} and brx\mathcal{M}_{b}^{\text{rx}} denote the local coordinate matrices for the transmit and receive ports, respectively. These coordinates are jointly determined by the selected port indices and the surface orientation defined by rotation angles (θb,ϕb)(\theta_{b},\phi_{b}) [17]. This geometric configuration is defined as at={b,ttx,b,trx,θb,t,ϕb,t}ba_{t}=\{\mathcal{M}_{b,t}^{\text{tx}},\mathcal{M}_{b,t}^{\text{rx}},\theta_{b,t},\phi_{b,t}\}_{b\in\mathcal{B}}, which allows for dynamic alignment of the FAS antenna position at each time slot tt. Based on the FAS configuration ata_{t}, the BSs transmit pilot signals and compute the downlink precoding matrix 𝐖(at)\mathbf{W}(a_{t}) for subcarriers n𝒩n\in\mathcal{N} according to the received feedback, then evaluate the ISAC performance.

The online decision process is structured into discrete time intervals of duration TsT_{\mathrm{s}}, where each slot comprises a switching stage for port switching and a dwell stage for stable communication and sensing. Let vFAv_{\mathrm{FA}} denote the speed of liquid metal. Then, the total switching time Tsw(at|at1)T_{\mathrm{sw}}(a_{t}|a_{t-1}) is defined as

Tsw(at|at1)=maxb,mb𝒩bdpath(b,m)(at1,at)vFA,T_{\mathrm{sw}}(a_{t}|a_{t-1})=\max_{b\in\mathcal{B},m\in\mathcal{M}_{b}\cup\mathcal{N}_{b}}\frac{d_{\mathrm{path}}^{(b,m)}(a_{t-1},a_{t})}{v_{\mathrm{FA}}}, (1)

where dpath(b,m)(at1,at)d_{\mathrm{path}}^{(b,m)}(a_{t-1},a_{t}) denotes the distance traveled by the mm-th port at BS bb during the switching interval [12].

II-B ISAC Model

The communication and sensing spatial responses are modeled by the field response matrix (FRM) [1]. For a spatial path with direction 𝐮(θ,ϕ)\mathbf{u}(\theta,\phi), the transmit steering vector is determined by the selected antenna positions:

𝐚b(b,ttx,θb,t,ϕb,t)=[ej2πλ𝐪b,1T𝐮,,ej2πλ𝐪b,MbT𝐮]T,\mathbf{a}_{b}(\mathcal{M}_{b,t}^{\text{tx}},\theta_{b,t},\phi_{b,t})=\left[e^{-j\frac{2\pi}{\lambda}\mathbf{q}_{b,1}^{T}\mathbf{u}},\dots,e^{-j\frac{2\pi}{\lambda}\mathbf{q}_{b,M_{b}}^{T}\mathbf{u}}\right]^{T}, (2)

where {𝐪b,m}m=1Mb\left\{\mathbf{q}_{b,m}\right\}^{M_{b}}_{m=1} represents the three-dimensional global positions of the mm-th transmit port. The receive steering vector 𝐛b(brx,θ,ϕ)Nbrx×1\mathbf{b}_{b}(\mathcal{M}_{b}^{\mathrm{rx}},\theta,\phi)\in\mathbb{C}^{N_{b}^{\mathrm{rx}}\times 1} is defined analogously. With bb, kk, and nn denoting the BS index, user index, and orthogonal frequency division multiplexing (OFDM) subcarrier index, respectively, the communication channel vector 𝐡b,k,n\mathbf{h}_{b,k,n} from BS bb to user kk over subcarrier nn is obtained from the FRM 𝐇b,k,n\mathbf{H}_{b,k,n} as

𝐡b,k,n=𝐇b,k,n(btx)𝝈b,k,n,\mathbf{h}_{b,k,n}=\mathbf{H}_{b,k,n}(\mathcal{M}_{b}^{\mathrm{tx}})\boldsymbol{\sigma}_{b,k,n}, (3)

where 𝐇b,k,n(btx)\mathbf{H}_{b,k,n}(\mathcal{M}_{b}^{\mathrm{tx}}) denotes the submatrix of 𝐇b,k,n\mathbf{H}_{b,k,n} corresponding to the selected transmit ports.

Refer to caption
Figure 1: Structure of the considered FAS-ISAC systems.

Adopting the OFDM scheme, the transmitted signal vector 𝐱b,i,n\mathbf{x}_{b,i,n} from BS bb for the ii-th symbol and nn-th subcarrier is given by 𝐱b,i,n=k𝒦𝐰b,k,n(at)sk,i,n\mathbf{x}_{b,i,n}=\sum_{k\in\mathcal{K}}\mathbf{w}_{b,k,n}(a_{t})s_{k,i,n}, where 𝐰b,k,n(at)\mathbf{w}_{b,k,n}(a_{t}) is the precoding vector corresponding to the configuration ata_{t}, and sk,i,ns_{k,i,n} is the data symbol. The received signal at user kk for the nn-th subcarrier is expressed as yk,i,n=b𝐡b,k,nH(at)𝐱b,i,n+zk,i,ny_{k,i,n}=\sum_{b\in\mathcal{B}}\mathbf{h}_{b,k,n}^{H}(a_{t})\mathbf{x}_{b,i,n}+z_{k,i,n}, where zk,i,nz_{k,i,n} denotes the additive white Gaussian noise (AWGN). The instantaneous signal-to-interference-plus-noise ratio (SINR) for user kk on subcarrier nn is formulated as:

γk,n(at)=|b𝐡b,k,nH(at)𝐰b,k,n(at)|2kk|b𝐡b,k,nH(at)𝐰b,k,n(at)|2+σk2.\gamma_{k,n}(a_{t})=\frac{\left|\sum_{b\in\mathcal{B}}\mathbf{h}_{b,k,n}^{H}(a_{t})\mathbf{w}_{b,k,n}(a_{t})\right|^{2}}{\sum_{k^{\prime}\neq k}\left|\sum_{b\in\mathcal{B}}\mathbf{h}_{b,k,n}^{H}(a_{t})\mathbf{w}_{b,k^{\prime},n}(a_{t})\right|^{2}+\sigma_{k}^{2}}. (4)

Accordingly, the steady-state sum-rate for the dwell stage is defined as Rd(at)=1|𝒩|n𝒩k𝒦log2(1+γk,n(at)).R_{\mathrm{d}}(a_{t})=\frac{1}{|\mathcal{N}|}\sum_{n\in\mathcal{N}}\sum_{k\in\mathcal{K}}\log_{2}(1+\gamma_{k,n}(a_{t})).

We consider a fully cooperative sensing mode, where all BSs in \mathcal{B} simultaneously act as sensing transmitters and receivers. For a given FAS configuration ata_{t}, the sensing observation collected at each BS rr\in\mathcal{B} is expressed as:

𝐲r=𝐀r(at)𝐠s+𝐇SI,r(at)𝐱r+b{r}𝐇CI,r,b(at)𝐱b+𝐧r,\mathbf{y}_{r}=\mathbf{A}_{r}(a_{t})\mathbf{g}_{\mathrm{s}}+\mathbf{H}_{\mathrm{SI},r}(a_{t})\mathbf{x}_{r}+\sum_{b\in\mathcal{B}\setminus\{r\}}\mathbf{H}_{\mathrm{CI},r,b}(a_{t})\mathbf{x}_{b}+\mathbf{n}_{r}, (5)

where 𝐠sL×1\mathbf{g}_{\mathrm{s}}\in\mathbb{C}^{L\times 1} denotes the target response vector for the LL targets, and 𝐀r(at)\mathbf{A}_{r}(a_{t}) is the sensing dictionary at receiver rr, which incorporates the aggregate echoes from all transmitting BSs. The terms 𝐇SI,r(at)\mathbf{H}_{\mathrm{SI},r}(a_{t}) and 𝐇CI,r,b(at)\mathbf{H}_{\mathrm{CI},r,b}(a_{t}) represent the residual self-interference (SI) channel at BS rr and the cross-interference (CI) channel from BS bb to BS rr, respectively. By stacking the observations from all BSs, the joint sensing model is given by:

𝐲s=𝐀s(at)𝐠s+𝐳s(at)+𝐧s,\mathbf{y}_{\mathrm{s}}=\mathbf{A}_{\mathrm{s}}(a_{t})\mathbf{g}_{\mathrm{s}}+\mathbf{z}_{\mathrm{s}}(a_{t})+\mathbf{n}_{\mathrm{s}}, (6)

where 𝐀s(at)=[𝐀1T,,𝐀BT]T\mathbf{A}_{\mathrm{s}}(a_{t})=[\mathbf{A}_{1}^{T},\dots,\mathbf{A}_{B}^{T}]^{T} is the global sensing dictionary, 𝐳s(at)\mathbf{z}_{\mathrm{s}}(a_{t}) denotes the aggregate residual SI and CI, and 𝐧s\mathbf{n}_{\mathrm{s}} denotes the stacked sensing noise. After interference suppression, the effective disturbance is modeled by the covariance matrix 𝐑e(at)=σs2𝐈+𝐑inf(at)\mathbf{R}_{\mathrm{e}}(a_{t})=\sigma_{\mathrm{s}}^{2}\mathbf{I}+\mathbf{R}_{\mathrm{inf}}(a_{t}). Assuming 𝐠s𝒞𝒩(𝟎,𝐑g)\mathbf{g}_{\mathrm{s}}\sim\mathcal{CN}(\mathbf{0},\mathbf{R}_{\mathrm{g}}), the configuration-level sensing mutual information (MI) is defined as:

s(at)=log2det(𝐈+𝐑e1(at)𝐀s(at)𝐑g𝐀sH(at)).\mathcal{I}_{\mathrm{s}}(a_{t})=\log_{2}\det\left(\mathbf{I}+\mathbf{R}_{\mathrm{e}}^{-1}(a_{t})\mathbf{A}_{\mathrm{s}}(a_{t})\mathbf{R}_{\mathrm{g}}\mathbf{A}_{\mathrm{s}}^{H}(a_{t})\right). (7)

III Cost-aware FAS Switching via OCBO Method

This section develops the cost-aware FAS-ISAC switching framework. First, we derive cost-aware performance metrics that characterize the slot-level communication and sensing objectives under both port movement and stable dwell stages. Then, we formulate the resulting stay-or-switch decision as a multi-objective online optimization problem. Finally, we propose the OCBO algorithm to enable robust FAS reconfiguration.

Refer to caption
Figure 2: Cost-aware slot structure for FAS-ISAC port switching

III-A Cost-Aware Performance Metrics

The movement of FAS ports introduces a time-varying system geometry during the switching interval, making the slot-level performance dependent on the switching trajectory between the preceding configuration at1a_{t-1} and the target configuration ata_{t}. Here, the slot-level communication utility is defined as the weighted average of the rates achieved in each stage:

Rc(at|at1)=TswTsRmv(at|at1)+TdTsRd(at,𝐖(at)),R_{\mathrm{c}}(a_{t}|a_{t-1})=\frac{T_{\mathrm{sw}}}{T_{\mathrm{s}}}R_{\mathrm{mv}}(a_{t}|a_{t-1})+\frac{T_{\mathrm{d}}}{T_{\mathrm{s}}}R_{\mathrm{d}}(a_{t},\mathbf{W}(a_{t})), (8)

where Rd(at,𝐖(at))R_{\mathrm{d}}(a_{t},\mathbf{W}(a_{t})) is the steady-state rate at the target configuration. Td=TsTsw(at|at1)T_{\mathrm{d}}=T_{\mathrm{s}}-T_{\mathrm{sw}}(a_{t}|a_{t-1}) represents the subsequent dwell stage. RmvR_{\mathrm{mv}} is characterized as follows.

Proposition 1.

In the high-port-density regime, the switching-stage rate can be approximated by the average of the instantaneous rates evaluated at the available ports along the movement path from at1a_{t-1} to ata_{t},

Rmv(at|at1)1|𝒫|ap𝒫Rcins(ap,𝐖t1),R_{\mathrm{mv}}(a_{t}|a_{t-1})\approx\frac{1}{|\mathcal{P}|}\sum_{a_{p}\in\mathcal{P}}R_{\mathrm{c}}^{\mathrm{ins}}(a_{p},\mathbf{W}_{t-1}), (9)

where 𝒫\mathcal{P} denotes the set of available ports along the movement path, and Rcins(an,𝐖t1)R_{\mathrm{c}}^{\mathrm{ins}}(a_{n},\mathbf{W}_{t-1}) denotes the instantaneous sum-rate evaluated at port configuration ana_{n} using 𝐖t1\mathbf{W}_{t-1}, i.e., Rcins(an,𝐖t1)=1|𝒩|n𝒩k𝒦log2(1+γk,n(an,𝐖t1))R_{\mathrm{c}}^{\mathrm{ins}}(a_{n},\mathbf{W}_{t-1})=\frac{1}{|\mathcal{N}|}\sum_{n\in\mathcal{N}}\sum_{k\in\mathcal{K}}\log_{2}(1+\gamma_{k,n}(a_{n},\mathbf{W}_{t-1})).

Proof.

Let a(τ)a(\tau) denote the continuous FAS configuration during the switching interval τ[0,Tsw]\tau\in[0,T_{\mathrm{sw}}]. During the switching, the beamformer is fixed as 𝐖t1=𝐖(at1)\mathbf{W}_{t-1}=\mathbf{W}(a_{t-1}), and the switching-stage rate is

Rmv(at|at1)=1Tsw0TswRcins(a(τ),𝐖t1)𝑑τ.R_{\mathrm{mv}}(a_{t}|a_{t-1})=\frac{1}{T_{\mathrm{sw}}}\int_{0}^{T_{\mathrm{sw}}}R_{\mathrm{c}}^{\mathrm{ins}}(a(\tau),\mathbf{W}_{t-1})d\tau. (10)

For FAS channels, the spatial correlation between two nearby positions separated by distance dd can be characterized by ρh(d)=J0(2πd/λ)\rho_{h}(d)=J_{0}(2\pi d/\lambda), where J0()J_{0}(\cdot) is the zero-order Bessel function. In the high-port-density regime, d/λd/\lambda is small, and using J0(x)=1x2/4+O(x4)J_{0}(x)=1-x^{2}/4+O(x^{4}) gives ρh(d)=1π2d2/λ2+O(d4/λ4)\rho_{h}(d)=1-\pi^{2}d^{2}/\lambda^{2}+O(d^{4}/\lambda^{4}). ρh(d)\rho_{h}(d) approaches one as d/λ0d/\lambda\rightarrow 0, indicating strong channel correlation within each local port neighborhood. Therefore, the continuous rate around each local segment can be represented by the rate evaluated at the corresponding port. Accordingly, the integral in (10) can be approximated by a weighted sum over the ports along the movement path,

Rmv(at|at1)an𝒫(at1,at)ωnRcins(an,𝐖t1).R_{\mathrm{mv}}(a_{t}|a_{t-1})\approx\sum_{a_{n}\in\mathcal{P}(a_{t-1},a_{t})}\omega_{n}R_{\mathrm{c}}^{\mathrm{ins}}(a_{n},\mathbf{W}_{t-1}). (11)

With uniformly spaced ports and constant movement speed, the local weights are approximately identical, i.e., ωn1/|𝒫(at1,at)|\omega_{n}\approx 1/|\mathcal{P}(a_{t-1},a_{t})|. Thus,

Rmv(at|at1)1|𝒫(at1,at)|an𝒫Rcins(an,𝐖t1).R_{\mathrm{mv}}(a_{t}|a_{t-1})\approx\frac{1}{|\mathcal{P}(a_{t-1},a_{t})|}\sum_{a_{n}\in\mathcal{P}}R_{\mathrm{c}}^{\mathrm{ins}}(a_{n},\mathbf{W}_{t-1}). (12)

By writing Rcins(an)=Rcins(an,𝐖t1)R_{\mathrm{c}}^{\mathrm{ins}}(a_{n})=R_{\mathrm{c}}^{\mathrm{ins}}(a_{n},\mathbf{W}_{t-1}) during the switching-stage, (9) follows. ∎

Lemma 1.

In the high-port-density regime, the switching-stage sensing MI can be approximated by the average of the configuration-level sensing MI evaluated at the available ports along the movement path:

Imv(at|at1)1|𝒫(at1,at)|an𝒫(at1,at)s(an,𝐖t1),I_{\mathrm{mv}}(a_{t}|a_{t-1})\approx\frac{1}{|\mathcal{P}(a_{t-1},a_{t})|}\sum_{a_{n}\in\mathcal{P}(a_{t-1},a_{t})}\mathcal{I}_{\mathrm{s}}(a_{n},\mathbf{W}_{t-1}), (13)

where 𝒫(at1,at)\mathcal{P}(a_{t-1},a_{t}) denotes the set of available ports along the movement path, and s(an,𝐖t1)\mathcal{I}_{\mathrm{s}}(a_{n},\mathbf{W}_{t-1}) denotes the instantaneous sensing MI evaluated at port configuration ana_{n} using 𝐖t1\mathbf{W}_{t-1}.

Similarly, the slot-level sensing utility is modeled as

Is(at|at1)=TswTsImv(at|at1)+TdTsId(at,𝐖(at)),I_{\mathrm{s}}(a_{t}|a_{t-1})=\frac{T_{\mathrm{sw}}}{T_{\mathrm{s}}}I_{\mathrm{mv}}(a_{t}|a_{t-1})+\frac{T_{\mathrm{d}}}{T_{\mathrm{s}}}I_{\mathrm{d}}(a_{t},\mathbf{W}(a_{t})), (14)

where Id(at,𝐖(at))=s(at,𝐖(at))I_{\mathrm{d}}(a_{t},\mathbf{W}(a_{t}))=\mathcal{I}_{\mathrm{s}}(a_{t},\mathbf{W}(a_{t})). According to Lemma 1, the switching-stage sensing MI can be evaluated by the port-sampled approximation in (13).

The switching energy is computed according to the total movement time of all activated FA elements:

Esw(at|at1)=b,mPmvdpath(b,m)(at1,at)vFA,E_{\mathrm{sw}}(a_{t}|a_{t-1})=\sum_{b,m}P_{\mathrm{mv}}\frac{d_{\mathrm{path}}^{(b,m)}(a_{t-1},a_{t})}{v_{\mathrm{FA}}}, (15)

where PmvP_{\mathrm{mv}} denotes the movement power.

III-B Problem Formulation

Based on the above discussion, the reward vector capturing the multi-objective trade-off is defined as:

𝐅(at|at1)=[Rc(at|at1),Is(at|at1),Esw(at|at1)]T,\mathbf{F}(a_{t}|a_{t-1})=\left[R_{\mathrm{c}}(a_{t}|a_{t-1}),I_{\mathrm{s}}(a_{t}|a_{t-1}),-E_{\mathrm{sw}}(a_{t}|a_{t-1})\right]^{T}, (16)

where the negative sign indicates that the switching energy is minimized while communication and sensing objectives are maximized. Then, the optimization problem is formulated as:

(P1) max{at}t=1Tt=1T𝐅(at|at1)\displaystyle\max_{\{a_{t}\}_{t=1}^{T}}\quad\sum_{t=1}^{T}\mathbf{F}(a_{t}|a_{t-1})
s.t. b,ttx𝒞b,b,trx𝒞b,b,\displaystyle\mathcal{M}_{b,t}^{\mathrm{tx}}\subseteq\mathcal{C}_{b},\quad\mathcal{M}_{b,t}^{\mathrm{rx}}\subseteq\mathcal{C}_{b},\quad\forall b\in\mathcal{B}, (17a)
|b,ttx|=Mb,|b,trx|=Nb,b,\displaystyle|\mathcal{M}_{b,t}^{\mathrm{tx}}|=M_{b},\quad|\mathcal{M}_{b,t}^{\mathrm{rx}}|=N_{b},\quad\forall b\in\mathcal{B}, (17b)
b,ttxb,trx=,b,\displaystyle\mathcal{M}_{b,t}^{\mathrm{tx}}\cap\mathcal{M}_{b,t}^{\mathrm{rx}}=\emptyset,\quad\forall b\in\mathcal{B}, (17c)
Tsw(at|at1)Ts,t𝒯,\displaystyle T_{\mathrm{sw}}(a_{t}|a_{t-1})\leq T_{\mathrm{s}},\quad\forall t\in\mathcal{T}, (17d)
𝐖(at)=(𝐇^(at)),t𝒯,\displaystyle\mathbf{W}(a_{t})=\mathcal{B}(\widehat{\mathbf{H}}(a_{t})),\quad\forall t\in\mathcal{T}, (17e)
Tr(𝐖b,n(at)𝐖b,nH(at))Pb,max,b,n,\displaystyle\operatorname{Tr}\left(\mathbf{W}_{b,n}(a_{t})\mathbf{W}_{b,n}^{H}(a_{t})\right)\leq P_{b,\max},\quad\forall b,n, (17f)

where (17a)–(17c) define the feasible port selection and exclusivity constraints, (17d) ensures the reconfiguration completes within one slot, and (17e)–(17f) specify the precoding policy and power constraints. Since 𝐅(at|at1)\mathbf{F}(a_{t}|a_{t-1}) is a vector-valued reward, the maximization in (P1) is pursued in the Pareto sense. Here, 𝒞b\mathcal{C}_{b} is the candidate port set on its FAS. The constants MbM_{b} and NbN_{b} represent the numbers of activated transmit and receive ports at BS bb, respectively, and 𝒯\mathcal{T} denotes the set of time slots. The precoding matrix 𝐖(at)\mathbf{W}(a_{t}) is obtained from the precoding rule (𝐇^(at))\mathcal{B}(\hat{\mathbf{H}}(a_{t})) based on the pilot feedback, and Pb,maxP_{b,\max} denotes the per-subcarrier transmit power budget of BS bb.

The ISAC objectives depend on unknown instantaneous channels, target responses, and residual interference, which can be represented as implicit mappings Rc(at|at1)=Φc(at1,at,𝐇t)R_{\mathrm{c}}(a_{t}|a_{t-1})=\Phi_{\mathrm{c}}(a_{t-1},a_{t},\mathbf{H}_{t}) and Is(at|at1)=Φs(at1,at,𝐀s,t,𝐑SI,t)I_{\mathrm{s}}(a_{t}|a_{t-1})=\Phi_{\mathrm{s}}(a_{t-1},a_{t},\mathbf{A}_{\mathrm{s},t},\mathbf{R}_{\mathrm{SI},t}). Therefore, part of the objectives in (P1) have a gray-box structure, which cannot be calculated in closed form by the white box methods before the configuration is evaluated.

III-C Cost-aware GP Surrogates

The whole slot structure is shown in Fig. 2. In the proposed FAS switching process, each action corresponds to a port switching from the previous FAS configuration at1a_{t-1} to the target configuration ata_{t}. After executing this switching, the system observes noisy feedback from the realized ISAC objectives:

yt(m)=Fm(at|at1)+ξt(m),m{1,2},y_{t}^{(m)}=F_{m}(a_{t}|a_{t-1})+\xi_{t}^{(m)},\quad m\in\{1,2\}, (18)

where m=1m=1 and m=2m=2 correspond to the observations of RcR_{\mathrm{c}} and IsI_{\mathrm{s}}, respectively, and ξt(m)\xi_{t}^{(m)} denotes measurement noise. The feedback is then used by the proposed OCBO framework to update the surrogate models and guide the online search. Specifically, two independent Gaussian Process (GP) surrogates are deployed to learn the two cost-aware gray-box physical objectives (m{1,2}m\in\{1,2\}). Since the slot-level feedback depends on the FAS port switching from at1a_{t-1} to aa, the surrogate input is defined over the switching pair (at1,a)(a_{t-1},a) rather than the target port alone. The surrogate can be modeled as:

Fm(a|at1)𝒢𝒫(μmg(a|at1,𝐛t),km((a,at1),)).F_{m}(a|a_{t-1})\sim\mathcal{GP}\left(\mu_{m}^{\mathrm{g}}(a|a_{t-1},\mathbf{b}_{t}),k_{m}((a,a_{t-1}),\cdot)\right). (19)

Given the historical evaluation dataset 𝒟t(m)={((aτ,aτ1,sτ),yτ(m))}τ=1t\mathcal{D}_{t}^{(m)}=\{((a_{\tau},a_{\tau-1},s_{\tau}),y_{\tau}^{(m)})\}_{\tau=1}^{t}, the GP predictive distribution for a candidate switching is [3]:

p(y(m)|a,at1,𝒟t(m))=𝒩(μm,σm2).p(y^{(m)}|a,a_{t-1},\mathcal{D}_{t}^{(m)})=\mathcal{N}\left(\mu_{m},\sigma_{m}^{2}\right). (20)

In the proposed cost-aware FAS switching architecture, the feedback associated with a target configuration ata_{t} is not solely determined by its dwell performance, but also by the preceding configuration at1a_{t-1}. Consequently, the observed utility may exhibit action-dependent fluctuations and temporal drift that are not fully captured by the standard GP posterior variance. This can make the GP surrogate overconfident in dynamically changing regions, thereby misleading the acquisition function. To mitigate this issue, we introduce conformal calibration to adjust the predictive uncertainty using residual feedback, enabling more reliable stay-or-switch decisions.

Algorithm 1 The OCBO-Based FAS Port Switching Algorithm
1:Input: Feasible FAS configuration space 𝒜\mathcal{A}; initial FAS configuration a0a_{0}; miscoverage level α\alpha; energy penalty weight ηe\eta_{\rm e}; time horizon TT.
2:Initialization: Fit cost-aware GP surrogates using initial evaluations 𝒟0\mathcal{D}_{0}.
3:for each time slot t=1,2,,Tt=1,2,\dots,T do
4:  Observe the previous configuration at1a_{t-1} and construct the feasible candidate set 𝒞t𝒜\mathcal{C}_{t}\subseteq\mathcal{A} according to the constraints in (17a)–(17d).
5:  Compute the OCBO acquisition score Qt(a)Q_{t}(a) for each feasible candidate a𝒞ta\in\mathcal{C}_{t} according to (23).
6:  Select target configuration by maximizing (23).
7:  Deploy ata_{t} and observe 𝐲t=[yt(1),yt(2)]T\mathbf{y}_{t}=[y_{t}^{(1)},y_{t}^{(2)}]^{T} according to (18).
8:  for m{1,2}m\in\{1,2\} do
9:   Update dataset 𝒟t(m)𝒟t1(m){(ϕt(at),yt(m))}\mathcal{D}_{t}^{(m)}\leftarrow\mathcal{D}_{t-1}^{(m)}\cup\{(\phi_{t}(a_{t}),y_{t}^{(m)})\}.
10:   Fit the GP surrogate using 𝒟t(m)\mathcal{D}_{t}^{(m)} and obtain the posterior in (20).
11:   Update 𝒮t,m\mathcal{S}_{t,m} using residual st,ms_{t,m} in (21).
12:   Update qt,mq_{t,m} according to (22).
13:  end for
14:end for
15:Output: Deployed configuration sequence {at}t=1T\{a_{t}\}_{t=1}^{T} and collected observations.

III-D Online Conformal Calibration

To robustify the uncertainty estimates against model misspecification caused by FAS port switching, beamformer mismatch, and environmental drift, we apply online conformal calibration [11]. After observing the noisy feedback yt(m)y_{t}^{(m)} and computing the GP prediction mean μm(at|at1)\mu_{m}(a_{t}|a_{t-1}) and standard deviation σm(at|at1)\sigma_{m}(a_{t}|a_{t-1}), we define the residual score:

st,m=|yt(m)μm(at|at1)|σm(at|at1)+ϵ,s_{t,m}=\frac{|y_{t}^{(m)}-\mu_{m}(a_{t}|a_{t-1})|}{\sigma_{m}(a_{t}|a_{t-1})+\epsilon}, (21)

where ϵ\epsilon is a small positive constant. We maintain a recent calibration buffer 𝒮t,m={stW,m,,st1,m}\mathcal{S}_{t,m}=\{s_{t-W,m},\dots,s_{t-1,m}\}. Based on this online buffer, we compute the conformal calibration factor as the empirical quantile for a target miscoverage risk α\alpha as:

qt,m=Quantile1α(𝒮t,m),q_{t,m}=\mathrm{Quantile}_{1-\alpha}(\mathcal{S}_{t,m}), (22)

where Quantile1α(𝒮t,m)\mathrm{Quantile}_{1-\alpha}(\mathcal{S}_{t,m}) returns the value below which approximately a fraction 1α1-\alpha of the residual scores in 𝒮t,m\mathcal{S}_{t,m} fall. The GP predictive uncertainty is then robustly scaled as σ~t,m(a)=qt,mσm(a|at1)\tilde{\sigma}_{t,m}(a)=q_{t,m}\sigma_{m}(a|a_{t-1}).

We adopt the expected hypervolume improvement (EHI) as the acquisition function to guide the multi-objective FAS switching search [2]. Let 𝒫t\mathcal{P}_{t} denote the current approximated Pareto set and 𝐫\mathbf{r} be the reference point. For a predicted objective sample 𝐅~(a|at1)\tilde{\mathbf{F}}(a|a_{t-1}), the hypervolume improvement is HVIt(a)=HV(𝒫t{𝐅~(a|at1)};𝐫)HV(𝒫t;𝐫)\mathrm{HVI}_{t}(a)=\mathrm{HV}\!\left(\mathcal{P}_{t}\cup\{\tilde{\mathbf{F}}(a|a_{t-1})\};\mathbf{r}\right)-\mathrm{HV}\!\left(\mathcal{P}_{t};\mathbf{r}\right), where HV(;𝐫)\mathrm{HV}(\cdot;\mathbf{r}) denotes the hypervolume dominated by a Pareto set and bounded by 𝐫\mathbf{r}. The EHI score is then given by EHIt(a|at1)=𝔼𝐅~[HVIt(a)]\mathrm{EHI}_{t}(a|a_{t-1})=\mathbb{E}_{\tilde{\mathbf{F}}}\left[\mathrm{HVI}_{t}(a)\right], where 𝐅~\tilde{\mathbf{F}} is sampled from the conformal-calibrated predictive distribution. Based on the calibrated EHI score, the deterministic FAS port movement energy is further introduced as an explicit penalty, yielding the following cost-aware acquisition function:

𝒬t(a)=EHIt(a|at1)ηeEmv(a|at1),\mathcal{Q}_{t}(a)=\mathrm{EHI}_{t}(a|a_{t-1})-\eta_{\mathrm{e}}E_{\mathrm{mv}}(a|a_{t-1}), (23)

where ηe\eta_{\mathrm{e}} controls the trade-off between the expected physical performance improvement and the switching energy cost. The next FAS configuration is selected by maximizing 𝒬t(a)\mathcal{Q}_{t}(a). For the stay action (a=at1a=a_{t-1}), its cost-aware score is simply EHIt(at1|at1)\mathrm{EHI}_{t}(a_{t-1}|a_{t-1}). By combining the known analytical energy cost with the conformal-calibrated physical uncertainty, the framework actively avoids risky or overly frequent FAS port switchings. The procedure is summarized in Algorithm 1.

Refer to caption
Figure 3: Running average online cost-aware utility over 200200 time slots.
Refer to caption
Figure 4: Average online cost-aware utility versus switching power.
Refer to caption
Figure 5: Average online cost-aware utility versus FA movement speed.

IV Simulations

In this section, we evaluate the performance of the proposed cost-aware FAS-ISAC switching algorithm through numerical simulations. The simulation scenario is considered within a 3D space of 200×200×80200\times 200\times 80 m3, where B=3B=3 fixed BSs are deployed. The system serves K=3K=3 dynamic communication users moving at speeds up to 1010 m/s, while simultaneously monitoring L=2L=2 dynamic point targets (UAVs) via a fully cooperative networked sensing architecture. These UAV targets move at velocities of approximately 1515 m/s, necessitating real-time configuration adaptation. Each BS is equipped with a fluid antenna surface featuring a 16×1616\times 16 candidate port grid (256 ports in total) with a port spacing of 0.25λ0.25\lambda. In each time slot tt, each BS activates Mb=2M_{b}=2 ports for communication transmission and Nb=2N_{b}=2 ports for sensing reception. The system operates at a carrier frequency of 2828 GHz using an OFDM waveform with a 128128-point FFT and |𝒩|=32|\mathcal{N}|=32 active subcarriers. The transmit power is set to 4040 dBm. The channel model incorporates Rician fading with a KK-factor of 1010 dB and a pathloss model with an exponent of 2.52.5. The online decision-making process spans T=200T=200 slots with a slot interval Ts=0.02T_{\mathrm{s}}=0.02 s, totaling a duration of 44 s.

We compare the proposed cost-aware FAS switching architecture with four baselines. Random selects the FAS configuration uniformly from the candidate set. Fixed keeps the initial port configuration unchanged during all slots. MOBO adopts standard multi-objective Bayesian optimization without conformal calibration. Dwell-only uses multi-objective search strategy but evaluates only the dwell-stage objectives, ignoring the switching-stage contribution. These baselines are used to verify the benefits of cost-aware performance modeling and conformal uncertainty calibration.

Fig. 3 illustrates the running average online cost-aware utility across 200200 slots. It can be observed that the proposed OCBO method significantly outperforms all other baselines, reaching a superior long-term utility of approximately 23.523.5. While MOBO initially tracks closely with the proposed method, its performance gap widens as the environment evolves, primarily because standard GP surrogates struggle with miscalibration under the dynamic mismatches induced by user mobility. In contrast, the conformal calibration in OCBO robustly adjusts uncertainty estimates, preventing the optimizer from making overly aggressive or suboptimal switching decisions. The Dwell-only baseline, which ignores switching-stage contributions, exhibit a noticeable performance degradation. This confirms that modeling the unstable yet usable signals during antenna movement is crucial for maximizing long-term reward.

Fig. 4 shows the average online cost-aware utility under different switching power levels. As PswP_{\mathrm{sw}} increases, the movement energy penalty becomes more significant, leading to a gradual utility reduction for switching-based schemes. The proposed OCBO consistently achieves the highest utility across all tested power levels, indicating that the cost-aware acquisition function effectively suppresses low-benefit switching when the mechanical cost becomes large. In contrast, Random suffers a sharp degradation because it frequently triggers unnecessary port movements without considering the energy overhead. The Dwell-only variants also exhibit lower performance since their decisions are made without explicitly accounting for the usable switching-stage performance, resulting in a suboptimal performance-cost tradeoff.

Fig. 5 evaluates the impact of the FA movement speed vFAv_{\mathrm{FA}}. A higher movement speed shortens the switching duration and reduces the switching overhead, thereby improving the average utility of most adaptive schemes. The proposed OCBO maintains the best performance over the entire speed range, showing its robustness to different hardware movement capabilities. Compared with dwell-only MOBO and dwell-only OCBO, the gain of the proposed method confirms the benefit of incorporating switching-stage communication and sensing objectives into the slot-level evaluation. The Fixed baseline remains almost unchanged because no port movement is performed, while the Random baseline improves with speed but remains inferior due to its lack of informed switching decisions.

V Conclusions

This paper investigated a smart switching mechanism for FAS to achieve improved ISAC performance. A cost-aware framework was developed that incorporates both switching-stage and dwell-stage ISAC objectives. The online conformal Bayesian optimization was adopted to learn the unknown physical performance while calibrating surrogate uncertainty under environmental drift. An acquisition function was further tailored to balance the calibrated utility improvement against deterministic movement energy. Simulation results demonstrate that the proposed switching scheme achieves a better long-term performance–cost trade-off than conventional cost-agnostic schemes, by exploiting usable switching-stage objectives and reducing unnecessary port movements. Future work will address complex dynamic environments with abrupt distribution shifts, along with extended-target sensing and tighter communication–sensing integration for more consistent ISAC performance.

References

  • [1] J. Chen, T. Cheng, K. Wong, and H. Shin (2025-06) Improved Joint Transmit and Receive Port Selection for Capacity Maximization in Fluid-MIMO Systems. IEEE Wireless Communications Letters 14 (6), pp. 1693–1697. External Links: Document Cited by: §II-B.
  • [2] S. Daulton, M. Balandat, and E. Bakshy (2021) Parallel bayesian optimization of multiple noisy objectives with expected hypervolume improvement. In Advances in Neural Information Processing Systems, Vol. 34, pp. 2187–2200. Cited by: §III-D.
  • [3] S. Daulton, D. Eriksson, M. Balandat, and E. Bakshy (2022-08) Multi-Objective Bayesian Optimization over High-Dimensional Search Spaces. In Proceedings of the 38th Conference on Uncertainty in Artificial Intelligence, Vol. 180, Eindhoven, pp. 507–517. Cited by: §III-C.
  • [4] D. Dinis and R. Wichman (2026-01) Spatio-Temporal Port Selection in Fluid Antennas Under Switching Delays. IEEE Communications Letters PP (99), pp. 1–1. External Links: Document Cited by: §I.
  • [5] S. Gao et al. (2026-Mar.) Integrated Sensing, Communication, and Computation for Low-Altitude Networks Towards Seamless Connectivity and Connected Intelligence. IEEE Internet of Things Magazine, pp. 1–9. External Links: Document Cited by: §I.
  • [6] N. González-Prelcic et al. (2024-07) The Integrated Sensing and Communication Revolution for 6G: Vision, Techniques, and Applications. Proceedings of the IEEE 112 (7), pp. 676–723. External Links: Document Cited by: §I.
  • [7] G. Hu et al. (2026) Fundamental Tradeoff in Movable Antenna Systems: How Long to Move Before Transmission?. arXiv preprint arXiv:2604.20386. Cited by: §I.
  • [8] J. Liu et al. (2026-03) Switching-Cost-Aware Deep Reinforcement Learning for Dynamic Port Selection in Fluid Antenna Systems. IEEE Communications Letters 30, pp. 1548–1552. External Links: Document Cited by: §I.
  • [9] S. Lu et al. (2024-06) Integrated Sensing and Communications: Recent Advances and Ten Open Challenges. IEEE Internet of Things Journal 11 (11), pp. 19094–19120. External Links: Document Cited by: §I.
  • [10] W. K. New et al. (2025-Aug.) A Tutorial On Fluid Antenna System For 6G Networks: Encompassing Communication Theory, Optimization Methods And Hardware Designs. IEEE Communications Surveys & Tutorials 27 (4), pp. 2325–2377. Cited by: §I.
  • [11] S. Stanton, W. Maddox, and A. G. Wilson (2023-04) Bayesian Optimization with Conformal Prediction Sets. In Proceedings of the 26th International Conference on Artificial Intelligence and Statistics, Vol. 206, Valencia, pp. 959–986. Cited by: §III-D.
  • [12] J. Wang, Y. Mao, X. Yu, and Y. A. Zhang (2026) Energy-efficient velocity profile optimization for movable antenna-enabled sensing systems. arXiv preprint arXiv:2603.27540. Cited by: §II-A.
  • [13] T. Wu et al. (2025-12) Fluid Antenna Systems Enabling 6G: Principles, Applications, and Research Directions. IEEE Wireless Communications, pp. 1–9. Note: Early Access External Links: Document Cited by: §I.
  • [14] Z. Yang, S. Gao, X. Cheng, and L. Yang (2025-Sep.) Synesthesia Of Machines (SoM)-Enhanced ISAC Precoding For Vehicular Networks With Double Dynamics. IEEE Transactions on Communications 73 (9), pp. 7967–7984. Cited by: §I.
  • [15] L. Zhang, Y. Zhao, H. Yang, G. Liang, and J. Hu (2025-09) Energy-Efficient Port Selection and Beamforming Design for Integrated Data and Energy Transfer Assisted by Fluid Antennas. IEEE Journal on Selected Areas in Communications 44, pp. 1480–1494. External Links: Document Cited by: §I.
  • [16] Z. Zhang, K. Wong, J. Dang, Z. Zhang, and C. Chae (2026-01) On fundamental limits for fluid antenna-assisted integrated sensing and communications for unsourced random access. IEEE Journal on Selected Areas in Communications 44, pp. 136–149. External Links: Document Cited by: §I.
  • [17] B. Zheng et al. (2025-10) Rotatable Antenna Enabled Wireless Communication and Sensing: Opportunities and Challenges. IEEE Wireless Communications, pp. 1–8. Note: Early Access External Links: Document Cited by: §II-A.
  • [18] L. Zhu et al. (2024-09) Movable Antenna Enhanced Multiuser Communication via Antenna Position Optimization. IEEE Transactions on Wireless Communications 23 (9), pp. 11814–11830. External Links: Document Cited by: §I.
  • [19] J. Zou, H. Xu, C. Wang, L. Xu, S. Sun, and K. T. Meng (2024-12) Shifting the isac trade-off with fluid antenna systems. IEEE Wireless Communications Letters 13 (12), pp. 3479–3483. External Links: Document Cited by: §I.