Skip to main content
archive
Search Submit Donate Log in
Press Enter to search · Advanced search

Signal Processing

  • New submissions
  • Cross-lists
  • Replacements

See recent articles

Showing new listings for Thursday, 30 July 2026

Total of 32 entries
Showing up to 2000 entries per page: fewer | more | all

New submissions (showing 15 of 15 entries)

[1] arXiv:2607.26209 [pdf, html, other]
Title: Evolutionary AP Switch ON/OFF Techniques for Energy-efficient Cell-free Massive MIMO Networks
Jan García-Morales, Alejandro de la Fuente, David Gualda, Leopoldo Carro-Calvo, Felip Riera-Palou, Guillem Femenias
Comments: Author Accepted Manuscript. Published in IEEE Open Journal of the Communications Society. Available at this https URL. Copyright 2026 IEEE
Subjects: Signal Processing (eess.SP)

Cell-free massive multiple input multiple output (CF-mMIMO) is an emerging technology for next-generation wireless systems, where dynamically adapting the set of active access points (APs) is crucial to balance quality of service (QoS) requirements and network energy consumption under highly time-varying and spatially non-uniform traffic loads. Existing AP ON/OFF mechanisms--typically based on worst-case dimensioning or greedy heuristics--explore the combinatorial activation space inadequately, leading to suboptimal energy-efficiency outcomes. This paper introduces two evolutionary AP-selection strategies tailored to CF-mMIMO networks. The first, a constrained genetic algorithm (CGA), identifies the near-optimal subset of active APs for any fixed activation cardinality, while an outer search determines the globally optimal operating point. The second, a Pareto-driven genetic algorithm (PDGA), jointly optimizes spectral and energy efficiency by evolving a Pareto front over all feasible activation patterns. A detailed computational-complexity analysis is provided for both techniques. Simulations conducted under realistic spatially inhomogeneous traffic and considering both conjugate beamforming (CB) and minimum mean square error (MMSE) processing confirm consistent performance gains. The proposed methods consistently outperform state-of-the-art greedy benchmarks, delivering noticeable improvements in energy efficiency for both CB and MMSE schemes, while simultaneously enhancing the energy-spectral efficiency tradeoff, which is typically difficult to improve without incurring penalties elsewhere. These results highlight the strong potential of evolutionary optimization as a powerful and reliable approach for energy-efficient CF-mMIMO deployments.

[2] arXiv:2607.26303 [pdf, other]
Title: SWIPT-Enhanced Cell-Free Massive MIMO Networks
Guillem Femenias, Jan García-Morales, Felip Riera-Palou
Comments: Author Accepted Manuscript. Published in IEEE Transactions on Communications. The final published version is available at: this https URL. Copyright © 2021 IEEE
Journal-ref: G. Femenias, J. Garc\'ia-Morales and F. Riera-Palou, "SWIPT-Enhanced Cell-Free Massive MIMO Networks," in IEEE Transactions on Communications, vol. 69, no. 8, pp. 5593-5607, Aug. 2021
Subjects: Signal Processing (eess.SP)

Simultaneous wireless information and power transfer (SWIPT) has been advocated as a highly promising technology to provide near-perpetual operation to low-powered wireless devices in Internet-of-Things (IoT)-based wireless networks. In this paper, a SWIPT-enhanced cell-free massive MIMO network is proposed. In such a network, a large set of spatially distributed access points (APs) interconnected via a central processing unit (CPU) can collaboratively serve a large number of both energy harvesting mobile stations (MSs) (requiring wireless energy transfer) and conventional MSs (not requiring wireless energy transfer) on the same time-frequency resources. We consider spatially correlated Rician fading channels and the use of different precoding schemes that are based on different channel estimators differing on the assumed knowledge of the line-of-sight component. Mathematically manageable expressions are derived for the harvested energy during the downlink (DL) energy harvesting phase and the achievable spectral and energy efficiencies during the uplink (UL) payload transmission phase. A coupled UL/DL optimization problem is formulated aiming at finding the power control coefficients that maximize the minimum of the weighted achievable UL signal-to-interference-plus-noise ratios (SINRs) of all MSs. Extensive numerical results are presented that serve to highlight the existing trade-offs among the achievable spectral and energy efficiencies, the harvested energy, the energy dedicated to UL pilot transmission or the system configuration.

[3] arXiv:2607.26325 [pdf, html, other]
Title: Joint Beamforming, Energy Management, and Trajectory Optimization for Figure-Eight Loitering in Solar-Powered HAPS-Enabled ISAC Systems
Xue Zhang, Bang Huang, Mohamed-Slim Alouini
Subjects: Signal Processing (eess.SP)

Solar-powered high-altitude platform stations (HAPSs) provide a promising platform for integrated sensing and communication (ISAC) owing to their wide-area coverage and long-endurance operation. This paper proposes a solar-powered HAPS-enabled ISAC framework for sustainable day-night operation, where a figure-eight loitering architecture is adopted to provide persistent ISAC services over geographically separated regions while harvesting solar energy. A unified communication-sensing-energy model is developed by jointly characterizing solar energy harvesting, battery dynamics, propulsion power consumption, communication transmission, and synthetic aperture radar (SAR) imaging. Based on this model, coupled optimization problems are formulated for daytime operation (DTO) and nighttime operation (NTO), where the battery state bridges the two operational phases through a long-term energy budget. The proposed framework jointly optimizes communication, sensing, mobility, and energy management to maximize daytime communication performance while minimizing nighttime propulsion energy consumption. Efficient iterative algorithms are developed to solve the resulting non-convex optimization problems. Simulation results verify the effectiveness of the proposed communication-sensing-energy co-design and demonstrate that the proposed framework effectively supports sustainable day-night ISAC operation.

[4] arXiv:2607.26421 [pdf, html, other]
Title: MVLA-GR: A Phase-Free Multipath-Based Geometry Reconstruction Method via Multi-View Likelihood Accumulation for ISAC
Bowei Xing, Yuxiang Zhang, Jianhua Zhang, Yifeng Xiong, Hongbo Xing, Li Yu, Guangyi Liu
Comments: 13 pages, 10 figures. Submitted to IEEE Transactions on Wireless Communications
Subjects: Signal Processing (eess.SP)

Integrated sensing and communication (ISAC) enables wireless systems to reuse communication signals for environmental sensing, where reconstructing the geometry of surrounding objects is a representative sensing task. However, many conventional methods rely on coherent processing and require accurate phase information, which is often hard to guarantee in practical communication systems, particularly at high carrier frequencies. To address this problem, this paper proposes a Multi-View Likelihood Accumulation Geometry Reconstruction (MVLA-GR) method based on channel impulse response (CIR) measurements, which uses only delay and power observations without requiring phase information. The method extracts dominant multipath components from each observation, and for each candidate spatial location, accumulates components across views whose propagation distances match the location as supporting evidence. A soft distance-matching kernel is introduced to tolerate range estimation errors and viewpoint-dependent scattering migration, and the received power of each component is used as a reliability weight. A joint thresholding strategy combining response magnitude and angular support continuity then converts the continuous support map into a binary geometry estimate. Ray-tracing simulations on canonical and complex targets, as well as real-world vehicle measurements at 36 GHz, demonstrate that MVLA-GR can effectively recover target geometry, providing a low-complexity phase-free solution for ISAC.

[5] arXiv:2607.26449 [pdf, html, other]
Title: Energy-Efficient Access-Point Sleep-Mode Techniques for Cell-Free mmWave Massive MIMO Networks With Non-Uniform Spatial Traffic Density
Jan García-Morales, Guillem Femenias, Felip Riera-Palou
Comments: Author Accepted Manuscript. Published in IEEE Access, vol. 8, pp. 137587-137605, 2020. The final published version is available at: this https URL
Journal-ref: J. Garc\'ia-Morales, G. Femenias and F. Riera-Palou, "Energy-Efficient Access-Point Sleep-Mode Techniques for Cell-Free mmWave Massive MIMO Networks With Non-Uniform Spatial Traffic Density," in IEEE Access, vol. 8, pp. 137587-137605, 2020
Subjects: Signal Processing (eess.SP)

Cell-free massive multiple-input multiple-output (MIMO) is a novel beyond 5G (B5G) and 6G paradigm that, through the use of a common central processing unit (CPU), coordinates a large number of distributed access points (APs) to coherently serve mobile stations (MSs) on the same time/frequency resource. By exploiting the characteristics of new less-congested millimeter wave (mmWave) frequency bands, these networks can improve the overall system spectral and energy efficiencies by using low-complexity hybrid precoders/decoders. For this purpose, the system must be correctly dimensioned to provide the required quality of service (QoS) to MSs under different traffic load conditions. However, only heavy traffic load conditions are usually taken into account when analysing these networks and, thus, many APs might be underutilized during low traffic load periods, leading to an inefficient use of resources and waste of energy. Aiming at the implementation of energy-efficient AP switch on/off strategies, several approaches have been proposed in the literature that only consider rather unrealistic uniform spatial traffic distribution in the whole coverage area. Unlike prior works, this paper proposes energy efficient AP sleep-mode techniques for cell-free mmWave massive MIMO networks that are able to capture the inhomogeneous nature of spatial traffic distribution in realistic wireless networks. The proposed framework considers, analyzes and compares different AP switch ON-OFF (ASO) strategies that, based on the use of goodness-of-fit (GoF) tests, are specifically designed to dynamically turn on/off APs to adapt to both the number and the statistical distribution of MSs in the network. Numerical results show that the use of properly designed GoF-based ASO strategies under a non-uniform spatial traffic distribution can serve to considerably improve the achievable energy efficiency.

[6] arXiv:2607.26501 [pdf, html, other]
Title: Calibrating the Digital Twin Channel: Statistics-Consistent Sim-to-Lab Adaptation for W-Band Industrial OFDM Links
Pulok Tarafder, Abigail O. Oyekola, Jasni Areepatta Mannil, Imtiaz Ahmed, Zoheb Hassan, Danda B. Rawat, Wenjie Che
Comments: Submitted for possible publication to IEEE. Paper currently under review. The contents of this paper may change at any time without notice
Subjects: Signal Processing (eess.SP)

Digital twins (DTs) can reduce over-the-air validation cost in industrial wireless networks, but their utility depends on the fidelity of the underlying channel twin (CT). At W-band, site-specific ray tracing captures deterministic propagation geometry, yet its channel frequency responses (CFRs) do not reproduce the small-scale impairments and capture-to-capture variability observed in laboratory orthogonal frequency-division multiplexing (OFDM) measurements above 90 GHz. This paper proposes Statistics-Consistent Sim-to-Lab Adaptation (SC-SLA), a calibration framework that improves the fidelity of a 95 GHz Sionna ray-traced CT toward that of the testbed by aligning the mean power delay profile (PDP), the distribution of root-mean-square delay spread ($\tau_{\mathrm{rms}}$), and per-subcarrier statistics at 50 MHz sampling bandwidth. SC-SLA uses a generative adversarial network (GAN)-inspired, cycle-consistent architecture with ResNet generators and batch-level channel-statistics losses on the PDP, sub-band PDP, $\tau_{\mathrm{rms}}$ moments and quantiles, and normalized mean-square error (NMSE) of the mean CFR-magnitude profile. The framework is non-adversarial and requires neither paired simulated/measured samples nor discriminators. On held-out 95 GHz data, SC-SLA reduces the $\tau_{\mathrm{rms}}$ Kolmogorov-Smirnov (KS) statistic from 0.86 to 0.050 relative to the impairment-augmented ray-traced input, and by 44% (from 0.089 to 0.050) relative to the strongest of four supervised baselines (FCNN, CNN1D, BiLSTM, and UNet1D). Without retraining, the same checkpoint also generalizes to 92-94 GHz carriers, where it achieves the lowest PDP and CFR-magnitude errors among all baselines while reducing the $\tau_{\mathrm{rms}}$-distribution mismatch relative to the uncalibrated twin.

[7] arXiv:2607.26519 [pdf, html, other]
Title: A Data-Driven Vibration Analysis Framework for Micro-Motor Fault Diagnosis and Quality Control
Xuan Chen, Xinjun Zuo, Yancheng Bi, Shunli Yu, Yunqi Cao
Comments: 6 pages, 4 figures, and 1 table. Accepted by the 37th Chinese Process Control Conference (CPCC 2026); to appear in the conference proceedings
Subjects: Signal Processing (eess.SP)

The reliability of the internal micro-motors is crucial for the performance and lifespan of electric toothbrushes. In this paper, a vibration-based fault detection method is proposed to identify micro-motor defects in electric toothbrushes. A dedicated signal acquisition device was designed and developed to capture the vibration signals of micro-motors using a high-precision accelerometer. To effectively characterize the micro-motor conditions, comprehensive features were extracted from the raw vibration data in both the time and frequency domains. A random forest (RF) algorithm was then employed to evaluate the importance of all extracted features. To better interpret the extracted features based on fault mechanisms, and to reduce dimensionality and computational overhead while avoiding overfitting, the top three features with the highest importance scores were selected to form the optimal feature subset. Finally, a support vector machine (SVM) model was utilized to classify the motor states based on the selected features. Experimental results demonstrate that the proposed method, combining RF-based feature selection and SVM classification, achieves outstanding diagnostic performance. Specifically, the model yields a balanced accuracy of 94.44%, a defect recall of 88.89%, a defect F1-score of 94.12%, a Matthews correlation coefficient of 93.74%, a geometric mean of 94.28%, and an area under the receiver operating characteristic curve of 100.00%. These robust metrics confirm that the proposed approach can accurately and efficiently detect micro-motor faults in electric toothbrushes, providing a practical and reliable solution for quality control and condition monitoring in manufacturing.

[8] arXiv:2607.26531 [pdf, html, other]
Title: Secure Relay Low-Altitude Networks via Hybrid Fixed-Position and Rotatable Antenna Arrays
Maolin Li, Qi Zhang, Riqing Chen, Wei Gao, Feng Shu, Liang Yang, Cunhua Pan
Subjects: Signal Processing (eess.SP)

In this paper, a relay network with hybrid fixed-position and rotatable antenna arrays is proposed. The deployment of rotatable arrays in conventional relay networks is considered to provide more secure communications for low-altitude economy applications. Specifically, both the base station and the relay station are equipped with fixed-position antenna arrays and rotatable arrays to serve ground users and aerial users, respectively. To address the challenge of multi-user interference, a low-cost reconfigurable intelligent surface is exploited as a candidate path. Accordingly, under constraints on transmit power, user quality of service, rotatable range, and path selection, the objective is to maximize the worst-case secrecy rate (SR) through joint beamforming, power allocation, and rotatable antenna orientation design. First, the SR performance in the single-user scenario is investigated, and a step-by-step leakage-based scheme is proposed. Then, the general multi-user scenario is studied, and a Distributional Soft Actor-Critic with Three refinements (DSAC-T)-based learning scheme, which supports hybrid discrete and continuous actions, is proposed to maximize the worst-case SR. Simulation results validate the effectiveness of the proposed schemes. The proposed schemes achieve approximately a twofold improvement in SR performance compared to isotropic antennas. The proposed system achieves approximately 71.4\% power saving, 55\% antenna saving, and can serve more users.

[9] arXiv:2607.26547 [pdf, html, other]
Title: Stay or Switch: Online Conformal Bayesian Optimization Guided Fluid Antenna Configuration
Gangyong Zhu, Jia Yan, Shijian Gao
Comments: 6 pages, 5 figures, conference
Subjects: Signal Processing (eess.SP)

Fluid antenna systems (FAS) introduce additional spatial degrees of freedom to enable integrated sensing and communication (ISAC) in air-ground networks. However, conventional studies often overlook or simplify the physical overheads and switching costs of FAS. In practice, port switching incurs non-negligible time, during which communication and sensing may continue but with potentially degraded slot-level performance. This leads to two key challenges: (1) the characterization of a slot-level, cost-aware ISAC metric is difficult, and (2) the large port space and accompanying abrupt environmental variations demand more reliable online decision-making. To address these challenges, a cost-aware multi-objective FAS switching problem is formulated, jointly considering slot-level ISAC performance and switching energy. The online conformal Bayesian optimization (OCBO) algorithm is then proposed to learn the unknown gray-box ISAC objectives and calibrate surrogate uncertainty for robust stay-or-switch decisions. Simulation results demonstrate that the proposed cost-aware optimization framework achieves substantially improved long-term ISAC performance compared to existing baselines.

[10] arXiv:2607.26605 [pdf, html, other]
Title: Multi-User Localization via Active Sensing with Electromagnetically Reconfigurable Antennas
Ruizhi Zhang, Yuchen Zhang, Ying Zhang, Henk Wymeersch
Subjects: Signal Processing (eess.SP)

This paper investigates multi-user localization in uplink wireless systems assisted by electromagnetically reconfigurable antennas (ERAs). Unlike traditional localization schemes, we formulate an active sensing problem where a base station (BS) exploits historical pilot observations accumulated over previous sensing stages to adapt the shared ERA configuration and progressively refine position estimates. To capture both theoretical flexibility and practical hardware constraints, we establish a unified wideband geometric signal model accommodating two complementary ERA paradigms: a synthesis-based model utilizing spherical-harmonic basis functions, and a finite-state model based on measured radiation codebooks. Because analytically solving the resulting joint design problem is highly intractable due to the high-dimensional observation and the shared-aperture coupling among multiple users, we develop a learning-based active sensing framework. Specifically, pilot-matched wideband observations are compressed into compact user-wise features and sequentially accumulated by a long short-term memory (LSTM) module. These temporal features are then processed by a graph neural network (GNN) to capture multi-user shared-aperture coupling. Model-specific output heads generate either continuous synthesis coefficients or finite-state ERA selections, while a localization head produces stage-wise position estimates. Numerical results under a specific channel distribution show that the proposed ERA-assisted active sensing framework achieves progressive localization refinement across sensing stages and obtains better performance than conventional non-reconfigurable arrays and representative ablation baselines.

[11] arXiv:2607.26682 [pdf, html, other]
Title: An Informativeness-based Clustered Federated Learning Method for Reliable Traffic Prediction in Managed Wi-Fi Networks
Luca Barbieri, Gianluca Fontanesi, Lorenzo Galati Giordano, Alfonso Fernandez Duran, Thorsten Wild
Comments: submitted to IEEE for possible publication
Subjects: Signal Processing (eess.SP); Machine Learning (cs.LG)

Centrally-managed Wi-Fi solutions are increasingly leveraging Distributed Artificial Intelligence (AI) to predict key operational statistics of Access Points (APs) and proactively optimize network performance. In this context, Clustered Federated Learning (CFL) represents a fitting methodology, enabling the generation of multiple AI models that account for diverse statistical properties of the APs data distribution. However, identifying informative clusters for grouping APs models remains a significant challenge. In this paper, we address this problem by proposing a novel CFL tool integrating a two step clustering procedure. Initially, multiple clustering solutions are generated and filtered based on a minimum set of desired clustering criteria. Subsequently, if no solutions meet sufficient quality metrics, a global model is produced by aggregating all AP models. Otherwise, the final clustering solution is selected as the one that maximizes the informativeness (quantified via differential entropy) for the smallest cluster. Our results, focusing on a Wi-Fi traffic prediction problem, demonstrate that the developed CFL tool achieves the best predictive performance among all evaluated distributed strategies and the lowest communication and energy footprint among the clustered ones, exceeding the cost of single-model FL only in the regimes where it markedly improves accuracy.

[12] arXiv:2607.26738 [pdf, html, other]
Title: A Dual-Mode FM/AM Modulator Based on a Time-Varying Inverting Integrator
Azalía G. Gil, Alfonso T. Muriel-Barrado, Mario Pérez-Escribano, Carlos Molero, Antonio Alex-Amor
Subjects: Signal Processing (eess.SP); Applied Physics (physics.app-ph)

This paper presents the analysis, design, fabrication, and experimental validation of a dual-mode frequency/amplitude modulator based on a time-modulated varactor diode. By exploiting the varactor as a time-varying capacitor in combination with an operational amplifier configured as an inverting integrator and a passband filter, the proposed circuit generates frequency-modulated (FM) signals in an efficient manner. Amplitude-modulated (AM) signals can also be obtained with a simple modification. The implementation, realized in microstrip technology, leverages the unique properties of time-modulated electronic components, particularly their inherent frequency-mixing capability. Analytical expressions are derived to predict the characteristics of the generated waveforms, and their accuracy is verified through numerical simulations performed in Keysight ADS. A microstrip PCB prototype is then fabricated and experimentally characterized. The measured results show excellent agreement with both the theoretical predictions and the numerical simulations. The proposed approach demonstrates the potential of time-varying capacitors as an attractive alternative to conventional FM techniques for telecommunications and radar applications.

[13] arXiv:2607.26959 [pdf, html, other]
Title: Efficient Channel Prediction based on Gram-Square-Root Factorization using GMMs
Kathrin Klein, Amar Kasibovic, Michael Joham, Shachar Shayovitz, Wolfgang Utschick
Comments: This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible
Subjects: Signal Processing (eess.SP)

Accurate channel state information (CSI) is critical for downlink (DL)-multi-user (MU)-multiple-input multiple-output (MIMO) systems, where feedback delays and mobility can degrade precoding performance. To ensure reliable beamforming and interference mitigation, CSI prediction is required. In practical systems, full CSI feedback is often infeasible due to signaling overhead, so transmitters rely on partial CSI reported by the receivers. In this work, we propose a Gaussian mixture model (GMM)-based prediction framework for MIMO-orthogonal frequency-division multiplexing (OFDM) channels under partial feedback using Gram-square-root factorization. To address the high dimensionality, we introduce an efficient parameter reduction technique that exploits structured covariance matrices, significantly lowering complexity without noticeable performance degradation. This reduction is based on the Gram-square-root factorization and remains of interest even when full CSI is available. Simulation results demonstrate that GMMs achieve the highest prediction accuracy and correctly capture the underlying channel subspaces, which is essential for effective MU-precoding. The proposed method outperforms classical baselines such as zero-order hold (ZOH), first-order hold (FOH), and linear minimum mean squared error (LMMSE) predictors, and an advanced neural network (NN)-based predictor. Notably, the parameter-reduced partial CSI GMM achieves performance comparable to that of full CSI prediction, highlighting its ability to efficiently model the channel structure under limited feedback.

[14] arXiv:2607.26994 [pdf, html, other]
Title: Probabilistic Denoising-Enhanced ISAC for Stochastic Cluttered Mobile Environments
Nghia Thinh Nguyen, Tri Nhu Do
Subjects: Signal Processing (eess.SP)

In this paper, we propose Probabilistic Denoising ISAC (PDISAC), a framework built on a multi-bit slot-partitioned ISAC waveform: by partitioning each maximal-length sequence into alternating pilot and data slots, we embed multiple bits per sequence through symbol-level spreading, multiplying the data rate while every chip retains the deterministic radar code. The added throughput, however, injects data-dependent, non-white sidelobes into the range-Doppler (RD) heatmap that degrade matched-filter (MF) sensing. Rather than modifying the MF receiver, we develop RDPDNet, a lightweight probabilistic denoising network inserted between RD-map formation and constant-false-alarm-rate detection; training it with an adversarial frequency-mixup mechanism, we suppress the data-induced sidelobes and thermal noise without knowledge of the embedded symbols. We further characterize the statistics of the geometry-determined channel. The fundamental performance limits of the design are then analyzed through an analytical lower bound, a semi-analytical bit error rate (BER), and an average capacity that tie the slot allocation and sequence length to the sensing-communication trade-off. Through analytical and numerical results over a realistic urban geometry, we show that RDPDNet absorbs most of the data-embedding sensing penalty and markedly lowers the RMSE at low SNR, while the conventional data-free chain attains the bias-adjusted benchmark at high SNR. Moreover, increasing the slot allocation raises the data rate at the expense of a higher BER, exposing a tunable sensing--communication trade-off.

[15] arXiv:2607.27151 [pdf, html, other]
Title: Two-Filter Adaptive Gaussian Mixture Smoothing for Nonlinear Systems
Benjamin Schneiderheinze, Andrea De Vittori, Keith A. LeGrand, Jill Bruer
Subjects: Signal Processing (eess.SP)

Space object tracking poses challenging estimation problems due to the significantly non-Gaussian distributions that can arise, particularly under highly nonlinear dynamics or during periods of measurement unavailability. Adaptive Gaussian mixture filters can dynamically adjust their mixture resolution to systematically approximate these non-Gaussian distributions, but challenging estimation problems can still produce highly uncertain or inaccurate estimates, especially during prolonged observation gaps. Smoothing algorithms can significantly improve filtered estimates by incorporating future measurement information for applications where immediacy is not required, but smoothing in nonlinear, non-Gaussian settings poses additional theoretical and computational challenges. This work develops a new recursive Bayesian smoothing algorithm for nonlinear systems that refines Gaussian mixture posteriors produced by a forward adaptive Gaussian mixture filter. A two-filter smoothing approach approximates the future measurement information by an information-form Gaussian mixture in the state variable. Techniques from nonlinear Gaussian mixture filtering including splitting, merging, and recursive measurement updating are also incorporated to improve the accuracy and computational efficiency of this approximation. The proposed smoother's estimation capabilities are demonstrated on space object tracking problems for Molniya and Earth-Moon halo orbits and shown to significantly reduce estimation error and uncertainty compared to the forward filter.

Cross submissions (showing 9 of 9 entries)

[16] arXiv:2607.26401 (cross-list from cs.HC) [pdf, html, other]
Title: Sensor-Placement-Agnostic Sonomyography: Toward Continuous High-Dimensional Control by Users with Tetraplegia
Gavin Sueltz, Vikram Athithan, Emma Ferran, Maria Herrera, Carson J. Wynn, Laura A. Hallock
Comments: Vikram Athithan, Emma Ferran, Maria Herrera, and Carson J. Wynn contributed equally to this work. Copyright 2026 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media
Subjects: Human-Computer Interaction (cs.HC); Robotics (cs.RO); Signal Processing (eess.SP)

Sonomyography (SMG) enables continuous device control via ultrasound-measured muscle deformation signals, but existing SMG interfaces generally require substantial user- and sensor-location-specific training data and provide only one proportional signal or task-specific classification. We present a real-time, sensor-placement-agnostic SMG control system based on sparse optical flow tracking that enables continuous 1-DOF control after minimal calibration (3 pose definitions). We also present a preliminary expansion of this method that augments this algorithm with a short computer-aided calibration to enable 2-DOF control.
We evaluate both 1- and 2-DOF systems' performance for a preliminary cohort of 3 cervical spinal cord injury survivors and 6 uninjured individuals across 6 sensor placements spanning the arm, neck, and upper torso. As assessed by a cursor trajectory tracking task, all participants achieved continuous 1-DOF control at all tested sensor locations (even those that relied on passive tissue motions), with all participants achieving <5.5% tracking error using at least one placement (and many <4% across many). All participants were also able to modulate 2D cursor position via the 2-DOF system, with varying levels of control authority, and several were able to complete a 2D drawing task, constituting the first (to our knowledge) demonstration of location-agnostic multi-DOF continuous SMG-based control. These results highlight the promise of SMG to enable rapidly calibratable, high-dimensional, sensor-placement-agnostic device control by users with tetraplegia, and also illuminate key challenges in both signal processing and practical system deployment. To enable further development by scientific and user communities, developed algorithms have been open-sourced as part of the OpenMyoControl project on SimTK (this http URL).

[17] arXiv:2607.26481 (cross-list from cs.LG) [pdf, html, other]
Title: Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations
Seunghun Yu, Meiyi Zhu, Petar Popovski, Joonhyuk Kang, Osvaldo Simeone
Comments: 34 pages, 6 figures
Subjects: Machine Learning (cs.LG); Signal Processing (eess.SP)

Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monitoring of telecommunication networks, robotic platforms, security infrastructure, and multi-agent systems. In safety- and mission-critical deployments, such decisions must be accompanied by statistical reliability guarantees rather than by point estimates alone. Conformal changepoint localization (CONCH) and conformal root cause analysis (CROC) meet this need by returning confidence sets that contain the true changepoint, or the true root-cause stream, with a user-specified probability, without parametric assumptions on the data-generating process. In practice, however, observations are frequently corrupted, e.g., by outliers, sensor faults, or adversarial perturbations. While the finite-sample coverage of these procedures is preserved under contamination, the resulting confidence sets can become uninformatively large. Adopting a Huber-type contamination model, this paper proposes weighted CONCH (W-CONCH) and weighted CROC (W-CROC), which downweight observations that are likely to be corrupted with the goal of reducing confidence set size when data may be corrupted. The weighting mechanism, derived from a formal bound on the unknown corrupted data densities, leverages pre-existing second-order classifier-based uncertainty signals, such as those produced by evidential deep learning or Bayesian learning. W-CONCH and W-CROC are further generalized by introducing a meta-learning procedure for the weights that optimizes a differentiable surrogate of the confidence set size. Experiments on image-based and real-world changepoint and root-cause benchmarks show that uncertainty-based weighting substantially reduces confidence set size while maintaining the target coverage.

[18] arXiv:2607.26575 (cross-list from eess.AS) [pdf, html, other]
Title: Unfolded Recursive Expectation-Maximization Neural Network For Speaker Tracking
Rina Veler, Sharon Gannot
Comments: proceedings of IWAENC 2026
Subjects: Audio and Speech Processing (eess.AS); Signal Processing (eess.SP)

We propose a deep unfolded REM network for robust tracking of a single moving speaker in mild reverberant environments. Unlike classical REM algorithms, which rely on fixed-step-size decay schedules, the proposed architecture learns an adaptive update policy by unfolding the iterative procedure into differentiable layers. We introduce a Step Size Network that leverages FiLM and PE to dynamically adjust the recursion weights based on temporal context and convergence state. Experimental results for tracking a single speaker under reverberant conditions demonstrate that the proposed unfolded network outperforms the classical CREM baseline, which employs a spatial grid search to map the estimated centroids to physical positions. In the single-speaker tracking task, the proposed method achieves a lower RMSE than the CREM baseline, highlighting its potential for dynamic acoustic scenarios.

[19] arXiv:2607.26658 (cross-list from cs.NI) [pdf, html, other]
Title: Active Movable-Element RIS Assisted Vehicular Semantic Communications: Modeling and Optimization
Maoxin Ji, Qiong Wu, Jingbo Zhang, Pingyi Fan, Kezhi Wang, Wen Chen, Guoqiang Mao, Khaled B. Letaief
Comments: This paper has been accepted by IEEE TWC
Subjects: Networking and Internet Architecture (cs.NI); Signal Processing (eess.SP)

Severe signal blockage and fast-varying channels in vehicular environments pose critical challenges to reliable semantic communication. To address these, this paper proposes a novel Row-Movable Active Reconfigurable Intelligent Surface (RM-A-RIS) assisted vehicular semantic communication system. This architecture uniquely combines active signal amplification with element mobility to compensate for multiplicative fading and reconstruct channel geometry, thereby enhancing spatial diversity. We formulate a joint optimization problem to maximize Semantic Spectral Efficiency (SSE) by coordinating RIS element positions, active reflection coefficients, and semantic symbol length. An efficient Alternating Optimization (AO) algorithm is developed to tackle the coupled non-convexity. Simulation results demonstrate that the proposed scheme substantially outperforms existing benchmarks, achieving up to 132.9%, 9.2%, and 35.2% improvements in Sum-Semantic Spectral Efficiency (Sum-SSE) compared to the passive RIS, fixed-position active RIS, and QPSO baselines, respectively.

[20] arXiv:2607.26746 (cross-list from eess.IV) [pdf, html, other]
Title: An Attention-Based Framework for Alzheimers Disease Classification Using Resting-State fMRI
Harshiddhi Pathak, Gowtham Reddy N, Mrinal Acharya, Manjunatha Mahadevappa
Comments: 7 pages, 5 figures, 2 tables, accepted at 48th Annual International Conference of the IEEE Engineering in Medicine and Biology Society (IEEE EMBC 2026)
Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Human-Computer Interaction (cs.HC); Machine Learning (cs.LG); Signal Processing (eess.SP)

Accurate identification of Alzheimers disease (AD) using resting-state functional magnetic resonance imaging (rs-fMRI) remains challenging due to the high dimensionality, noise, and complex inter-regional dependencies inherent in functional brain connectivity, which limit the effectiveness of traditional approaches based on handcrafted connectivity features or conventional machine learning models. In this work, we present an attention-based deep learning framework for Alzheimers disease classification that operates directly on rs-fMRI functional connectivity matrices by treating brain regions as tokens and employing a Transformer-inspired self-attention mechanism to model long-range and global functional dependencies across distributed brain networks. The proposed framework learns discriminative functional representations without reliance on manual feature engineering and is evaluated on a longitudinal cohort from the Alzheimers Disease Neuroimaging Initiative (ADNI) comprising cognitively normal and Alzheimers disease subjects with multiple visits. A subject-wise evaluation protocol is adopted to prevent information leakage across visits, and class-weighted optimization is incorporated to address mild class imbalance. Experimental results for binary AD versus cognitively normal classification demonstrate that the proposed attention- based rs-fMRI model achieves an accuracy of 88.95% and a ROC-AUC of 0.90, along with a favorable precision-recall balance, highlighting the effectiveness of self-attention-driven functional connectivity modeling as a robust and interpretable approach for Alzheimers disease detection using resting-state fMRI.

[21] arXiv:2607.26751 (cross-list from cs.CL) [pdf, html, other]
Title: Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text
Lucas Zamora Vera, Jose A. Gonzalez-Lopez
Comments: 6 pages, 1 figure, 6 tables, submitted to IberSPEECH 2026
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Signal Processing (eess.SP)

State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain underexplored: whether selective state-space models (Mamba) improve on recurrent decoders, and how the output target (phonetic vs.\ character) interacts with that choice. On the public Brain-to-Text '25 benchmark, we study a controlled 2x2 grid (GRU vs.\ hybrid Mamba decoder; phonetic vs.\ character targets) trained with a CTC objective under one reproducible protocol. The recurrent baseline remains strongest: the best phonetic GRU reaches 12.62\% PER and 21.19\% WER, while the best textual GRU after LM rescoring reaches 13.39\% CER and 26.28\% WER. The Mamba hybrid is competitive but does not surpass it. Ablations isolate architectural contributions, and error analysis shows representation-dependent failures: articulatory-like phoneme confusions vs.\ lexical and word-boundary errors.

[22] arXiv:2607.26815 (cross-list from cond-mat.mtrl-sci) [pdf, other]
Title: End-to-End Modeling of a Volatile TiO2 Memristor for Neuromorphic Circuit Simulation
Lukas Endres, Hannes Töpfer, Michaela Blum, Hauke Honig, Peter Schaaf
Comments: 10 pages, 13 figures, 14 references
Subjects: Materials Science (cond-mat.mtrl-sci); Signal Processing (eess.SP)

Memristors are promising devices for applications such as non-volatile memory, neuromorphic computing, logic circuits, and analog signal processing. The development of such systems requires accurate simulations based on models that reproduce the electrical behavior of real devices under both continuous and pulsed excitation. This work presents the development of a simulation environment for a volatile TiO2-based memristor. Experimental measurement data are analyzed to verify the memristive behavior of the device and to identify a suitable model. The model parameters are then optimized to match the measured characteristics. The resulting model is implemented in SPICE and validated by comparing simulation results with measurement data. The comparison shows a good agreement between simulation and experiment, demonstrating that the developed model is suitable for reproducing the electrical behavior of the investigated memristor and can be applied in circuit-level simulations, as demonstrated by a leaky integrate-and-fire neuron.

[23] arXiv:2607.26980 (cross-list from cs.RO) [pdf, html, other]
Title: Dense Soft Weighting for Radar Ego-Velocity Estimation
Atar Babgei, Chenyu Zhao, Michael Breza, Julie A. McCann
Comments: Submitted to
Subjects: Robotics (cs.RO); Signal Processing (eess.SP)

Sensing ego-velocity estimation is fundamental to state estimation in visually degraded environments, where camera- and LiDAR-based pipelines can become unreliable. Millimetre-wave radar is well suited to these conditions because it provides direct Doppler velocity sensing and remains robust to poor illumination, textureless scenes, and airborne particulates. However, conventional radar ego-velocity pipelines typically apply constant false alarm rate (CFAR) thresholding to convert dense radar spectra into sparse point clouds, prematurely discarding sub-threshold returns that may still retain useful Doppler motion cues. We present Dense Soft Weighting, an analytic radar front-end that maps every range-Doppler cell to a continuous confidence metric rather than enforcing a binary detection threshold. Ego-velocity is then estimated using a deterministic robust weighted least-squares formulation, while the same weighted measurements provide a closed-form, measurement-derived velocity covariance for integration with a shared inertial back-end. The method requires no platform-specific training data or learning-based uncertainty model, supporting transfer across single-chip radar configurations. Across two public datasets and one self-collected dataset, Dense Soft Weighting reduces mean absolute pose error by 31-45% relative to the strongest CFAR point-cloud baseline under an identical inertial back-end, while running in real time on embedded hardware.

[24] arXiv:2607.27076 (cross-list from cs.LG) [pdf, html, other]
Title: Single-Beat Cuffless Blood Pressure Estimation Using Ear-PPG and ECG with a Lightweight Hybrid Learning Framework
Kindeep K. Dhatt, Tengyue Wu, Hanbang Hua, Yayun Du
Comments: 7 pages, 5 figures
Subjects: Machine Learning (cs.LG); Signal Processing (eess.SP); Systems and Control (eess.SY)

Continuous cuffless blood pressure (BP) monitoring remains challenging due to motion artifacts, physiological variability, and the limited robustness of conventional pulse transit time (PTT) models under dynamic conditions. Many prior approaches rely on multi-second windows to stabilize estimation, an assumption that is frequently violated during real-world monitoring with intermittent signal corruption. Here, we show that discriminative BP-related information is preserved at the single-beat level and present a lightweight multi-modal wearable framework for continuous BP estimation. The system integrates synchronized chest electrocardiography (ECG) and ear-clip reflectance photoplethysmography, each co-located with a 6-axis inertial measurement unit to provide motion context. We introduce a hybrid learning architecture in which a one-dimensional convolutional neural network extracts a 64-dimensional embedding from individual PPG beats and fuses it with 30 physiology-grounded features, including PTT statistics and heart rate variability, followed by LightGBM regression. The method was evaluated using a multi-phase stress protocol ($n=10$) and the PulseDB public dataset with subject-disjoint validation. Across 30 independent runs, the model achieved mean absolute errors of $4.02 \pm 0.21$~mmHg for systolic BP and $1.79 \pm 0.05$~mmHg for diastolic BP, corresponding to a 28.2\% reduction in combined MAE relative to baseline models. By enabling beat-wise estimation without long temporal context, this framework supports computationally efficient cuffless BP monitoring suitable for wearable deployment under practical resource constraints. The source code for this work is available at this https URL.

Replacement submissions (showing 8 of 8 entries)

[25] arXiv:2504.03772 (replaced) [pdf, html, other]
Title: Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare
Anton Lambrecht, Stijn Luchie, Jaron Fontaine, Ben Van Herbruggen, Adnan Shahid, Eli De Poorter
Journal-ref: IEEE SENSORS JOURNAL, VOL. 26, NO. 14, 15 JULY 2026
Subjects: Signal Processing (eess.SP); Machine Learning (cs.LG)

Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing these ailments. To this end, a low-cost commercial off-the-shelf (COTS), IEEE 802.15.4z standard compliant impulse-radio ultra-wideband (IR-UWB) radar system is used to estimate human respiration rates. We propose a convolutional neural network (CNN) specifically adapted to predict breathing rates from ultra-wideband (UWB) channel impulse response (CIR) data, and compare its performance with both other rule-based algorithms and model-based solutions. The study uses a diverse dataset, incorporating various real-life environments to evaluate system robustness. To facilitate future research, this dataset will be released as open source. Results show that the CNN achieves a mean absolute error (MAE) of 1.73 breaths per minute (BPM) in unseen situations, significantly outperforming rule-based methods (3.40 BPM). By incorporating calibration data from other individuals in the unseen situations, the error is further reduced to 0.84 BPM. In addition, this work evaluates the feasibility of running the pipeline on a low-cost embedded device. Applying 8-bit quantization to both the weights and input/output tensors, reduces memory requirements by 67% and inference time by 62% with only a 3% increase in MAE. As a result, we show it is feasible to deploy the algorithm on an nRF52840 system-on-chip (SoC) requiring only 46 KB of memory and operating with an inference time of only 199 ms. Once deployed, an analytical energy model estimates that the system, while continuously monitoring the room, can operate for up to 268.2 days without recharging when powered by a 20 000 mAh battery pack. For breathing monitoring in bed, the sampling rate can be lowered, extending battery life to 313.8 days, making the solution highly efficient for real-world, low-cost deployments.

[26] arXiv:2510.11214 (replaced) [pdf, other]
Title: CSI Prediction Using Diffusion Models
Mehdi Sattari, Javad Aliakbari, Alexandre Graell i Amat, Tommy Svensson
Comments: Accepted, IEEE Transactions on Wireless Communications
Subjects: Signal Processing (eess.SP); Information Theory (cs.IT)

Acquiring accurate channel state information (CSI) is critical for reliable and efficient wireless communication, but challenges such as high pilot overhead and channel aging hinder timely and accurate CSI acquisition. CSI prediction, which forecasts future CSI from historical observations, offers a promising solution. Recent deep learning approaches, including recurrent neural networks and Transformers, have achieved notable success but typically learn deterministic mappings, limiting their ability to capture the stochastic and multimodal nature of wireless channels. In this paper, we introduce a novel probabilistic framework for CSI prediction based on diffusion models, offering a flexible design that supports integration of diverse prediction schemes. We decompose the CSI prediction task into two components: a temporal encoder, which extracts channel dynamics, and a diffusion-based generator, which produces future CSI samples. We investigate two inference schemes-autoregressive and sequence-to-sequence- and explore multiple diffusion backbones, including U-Net and Transformer-based architectures. Furthermore, we examine a diffusion-based approach without an explicit temporal encoder and utilize the DDIM scheduling to reduce model complexity. Extensive simulations demonstrate that our diffusion-based models significantly outperform state-of-the-art baselines.

[27] arXiv:2602.08163 (replaced) [pdf, html, other]
Title: AFDM: Evolving OFDM Towards 6G+
Hyeon Seok Rou, Vincent Savaux, Zeping Sui, Giuseppe Thadeu Freitas de Abreu, Zilong Liu
Comments: Submitted to IEEE Journal
Subjects: Signal Processing (eess.SP)

As sixth generation (6G) standardization accelerates, there is growing consensus in favor of evolutionary waveforms that add new capabilities while preserving compatibility with the orthogonal frequency division multiplexing (OFDM) core of 4G and 5G. This article positions affine frequency division multiplexing (AFDM) as such a candidate, providing structural robustness for high-mobility communications and integrated sensing and communication (ISAC) over doubly dispersive channels while remaining backward-compatible with the legacy OFDM air interface. We first develop a generalized fractional-delay-fractional-Doppler (FDFD) channel model that accounts for practical pulse-shaping filters and the resulting inter-sample coupling. Building on this model, we show that the AFDM transceiver reuses nearly the entire OFDM chain, adding only lightweight digital pre- and post-processing. We then analyze the impact of hardware impairments such as phase noise and carrier frequency offset, and examine the advanced functionalities enabled by the chirp-parameter domain, including index modulation and physical-layer security. Assessing reusability across the radio-frequency, physical, and higher layers, we conclude that AFDM offers an efficient path toward high-fidelity later versions of 6G and beyond (6G+) communications.

[28] arXiv:2604.05175 (replaced) [pdf, html, other]
Title: Graph Signal Diffusion Models for Wireless Resource Allocation
Yigit Berkay Uslu, Samar Hadou, Shirin Saeedi Bidokhti, Alejandro Ribeiro
Comments: Accepted for presentation at 2026 IEEE SPAWC (Signal Processing Advances in Wireless Communications)
Subjects: Signal Processing (eess.SP); Information Theory (cs.IT); Machine Learning (cs.LG)

We consider constrained ergodic resource optimization in wireless networks with graph-structured interference. We train a diffusion model policy to match expert conditional distributions over resource allocations. By leveraging a primal-dual (expert) algorithm, we generate primal iterates that serve as draws from the corresponding expert conditionals for each training network instance. We view the allocations as stochastic graph signals supported on known channel state graphs. We implement the diffusion model architecture as a U-Net hierarchy of graph neural network (GNN) blocks, conditioned on the channel states and additional node states. At inference, the learned generative model amortizes the iterative expert policy by directly sampling allocation vectors from the near-optimal conditional distributions. In a power-control case study, we show that time-sharing the generated power allocations achieves near-optimal ergodic sum-rate utility and near-feasible ergodic minimum-rates, with strong generalization and transferability across network states.

[29] arXiv:2605.23682 (replaced) [pdf, html, other]
Title: Tri-Domain Multiuser MIMO Precoding Optimization and Channel Estimation with Spatial-EM Reconfigurable Antenna
Yining Li, Ziwei Wan, Zhen Gao, Keke Ying, Lipeng Zhu, Rui Zhang
Subjects: Signal Processing (eess.SP)

In this paper, we propose a tri-domain reconfigurable multiuser multiple-input multiple-output (MIMO) communication system that integrates the electromagnetic (EM) reconfigurable antenna (EMRA) with the spatially movable antenna (SMA), termed the spatial-EM reconfigurable antenna (SEMRA). The proposed system offers EM, spatial, and digital domain degrees of freedom (DoFs) for joint channel reconfiguration, yet introduces new challenges in channel estimation (CE) and precoding optimization. Specifically, for multiuser orthogonal frequency division multiplexing (OFDM) downlink, the precoding design is formulated as a tri-domain optimization problem over antenna positions, EM-domain radiation-pattern weights, and digital precoders. We first develop a zero-forcing (ZF)-based baseline algorithm to decouple the design of spatial reconfiguration, and then propose a weighted minimum mean square error (WMMSE)-based tri-domain joint optimization algorithm for further improving the spectral efficiency (SE). Furthermore, we propose a low-overhead movement-aided channel estimation scheme in which coordinated antenna repositioning across pilot slots synthesizes a denser virtual array, enabling more accurate angle-of-departure (AoD) estimation and EM-domain channel state information (eCSI) reconstruction under the same per-user pilot overhead as the EMRA baseline. The resulting parametric representation enables eCSI assembly at desired antenna positions without additional pilots. Simulation results show that the proposed CE scheme improves eCSI estimation accuracy and the proposed SEMRA achieves higher SE than the EMRA baseline under the same pilot overhead.

[30] arXiv:2607.21838 (replaced) [pdf, html, other]
Title: Noncoherent Detection and Interference Nulling for Terrestrial-Satellite Downlink Coexistence in the Upper Mid-Band
Shizhen Jia, C. Nicolas Barati, Marco Mezzavilla, Sundeep Rangan
Comments: Accepted by IEEE SPAWC 2026
Subjects: Signal Processing (eess.SP); Systems and Control (eess.SY)

Terrestrial--satellite coexistence in the upper mid-band is challenging when a terrestrial base station has limited prior information about non-terrestrial receivers and their uplink transmissions. This paper studies noncoherent victim sensing and interference nulling, where uplink sensing snapshots are observed while the transmitted waveform is treated as unknown. For a single victim, we show that the generalized likelihood-ratio test reduces to a principal-eigenvector estimator of the sample covariance. For multiple victims, we combine MDL-based model-order selection with MUSIC to recover anonymous direction and power information that is sufficient for beam design without pilot knowledge or user identities. These estimates are then used in a nulling beamformer that preserves the intended terrestrial link while reducing leakage toward detected victims. We further analyze the single-victim estimator in the large-matrix regime and show that estimation accuracy improves with sensing SNR, which reveals the observed interplay between path loss and estimation quality. Site-specific ray-tracing results show significant reduction of NTN INR with only modest degradation of TN SINR.

[31] arXiv:2607.25234 (replaced) [pdf, html, other]
Title: WHTMix: Efficient Stereo Depth Estimation via Walsh-Hadamard Token Mixing
Prathyush Sajith, Emadeldeen Hamdan, Ahmet Enis Cetin
Subjects: Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV); Signal Processing (eess.SP)

Stereo depth estimation for driving, robotics and augmented reality must run at high resolution under tight latency budgets, yet in transformer-based matchers the global self-attention that aggregates scene context grows quadratically with the number of pixels and comes to dominate runtime. We show that the joint self-attention stage of a stereo transformer, whose role is to spread context across both views, can be replaced by a data-independent Walsh-Hadamard token mixer that mixes tokens globally in the transform domain at log-linear cost, while the data-dependent cross-attention that performs left-right correspondence is retained. On synthetic driving data the mixer matches the attention baseline in end-point error while reducing model compute by a factor of 2.46 and single-image inference latency by a factor of 2.65. A complexity analysis shows the benefit is governed by the ratio of sequence length to channel width, which explains why high-resolution stereo matching is a particularly favorable setting and why classification transformers are not; we confirm this token-to-channel scaling on non-stereo long-sequence benchmarks. Furthermore, we introduce a hybrid log-disparity loss function designed to up-weight small-disparity pixels corresponding to long-range objects. This approach reduces the error on distant objects without incurring any additional computational overhead.

[32] arXiv:2607.25971 (replaced) [pdf, html, other]
Title: SplatStream: Fine Granular Scalable Gaussian Splatting for Adaptive 3D Scene Streaming
Muhammad Talha, William Gordon, Sajid Umair, Zhu Li, Anique Akhtar, Joel Jung
Comments: Accepted in Asilomar Conference on Signals, Systems, and Computers 2026
Subjects: Image and Video Processing (eess.IV); Signal Processing (eess.SP)

Dynamic 3D Gaussian Splatting (GS) enables high quality real-time rendering for immersive media, but its large representation size and frame-wise redundancy create significant challenges for adaptive streaming. This paper presents SplatStream, a fine granular scalable Gaussian splatting framework for dynamic 3D scene delivery. The proposed method decompose the GS scenes into quality and resolution layers, and introduces inter-layer predictive coding to achieve scalability. For temporal direction, B-frames are introduced to have temporal quality scalability. A lightweight cross-layer transformer based predictor is utilized for both cross layer and temporal predictions. In addition, a volume-opacity based importance measure is used for fine-grained Gaussian packetization, allowing visually important primitives to be transmitted earlier for progressive refinement. Finally, the scalable GS bitstream is mapped to an MPEG-DASH compatible sub-representation structure, enabling fine granular adaptive, low-latency delivery of dynamic Gaussian splatting content under bandwidth-varying conditions.

Total of 32 entries
Showing up to 2000 entries per page: fewer | more | all
We gratefully acknowledge support from our major funders, member institutions, , and all contributors.
About · Help · Contact · Subscribe · Copyright · Privacy · Accessibility · Operational Status (opens in new tab)
Major funding support from
Simons Foundation Simons Foundation International Schmidt Sciences