Signal Processing

2026-08-19 | | Total: 21

#1 Electromagnetic World Model for 6G: A Unified Framework for Joint Environment Reconstruction and Channel Prediction [PDF] [Copy] [Kimi] [REL]

Authors: Yizhu Zhao, Li Yu, Jianhua Zhang, Yuxiang Zhang, Zhen Zhang, Guangyi Liu

The integration of sensing, communication, and intelligence is becoming a key enabler for sixth generation (6G) wireless systems, where intelligent terminals are expected to simultaneously support efficient link establishment and reliable environmental sensing. However, existing studies mainly exploit sensing information or communication information to address a single task, such as channel prediction or environment reconstruction. Motivated by the shared dependence of optical and radio-frequency signals on the surrounding environment, we propose the electromagnetic world model (EMWM), the first unified framework for joint environment reconstruction and channel prediction. EMWM learns a common electromagnetic representation with the potential to provide a modeling foundation for 6G tasks. Specifically, partial channel state information (CSI) and multi-view red-green-blue (RGB) images are encoded into CSI and visual tokens and jointly processed by a hierarchical world-model backbone with local and global aggregation. Based on the learned representation, a mixture-of-experts (MoE)-based CSI prediction head reconstructs the complete CSI, while a depth prediction head estimates multi-view depth maps that are further converted into three-dimensional (3D) point clouds. Moreover, a large-scale multi-modal dataset is constructed based on a campus digital twin. Experimental results show that EMWM outperforms conventional neural network and large language model (LLM) baselines in both CSI prediction and environment reconstruction, achieving a squared generalized cosine similarity (SGCS) of 0.9699 for CSI prediction while demonstrating robustness across different signal-to-noise ratio (SNR) conditions and zero-shot generalization at 28 GHz.

Subject: Signal Processing

Publish: 2026-08-18 13:32:17 UTC


#2 M-QAM MIMO Maximum-Likelihood Detection with QAOA: ML-Rate Offline Angle Design and Correlated Infinite-Size Spin-Glass Models [PDF] [Copy] [Kimi] [REL]

Author: Burhan Gülbahar

The quantum approximate optimization algorithm (QAOA) targets NP-hard maximum-likelihood (ML) detection in multiple-input multiple-output (MIMO) systems. Existing $M$-ary quadrature amplitude modulation (M-QAM) detectors design angles by expected Ising energy: online per instance, warm-started, or ramped, while train-once designs remain B/QPSK-only or block-local, leaving M-QAM without a size-scalable benchmark. Their infinite-size spin-glass theory assumes independent disorder, matching the retained covariances at B/QPSK but not M-QAM's correlated couplings and fields. We develop a correlated infinite-size multi-species spin-glass framework whose covariance-matched evaluators make that energy an offline objective with a size-scalable benchmark. In addition, the ML rate, the exponential rate of sampling the ML string, is for the first time exploited for QAOA angle design in MIMO detection. The energy evaluator is $q$-free at $O(p\,4^p)$ cost while the ML rate transfers angles from a fixed $q_{\rm ref}$-qubit reference. Tests reach 4096-QAM, 128 antennas, $p=30$ and per-symbol SNR 0-45 dB. In simulations, ML rates fall as a power law $r_0\,p^{-α}$, with larger exponents for the sampling design, which tracks exact ML at $5\times5$ 16-QAM (0-20 dB) and $3\times3$ 64-QAM (8-28 dB) while its bit-error rate (BER) advantage widens with SNR to two orders of magnitude. The approach points toward near-optimum decoding on deeper noiseless fault-tolerant quantum (FTQ) circuits.

Subjects: Signal Processing , Information Theory , Quantum Physics

Publish: 2026-08-18 12:45:47 UTC


#3 Statistical Characterization and Block-EM Estimation of Frequency-Domain NSI for OFDM Systems in Bursty Impulsive Noise [PDF] [Copy] [Kimi] [REL]

Authors: Chin-Hung Chen, Wim van Houtum, Yan Wu, Alex Alvarado

Impulsive noise (IN), characterized by its high power and non-Gaussian distribution, poses a critical challenge in modern orthogonal frequency-division multiplexing (OFDM) systems, driven by the proliferation of electronic devices. Current IN mitigation techniques rely heavily on time-domain processing. These methods apply before the discrete Fourier transform (DFT), introducing additional complexity, failing to align with OFDM's inherent frequency-domain processing flow, and risking the destruction of subcarrier orthogonality due to imperfect IN subtraction. To address these limitations, we propose a frequency-domain, block-based framework for mitigating IN. The statistical representation of IN in the frequency domain is first derived using a transformed Gaussian mixture model. Based on this model, we develop an optimal receiver that leverages perfect noise state information (NSI), thereby identifying scenarios in which NSI is critical. We then propose an unsupervised block-based expectation-maximization (EM) framework for NSI estimation and develop three variants for evaluation. These include a simple symbol-by-symbol variance-updated EM, a sequence-based transition-updated EM, and a MAP-based EM that exploits a sparsity-promoting prior to automatically prune the number of states. Our frequency-domain design operates after the DFT, seamlessly integrates with the OFDM processing chain, preserves subcarrier orthogonality, and leverages the known IN block structure to achieve substantial performance gains without the immense complexity of time-domain impulse reconstruction.

Subjects: Signal Processing , Information Theory

Publish: 2026-08-18 11:56:46 UTC


#4 Empirical mode decomposition and interpretable machine learning for preterm birth classification from electrohysterography [PDF] [Copy] [Kimi] [REL]

Authors: Umesha Tilakarathna, Senith Jayakody, Kalana Jayasooriya, Roshan Godaliyadda, Parakrama Ekanayake, Isuru Nawinne, Chathura Rathnayake

Preterm birth (PTB) remains a major global health problem, and reliable non-invasive risk assessment remains difficult. Electrohysterography (EHG) records uterine electrical activity from the maternal abdomen and may support PTB assessment, but performance can be inflated when segments from the same recording are split across training and validation folds. We evaluated empirical mode decomposition (EMD) for term-versus-preterm classification using 26 pregnancy recordings (13 preterm, 13 term) from the public TPEHGT dataset. Annotated intervals and non-overlapping fixed 3-minute windows were compared, and the first four intrinsic mode functions (IMFs) were evaluated. Fourteen features from each of three EHG channels were assessed with nine classifiers using repeated five-fold recording-grouped cross-validation and recording-level aggregation. IMF1 gave the strongest mean performance. With fixed 3-minute IMF1 features, Random Forest achieved mean accuracy 0.8308, F1 0.7969, balanced accuracy 0.8308, MCC 0.6998, ROC-AUC 0.8157, and average precision 0.8877. IMF1 also outperformed matched filtered time-domain features across all reported mean metrics. Preterm recordings showed smaller, more regularly spaced peak-like events, lower temporal-energy measures, and higher entropy. These findings support further evaluation of IMF1-based EHG classification in larger independent cohorts.

Subjects: Signal Processing , Quantitative Methods

Publish: 2026-08-18 11:00:25 UTC


#5 On the Probability of Network States with Gaussian Connectivity Functions [PDF] [Copy] [Kimi] [REL]

Authors: Amy S. Inwood, Peter J. Smith, Pawel Dmochowski, Carl P. Dettmann, Justin P. Coon, Michail Matthaiou

In this paper, we consider the connectivity of a random network of N mobile devices in three dimensions (3D), where the location of each device or node has a Gaussian distribution in each dimension. For each pair of nodes, the probability of connectivity is related to the nodes' separation by a Gaussian connectivity function. The fundamental analytical tool for studying such systems is the probability of a given network state, derived and expressed in terms of its graph Laplacian. Leveraging this result, we obtain results for the connectivity of small networks, the probability of a complete network (where all nodes are connected to all other nodes), and the probability of an isolated node, which gives an approximation to the connectivity probability of larger networks. The general results are then simplified for special cases and limiting scenarios.

Subjects: Signal Processing , Probability

Publish: 2026-08-18 10:37:30 UTC


#6 Geometry-Aware DRL for Multi-Subband Scheduling in Satellite-Assisted UAM Networks [PDF] [Copy] [Kimi] [REL]

Authors: Hyung-Joo Moon, Sangha Park, Chan-Byoung Chae, Robert W. Heath

In this paper, we investigate downlink scheduling for urban air mobility (UAM) in a cooperative space-air-ground integrated network. Multiple ground stations (GSs) employ narrow three-dimensional beams and share spectrum across multiple subbands, while a satellite provides an orthogonal-band service option. Rapidly time-varying geometry and directional interference require joint decisions on base station association, GS subband assignment, and transmit powers. We formulate a finite-horizon mixed discrete-continuous problem that maximizes sum rate while penalizing handovers and GS overload, using only UAM positions and velocities. To address the combinatorial scheduling problem, we propose GeoSetPPO, a geometry-aware set-attention proximal policy optimization (PPO) method that outputs per-UAM discrete association and subband decisions with permutation-invariant representations. Conditioned on each schedule, GS powers are computed by a per-slot successive convex approximation (SCA) module under per-GS power budgets and minimum signal-to-interference-plus-noise ratio (SINR) constraints. To reduce training cost and improve stability, we adopt a two-stage training strategy that transitions reward evaluation from uniform power to SCA-based power allocation. Simulations demonstrate stable convergence, higher returns than multi-layer perceptron (MLP)- and Transformer-based PPO under the considered training setting, and favorable reward and schedule-feasibility performance relative to algorithm-based and distance-based schedulers. In the larger evaluated network, GeoSetPPO also reduces the scheduling latency from 40.84 ms to 2.90 ms relative to the previous algorithm-based method.

Subject: Signal Processing

Publish: 2026-08-18 10:22:28 UTC


#7 Channel2World: A Wireless Foundation Model for RF Environment Representation [PDF] [Copy] [Kimi] [REL]

Authors: Hyung-Joo Moon, Joonkyu Jang, Kwang Soon Kim, Seong-Lyun Kim, Robert W. Heath, Chan-Byoung Chae

Wireless channels are commonly treated as link-specific observations, although their multipath structure is governed by the surrounding radio-frequency (RF) environment. In this paper, we propose Channel2World, a wireless foundation model that learns a reusable environment-level representation from multiple-input multiple-output (MIMO) channel-position observations. The model aggregates channels collected within the same base-station-centered environment into a wireless world embedding using a Transformer-based encoder. The encoder is pretrained through context-query prediction, where context channels condition user equipment (UE) position and relative path-gain prediction for disjoint query channels. After pretraining, the encoder is frozen and used as a task-agnostic environment-conditioning module for downstream wireless models, enabling adaptation to unseen environments without site-specific fine-tuning. To learn an environment-level latent space that generalizes across deployments, we pretrain Channel2World using ray-tracing data from 26,000 environments, with approximately 5,000 channel measurements per environment. Evaluations on UE localization, beam-domain channel state information (CSI) reconstruction, and RF-observable geometry reconstruction show that the learned embeddings provide effective conditioning in unseen environments. For localization and CSI reconstruction tasks, embedding-based conditioning outperforms or remains competitive with site-specific fine-tuning, although fine-tuning requires task-specific labeled data and additional gradient-based adaptation. The embeddings also support the reconstruction of dominant reflector structures, indicating their utility as reusable environmental priors across tasks.

Subject: Signal Processing

Publish: 2026-08-18 09:04:12 UTC


#8 Channel Modeling for Phase-Based Ranging [PDF] [Copy] [Kimi] [REL]

Authors: Till Droemmer, Markus Gardill

This paper presents a Python-based simulation framework for the physical layer of Bluetooth Channel Sounding in accordance with Bluetooth Core Specification v6.2. The simulator implements Mode 3 phase-based ranging across 72 active tones in the 2.4 GHz ISM band and supports multiple propagation conditions, including free-space, static multipath, and IEEE 802.15.4a stochastic multipath channels. In addition, it models relevant hardware and interference effects such as phase noise, phase ramp, carrier frequency offset, IQ imbalance, ADC quantization, and narrowband interference from co-located Wi-Fi systems, with parameters derived from commercial SoC datasheets and field measurements. The framework enables controlled and repeatable analysis of physical layer effects that are difficult to isolate in over-the-air experiments and are not sufficiently represented in existing higher-level simulation tools. This work focuses exclusively on physical layer signal generation, channel modeling, and impairment injection, providing a foundation for future studies on Bluetooth Channel Sounding ranging and localization algorithms.

Subject: Signal Processing

Publish: 2026-08-18 08:24:26 UTC


#9 Super-resolution ranging using a sub-terahertz self-injection-locked frequency-modulated radar [PDF] [Copy] [Kimi] [REL]

Authors: Hossein Naghavi, Zainulabideen Khalifa, Hamad Alotaibi, Farzad Khoeini, James Gruber, Morteza Tavakoli Taba, Aditya Varma Muppala, Saghar Adler, Ali Mostajeran, Mohammed Aseeri, Andreia Cathelin, Ehsan Afshari

Sub-terahertz (sub-THz) and terahertz (THz) frequency-modulated continuous-wave (FMCW) radars have opened a plethora of scientific and industrial applications, especially in the imaging field. While strong candidates for sub-THz/THz FMCW radar imagers are implemented using photonic methods, there is a desire to achieve the full integration and portability that only electronics can offer. However, integrated electronic sub-THz/THz FMCW radars have significantly lower bandwidth (< 100 GHz) than photonic-based radars, restricting the radar range resolution to the millimeter scale (> 1.5 mm). In addition, the electronic FMCW radar's broad bandwidth comes with increased transmitter phase noise, consequently degrading the radar range accuracy. Here, we present a sub-THz fully-integrated autodyne frequency-modulated (AFM) radar utilizing a self-injection locking (SIL) mechanism that fundamentally overcomes the aforementioned challenges of FMCW radars. The AFM radar supports an exceptionally wide effective bandwidth extending into the terahertz sweep range by forming an intermediate frequency comb spectrum in a quadratic receiver, unlocking the path for super-resolution ranging. Furthermore, SIL significantly reduces the transmitter's phase noise, allowing high-accuracy range measurements. We theoretically describe and experimentally demonstrate the SIL operation of the AFM radar. The proposed radar experimentally achieves sub-millimeter range resolution and a range accuracy of < 0.002%, enabling the imaging of covered printed letters with micrometer features.

Subject: Signal Processing

Publish: 2026-08-18 08:05:45 UTC


#10 Multi-Sensor Edge Angle Detection for Performance Analysis in Ski Jumping [PDF] [Copy] [Kimi] [REL]

Authors: Ivan Simeonov, Lukas Schulthess, Hanna Mueller, Marc Nölke, Michele Magno, Luca Benini, Christoph Leitner

In ski jumping, performance during the gliding phase depends on achieving an aerodynamic posture that maximizes the lift-to-drag ratio. In the V-style technique, the ski edge angle is a key determinant. Reducing the edge angle flattens the skis, increases their effective surface area, and improves aerodynamic lift, ultimately contributing to longer flight distances. Ski edge angles are biomechanically constrained by the limited range of ankle inversion. Current sensing solutions widely quantify these angles using multi-system approaches that combine sensor signals through geometric relations. Such configurations require instrumentation on both the boot and the ski, altering mass distribution, affecting balance during flight, and increasing system complexity. To overcome these limitations, this work presents a wearable sensing system that measures both boot inclination and ski edge angle without modifying the ski surface. Two ultrasonic Time of Flight (ToF) sensors and an in-shoe Inertial Measurement Unit (IMU) are integrated into a single boot-mounted unit. Edge angles are estimated by combining ultrasonic distance measurements with IMU data through geometric reconstruction of the boot-ski configuration. Laboratory experiments demonstrate an angle resolution of 0.4500°, a Mean Absolute Error (MAE) of 0.2640°, and a coefficient of determination exceeding 99\% when compared with reference measurements, indicating strong linear agreement between the two modalities. The system achieves an end-to-end latency of 30.31 ms, enabling real-time feedback suitable for athlete training, while consuming 1.28 mW of power. With a total weight of only 18.6 g the proposed system enables unobtrusive measurement of ski edge angle and boot orientation.

Subjects: Signal Processing , Systems and Control

Publish: 2026-08-18 07:30:16 UTC


#11 Finite-range Lattice Momentum Operators for Quantum Field Theory [PDF] [Copy] [Kimi] [REL]

Authors: Jan C Olivier, Etienne Barnard

We propose a Z-transform framework for the analysis and synthesis of finite range lattice momentum operators in quantum field theory. In this formulation, translation-invariant lattice operators are represented as functions of the complex variable $z$ in the unit circle, allowing their spectral properties to be analyzed using tools from digital signal processing and rational approximation theory. Within this framework, the fermion doubling problem is reinterpreted as the appearance of unwanted zeros of the discrete momentum operator on the unit circle --- an aliasing phenomenon in the sense of the Nyquist sampling theorem --- and the conditions for ghost suppression are expressed as precise constraints on the zero structure of the operator's transfer function. It is proven that no rational function can satisfy all required conditions simultaneously, motivating the finite impulse response approach developed here. This reframing naturally suggests a class of finite-range momentum operators, constructed by solving a least-squares approximation problem in the frequency domain. The resulting finite impulse response (FIR) operator approximates the continuum derivative across the full Brillouin zone, with ghost suppression achieved through the accuracy of the spectral approximation rather than through the addition of a symmetry-breaking Wilson term or the infinite-range nonlocal SLAC derivative. Numerical investigation confirms that near $θ= π$ only plane waves propagate coherently, and these exhibit group velocities far exceeding the speed of light, further distinguishing them from physical low-energy excitations. No ghost wave packet solutions exist near $θ= π$.

Subjects: Signal Processing , High Energy Physics - Lattice , Quantum Physics

Publish: 2026-08-18 03:38:47 UTC


#12 Channel Knowledge Map Enabled Low-Complexity Dynamic Radio Environment Reconstruction [PDF] [Copy] [Kimi] [REL]

Authors: Yujun Lin, Zhiqiang Xiao, Hao Wu, Xiaoqiang Qiao, Fayu Wan, Tao Zhang

Accurate and timely radio environment reconstruction is important but challenging under particularly dynamic transmitter configurations. The conventional methods such as compressed sensing (CS), Kriging method or U-Net typically require environment measurements and reconstruction overhead for radio environment updating as the transmitter locations or radiation patterns change. In this paper, we propose a novel channel knowledge map (CKM)-enabled dynamic radio environment reconstruction method for efficient radio map updating. Specifically, the recently proposed CKM can store reusable path-level propagation knowledge that is decoupled from the transmitter-side radiation characteristics. We can leverage CKM for lightweight forward radio map generation as the transmitter locations and radiation patterns are known, without requiring new target-map measurements. Simulation results show that the proposed method outperforms CS, Kriging, and U-Net in reconstruction accuracy and exhibits strong robustness performance under dynamic transmitter configurations, which demonstrates the potential of the proposed method for flexible and efficient radio environment reconstruction in dynamic wireless networks.

Subject: Signal Processing

Publish: 2026-08-18 03:29:50 UTC


#13 Task-Based Evaluation of Raw Radar Data Compression: A Pre-Registered Study of Where Classical Codecs Fail to Preserve Target Detection, and Why [PDF] [Copy] [Kimi] [REL]

Author: Eric Michael Chrabot

Synthetic aperture radar (SAR) systems collect raw I/Q echo data at rates that exceed downlink capacity; fielded systems compress onboard with block-adaptive quantization (BAQ/FDBAQ). Proposed replacements, including learned compression, are typically evaluated with image-quality metrics (PSNR, SSIM, SQNR); none measures whether the data still supports its operational use. We introduce a pre-registered, task-based evaluation methodology for raw radar compression: codecs are scored after SAR focusing against a two-sided detection criterion -- a CFAR detection-agreement floor and a false-alarm budget at matched threshold -- using frozen task models trained once on uncompressed data. On Sentinel-1 stripmap Level-0 data the evaluation reproduces the operating point of the fielded FDBAQ codec (3-bit BAQ sustains utility at 6.12 bits per complex sample), with a previously reported reconstruction-lattice artifact tested for and ruled out. No classical configuration tested sustains utility below 4.86 bits per complex sample (roughly 13:1) on this scene, and each failure has an identifiable mechanism: raw echoes offer no transform-coding gain, while energy concentration after focusing delivers the predicted detection gain (Pd 0.888 vs 0.691 at 2 bits per complex sample) but converts it into false alarms through wavelet ringing; detection probability alone would have certified a codec producing roughly 53 times the tolerable false-alarm count. Applying the same protocol to AFRL Gotcha GMTI data (airborne, never compressed onboard) replicates both findings independently (frontier at 7.99 bits per complex sample). We release the evaluation harness, a unitary dechirp/focus transform with verified round-trip invertibility (relative error ~1e-7), and the complete pre-registration trail, including one retracted overclaim and a bounded negative result for a small learned codec, reported as a lower bound.

Subject: Signal Processing

Publish: 2026-08-18 01:51:35 UTC


#14 Multi-Tag Collision Recovery in UHF-RFID Using Self-Attention Decoding [PDF] [Copy] [Kimi] [REL]

Authors: Talha Akyildiz, Siva Aditya Gooty, Hessam Mahdavifar, Najme Ebrahimi

Passive ultra high frequency (UHF) radio frequency identification (RFID) enables battery-free tags to communicate with a reader through backscatter. When multiple tags respond in the same time slot, their waveforms overlap at the reader, and a conventional reader that follows framed slotted ALOHA (FSA) discards the resulting collided slot. This limits the throughput of the overall protocol even though the received signal still contains recoverable information about the responding tags. To address this limitation, we propose Self-Attention Tag Recovery (SATR), a transformer-based decoding algorithm that operates directly on the baseband in-phase and quadrature (I/Q) samples received during a standard tag response. SATR uses self-attention to model the temporal structure of the modulated waveform and learns candidate tag representations. It jointly estimates the number of responding tags and, more importantly, decodes the bit sequence of each detected tag. We numerically evaluate the decoding and throughput performance of SATR over a range of collision sizes and recovery configurations, and validate it with measurements of commercial UHF-RFID tags. The results show that, with proper design and training, SATR can reliably decode collisions of up to four tags. It achieves a throughput of approximately $0.815$ tags per slot under single acknowledgment and $1.87$ tags per slot under full recovery, corresponding to $2.2$ and $5.1$ times the conventional FSA limit of $1/e \approx 0.368$ tags per slot, while approaching optimal decoding performance and outperforming existing collision recovery methods.

Subject: Signal Processing

Publish: 2026-08-17 23:59:46 UTC


#15 Physically Consistent Channel Modeling and Signal Processing for Reconfigurable Wireless Systems [PDF] [Copy] [Kimi] [REL]

Authors: Ahmad Dkhan, Simon Tarboush, Hadi Sarieddeen, Robert W. Heath, Hakan Bagci, Tareq Y. Al-Naffouri

Reconfigurable antennas are increasingly integrated into multi-antenna communication systems to exploit large apertures while reducing the hardware complexity, energy consumption, and implementation costs of classical massive arrays. Their reconfigurable electromagnetic (EM) properties, including dynamically varying radiation patterns and state-dependent mutual coupling, challenge the fixed-antenna and decoupled-port assumptions of conventional channel models. This motivates a physically consistent framework connecting Maxwell's equations, circuit theory, and information theory. In this tutorial, we develop a unified framework spanning three coupled dimensions: (i) reconfigurable antenna and transceiver architectures, (ii) physically consistent channel modeling, and (iii) physically consistent signal processing. We first establish a taxonomy covering tunable antennas, reconfigurable transceivers, and emerging array architectures, highlighting their reconfiguration mechanisms and hardware-performance trade-offs. We then develop modeling approaches based on Maxwell's equations, wavenumber-domain representations, multiport network theory, and computational electromagnetics, and use them to construct end-to-end channel and noise models that capture near-field propagation, mutual coupling, and circuit-level impairments. Building on these models, we examine architecture-aware channel estimation, beamforming, data detection, and channel decoding, emphasizing how physical structure reshapes algorithm design and performance-complexity trade-offs. Overall, the tutorial treats physical architecture, channel and noise models, and communication algorithms as coupled components of an end-to-end design, providing a unified foundation for physically consistent reconfigurable wireless systems.

Subject: Signal Processing

Publish: 2026-08-17 21:00:31 UTC


#16 Primitive Representation Learning for Unsupervised Dynamic Contrast Enhanced MRI Reconstruction [PDF] [Copy] [Kimi] [REL]

Authors: Veronika Spieker, Wenqi Huang, Cemre Ariyurek, Liam Timms, Daniel Rueckert, Onur Afacan, Julia A. Schnabel, Sila Kurugol

Reliable quantitative analysis of dynamic contrast-enhanced MRI requires high-quality spatiotemporal reconstructions at high undersampling rates. Scan-specific reconstructions using Gaussian and Gabor primitives have shown promising results without the need for large training datasets, but have not addressed the additional dimension of dynamic contrast. We propose a multi-dimensional, primitive based framework for dynamic contrast-enhanced MRI reconstruction that disentangles the underlying anatomy, the dynamic contrast enhancement, and residual motion into separate temporal basis functions, thereby enabling a geometrical interpretation of the representation. We show that this architecture achieves performance competitive with conventional reconstruction methods, both in reconstruction quality and in the accuracy of extracted aorta and kidney enhancement curves. The modular tier design extends naturally to additional dynamic factors and higher acceleration rates. Code available at https://github.com/compai-lab/ 2026-GaborDCE-spieker.

Subjects: Image and Video Processing , Computer Vision and Pattern Recognition , Machine Learning , Signal Processing , Medical Physics

Publish: 2026-08-18 17:48:22 UTC


#17 The Zonotopic Mixture Filter [PDF] [Copy] [Kimi] [REL]

Authors: Rodrigo A. González, Angel L. Cedeño, Vicenç Puig

State estimation is commonly posed in either a probabilistic or an unknown-but-bounded framework. The former requires a fully specified noise distribution, typically with unbounded support, while the latter yields guaranteed enclosures that carry no probabilistic weighting. Bridging these noise descriptions, this paper proposes a zonotopic mixture noise model, in which the noise is generated by drawing a zonotope from a finite collection according to fixed probabilities and then realizing an arbitrary element of it. For this noise model, we derive the zonotopic mixture filter, which propagates a bank of zonotopic Kalman filters over mode histories, discards the histories falsified by the data, and weights the surviving ones by their relative probability. The resulting state enclosures yield guaranteed coverage probabilities and remain valid for every noise realization compatible with the bounds, and a greedy mixture reduction scheme preserves these statistical guarantees while keeping the representation tractable. Numerical examples illustrate the proposed approach and its potential benefits over related state estimation methods.

Subjects: Systems and Control , Signal Processing

Publish: 2026-08-18 15:29:36 UTC


#18 Edge-Native Embodied Intelligence for Action-Aware Wireless Edge Networks [PDF] [Copy] [Kimi] [REL]

Authors: Yiru Wang, Chuanao Jiang, Jiahui Cui, Zide Fan, Lei Wang, Zehui Xiong, Dong In Kim

Embodied intelligence is shifting artificial intelligence from passive digital perception toward active physical interaction. However, foundation-model-enabled embodied agents face a fundamental tension between open-world cognition and resource-constrained deployment. On-device models are limited by computation, memory, and energy budgets, whereas cloud-centric solutions introduce latency and reliability risks over dynamic wireless links. Edge general intelligence provides a promising cognitive backbone, but existing frameworks still lack physical grounding, action awareness, and mechanisms for actively acquiring useful physical experience. To address these limitations, this article introduces edge-native embodied intelligence (ENEI), an action-aware wireless edge framework that integrates embodied agents, the 6G communication and networking fabric, and edge cognitive services into a 6G-mediated bidirectional edge-embodiment loop. Along the edge-to-embodiment axis, confidence-aware assistance and edge-driven generative adaptation enhance local autonomy under out-of-distribution (OOD) conditions. Along the embodiment-to-edge axis, value-of-experience guided active embodied federated learning enables physical actions to generate informative experience for continuous edge model evolution. The 6G fabric supports both directions through goal-oriented transmission and programmable radio-resource allocation. Two case studies on OOD drone navigation and mobility-driven federated learning illustrate the feasibility and communication efficiency of the proposed mechanisms. ENEI provides a unified perspective in which edge cognition strengthens embodied action, while embodied agency actively enriches edge cognition, laying the foundation for scalable, adaptive, and self-evolving embodied wireless systems.

Subjects: Systems and Control , Signal Processing

Publish: 2026-08-18 13:39:12 UTC


#19 Learnware for CSI Feedback: Scene-specific Small Models Can Do Big [PDF] [Copy] [Kimi] [REL]

Authors: Xiangyi Li, Jiajia Guo, Chao-Kai Wen, Xin Geng, Shi Jin, Zhi-Hua Zhou

Intelligent channel state information (CSI) feedback is essential for realizing the high capacity and spectral efficiency goals of future 6G systems, yet existing deep learning solutions face a trade-off between model generalization and scenario-specific performance. Large neural networks generalize well but incur high computational and tuning costs, while small models excel in particular environments but require repetitive costly end-to-end training for each base station (BS). To address these challenges, we introduce a model repository-based deployment framework in which a centralized AI data center maintains a catalog of scene-specific CSI models. The repository is enhanced with a Learnware-based framework, where each model is associated with a specification including semantic part (network architecture parameters) and statistical part (codeboo-fingerprint embeddings of training-data distributions). A BS submits only its local statistical specifications to retrieve the most relevant pre-trained model, enhancing data privacy by avoiding raw CSI transmission and drastically reducing retrieval latency and communication overhead. We further develop a data-driven search strategy that matches codebook fingerprints to model performance, achieving over 90% selection accuracy. In simulations, our scheme yields 18.8% and 57.7% performance improvements over the General Model in LOS and NLOS scenarios, respectively while reducing local fine-tuning by up to 1000 samples and 100 epochs. This Learnware-based approach minimizes redundant training, maximizes model reuse, and supports rapid,privacy-enhancing deployment of CSI feedback models.

Subjects: Information Theory , Artificial Intelligence , Signal Processing

Publish: 2026-08-18 13:26:34 UTC


#20 Causal Discovery in Equal Variance Linear Gaussian DAGs via SURE-Tuned Ridge Regression [PDF] [Copy] [Kimi] [REL]

Authors: Sambit Mishra, Urbashi Mitra

Recovering the directed acyclic graph (DAG) of a structural equation model (SEM) from observational data is a central problem in causal discovery. The iterative gradient descent and per-problem hyperparameter tuning of continuous-optimization methods are poorly suited to two practically important regimes: the sample-limited regime, where the number of samples is comparable to or smaller than the number of nodes in the DAG, and the compute-limited regime. This work proposes SURE-Ridge, a non-iterative, closed-form estimator for equal variance linear Gaussian SEM. The method performs parallel node-wise regressions with regularization parameters chosen adaptively by Stein's unbiased risk estimate (SURE), and applies an adaptive thresholding procedure to extract a DAG from the resulting soft adjacency matrix. Numerical results show that SURE-Ridge achieves the lowest structural Hamming distance in the small-sample regime and the lowest run time across all sample sizes tested, compared with NOTEARS, DAGMA, and GBNSL baselines.

Subjects: Machine Learning , Signal Processing , Machine Learning

Publish: 2026-08-17 21:07:06 UTC


#21 GeoGS-CE: Learning Delay--Beam Channel Priors with 3D Gaussians for High-Mobility Scenarios [PDF] [Copy] [Kimi] [REL]

Authors: Yumeng Zhang, Jiajia Guo, Chaozheng Wen, Chenghong Bian, Jun Zhang

Wideband channel estimation (CE) in high-mobility scenarios remains challenging because channel responses vary rapidly, while practical systems can allocate only sparse pilots to accommodate dense users. Fortunately, many high-mobility environments, such as high-speed railways, exhibit scheduled trajectories, predictable velocities, and a limited number of dominant propagation paths. These properties induce a delay--beam power spectrum that is more stable than the instantaneous complex channel frequency response (CFR), less sensitive to the random phase coherence, and rich in geometric information. To exploit such environmental properties, we propose GeoGS-CE, a two-stage channel estimation framework for sparse-pilot high-mobility scenarios. In the offline stage, GeoGS-CE jointly models: 1) a scene-level 3D Gaussian representation that captures the non-line-of-sight (NLoS) geometric scattering support, and 2) a leakage-aware differentiable wireless rendering process that maps the NLoS Gaussians, together with an explicit virtual line-of-sight (LoS) component, to the measured delay--beam power spectrum, while accounting for practical OFDM delay and array leakage effects. In the online stage, the delay--beam power spectrum is predicted for each user location and used as a strong covariance prior, enabling accurate full-band and full-array CFR reconstruction and tracking through a linear MMSE estimator. Simulations based on channels generated from a segment of the Guangshen high-speed railway show that the proposed geometric prior substantially improves CFR reconstruction over pilot-only and non-geometric baselines.

Subjects: Information Theory , Artificial Intelligence

Publish: 2026-05-15 15:49:42 UTC