Search arXivSearch

subject

eess.SP

eess.SP: explore 26 source-linked works published from 2024 to 2026, with original documents and citations.

This collection is a preview while coverage and quality are evaluated.

Search within this collection

Coverage and selection

Includes records with this source-supplied label or an explicit phrase match in their metadata. Matches indicate a mention, not proof that a paper uses a method or tests a material. Source versions are consolidated by DOI.

Sources: arxiv. Collection updated 2026-09-15. Counts describe this index, not the complete source archives.

Model Selection and Parameter Estimation of One-Dimensional Gaussian Mixture Models

In this paper, we study the problem of learning one-dimensional Gaussian mixture models (GMMs) with a specific focus on estimating both the model order and the mixing distribution from independent and identically distributed (i.i.d.) samples. This paper establishes the optimal sampling complexity for model order estimation in one-dimensional Gaussian mixture models. We prove a fundamental lower bound on the number of samples required to correctly identify the number of components with high probability, showing that this limit depends critically on the separation between component means and the total number of components. We then propose a Fourier-based approach to estimate both the model order and the mixing distribution. Our algorithm utilizes Fourier measurements constructed from the samples, and our analysis demonstrates that its sample complexity matches the established lower bound, thereby confirming its optimality. Numerical experiments further show that our method outperforms conventional techniques in terms of efficiency and accuracy.

stat.ML

Beat-Synchronous Tokenization for ECG Transformers

Transformer-based electrocardiogram (ECG) models commonly tokenize waveforms into fixed temporal patches. Though convenient, fixed patching can split heartbeat structures across token boundaries. We study beat-synchronous tokenization as a physiologically grounded alternative, comparing fixed patches with three beat-aligned strategies: resampled beats, adaptive pooled beats, and resampled beats augmented with R--R interval information. Experiments span two settings: 10-second 12-lead diagnostic classification on PTB-XL after MIMIC-IV-ECG masked pretraining, and 60-second single-lead rhythm classification on Icentia11k after patient-level contrastive pretraining. On PTB-XL, resampled beat tokens achieve the highest mean macro Area Under the ROC Curve (AUROC; 0.8945) and nearly match the best fixed-patch macro Area Under the Precision-Recall Curve (AUPRC; 0.7414), reducing average sequence length from 100 to 11.2 tokens. On Icentia11k, beat-synchronous tokenizers obtain comparable AUPRC to fixed patching with better stability across runs. These results suggest morphology-preserving beat tokenization is a compact, competitive alternative to fixed temporal patching.

cs.LG

Benchmarking External Generalization of SPD Matrix Learning for Resting-State fMRI Connectome Prediction

Resting-state functional magnetic resonance imaging (rs-fMRI) functional connectivity (FC) matrices are widely used for individual-level prediction, but strong performance within one cohort may not generalize to a new cohort. We ask whether within-dataset performance remains when the test data come from an entirely held-out rs-fMRI dataset. Each scan is represented as a regularized symmetric positive definite (SPD) correlation connectome, which allows methods to use the geometry of the SPD manifold. We introduce a reproducible age-prediction benchmark across six rs-fMRI datasets: COBRE, ADNIDOD, Cam-CAN, ABIDE, OASIS-3, and ADNI. The benchmark compares a vectorized correlation baseline, Tangent-Space Ridge, SPDNet, and split-wise Riemannian harmonization under within-dataset GroupKFold, pooled GroupKFold, and leave-one-dataset-out (LODO) evaluation. Within-dataset and pooled GroupKFold results are substantially more favorable than LODO results. When an entire dataset is held out, prediction error increases, differences among methods narrow, and performance is strongly affected by age-range mismatch and cohort heterogeneity. The benchmark provides common inputs, model settings, data splits, and analysis scripts so that future SPD matrix learning methods can be evaluated under the same external-validation protocol.

eess.SP

Improving the decoding performance of CA-polar codes

We investigate the use of modern code-agnostic decoders to convert CA-SCL from an incomplete decoder to a complete one. When CA-SCL fails to identify a codeword that passes the CRC check, we apply a code-agnostic decoder that identifies a codeword that satisfies the CRC. We establish that this approach gives gains of up to 0.2 dB in block error rate for CA-polar codes from the 5G New Radio standard. If, instead, the message had been encoded in a systematic CA-polar code, the gain improves to more than 1.5 dB. Leveraging recent developments in blockwise soft output, we additionally establish that it is possible to control the undetected error rate even when using the CRC for error correction.

cs.IT

Grassmannian-Coded Beamforming for mmWave Channel Sensing with Unknown Complex Path Gain

This paper introduces a subspace-coding perspective to millimeter-wave channel sensing with a single RF chain when the complex channel gain is unknown. We show that in this case, candidate directions-of-arrival (DoAs) map naturally to subspaces through their beamspace responses, revealing an intrinsic Grassmannian geometry. This motivates beamspace Grassmannian codes (BGCs), designed to reduce DoA error by maximizing the minimum subspace distance of the joint beamformer-array response. We identify two regimes: one in which existing Grassmannian packings are exactly realizable as BGCs when the angular grid matches the array size, and another in which realizability for finer grids is constrained by the array geometry. Our analysis establishes the joint roles of subspace distance and beamforming gain in sensing performance and motivates two complementary beamformer designs. Without prior DoA information, we develop spatially isotropic beamformers based on algebraic Grassmannian packings and modulation-based channel codes. With a known DoA region of interest, we design convolutional beamspaces that combine directional gain with favorable subspace distance. Numerical results demonstrate robust BGC performance for both on-grid and off-grid DoAs, supporting the effectiveness of the proposed Grassmannian framework for mmWave channel sensing.

eess.SP

AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G

The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical-layer design. However, existing models often operate on channel state information (CSI) in the spatial-temporal-frequency (STF) domain, where multipath components are superimposed and structurally entangled. This hinders the learning of a universal channel representation. Their reliance on global attention also incurs prohibitive overhead. In this paper, we propose AirFM-DDA, an Air-interface Foundation Model in the Delay-Doppler-Angle (DDA) domain. AirFM-DDA reparameterizes CSI into the DDA domain to resolve multipath components along physically meaningful axes and employs window-based attention with frame-structure-aware positional encoding. Extensive experiments demonstrate transferability across scenarios, tasks, datasets, and antenna configurations. For channel prediction and estimation, AirFM-DDA generalizes zero-shot to unseen cities, achieving average normalized mean-square error (NMSE) gains of 4.9-8.5 dB over the strongest baselines. With only 10% labeled data, it achieves average gains of 12.0 percentage points in Top-1 accuracy for beam prediction and 3.4 percentage points in F1 score for line-of-sight (LoS) identification. It further transfers across simulated datasets and adapts to measured data and different antenna arrays. Compared with global attention, window-based attention reduces training and inference costs by nearly an order of magnitude.

cs.LG

Learning the Channel Gain from Anywhere to Anywhere via Cross-environment Transformer Estimators

Channel-gain maps provide the channel gain between any two locations in a geographical region. They find numerous applications, from resource allocation and interference control to path planning for autonomous vehicles. Channel-gain map estimation (CGME) is considerably more challenging than conventional radio map estimation (RME) because channel-gain maps are functions over a 6-dimensional input space. This calls for specialized methods, which currently rely on the (inaccurate) radio tomographic model or require a prohibitively large number of measurements since they do not exploit any spatial structure. This paper overcomes this issue by leveraging spatial patterns that channel-gain maps exhibit across environments, as dictated by the laws of physics and typical environmental characteristics (e.g. building materials and layouts). Adopting a metalearning perspective, a transformer-based estimator is proposed to implicitly learn this common structure from measurements collected in multiple environments. This enables CGME in new environments from significantly fewer measurements (five times less in our experiments). To maximize learning efficiency, the transformer is composed with a feature map that enforces the invariances of CGME, such as those following from reciprocity. Numerical experiments corroborate the merits of the proposed estimator relative to existing methods.

eess.SP

Adaptive Finite-Time Position-Force Control of Teleoperation Systems With Time-Varying Delays Using a Liquid State Machine Uncertainty Estimator

Teleoperation systems are increasingly used in medical, rehabilitation, and remote manipulation applications, where accurate position/force tracking and stable interaction are essential. In such applications, the remote environment may exhibit viscoelasticity, frictional memory, contact transitions, and other dynamic interaction effects, causing the system response to depend not only on the current state but also on its previous evolution. This history dependence, together with communication delays and uncertain nonlinear dynamics, makes accurate uncertainty compensation particularly challenging. Conventional feedforward neural approximators do not inherently retain temporal information, while fully recurrent architectures may introduce additional computational and online training complexity. To address this limitation, this article introduces the first application of a liquid state machine (LSM) to bilateral teleoperation control. A finite-time adaptive controller is developed using a hybrid position/force auxiliary error system with velocity and force filters, while the LSM is employed to estimate uncertain dynamics by exploiting its intrinsic temporal processing and fading-memory capabilities with a simple adaptation mechanism. Closed-loop stability and finite-time convergence are established through a Lyapunov--Krasovskii framework. Simulations in spring--damper and generalized Maxwell viscoelastic environments demonstrate improved position and force tracking and lower mean execution time compared with an RBFNN-based controller.

eess.SY

Identification of $dq$-Asymmetric Impedances as Complex Transfer Functions Using a Single Arbitrary Excitation

Cross-coupling between the $dq$ coordinates makes the identification of asymmetric grid impedances a challenging problem, particularly near the fundamental frequency where the asymmetric coupling is strongest. Existing schemes usually handle it either by perturbing the two coordinates sequentially, which lengthens the measurement, or by using a time-domain method with a global parametric model whose order must be tuned. This paper develops a single-shot active non-parametric frequency-domain method that avoids both. The equivalent impedance is parameterized by a pair of single-input single-output complex transfer functions. Each spectral line is fitted with a local rational model; the leakage and transient contributions are estimated, so that neither periodic steady-state excitation nor repeated excitation cycles are required. We give the exact finite-time discrete Fourier transform relation for the conjugate-coupled complex-signal model, and analyse the distortion that a stationary-frame filter placed ahead of the Park transform imposes on the identified pair. The method is validated on a controller hardware-in-the-loop platform against an analytically derived small-signal model, for a symmetric grid and for the same grid with an added grid-following converter that renders it asymmetric. Both complex transfer functions and all four real transfer functions of the $dq$ impedance are recovered over a wide band from a single one-second record of a random excitation, at 1 Hz resolution.

eess.SP

Large-Scale Bayesian Tensor Reconstruction via Approximate Message Passing

While CANDECOMP/PARAFAC (CP) decomposition (CPD) is fundamental for tensor reconstruction, Bayesian CPD often scales poorly because variational updates require repeated matrix inversions. We develop CP generalized approximate message passing (CP-GAMP) for incomplete noisy Bayesian CPD. The algorithm uses Gaussian message approximations to avoid high-dimensional inversions, and it combines a Bernoulli-Gaussian prior with expectation-maximization updates to estimate effective CP rank and noise variance. We also give a formal state evolution (SE) recursion and relate its fixed points to replica-symmetric saddle points, so CP-GAMP's SE-predicted error can be compared with the formal replica-symmetric minimum mean-squared error (MMSE) benchmark in the matched limit. Synthetic and image-inpainting experiments show that CP-GAMP substantially reduces runtime relative to variational Bayesian CPD while maintaining competitive reconstruction accuracy.

cs.LG

Secrecy Outage Analysis over Correlated Composite Generalized-Gamma Fading Channels

This paper investigates physical-layer security (PLS) over correlated composite generalized-Gamma (GG)/GG fading channels, where both shadowing and small-scale fading follow GG distributions. Using Mellin transforms and Fox-H functions, closed-form expressions are derived for the single-link probability density function (PDF), joint distribution, survival function, and zero-rate secrecy outage probability (SOP)/probability of non-zero secrecy capacity (PNZSC). The general-rate SOP is expressed as an exact double series with one residual onedimensional integral per term. The model includes the Nakagamim/GG and Nakagami-m/Gamma channels as special cases. Numerical results validate the analysis and demonstrate the impact of the fading parameters on secrecy performance.

cs.IT

An Attention-Assisted AI Model for Real-Time Underwater Sound Speed Estimation Leveraging Remote Sensing Sea Surface Temperature Data

The estimation of underwater sound velocity distribution serves as a critical basis for facilitating effective underwater communication and precise positioning, given that variations in sound velocity influence the path of signal transmission. Conventional techniques for the direct measurement of sound velocity, as well as methods that involve the inversion of sound velocity utilizing acoustic field data, necessitate on--site data collection. This requirement not only places high demands on device deployment, but also presents challenges in achieving real-time estimation of sound velocity distribution. In order to construct a real-time sound velocity field and eliminate the need for underwater onsite data measurement operations, we propose a self-attention embedded multimodal data fusion convolutional neural network (SA-MDF-CNN) for real-time underwater sound speed profile (SSP) estimation. The proposed model seeks to elucidate the inherent relationship between remote sensing sea surface temperature (SST) data, the primary component characteristics of historical SSPs, and their spatial coordinates. This is achieved by employing CNNs and attention mechanisms to extract local and global correlations from the input data, respectively. The ultimate objective is to facilitate a rapid and precise estimation of sound velocity distribution within a specified task area. The comparative analysis demonstrates that the proposed approach achieves superior performance in terms of both accuracy and stability, exhibiting reduced error rates and enhanced resistance to disturbances when benchmarked against existing advanced techniques.

eess.SP

ERP-XTTN: Interpretable Prototype-Guided Cross-Attention for Cross-Subject ERP Classification

Interpretable brain-computer interface classifiers that generalize across subjects without calibration remain an open challenge. We evaluated whether prototype-based cross-attention can provide competitive, inherently interpretable ERP classification across paradigms under deployment-compatible conditions. We propose ERP-XTTN (ERP Cross-Attention), a cross-attention architecture that routes input EEG peaks to fixed difference-wave prototypes via query-key-only cross-attention with no value projection. Classification is based directly on prototype similarity and a separate measure of component amplitude, so that prototype content contributes to every decision by construction. Prototypes are derived automatically from prominent extrema in the training-fold grand-average difference wave. We evaluated across three public sources (BNCI Horizon 2020, HRI Cursor, and ERP CORE) encompassing eight ERP components (ERN, LRP, ErrP, N170, P300, N2pc, MMN, N400). Evaluations used LOSO cross-validation with causal filtering at a three-channel montage, compared against EEGNet, EEG-Deformer, EPMN, and xDAWN with Riemannian geometry. The mean performance gap between the best baseline and ERP-XTTN was 0.025 AUROC. Prototype interventions confirmed that decisions depend on prototype content rather than on the routing attention pattern alone. False positives morphologically resembled true positives more than true negatives did, so classification errors are neurophysiologically explicable. ERP-XTTN generalizes across diverse ERP morphologies under causal, calibration-free conditions, while retaining competitive performance and decisions that depend directly on physiological prototype content. Unlike post-hoc explanation methods for black-box models, the basis of each decision is directly observable in the trained model. To our knowledge, this is the first epoch-level LOSO benchmark on ERP CORE.

cs.LG

Observability Analysis for Fusion of Doppler Measurements in Multistatic Radar Near the Tx-Rx Baseline

This paper studies multistatic measurement fusion when a target lies within the Tx--Rx (Transmitter-Receiver) baseline ambiguity zone, with particular emphasis on configurations involving two closely spaced stationary Tx--Rx pairs. Such configurations provide overlapping detectable regions and extend the effective detection range compared with sparsely spaced multistatic systems. However, in this region, the accuracy of range and bearing measurements degrades rapidly, and Doppler measurements often remain the only reliable information source. As a result, target trajectory estimation becomes highly challenging, with observability being marginal or even completely lost. To address this problem, the observability of target trajectories is analyzed under various conditions, enabling system designers to assess system performance in advance. A Doppler-only measurement fusion approach is then developed, employing a multiple-initial-point Maximum Likelihood (ML) nonlinear estimator for initial state estimation, followed by dynamic state updates using an Extended Kalman Filter (EKF). Simulation results are presented and shown to be consistent with the observability analysis.

eess.SY

ODMA-based MIMO Massive Unsourced Random Access with Soft-Output Polar Codes

This paper investigates the design of the on-off division multiple access (ODMA) transmission scheme for multiple-input multiple-output (MIMO) massive unsourced random access (URA) systems with soft-output (SO) polar codes. First, a three-segment pilot-uncoupled coding scheme is introduced under the ODMA framework, which reduces the coding rate of the data segment without increasing the transmission overhead, improving the overall system performance. Building upon this architecture, a hierarchical pattern detection framework is developed. Specifically, a coarse-grained candidate set of transmission patterns is first identified through correlation operations. Based on this, a message-passing (MP)-based pattern detection algorithm is developed to iteratively estimate the posterior probabilities of transmission patterns, followed by the \textit{maximum a posteriori} (MAP) estimation to obtain the precise pattern detection result. Furthermore, a joint pattern detection and data decoding algorithm based on the bit-wise SO information of polar decoder is investigated, where the posterior probability information provided by the polar decoder is exploited to refine the pattern detection and contribute to an improved accuracy. In addition, by leveraging bit-wise SO information of the successive cancellation list polar decoder, an MP-based iterative decoding algorithm is developed to significantly enhance the decoding performance. The proposed scheme simultaneously exploits the transmission gain of uncoupled-ODMA framework, the coding gain of polar codes in the short-blocklength regime, and the iterative decoding gain enabled by SO information, while the computational complexity is significantly reduced through the hierarchical detection framework. Simulation results demonstrate that the proposed scheme achieves strong robustness ...

cs.IT

False-CSI Attacks in Power-Domain NOMA for 6G: A Threat Taxonomy and System-Level Impacts

Power-domain non-orthogonal multiple access (NOMA) remains a widely studied technique for improving spectral efficiency and supporting dense connectivity in beyond-5G and 6G networks. Its main operating mechanisms, however, depend on the integrity of channel-state information (CSI). Power allocation, user ordering, pairing, clustering, and beamforming can all be distorted when the CSI consumed by the base station is deliberately biased rather than merely noisy. This article examines false CSI as an attack surface in power-domain NOMA. We organize the threat space using a compact taxonomy with two primary axes: magnitude, which distinguishes underreporting from overreporting, and ordering effect, which distinguishes order-preserving, boundary, and order-reversing attacks. We then show how coordinated false- CSI behavior, group-changing attacks, direction forgery, pilot spoofing, training-phase injection, and RIS-induced channel manipulation extend this basic taxonomy. Finally, we map each attack family to system-level impacts on power allocation, SIC reliability, scheduler behavior, fairness, throughput, and secrecy. The central message is that false CSI should be treated not only as a channel-estimation problem, but also as a control-input integrity problem for 6G NOMA.

cs.CR

Optimal Adversarial Testing: Extracting Honest Test Results from Dishonest Test Takers

In applications, it is often required to test objects or people to determine their qualities in terms of certain metrics. However, besides being naturally noisy, the test results can be corrupted by adversarial behaviors of objects or people being tested (test takers). For example, dishonest test takers can cheat in the exams to distort the test results. With the development of AI technologies, such distortions driven by cheating using AI technologies are becoming more commonplace and severe. In this paper, we propose optimal testing strategies which can still recover needed test results even if there are cheaters polluting the results. The proposed testing strategies will optimally re-test selected group of test takers using different testing security measures. We determine the optimal testing strategies using a dynamic programming method.

cs.CR

Accurate Plate Reverb Parameter Estimation Using Two-Stage Evolutionary Search

We describe our submission to Task A of the 1st DAFx parameter estimation challenge. The task is to recover the six physical parameters of a simulated metal-plate reverberator -- its dimensions and material properties -- from a single impulse response (IR). We treat this as a black-box optimization: candidate parameter sets are fed to the simulator and scored by a loss against the target IR. The method has two stages. The first uses CMA-ES, an evolutionary optimizer, to recover five of the six parameters, comparing IRs under an amplitude-normalized loss. Amplitude normalization makes the search robust but discards the cue to the sixth parameter, the plate's surface density; a second stage therefore estimates it alone, with a ternary search on the un-normalized loss. As the choice of loss strongly affects the search, we select it beforehand, and analyze why compression in the common multi-scale spectral loss degrades recovery. Finally, we test our method on a validation set of 50 IRs, discuss a pathological failure mode, and ablate to justify having two different stages instead of a unified CMA-ES search.

eess.AS
Compare source metadata on this page
WorkPublishedSource identifierSource
Model Selection and Parameter Estimation of One-Dimensional Gaussian Mixture Models2026-08-312404.12613arxiv
Beat-Synchronous Tokenization for ECG Transformers2026-08-312608.30367arxiv
Benchmarking External Generalization of SPD Matrix Learning for Resting-State fMRI Connectome Prediction2026-08-312608.30418arxiv
Improving the decoding performance of CA-polar codes2026-08-302512.10223arxiv
Grassmannian-Coded Beamforming for mmWave Channel Sensing with Unknown Complex Path Gain2026-08-302604.19904arxiv
AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G2026-08-302605.00020arxiv
Learning the Channel Gain from Anywhere to Anywhere via Cross-environment Transformer Estimators2026-08-302605.08211arxiv
Adaptive Finite-Time Position-Force Control of Teleoperation Systems With Time-Varying Delays Using a Liquid State Machine Uncertainty Estimator2026-08-302608.29544arxiv
Identification of $dq$-Asymmetric Impedances as Complex Transfer Functions Using a Single Arbitrary Excitation2026-08-302608.29740arxiv
Large-Scale Bayesian Tensor Reconstruction via Approximate Message Passing2026-08-292505.16305arxiv
Secrecy Outage Analysis over Correlated Composite Generalized-Gamma Fading Channels2026-08-292608.29414arxiv
An Attention-Assisted AI Model for Real-Time Underwater Sound Speed Estimation Leveraging Remote Sensing Sea Surface Temperature Data2026-08-282502.12817arxiv
ERP-XTTN: Interpretable Prototype-Guided Cross-Attention for Cross-Subject ERP Classification2026-08-282606.02939arxiv
Observability Analysis for Fusion of Doppler Measurements in Multistatic Radar Near the Tx-Rx Baseline2026-08-282608.27838arxiv
ODMA-based MIMO Massive Unsourced Random Access with Soft-Output Polar Codes2026-08-282608.28085arxiv
False-CSI Attacks in Power-Domain NOMA for 6G: A Threat Taxonomy and System-Level Impacts2026-08-282608.28351arxiv
Optimal Adversarial Testing: Extracting Honest Test Results from Dishonest Test Takers2026-08-282608.28362arxiv
Accurate Plate Reverb Parameter Estimation Using Two-Stage Evolutionary Search2026-08-282608.28818arxiv

These are bibliographic comparisons, not experimental rankings. Follow the original document for methods and conditions.