Clay weekly context brief for the Systems category (ISO week 2026-W38). Clay tracks publications from the Systems feed list. Below are recent items from this category, each with its source and a short description of what the publication covers when one is available in the source feed. Recent publications: 1. Downstream-Task-Aware Unified Source Separation Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11092 Task-aware unified source separation (TUSS) enables a single model to handle diverse separation tasks by conditioning on input prompts. 2. XPos3R: Cross-Modal Transformer for Intraoperative 2D/3D Registration Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10733 Intraoperative 2D/3D registration, which aligns live X-ray images with preoperative volumes, is essential for image-guided interventions. 3. Load Balancing in Multi-Shell LEO Satellite Networks with Successive Interference Cancellation Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11033 Multi-shell low Earth orbit (LEO) networks can increase service opportunities, but altitude-dependent propagation can concentrate traffic on lower shells and create strong inter-shell interference under full frequency reuse. 4. Practical Zero-Trust for Mission-Critical Robotic Fleets via Hardware Attestation and Packet Timing Watermarking Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.05741 Autonomous unmanned vehicles are vital to tactical missions, mission-critical public-safety operations like search and rescue and disaster response. 5. Low-Latency State Space Voice Activity Detection with Robust Onset Time Evaluation Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11110 Voice Activity Detection (VAD) systems are commonly evaluated using metrics such as the area under the receiver operating characteristic curve (AUROC), but these metrics do not account for temporal responsiveness. 6. Scale-Aware 3D Deep Learning for Robust Brain Metastasis Detection in Multimodal MRI Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10825 Detecting brain metastases in magnetic resonance imaging (MRI) remains challenging because lesions vary widely in size and appearance, with very small metastases occupying only a minute fraction of a three-dimensional input. 7. Tensor Decomposition Based Mixed-Field Sensing for XL-MIMO AFDM Systems Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11051 Integrated sensing and communications enabled by extremely large-scale MIMO (XL-MIMO) and affine frequency division multiplexing (AFDM) is a highly promising paradigm for vehicular networks. 8. AgentHomeID - Agent-based modelling of building stock transformation: A multi-scale framework for policy assessment and infrastructure planning Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.05763 Decarbonising the building sector is central to meeting climate targets, yet existing models rarely capture the interaction between system-level transformation dynamics and heterogeneous individual investment decisions. 9. Exploring Second-Order Pattern Recognition in Speaker Recognition Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11182 In classical pattern recognition tasks, neural networks are trained to recognise human-defined patterns for model inputs. 10. Annotating anatomy and pathology in the National Lung Screening Trial computed tomography images Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10858 Large-scale public medical imaging datasets contribute critically to translational research. 11. Radio Map Construction with Post-Hoc Location Calibration under Quasi-Static Positioning Errors: Joint Estimation, Performance Bounds, and GNSS-Based Evaluation Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11142 Radio maps enable environment-aware wireless and Internet-of-Things applications and can be constructed from location-tagged received signal strength (RSS) measurements collected by mobile devices. 12. Bidirectional Wireless Communication for Weakly Coupled Implantable Brain-Computer Interfaces Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.05786 Implantable brain-computer interfaces (BCIs) promise transformative societal impact, from restoring lost motor, sensory, and speech function in patients with paralysis, stroke, sclerosis, and sensory deficits to serving as a high-bandwidth conduit between human cognition and machine intelligence. 13. AudioICL-Bench: A Benchmark for Large Audio Language Model In-Context Learning Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11252 In-context learning (ICL) promises training-free adaptation for audio, where labeling every new condition is costly. 14. Seamless Whole Slide Label-Free Virtual Staining Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10914 Label-free virtual staining offers a compelling, non-destructive alternative to standard histopathology; however, its clinical adoption is hindered by the computational bottlenecks inherent to processing gigapixel Whole Slide Images (WSIs). 15. Trellis-Based Noise Modulation with Soft-Decision Viterbi Detection Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11184 Noise modulation conveys information through the statistical properties of noise rather than conventional deterministic signal parameters. 16. Multimodal Large Language Model-guided Constrained Optimization for RAN Intelligent Control Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.06122 Artificial intelligence (AI)-based radio access network (RAN) controllers are commonly designed for predefined operating scenarios and optimization tasks, limiting their adaptability when network conditions and operator requirements change after deployment. 17. Not All Attacks Are Learned Equally in Speech Deepfake Detection Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11763 Speech deepfake detection (SDD) models are trained on multi-attack datasets containing diverse spoofing systems, such as text-to-speech (TTS) and voice conversion (VC). 18. Exponential Pixelating Integral transform with dual fractal features for enhanced chest X-ray abnormality detection Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10988 The heightened prevalence of respiratory disorders, particularly exacerbated by a significant upswing in fatalities due to the novel coronavirus, underscores the critical need for early detection and timely intervention. 19. X-RACE: XAI-assisted Recurrent neural network Attribution for Channel Estimation Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11211 Deep learning models, notably Long Short-Term Memory (LSTM), have demonstrated promising performance in channel estimation for high-mobility vehicular environments. 20. Fixed-Time Integral Reinforcement Learning for Saturated Nonlinear Multi-Agent Systems Under FDI Attacks Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.06163 The leader-follower formation control problem is investigated for nonlinear multi-agent systems with unknown dynamics, external disturbances, and false data injection (FDI) attacks on actuator channels. 21. Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural Understanding Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11772 Cross-cultural understanding has become increasingly important in today's highly connected, cross-national world. 22. Rethinking Handwritten Character Recognition Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.10572 Non-Latin handwritten character recognition (HCR) remains understudied. 23. Sparse Approximation via Polynomial Equations Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11215 We consider the problem of finding sparse solutions of an underdetermined linear system $Ax=b$. 24. Non-parametric Formal Synthesis of Unknown Stochastic Systems: Asymptotic Convergence Guarantees Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.06206 Data-driven techniques have shown promising potential for checking behavior of complex systems operating in safety-critical domains against safety and other temporal requirements. 25. RetroThinker: Enabling Retrospective Thinking in Speech LLMs Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2609.11864 Speech large language models (SpeechLLMs) offer reduced latency and retain paralinguistic nuances that are typically lost in cascaded automatic speech recognition (ASR) and text-based LM architectures. 26. Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.11310 We address few-shot object detection with vision-language models (VLMs) in out-of-domain settings such as aerial, industrial, and medical imagery, using only ten annotated images for supervision. 27. Rethinking Radiomap Blind Prediction with Limited Environment and Configuration Representations Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11255 Radiomap blind prediction infers radiomaps from observable representations of the propagation environment and base station (BS) configuration without field measurements. 28. Rethinking Safety for Generalist Robots Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.06326 Generalist robots promise to transform our society: the same system that prepares a meal or folds laundry might also repair a car, inspect infrastructure, or care for a loved one. 29. Cyclic MPDR Beamforming for Suppression of Almost-Cyclostationary Acoustic Interference Source: eess.AS (Audio and Speech Processing) Link: https://arxiv.org/abs/2510.18391 Conventional acoustic beamformers typically assume short-time stationarity and process frequency bins independently, ignoring inter-frequency correlations. 30. UBone3D: Physics-Rectified Conditional Flow Matching for Anatomical 3D Shape Completion from Ultrasound Source: eess.IV (Image and Video Processing) Link: https://arxiv.org/abs/2609.11506 Three-dimensional ultrasound (US) is a safe, radiation-free complementary modality to CT and X-rays for longitudinal monitoring, yet its segmentation-derived partial point clouds are extremely artifact-laden. 31. Exact Bayesian Tracking of Dynamic Network Topologies Source: eess.SP (Signal Processing) Link: https://arxiv.org/abs/2609.11263 Tracking the temporal evolution of network topologies is a fundamental challenge in social networks, epidemiology, and sensor systems, among others. 32. A Theory of Information Architecture for Networked Decisions: Freshness, Locality, and Coordination Source: eess.SY (Systems and Control) Link: https://arxiv.org/abs/2609.06351 Networked systems face a tradeoff between the scope of the information behind a decision and its freshness: a broader view of the system supports better coordination, but assembling and communicating it takes time, so it arrives older. Sources in this brief: eess.AS (Audio and Speech Processing); eess.IV (Image and Video Processing); eess.SP (Signal Processing); eess.SY (Systems and Control). Selected 32 of 409 available items for this weekly brief.