Health
1353 storiesPublic health and medicine worldwide — sourced from WHO News.
Seeing the World and the Self from Egocentric Video
arXiv:2609.01276v1 Announce Type: new Abstract: Complete 3D perception from egocentric video requires recovering the surrounding scene and the wearer's full-body motion in a shared metric frame. Existing methods typicall…
TimeSteer: Inference-Time Speech Scheduling in Joint Audio-Visual Diffusion Models
arXiv:2609.01277v1 Announce Type: new Abstract: Although pretrained joint audio-visual diffusion models offer rich control over \emph{what} to generate, they provide no explicit control over \emph{when} an utterance shou…
Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models
arXiv:2609.01279v1 Announce Type: new Abstract: Emotion is expressed in text along a wide spectrum, from surface lexical cues to inferences entangled with content. Most layer-wise analyses of emotion in LLMs use a single…
EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents
arXiv:2609.01281v1 Announce Type: new Abstract: Vision-language-action (VLA) models map visual observations and language instructions directly to robot actions, but long-horizon tasks require more than action prediction.…
HiLRP: Toward One Trustworthy Explanation for Vision Transformer: Conservation-Valid Attribution via Attention Primitives
arXiv:2609.01282v1 Announce Type: new Abstract: Vision Transformer (ViT) design has become increasingly diverse, with backbones combining convolutional stems, windowed, linear, or multi-axis attention, patch merging, and…
Sensitivity Oracles for Matroid Packing, Matroid Covering, and Matching Problems with Applications
arXiv:2609.01283v1 Announce Type: new Abstract: Sensitivity oracles preprocess a graph so that queries can be answered after any $f$ edge insertions and deletions, without recomputing from scratch. For structural optimiz…
Analog-DB: An Agent-First Analog Integrated Circuit Database, From Blocks to Systems
arXiv:2609.01286v1 Announce Type: new Abstract: Sharing analog integrated circuit designs remains difficult: foundry non-disclosure agreements restrict the process details a design depends on, and the testbenches behind …
Soft Posterior Speaker Injection for Multi-Talker Speech Recognition
arXiv:2609.01287v1 Announce Type: new Abstract: Multi-talker automatic speech recognition (MT-ASR) remains challenging under overlapping speech. Hard diarization-based segmentation introduces irreversible errors, whereas…
Accelerating the Improved Arrow--Hurwicz Iteration via the Anderson Algorithm for Steady-State Navier--Stokes Equations
arXiv:2609.01288v1 Announce Type: new Abstract: We apply Anderson acceleration to the improved Arrow--Hurwicz (IAH) method for the finite element solution of the steady-state incompressible Navier--Stokes equations. The …
Agentic Multimodal Models for Environmental Hyperspectral Unmixing
arXiv:2609.01289v1 Announce Type: new Abstract: Hyperspectral unmixing is a key task in remote sensing that aims to decompose mixed pixels in hyperspectral images into their constituent material signatures, or endmembers…
Relational Task Generation Language: A Declarative Specification Framework for Relational Deep Learning
arXiv:2609.01292v1 Announce Type: new Abstract: Relational Deep Learning (RDL) has become a powerful paradigm for learning from multi-tabular data. However, manually defining RDL prediction tasks is a laborious process t…
TriSLA: A Preventive and Closed-Loop SLA-Aware Architecture for Multidomain Decision-Making with Explainable Artificial Intelligence in 5G Networks
arXiv:2609.01293v1 Announce Type: new Abstract: Network slicing in multidomain 5G environments introduces critical challenges in guaranteeing Service Level Agreements (SLAs) under dynamic resource variability and heterog…
Explore Before Committing: Hypothesis-Guided Search for Deep Research Agents
arXiv:2609.01294v1 Announce Type: new Abstract: Deep-research agents answer complex questions by interacting with search and browsing tools, yet they often search along a single evolving trajectory. Our trajectory-level …
Lifted-Product QLDPC Codes in the Polynomial Domain
arXiv:2609.01305v1 Announce Type: new Abstract: This paper presents a finite-length polynomial-domain formulation of lifted-product quantum low-density parity-check (QLDPC) codes. We formulate the code construction over …
MeshSplatBench: A Unified Benchmark for Triangle-Based Neural Rendering
arXiv:2609.01306v1 Announce Type: new Abstract: Triangle-based neural rendering bridges neural scene representations and conventional graphics pipelines by optimizing explicit geometric primitives compatible with standar…
CMRVision: A Foundation Model for Cardiac MR Image Analysis
arXiv:2609.01308v1 Announce Type: new Abstract: Cardiac magnetic resonance (CMR) imaging provides complementary information on cardiac anatomy, function, and tissue characterization across multiple sequences and views. I…
One-Layer Transformer Provably Learns Multiclass One-Nearest Neighbor in Context
arXiv:2609.01311v1 Announce Type: new Abstract: We extend recent work establishing an equivalence between one-layer transformers and nearest-neighbor classifiers in the binary setting to the multiclass case. By leveragin…
A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation
arXiv:2609.01315v1 Announce Type: new Abstract: Building an omni-modal foundation model means evaluating it across text, image, video, and audio. Excellent evaluation toolkits exist for each modality, but their inference…
MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval
arXiv:2609.01316v1 Announce Type: new Abstract: Retrieval over visually rich documents has a representation problem: important content often lives in tables, charts, figures, and layout relations that plain OCR linearize…
Reliability Challenges in Diffusion Vision-Language Models
arXiv:2609.01318v1 Announce Type: new Abstract: Diffusion-based Large Vision-Language Models (dLVLMs) have recently emerged as a compelling alternative to autoregressive (AR) LVLMs, offering advantages in parallel decodi…