Health
1353 storiesPublic health and medicine worldwide — sourced from WHO News.
Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images
arXiv:2605.12413v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) show strong visual perception, yet remain limited in reasoning about space under changing viewpoints. We study this challenge a…
The Scientific Contribution Graph: Automated Literature-based Technological Roadmapping at Scale
arXiv:2605.15011v3 Announce Type: replace Abstract: Scientific contributions rarely develop in isolation, but instead build upon prior discoveries. We formulate the task of automated technological roadmapping as extracti…
Universal Approximation of Nonlinear Operators and Their Derivatives
arXiv:2605.15285v3 Announce Type: replace Abstract: Establishing Universal Approximation Theorems (UATs) for nonlinear operators and their derivatives is a foundational open problem in Operator Learning (OL) and raises d…
Towards Generalized Image Manipulation Localization via Score-based Model
arXiv:2605.16879v2 Announce Type: replace Abstract: With the rapid evolution of synthetic media, Image Manipulation Localization (IML) has emerged as a critical component in multimedia forensics for ensuring the integrit…
Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road
arXiv:2605.17026v2 Announce Type: replace Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through specialized fine-tun…
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
arXiv:2605.20382v3 Announce Type: replace Abstract: Language models are trained to follow instructions, but they are also powerful pattern completers. What happens when these two objectives conflict? We construct convers…
PromptNCE: Conditional Probabilities and PMI Using Only LLMs and Contrastive Estimation Prompts
arXiv:2605.21776v2 Announce Type: replace Abstract: Estimating mutual information from text usually requires training a task-specific critic, which limits its use in low-data settings. We ask whether large language model…
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
arXiv:2605.23651v4 Announce Type: replace Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of how human-like gen…
LLM-driven design of physics-constrained constitutive models: two agents are better than one
arXiv:2605.23754v2 Announce Type: replace Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics, machine learni…
Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior
arXiv:2605.26797v2 Announce Type: replace Abstract: We study Latent Recurrent Transformer (LRT), a lightweight augmentation of autoregressive transformers that reuses a high-level source-layer hidden state from the previ…
When the Strongest Teacher Is Not the Best Teacher: Student-Centric Answer Selection
arXiv:2605.26872v5 Announce Type: replace Abstract: LLM training increasingly relies on teacher-generated supervision, from synthetic responses to reasoning traces and tool-use demonstrations. Current practice often choo…
UniACE: A Unified Framework for Evaluating LLM Agentic Capabilities
arXiv:2605.27898v3 Announce Type: replace Abstract: Agent benchmarks are increasingly used to compare large language models (LLMs) across domains, yet a reported score reflects a complete model--harness--environment conf…
Argument Quality Assessment with Large Language Models: A Pairwise Bradley-Terry Approach
arXiv:2605.28313v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in tasks related to reasoning and judgment. However, assessing the quality of arguments requires …
Diffusion Large Language Models for Visual Speech Recognition
arXiv:2605.28456v2 Announce Type: replace Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually ambiguous token…
The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic
arXiv:2605.28700v3 Announce Type: replace Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-generated varian…
Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay
arXiv:2605.28782v2 Announce Type: replace Abstract: Discourse particles, such as well and kind of, are crucial components that enable LLMs to "speak" more like humans. They are used to convey emotions, intentions, and in…
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
arXiv:2605.29048v2 Announce Type: replace Abstract: In this paper, we introduce LLMBridge, a new LLM based system for the task of end-to-end referential bridging resolution in English. Our bridging resolution pipeline co…
MusTBench: Benchmarking and Advancing Temporal Grounding in Music LLMs
arXiv:2605.29300v2 Announce Type: replace Abstract: Recent Large Audio-Language Models (LALMs) have demonstrated promising abilities in understanding musical content. However, whether their responses are grounded in the …
PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning
arXiv:2605.29582v2 Announce Type: replace Abstract: Large Language Models (LLMs) show strong potential as educational tutors. Existing approaches typically train them to solve problems and provide correct answers, but th…
On the Application of Hybrid Mixed Domain Decomposition Methods to Permanent Magnet Synchronous Machines
arXiv:2605.31032v2 Announce Type: replace Abstract: In this work, we study the application of a hybrid mixed domain decomposition(HMDD) method for the rotor-stator coupling of a permanent magnet synchronous machine. For …