Health
1353 storiesPublic health and medicine worldwide — sourced from WHO News.
Don't Let the Model Write the YAML: Deterministic, Minimal-Diff GitOps Remediation from LLM-Proposed Field Changes
arXiv:2609.00227v1 Announce Type: new Abstract: LLM agents increasingly diagnose incidents and propose remediations. In a GitOps workflow, applying a fix means editing a version-controlled config file, and the obvious im…
Bridging Lexical Divergence: LLM-Assisted, Cost-Efficient, Zero-shot Scientific Entity Linking
arXiv:2609.00228v1 Announce Type: new Abstract: Scientific domain entity linking (EL) differs from general domain EL because mentions and entity names often lack lexical overlap. Another challenge is that specialized ter…
Beyond Language Priors: Diagnosing and Fixing Visual-Origin Hallucinations in Multimodal LLM
arXiv:2609.00231v1 Announce Type: new Abstract: Existing research on object hallucination in multimodal large language models (MLLMs) predominantly attributes the problem to language priors such as over-reliance on textu…
Beyond Blind Compliance: Benchmarking Task Verification in OCR Reasoning
arXiv:2609.00232v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on OCR-centric document understanding and text-rich visual reasoning benchmarks. Yet existing eval…
An Approach to Asynchronous Unsourced Random Access
arXiv:2609.00236v1 Announce Type: new Abstract: In this work we propose an approach to construct a fully asynchronous unsourced random access communications (URA) system. Besides the common URA features such as the absen…
Learning What to Retain: Gated-Memory Routing for Efficient Collaboration in Multi-Agent LLM Systems
arXiv:2609.00237v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems tackle complex reasoning by orchestrating how multiple agents are configured and how they collaborate. A central challe…
XVAE-WMT: Explainable Wavelet-Temporal Variational Autoencoder for Blind Source Separation of Heart and Lung Sounds
arXiv:2609.00238v1 Announce Type: new Abstract: The separation of cardiovascular sounds is a critical task in biomedical signal processing. In this paper, we introduce XVAE-WMT1, an unsupervised explainable generative AI…
Bounded Relative Boundary Implies Narrow DNF Approximation
arXiv:2609.00240v1 Announce Type: new Abstract: Friedgut conjectured that an increasing family in the $p$-biased discrete cube with bounded relative boundary can be approximated arbitrarily well by one whose minimal elem…
LOOMSUM:Weaving Quantitative and Narrative Evidence for Faithful Long Text-Table Summarization
arXiv:2609.00241v1 Announce Type: new Abstract: Long documents often distribute important information across extensive narrative passages and multiple tables, making faithful summarization particularly challenging. Exist…
CoLT-Drive: Counterfactual Long-Tail Benchmarking and Knowledge-Preserving Adaptation for Driving Affordance Prediction
arXiv:2609.00242v1 Announce Type: new Abstract: Long-tail autonomous driving failures are often framed as rare-object recognition errors. We argue that this view is incomplete: the decision-critical question is not only …
Invalidation Contracts for Cross-Episode Agent Memory
arXiv:2609.00243v1 Announce Type: new Abstract: LLM agents that cache recovery suggestions from API errors can skip re-derivation in later episodes, spending fewer tokens and fewer model calls on constraints they have al…
Beyond Locks and Thread IDs: Static Data Race Detection Off The Beaten Path (Extended Version)
arXiv:2609.00246v1 Announce Type: new Abstract: Maintaining an abstraction of the execution history of threads can improve the precision of data race detection in static analysis. Here, we extend the digest framework to …
Empirical Software Engineering in Practice: Insights from Google
arXiv:2609.00247v1 Announce Type: new Abstract: While it is fairly well known how empirical software engineering (ESE) is used in the academic world, we have limited knowledge of how ESE is practiced in industry. As part…
Authority Bias in Conversational Search Engines for Academic Paper Recommendation
arXiv:2609.00248v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used as conversational search engines for academic literature, yet whether they judge papers on content or on authority signal…
CompanionSim: Synthetic Data for Evaluating Anthropomorphism in Human-AI Relationships
arXiv:2609.00250v1 Announce Type: new Abstract: Many people now see AI systems as not just productivity tools but as social companions. Researchers are eager to study the consequences of AI companionship behaviors, such …
Hypotheses-Guided Self Distillation for Continual Personalization
arXiv:2609.00251v1 Announce Type: new Abstract: As people increasingly interact with LLM assistants in daily life, continually adapting to individual preferences has become essential for effective long-term interactions.…
Spec-Driven Development for Agentic Software Engineering: Harnessing Human-Agent Teamwork
arXiv:2609.00252v1 Announce Type: new Abstract: Context: Software engineering is moving from AI-assisted practices like vibe coding, in which assistants accelerate individual developers, towards Agentic Software Engineer…
NSIDDx: A Design Framework for Neuro-Symbolic, Practitioner-First Differential Diagnosis in Low-Resource Settings
arXiv:2609.00256v1 Announce Type: new Abstract: LLM-based diagnostic systems achieve high semantic accuracy on benchmarks, but open-ended evaluation on clinically uncommon presentations reveals a systematic gap between h…
DUPIN: Attack Learning Is Still Needed! Demonstrating Few-Shot after Unsupervised Pretraining Is A Nimble Forensics Learner
arXiv:2609.00259v1 Announce Type: new Abstract: We propose a novel approach to learning-based attack forensics called DUPIN. DUPIN performs unsupervised pre-training on an enormous amount of audit events in the form of p…
The Answer Is Not the Argument
arXiv:2609.00264v1 Announce Type: new Abstract: Chain-of-thought monitoring is proposed for AI oversight, yet evaluations often provide monitors with a trusted reference answer. We ask whether answer access improves reas…