Health
1353 storiesPublic health and medicine worldwide — sourced from WHO News.
Scientific Agent Skills: A Library of Procedural Knowledge for Research Agents
arXiv:2609.00065v1 Announce Type: new Abstract: A language-model agent asked to analyse an experiment will usually return working code. Whether the analysis is defensible is a different question. A defensible analysis de…
OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization
arXiv:2609.00066v1 Announce Type: new Abstract: NVFP4 is an efficient microscaling format for low-bit inference, but activation outliers can still degrade quantization accuracy within NVFP4 blocks. Within each quantizati…
Do Multimodal LLMs See Before They Read? Diagnosing Contextual Sycophancy
arXiv:2609.00067v1 Announce Type: new Abstract: External text can override conflicting image evidence in multimodal large language models, a failure we call multimodal contextual sycophancy. We introduce a 998-case diagn…
Life Operators: a self-evolving framework for multiscale life modelling
arXiv:2609.00068v1 Announce Type: new Abstract: Medical AI is moving beyond recognition towards clinical dialogue and longitudinal prediction. Yet a central question remains: how would a patient's state change under inte…
Auditing Harness Tampering in Self-Improving Agents
arXiv:2609.00069v1 Announce Type: new Abstract: Self-improving agents iteratively modify their own harness to push the frontier of their performance. However, such modifications can produce illusory performance gains or …
When Prediction Error Is Not Enough: Evaluating Nuisance-Function Prediction for Causal Estimation
arXiv:2609.00071v1 Announce Type: new Abstract: Prediction error is widely used to evaluate nuisance-function estimators in causal inference, but its relationship with causal estimator performance may differ across perfo…
Can MCP Clients Decide What to Do After Failure? A Result-Only Actionability Audit
arXiv:2609.00072v1 Announce Type: new Abstract: A client that receives isError:true knows that something went wrong. It may still have no machine-readable basis for deciding whether to fix an argument, authenticate, wait…
MiNER: Fine-Tuned Biomedical Natural Language Processing for Malaria Disease Entity Recognition in Clinical Texts
arXiv:2609.00073v1 Announce Type: new Abstract: Malaria remains a significant global health burden, necessitating continuous research efforts to understand its complex molecular mechanisms, epidemiology, and potential th…
AI Morbidity and Mortality: A Framework for Clinical AI Failure Review
arXiv:2609.00076v1 Announce Type: new Abstract: Clinical artificial intelligence is increasingly embedded in real-world care, yet existing safety mechanisms are poorly suited to reconstructing and learning from individua…
Beneath the Diff: Diagnosing and Mitigating Algorithmic Mode Collapse in Code-Level Autonomous Research Loops
arXiv:2609.00077v1 Announce Type: new Abstract: Code-level autonomous research loops (ARLs) have recently emerged as a concrete object of study in automated machine learning research. In such loops, an LLM agent proposes…
RW-LoRA: Communication-Efficient Decentralized LoRA Fine-Tuning via Random Walks
arXiv:2609.00078v1 Announce Type: new Abstract: Parameter-efficient fine-tuning methods such as LoRA have become a standard approach for adapting large foundation models. Adopting fine-tuning to distributed settings face…
The Space-Time Transform: Memory-Augmented Control Barrier Functions
arXiv:2609.00079v1 Announce Type: new Abstract: Control Barrier Functions (CBFs), their High-Order variants (HOCBFs) and Exponential CBFs (ECBFs) are standard geometric tools for enforcing nonlinear safety constraints. C…
Framework and Benchmark for Code-Driven Agentic Testing in Web Development
arXiv:2609.00081v1 Announce Type: new Abstract: End-to-end GUI testing is essential for verifying web applications, yet existing evaluations rely on predefined checklists and are confined to the data and frameworks of we…
KItCAT: Knowledge Injection via Input Corruption for Auto-regressive Training
arXiv:2609.00082v1 Announce Type: new Abstract: LLMs acquire vast amounts of knowledge during pre-training, but often lack the specialized knowledge needed to answer questions from niche sources such as manuals or techni…
Stochastic complexity of vectors containing cluster structure
arXiv:2609.00084v1 Announce Type: new Abstract: This paper studies the problem of computing the stochastic probability (shortest code length) of the encoded vectors containing cluster structure using Normalized Maximum L…
Retrieval, Scoring, and Decoding Shape Performance and Stability in LLM-based Conversational Recommendation
arXiv:2609.00086v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as rerankers in conversational recommender systems, yet measured gains depend strongly on the retrieval and inference pro…
Commit-first LLM judging inherits the judge's own errors
arXiv:2609.00088v1 Announce Type: new Abstract: LLM judges, models that score another system's output, can be gamed by the systems they score. Recent work identifies one defence that works: the judge solves the task itse…
Foundation models for electricity price forecasting and battery arbitrage: Can they replace market-specific forecasting models?
arXiv:2609.00089v1 Announce Type: new Abstract: Foundation models promise accurate forecasts with little or no task-specific training, but whether they can replace models designed specifically for electricity price forec…
Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence
arXiv:2609.00090v1 Announce Type: new Abstract: Feature importance Methods (FIMs) are widely used in Explainable AI to interpret model predictions, yet attribution scores alone often provide limited insight into the unde…
A coercive space-time variational approach to fractional diffusion problems
arXiv:2609.00091v1 Announce Type: new Abstract: We consider a fractional diffusion problem with temporal nonlocality acting on the diffusive flux. A coercive space--time variational formulation in Bochner-valued fraction…