Health
1353 storiesPublic health and medicine worldwide — sourced from WHO News.
Designing Proactive Thought Partners for Writing
arXiv:2609.01588v1 Announce Type: new Abstract: Writing involves diverse cognitive activities, from ideation to revision, and writers' needs vary across individuals and moments. Proactive AI promises to provide the right…
StudentSim: Training LLM-based Student Simulators
arXiv:2609.01591v1 Announce Type: new Abstract: AI tutors are most useful when they adapt to each student's strengths, weaknesses, and preferred guidance, but evidence about which guidance works for which student is spar…
Facet-0: A Robotic Foundation Model for Contact-Rich Precise Manipulation
arXiv:2609.01596v1 Announce Type: new Abstract: Real-world robotic assembly at sub-millimeter tolerances demands spatial precision, compliant interaction, and robustness to contact failures. We present Facet-0, a robotic…
The Rise of Verbal Reinforcement Learning
arXiv:2609.01597v1 Announce Type: new Abstract: Natural language is emerging as a primary feedback channel for improving language agents, capable of conveying intent, preferences, and causal structure in forms interpreta…
UI-VISA: U-Net Initialized Vascular Image Segmentation Architecture
arXiv:2609.01598v1 Announce Type: new Abstract: Accurate segmentation of vascular structures in digital subtraction angiography (DSA) images remains challenging due to the thin, elongated, and branching nature of blood v…
CordisBench: Can Language Models Reason About Component Lifecycles in Dynamic Agent Harnesses?
arXiv:2609.01600v1 Announce Type: new Abstract: Dynamic agent harnesses let language models change the software that shapes their own execution. This flexibility brings a new reasoning burden: a local plugin change can p…
Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation
arXiv:2609.01601v1 Announce Type: new Abstract: The repository-level code generation task requires synthesizing code that satisfies task requirements while remaining consistent with the target repository context. Since r…
Efficient SWE Agent Benchmarking via Trajectory-Aware Evaluation
arXiv:2609.01603v1 Announce Type: new Abstract: Evaluating software engineering agents on realistic benchmarks is costly, since each task may require multi-step code exploration, modification, and test execution. Existin…
Beyond Scores: Understanding LLM-as-a-Judge Mechanisms in Summarization Evaluation
arXiv:2609.01604v1 Announce Type: new Abstract: LLM-based evaluators of natural language generation (NLG) quality are widely deployed as scoring tools and as automated training signals, yet the internal procedure by whic…
Uncovering Understanding-Generation Synergy in Native Unified Multimodal Models: From Representation, Task to System
arXiv:2609.01607v1 Announce Type: new Abstract: While unified multimodal models (UMMs) jointly perform visual understanding and generation within a single model, functional unification does not guarantee learning synergy…
Structural Bias Beyond Homophily: A Study of Fairness in Link Prediction
arXiv:2602.11802v2 Announce Type: cross Abstract: Graph link prediction (LP) plays a critical role in socially impactful applications such as job recommendation and friendship formation, making fairness a critical concer…
A survey of AI-generated voices and their detection
arXiv:2608.15411v1 Announce Type: cross Abstract: The ability of artificial intelligence (AI) models to generate highly realistic human voices has advanced rapidly. These technologies power accessibility tools, virtual a…
Stochastic Estimation of Transduced Language Models
arXiv:2608.27428v1 Announce Type: cross Abstract: Transduced language models (TLMs) compose a pretrained \emph{source} language model with a functional finite-state transducer to induce a language model over \emph{target…
RealSWE: A Compositional Evaluation of Coding Agents under Realistic User Requests
arXiv:2608.27831v2 Announce Type: cross Abstract: Coding agents are now commonly evaluated on the SWE-bench family of benchmarks, whose tasks are built from curated GitHub issues: long, structured, and information-rich. …
InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information
arXiv:2608.29632v1 Announce Type: cross Abstract: Competitive programming is increasingly being used to evaluate the algorithmic reasoning capabilities of large language models (LLMs). However, existing benchmarks primar…
AMINA: The Inclusive and Accountable AI for Marginalized Immigrant Nonprofit Assistance
arXiv:2608.30084v1 Announce Type: cross Abstract: Immigrant-led nonprofit groups, particularly those operating in politically sensitive contexts, face exclusion from formal registries and digital platforms. This paper re…
TPR-Attention for Combinatorial Generalization
arXiv:2608.30124v1 Announce Type: cross Abstract: Systematic generalization remains a significant challenge in deep learning. In particular, combinatorial generalization - generalizing to new configurations of known fact…
Toward a social psychology of AI: language-model agents reproduce human-like minimal-group bias
arXiv:2609.00009v1 Announce Type: cross Abstract: Language-model agents now interact in groups, but evaluations that probe memorised stereotype content or use models to simulate people leave this social behaviour unmeasu…
AutoXRD: Autonomous LLM Agents and Comprehensive Evaluation for Powder Diffraction Analysis
arXiv:2609.00070v1 Announce Type: cross Abstract: Powder X-ray diffraction (XRD) is central to materials characterization, yet reliable end-to-end automation remains challenging. An XRD agent must interpret diffraction e…
Provably Efficient Federated Reinforcement Learning with Linear Function Approximation and Logarithmic Communication Cost
arXiv:2609.00193v1 Announce Type: cross Abstract: We study federated online reinforcement learning with linear function approximation. While recent multi-agent reinforcement learning algorithms achieve strong regret guar…