Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 57,481 to 57,490 of 227,388 articles

SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery Attacks

arXiv
Latent-based watermarks, integrated into the generation process of latent diffusion models (LDMs), simplify detection and attribution of generated images. However, recent black-box forgery attacks, where an attacker needs at least one watermarked ima... read more 

PsychePass: Calibrating LLM Therapeutic Competence via Trajectory-Anchored Tournaments

arXiv
While large language models show promise in mental healthcare, evaluating their therapeutic competence remains challenging due to the unstructured and longitudinal nature of counseling. We argue that current evaluation paradigms suffer from an unanch... read more 

MMSF: Multitask and Multimodal Supervised Framework for WSI Classification and Survival Analysis

arXiv
Multimodal evidence is critical in computational pathology: gigapixel whole slide images capture tumor morphology, while patient-level clinical descriptors preserve complementary context for prognosis. Integrating such heterogeneous signals remains c... read more 

Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models

arXiv
Text-to-image (T2I) models have achieved remarkable success in generating high-fidelity images, but they often fail in handling complex spatial relationships, e.g., spatial perception, reasoning, or interaction. These critical aspects are largely ove... read more 

TABED: Test-Time Adaptive Ensemble Drafting for Robust Speculative Decoding in LVLMs

arXiv
Speculative decoding (SD) has proven effective for accelerating LLM inference by quickly generating draft tokens and verifying them in parallel. However, SD remains largely unexplored for Large Vision-Language Models (LVLMs), which extend LLMs to pro... read more 

RAW-Flow: Advancing RGB-to-RAW Image Reconstruction with Deterministic Latent Flow Matching

arXiv
RGB-to-RAW reconstruction, or the reverse modeling of a camera Image Signal Processing (ISP) pipeline, aims to recover high-fidelity RAW data from RGB images. Despite notable progress, existing learning-based methods typically treat this task as a di... read more 

LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning

arXiv
Large Language Models (LLMs) can be fine-tuned on domain-specific data to enhance their performance in specialized fields. However, such data often contains numerous low-quality samples, necessitating effective data processing (DP). In practice, DP s... read more 

Let's Roll a BiFTA: Bi-refinement for Fine-grained Text-visual Alignment in Vision-Language Models

arXiv
Recent research has shown that aligning fine-grained text descriptions with localized image patches can significantly improve the zero-shot performance of pre-trained vision-language models (e.g., CLIP). However, we find that both fine-grained text d... read more 

Assembling the Mind's Mosaic: Towards EEG Semantic Intent Decoding

arXiv
Enabling natural communication through brain-computer interfaces (BCIs) remains one of the most profound challenges in neuroscience and neurotechnology. While existing frameworks offer partial solutions, they are constrained by oversimplified semanti... read more 

Exploiting the Final Component of Generator Architectures for AI-Generated Image Detection

arXiv
With the rapid proliferation of powerful image generators, accurate detection of AI-generated images has become essential for maintaining a trustworthy online environment. However, existing deepfake detectors often generalize poorly to images produce... read more