Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 28,931 to 28,940 of 219,647 articles

Structure-guided molecular design with contrastive 3D protein-ligand learning

arXiv
Structure-based drug discovery faces the dual challenge of accurately capturing 3D protein-ligand interactions while navigating ultra-large chemical spaces to identify synthetically accessible candidates. In this work, we present a unified framework ... read more 

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

arXiv
Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard outputs that are costly or impractical to obtain. Moreover, hallucination detection methods developed f... read more 

RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation

arXiv
Accurate medical image segmentation requires both long-range contextual reasoning and precise boundary delineation, a task where existing transformer- and diffusion-based paradigms are frequently bottlenecked by quadratic computational complexity and... read more 

SmartPhotoCrafter: Unified Reasoning, Generation and Optimization for Automatic Photographic Image Editing

arXiv
Traditional photographic image editing typically requires users to possess sufficient aesthetic understanding to provide appropriate instructions for adjusting image quality and camera parameters. However, this paradigm relies on explicit human instr... read more 

GRAFT: Geometric Refinement and Fitting Transformer for Human Scene Reconstruction

arXiv
Reconstructing physically plausible 3D human-scene interactions (HSI) from a single image currently presents a trade-off: optimization based methods offer accurate contact but are slow (~20s), while feed-forward approaches are fast yet lack explicit ... read more 

MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation

arXiv
Dynamic Scene Graph Generation (DSGG) aims to structurally model objects and their dynamic interactions in video sequences for high-level semantic understanding. However, existing methods struggle with fine-grained relationship modeling, semantic rep... read more 

CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers

arXiv
Graphic design images consist of multiple editable layers, such as text, background, and decorative elements, while most generative models produce rasterized outputs without explicit layer structures, limiting downstream editing. Existing graphic des... read more 

CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation

arXiv
Synthesizing human--object interaction (HOI) videos has broad practical value in e-commerce, digital advertising, and virtual marketing. However, current diffusion models, despite their photorealistic rendering capability, still frequently fail on (i... read more 

InHabit: Leveraging Image Foundation Models for Scalable 3D Human Placement

arXiv
Training embodied agents to understand 3D scenes as humans do requires large-scale data of people meaningfully interacting with diverse environments, yet such data is scarce. Real-world motion capture is costly and limited to controlled settings, whi... read more 

MedFlowSeg: Flow Matching for Medical Image Segmentation with Frequency-Aware Attention

arXiv
Flow matching has recently emerged as a principled framework for learning continuous-time transport maps, enabling efficient deterministic generation without relying on stochastic diffusion processes. While generative modeling has shown promise for m... read more