Latest AI and machine learning research in prescriptions for healthcare professionals.
The sense of agency, the experience of controlling one's actions and their consequences, is a fundamental component of human interaction with autonomous systems. As artificial intelligence increasingly mediates decision-making in domains such as autonomous driving, understanding and monitoring agency-related processes becomes critical for maintaining user engagement, trust, and appropriate levels ...
Editable 3D scene creation requires object instances and lights that can be inspected, moved, and imported into standard engines, yet existing single-image methods largely stop at room-scale geometry, baked/global illumination, or text-driven generation. We introduce Lumera (Light-aware Unified Engine-native Reconstruction and Assembly), a benchmark and reference pipeline for engine-native, light-...
Controllable video generation remains challenging due to the difficulty of specifying precise multi-object interactions using text prompts or motion-c...
Open-world video anomaly detection (OWVAD) is expected to detect events that match a user-specified definition of abnormality. This requirement is str...
Cells sense and integrate extracellular cues through intracellular signaling networks that reshape transcription factor activity to dictate cellular r...
Understanding instrument-tissue interactions is essential for context-aware surgical AI and autonomous robotic surgery. Pretrained vision-language mod...
Global visual localization of unmanned aerial vehicles (UAVs) using remote-sensing reference maps has attracted increasing attention. However, acquisi...
Existing human--object interaction (HOI) video generation methods are largely limited to offline short-video generation with complex driving condition...
Recovering scene-consistent 4D crowd motion from monocular video in large-scale scenes remains challenging due to severe depth ambiguity and complex s...
Bundle adjustment (BA) remains a critical refinement module for image-based 3D reconstruction and continues to improve geometric accuracy even in lear...
Multimodal AI agents increasingly rely on persistent long-term memory to ground generation in past visual and textual episodes. We show that unconditi...
We describe a model of perceptual inference in primary visual cortex (V1) equivalent to a minimal diffusion model whose function can be readily unders...
The increasing deployment of artificial intelligence (AI) assistive systems across healthcare, education, and organisational domains necessitates a de...
The cost of healthcare remains a concern in the United States and may have been influenced by disruptions associated with the COVID-19 pandemic. This ...
Inferring apparent personality from facial images is important in social scenarios for embodied agents in human-robot interaction. Unlike inferring in...
A surveillance camera is an image sensor whose silent physical degradation invalidates every downstream consumer of its data. In-situ integrity alarms...
Genomic foundation models such as Evo 2 learn rich sequence representations, but their value for biosecurity screening is largely unexplored. We ask h...
In operational 1:N face identification, a crucial question arises for each probe: is this person enrolled in the gallery or not? The stakes are high a...
Synthetic image attribution aims at identifying the generator responsible for a given AI-generated image. Training-free reference-based attribution me...
Motivation: Identifying bacterial antimicrobial resistance (AMR) is critical for diagnostics and treatment, but resistance is a complex trait arising ...