Latest AI and machine learning research in adhd/add for healthcare professionals.
We present a motion-adaptive temporal attention mechanism for parameter-efficient video generation built upon frozen Stable Diffusion models. Rather than treating all video content uniformly, our method dynamically adjusts temporal attention receptive fields based on estimated motion content: high-motion sequences attend locally across frames to preserve rapidly changing details, while low-motion ...
We present a training-free framework for continuous and controllable image editing at test time for text-conditioned generative models. In contrast to prior approaches that rely on additional training or manual user intervention, we find that a simple steering in the text-embedding space is sufficient to produce smooth edit control. Given a target concept (e.g., enhancing photorealism or changing ...
The progressive automation of transport promises to enhance safety and sustainability through shared mobility. Like other vehicles and road users, and...
The remarkable realism of images generated by diffusion models poses critical detection challenges. Current methods utilize reconstruction error as a ...
Brain imaging classification is commonly approached from two perspectives: modeling the full image volume to capture global anatomical context, or con...
Uncovering the genetic architecture of quantitative traits is challenging because polygenic control yields small individual gene effects and because g...
Vision transformers have demonstrated remarkable success in classification by leveraging global self-attention to capture long-range dependencies. How...
Estimating the 6D pose of objects from a single RGB image is a critical task for robotics and extended reality applications. However, state-of-the-art...
Inhibition is a core cognitive control function whose competence is distributed across the population, with more extreme impairments in psychiatric co...
The image purification strategy constructs an intermediate distribution with aligned anatomical structures, which effectively corrects the spatial mis...
Single-image 3D generation with part-level structure remains challenging: learned priors struggle to cover the long tail of part geometries and mainta...
Generating diagnostic text from histopathology whole slide images (WSIs) is challenging due to the gigapixel scale of the input and the requirement fo...
When learning to find the most beneficial course of action, the prefrontal cortex guides decisions by comparing estimates of the relative value of the...
Depression is a heterogeneous disorder, often diagnosed based on symptom co-occurrence. However, individuals may present with markedly different sympt...
Large language models (LLMs) continue to struggle with knowledge-intensive questions that require up-to-date information and multi-hop reasoning. Augm...
Self-supervised learning (SSL) methods based on Siamese networks learn visual representations by aligning different views of the same image. The multi...
We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs). JEPAs...
Multiple-instance Learning (MIL) is commonly used to undertake computational pathology (CPath) tasks, and the use of multi-scale patches allows divers...
Dopamine (DA) has been implicated in exploration-exploitation behaviour, i.e., exploring novel, potentiallybetter options vs. exploiting known, previo...
Data science agents promise to accelerate discovery and insight-generation by turning data into executable analyses and findings. Yet existing data sc...