State Required CME

Care of terminally ill / Palliative care

Latest AI and machine learning research in care of terminally ill / palliative care for healthcare professionals.

5,028 articles
Stay Ahead - Weekly Care of terminally ill / Palliative care research updates
Subscribe
Browse Categories
Showing 941-960 of 5,028 articles

Masked Modeling for Human Motion Recovery Under Occlusions

Human motion reconstruction from monocular videos is a fundamental challenge in computer vision, with broad applications in AR/VR, robotics, and digital content creation, but remains challenging under frequent occlusions in real-world settings. Existing regression-based methods are efficient but fragile to missing observations, while optimization- and diffusion-based approaches improve robustness ...

Jan 22 2026 2601.16079v2

Masked Modeling for Human Motion Recovery Under Occlusions

Human motion reconstruction from monocular videos is a fundamental challenge in computer vision, with broad applications in AR/VR, robotics, and digital content creation, but remains challenging under frequent occlusions in real-world settings.Existing regression-based methods are efficient but fragile to missing observations, while optimization- and diffusion-based approaches improve robustness a...

Jan 22 2026 2601.16079v1
LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR

We present \textbf{LightOnOCR-2-1B}, a 1B-parameter end-to-end multilingual vision--language model that converts document images (e.g., PDFs) into cle...

Jan 20 2026 2601.14251v1
Normal Breast Tissue (NBT)-Classifiers: Advancing Compartment Classification in Normal Breast Histology

Background: Cancer research emphasises early detection, yet quantitative methods for analysing normal tissue remain limited. Hematoxylin and eosin (H&...

Generative Scenario Rollouts for End-to-End Autonomous Driving

Vision-Language-Action (VLA) models are emerging as highly effective planning models for end-to-end autonomous driving systems. However, current works...

Jan 16 2026 2601.11475v1
Cell Behavior Video Classification Challenge, a benchmark for computer vision methods in time-lapse microscopy

The classification of microscopy videos capturing complex cellular behaviors is crucial for understanding and quantifying the dynamics of biological p...

Jan 15 2026 2601.10250v1
Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive Reasoning

Vision-Language Navigation aims to enable agents to navigate to a target location based on language instructions. Traditional VLN often follows a clos...

Jan 14 2026 2601.09111v1
LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models

Real-world License Plate Recognition (LPR) faces significant challenges from severe degradations such as motion blur, low resolution, and complex illu...

Jan 14 2026 2601.09116v1
Real-time and accurate stereo matching via tri-fusion volume for stereo vision.

In the field of real-time stereo matching, a concise and informative cost volume is crucial for achieving high efficiency and accuracy. To this end, i...

Sep 1 2025 40367718
A document is worth a structured record: Principled inductive bias design for document recognition

Many document types use intrinsic, convention-driven structures that serve to encode precise and structured information, such as the conventions gov...

IRAF-SLAM: An Illumination-Robust and Adaptive Feature-Culling Front-End for Visual SLAM in Challenging Environments

Robust Visual SLAM (vSLAM) is essential for autonomous systems operating in real-world environments, where challenges such as dynamic objects, low t...

ULC: A Unified and Fine-Grained Controller for Humanoid Loco-Manipulation

Loco-Manipulation for humanoid robots aims to enable robots to integrate mobility with upper-body tracking capabilities. Most existing approaches ad...

Enhancing Synthetic CT from CBCT via Multimodal Fusion and End-To-End Registration

Cone-Beam Computed Tomography (CBCT) is widely used for intraoperative imaging due to its rapid acquisition and low radiation dose. However, CBCT im...

DREAM: Document Reconstruction via End-to-end Autoregressive Model

Document reconstruction constitutes a significant facet of document analysis and recognition, a field that has been progressively accruing interest ...

Flippi: End To End GenAI Assistant for E-Commerce

The emergence of conversational assistants has fundamentally reshaped user interactions with digital platforms. This paper introduces Flippi-a cutti...

MatDecompSDF: High-Fidelity 3D Shape and PBR Material Decomposition from Multi-View Images

We present MatDecompSDF, a novel framework for recovering high-fidelity 3D shapes and decomposing their physically-based material properties from mu...

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization

Camera relocalization, a cornerstone capability of modern computer vision, accurately determines a camera's position and orientation (6-DoF) from im...

Towards Accurate and Efficient 3D Object Detection for Autonomous Driving: A Mixture of Experts Computing System on Edge

This paper presents Edge-based Mixture of Experts (MoE) Collaborative Computing (EMC2), an optimal computing system designed for autonomous vehicles...

RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs

The foundational capabilities of large language models (LLMs) are deeply influenced by the quality of their pre-training corpora. However, enhancing...

2024 NASA SUITS Report: LLM-Driven Immersive Augmented Reality User Interface for Robotics and Space Exploration

As modern computing advances, new interaction paradigms have emerged, particularly in Augmented Reality (AR), which overlays virtual interfaces onto...

Browse Categories