State Required CME

Identifying and Reporting Child abuse

Latest AI and machine learning research in identifying and reporting child abuse for healthcare professionals.

4,379 articles
Stay Ahead - Weekly Identifying and Reporting Child abuse research updates
Subscribe
Browse Categories
Showing 361-380 of 4,379 articles

InstructEngine: Instruction-driven Text-to-Image Alignment

Reinforcement Learning from Human/AI Feedback (RLHF/RLAIF) has been extensively utilized for preference alignment of text-to-image models. Existing methods face certain limitations in terms of both data and algorithm. For training data, most approaches rely on manual annotated preference data, either by directly fine-tuning the generators or by training reward models to provide training signals....

Probability Distribution Alignment and Low-Rank Weight Decomposition for Source-Free Domain Adaptive Brain Decoding

Brain decoding currently faces significant challenges in individual differences, modality alignment, and high-dimensional embeddings. To address individual differences, researchers often use source subject data, which leads to issues such as privacy leakage and heavy data storage burdens. In modality alignment, current works focus on aligning the softmax probability distribution but neglect the ...

All-in-Memory Stochastic Computing using ReRAM

As the demand for efficient, low-power computing in embedded and edge devices grows, traditional computing methods are becoming less effective for h...

ColorizeDiffusion v2: Enhancing Reference-based Sketch Colorization Through Separating Utilities

Reference-based sketch colorization methods have garnered significant attention due to their potential applications in the animation production indu...

iEBAKER: Improved Remote Sensing Image-Text Retrieval Framework via Eliminate Before Align and Keyword Explicit Reasoning

Recent studies focus on the Remote Sensing Image-Text Retrieval (RSITR), which aims at searching for the corresponding targets based on the given qu...

FontGuard: A Robust Font Watermarking Approach Leveraging Deep Font Knowledge

The proliferation of AI-generated content brings significant concerns on the forensic and security issues such as source tracing, copyright protecti...

Sample-level Adaptive Knowledge Distillation for Action Recognition

Knowledge Distillation (KD) compresses neural networks by learning a small network (student) via transferring knowledge from a pre-trained large net...

Recurrent Feature Mining and Keypoint Mixup Padding for Category-Agnostic Pose Estimation

Category-agnostic pose estimation aims to locate keypoints on query images according to a few annotated support images for arbitrary novel classes. ...

RomanTex: Decoupling 3D-aware Rotary Positional Embedded Multi-Attention Network for Texture Synthesis

Painting textures for existing geometries is a critical yet labor-intensive process in 3D asset generation. Recent advancements in text-to-image (T2...

LaMOuR: Leveraging Language Models for Out-of-Distribution Recovery in Reinforcement Learning

Deep Reinforcement Learning (DRL) has demonstrated strong performance in robotic control but remains susceptible to out-of-distribution (OOD) states...

A Study into Investigating Temporal Robustness of LLMs

Large Language Models (LLMs) encapsulate a surprising amount of factual world knowledge. However, their performance on temporal questions and histor...

Multi-Granular Multimodal Clue Fusion for Meme Understanding

With the continuous emergence of various social media platforms frequently used in daily life, the multimodal meme understanding (MMU) task has been...

Aerial Vision-and-Language Navigation with Grid-based View Selection and Map Construction

Aerial Vision-and-Language Navigation (Aerial VLN) aims to obtain an unmanned aerial vehicle agent to navigate aerial 3D environments following huma...

Natural Humanoid Robot Locomotion with Generative Motion Prior

Natural and lifelike locomotion remains a fundamental challenge for humanoid robots to interact with human society. However, previous methods either...

"Principal Components" Enable A New Language of Images

We introduce a novel visual tokenization framework that embeds a provable PCA-like structure into the latent token space. While existing visual toke...

GAS-NeRF: Geometry-Aware Stylization of Dynamic Radiance Fields

Current 3D stylization techniques primarily focus on static scenes, while our world is inherently dynamic, filled with moving objects and changing e...

Personalized Code Readability Assessment: Are We There Yet?

Unreadable code could be a breeding ground for errors. Thus, previous work defined approaches based on machine learning to automatically assess code...

SeCap: Self-Calibrating and Adaptive Prompts for Cross-view Person Re-Identification in Aerial-Ground Networks

When discussing the Aerial-Ground Person Re-identification (AGPReID) task, we face the main challenge of the significant appearance variations cause...

MPTSNet: Integrating Multiscale Periodic Local Patterns and Global Dependencies for Multivariate Time Series Classification

Multivariate Time Series Classification (MTSC) is crucial in extensive practical applications, such as environmental monitoring, medical EEG analysi...

V$^2$Dial: Unification of Video and Visual Dialog via Multimodal Experts

We present V$^2$Dial - a novel expert-based model specifically geared towards simultaneously handling image and video input data for multimodal conv...

Browse Categories