Audio Visual Training Online

Audio-Visual Target Speaker Extraction With Selective Auditory Attention

Abstract: Audio-visual target speaker extraction (AV-TSE) aims to extract the specific person's speech from the audio mixture given auxiliary visual cues. Previous methods usually search for the ...

IEEE

MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations

Abstract: Despite significant progress in Vision-Language Pre-training (VLP), current approaches predominantly emphasize feature extraction and cross-modal comprehension, with limited attention to ...

GitHub

Audio-Visual Instance Segmentation

In this paper, we propose a new multi-modal task, termed audio-visual instance segmentation (AVIS), which aims to simultaneously identify, segment and track individual sounding object instances in ...

Frontiers

The effects of stroboscopic visual training on human cognitive function and motor performance: a systematic review

1 College of Exercise and Health, Shenyang Sport University, Shenyang, China 2 Sports, Exercise and Brain Sciences Laboratory, Sports Coaching College, Beijing Sport University, Beijing, China This ...

Some results have been hidden because they may be inaccessible to you

Show inaccessible results