Visual Basic Component Object Model

T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy

Note: This model has been trained for approximately 2.7M steps (batch size = 1) and is still in the training process. I have attached a .ipynb file in the repository. You can refer to it to know how ...

IEEE

Stereoscopic Visual Attention Model Based on Bioinspiration

Abstract: A stereoscopic visual attention model predicts the regions that people focus on most when viewing stereoscopic images, holding significant application value in the fields of robot vision, ...

IEEE

Object-Aware Image Augmentation for Audio-Visual Zero-Shot Learning

Abstract: Audio-visual zero-shot learning (ZSL) leverages both video and audio information for model training, aiming to classify new video categories that were not seen during the training. However, ...

Some results have been hidden because they may be inaccessible to you

Show inaccessible results

T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy

Stereoscopic Visual Attention Model Based on Bioinspiration

Object-Aware Image Augmentation for Audio-Visual Zero-Shot Learning

Trending now