Sihan Chen

Results 3 repositories owned by Sihan Chen

VAST

194
Stars
12
Forks
Watchers

Code and Model for VAST: A Vision-Audio-Subtitle-Text Omni-Modality Foundation Model and Dataset

COSA

36
Stars
1
Forks
Watchers

Codes and Models for COSA: Concatenated Sample Pretrained Vision-Language Foundation Model

VALOR

233
Stars
14
Forks
Watchers

Codes and Models for VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset