1
stars
0
commits
10
repos using this model
linked in READMEs
Jun 7, 2024
updated
EvolvingLMMs-Lab/lmms-eval
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
4,399
orca-wm/Orca
Orca: The World is in Your Mind
1,020
facebookresearch/tuna-2
Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding…
755
wenhaochai/MovieChat
[CVPR 2024] MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
706
XiaomiMiMo/MiMo-Embodied
MiMo-Embodied
405
ML-GSAI/LLaDA-V
351
VARGPT-family/VARGPT
VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language…
326
jacklishufan/LaViDa
Official Implementation of LaViDa: :A Large Diffusion Language Model for Multimodal Understanding
229
ustcwhy/BitVLA
Official implementation for BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
165
SusungHong/MusicInfuser
Official implementation of the paper "MusicInfuser: Making Video Diffusion Listen and Dance"…
87
zhang-tao-whu/sa2va_eval
3