18
stars
0
commits
10
repos using this model
2
linked in READMEs
Dec 18, 2024
updated
OpenGVLab/InternVideo2-Chat-8B
26
ziHoHe/VidLaDA-8B
omni-research/Tarsier2-Recap-7b
39
omni-research/Tarsier2-7b-0115
9
chenjoya/videollm-online-8b-v1plus
30
Vchitect/Vchitect-XL-2B
40
stabilityai/sv3d
748
videocon/videocon-model
EvolvingLMMs-Lab/lmms-eval
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
4,399
HITsz-TMG/Uni-MoE
Uni-MoE: Lychee's Large Multimodal Model Family.
1,116
orca-wm/Orca
Orca: The World is in Your Mind
1,020
facebookresearch/tuna-2
Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding…
755
wenhaochai/MovieChat
[CVPR 2024] MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
706
XiaomiMiMo/MiMo-Embodied
MiMo-Embodied
405
ML-GSAI/LLaDA-V
351
VARGPT-family/VARGPT
VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language…
326
jacklishufan/LaViDa
Official Implementation of LaViDa: :A Large Diffusion Language Model for Multimodal Understanding
229
ustcwhy/BitVLA
Official implementation for BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
165
OpenGVLab/InternVideo
[ECCV2024] Video Foundation Models & Data for Multimodal Understanding
2,382
alphaXiv/internvideo-d2a11ea9