MCG-NJU/VideoChat3-4B
Video-Text-to-Text • 4B • Updated • 2.07k • 20
Computer Vision; Video Understanding; Action Recognition
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding