VLM lmms-lab/LLaVA-Video-7B-Qwen2 Video-Text-to-Text • 8B • Updated Oct 25, 2024 • 13.6k • 127 lmms-lab/VideoDetailCaption Viewer • Updated Jun 18, 2024 • 499 • 2.53k • 12 TIGER-Lab/VLM2Vec-LoRA Text Generation • Updated Jul 13, 2025 • 554 • 11 microsoft/Phi-3.5-vision-instruct Image-Text-to-Text • 4B • Updated Dec 10, 2025 • 747k • 739
VLM lmms-lab/LLaVA-Video-7B-Qwen2 Video-Text-to-Text • 8B • Updated Oct 25, 2024 • 13.6k • 127 lmms-lab/VideoDetailCaption Viewer • Updated Jun 18, 2024 • 499 • 2.53k • 12 TIGER-Lab/VLM2Vec-LoRA Text Generation • Updated Jul 13, 2025 • 554 • 11 microsoft/Phi-3.5-vision-instruct Image-Text-to-Text • 4B • Updated Dec 10, 2025 • 747k • 739