Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
wangyueqian
/
MMDuet
like
4
Video-Text-to-Text
PEFT
Safetensors
wangyueqian/MMDuetIT
English
llava-onevision
llava
multimodal
online video understanding
video understanding
arxiv:
2411.17991
License:
mit
Model card
Files
Files and versions
Community
Use this model
main
MMDuet
/
vocab.json
wangyueqian
Upload 9 files
b6413e7
verified
29 days ago
raw
Copy download link
history
contribute
delete
Safe
2.78 MB
File too large to display, you can
check the raw version
instead.