Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
wangyueqian
/
MMDuet
like
4
Video-Text-to-Text
PEFT
Safetensors
wangyueqian/MMDuetIT
English
llava-onevision
llava
multimodal
online video understanding
video understanding
arxiv:
2411.17991
License:
mit
Model card
Files
Files and versions
Community
Use this model
main
MMDuet
/
tokenizer_config.json
Commit History
Upload 9 files
b6413e7
verified
wangyueqian
commited on
30 days ago