UniVL is a video-language pretrain model. It is designed with four modules and five objectives for both video language understanding and generation tasks. It is also a flexible model for most of the multimodal downstream tasks considering both efficiency and effectiveness.
Features
- Finetune on YoucookII
- Documentation available
- Examples available
- Run caption task on YoucookII
- Pretrain on HowTo100M
- Licensed under the MIT License
License
MIT LicenseFollow UniVL
Other Useful Business Software
Streamline Azure Security with Palo Alto Networks VM-Series
Improve your security posture and reduce incident response time. Use the VM-Series to natively analyze Azure traffic and dynamically drive policy updates based on workload changes.
Rate This Project
Login To Rate This Project
User Reviews
Be the first to post a review of UniVL!