Implementation of Phenaki Video, which uses Mask GIT to produce text-guided videos of up to 2 minutes in length, in Pytorch. It will also combine another technique involving a token critic for potentially even better generations. A new paper suggests that instead of relying on the predicted probabilities of each token as a measure of confidence, one can train an extra critic to decide what to iteratively mask during sampling. This repository will also endeavor to allow the researcher to train on text-to-image and then text-to-video. Similarly, for unconditional training, the researcher should be able to first train on images and then fine tune on video.
Features
- Implementation of Phenaki Video
- Uses Mask GIT to produce text guided videos
- Videos of up to 2 minutes in length
- Combines techniques involving a token critic for potentially even better generations
- You can optionally train this critic for potentially better generations
- This repository will also endeavor to allow the researcher to train on text-to-image and then text-to-video
License
MIT LicenseFollow Phenaki - Pytorch
Other Useful Business Software
AI-generated apps that pass security review
Retool lets you generate dashboards, admin panels, and workflows directly on your data. Type something like “Build me a revenue dashboard on my Stripe data” and get a working app with security, permissions, and compliance built in from day one. Whether on our cloud or self-hosted, create the internal software your team needs without compromising enterprise standards or control.
Rate This Project
Login To Rate This Project
User Reviews
Be the first to post a review of Phenaki - Pytorch!