Giter Site home page Giter Site logo

xymfei / videocrafter Goto Github PK

View Code? Open in Web Editor NEW

This project forked from ailab-cvc/videocrafter

0.0 0.0 0.0 169.97 MB

高质量的文本2视频和图像2视频制模型-A Toolkit for Text-to-Video Generation and Editing

Home Page: https://youtu.be/SJ_TOVjn5zs

Shell 0.48% Python 99.52%

videocrafter's Introduction

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Discord GitHub

🔥🔥 The VideoCrafter1 for high-quality video generation are now released! Please Join us and create your own film on Discord/Floor33.

Floor33 | Film

IMAGE ALT TEXT HERE

🔆 Introduction

🤗🤗🤗 VideoCrafter is an open-source video generation and editing toolbox for crafting video content.
It currently includes the Text2Video and Image2Video models:

1. Generic Text-to-video Generation

Click the GIF to access the high-resolution video.

"A girl is looking at the camera smiling. High Definition." "an astronaut running away from a dust storm on the surface of the moon, the astronaut is running towards the camera, cinematic"
"A giant spaceship is landing on mars in the sunset. High Definition." "A blue unicorn flying over a mystical land"

2. Generic Image-to-video Generation

"a black swan swims on the pond" "a girl is riding a horse fast on grassland" "a boy sits on a chair facing the sea" "two galleons moving in the wind at sunset"

📝 Changelog

  • [2023.10.13]: 🔥🔥 Release the VideoCrafter1, High Quality Video Generation!

  • [2023.08.14]: Release a new version of VideoCrafter on Discord/Floor33. Please join us to create your own film!

  • [2023.04.18]: Release a VideoControl model with most of the watermarks removed!

  • [2023.04.05]: Release pretrained Text-to-Video models, VideoLora models, and inference code.


⏳ Models

Models Resolution Checkpoints
Text2Video 576x1024 Hugging Face
Image2Video 320x512 Hugging Face

⚙️ Setup

1. Install Environment via Anaconda (Recommended)

conda create -n videocrafter python=3.8.5
conda activate videocrafter
pip install -r requirements.txt

💫 Inference

1. Text-to-Video

  1. Download pretrained T2V models via Hugging Face, and put the model.ckpt in checkpoints/base_1024_v1/model.ckpt.
  2. Input the following commands in terminal.
  sh scripts/run_text2video.sh

2. Image-to-Video

  1. Download pretrained I2V models via Hugging Face, and put the model.ckpt in checkpoints/i2v_512_v1/model.ckpt.
  2. Input the following commands in terminal.
  sh scripts/run_image2video.sh

📋 Techinical Report

⏳⏳⏳ Comming soon. We are still working on it.💪

😉 Citation

The technical report is currently unavailable as it is still in preparation. You can cite the paper of our base model, on which we built our applications.

@article{he2022lvdm,
      title={Latent Video Diffusion Models for High-Fidelity Long Video Generation}, 
      author={Yingqing He and Tianyu Yang and Yong Zhang and Ying Shan and Qifeng Chen},
      year={2022},
      eprint={2211.13221},
      archivePrefix={arXiv},
      primaryClass={cs.CV}
}

🤗 Acknowledgements

Our codebase builds on Stable Diffusion. Thanks the authors for sharing their awesome codebases!

📢 Disclaimer

We develop this repository for RESEARCH purposes, so it can only be used for personal/research/non-commercial purposes.


videocrafter's People

Contributors

yzhang2016 avatar yingqinghe avatar scutpaul avatar vinthony avatar eltociear avatar mayuelala avatar

Recommend Projects

  • React photo React

    A declarative, efficient, and flexible JavaScript library for building user interfaces.

  • Vue.js photo Vue.js

    🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.

  • Typescript photo Typescript

    TypeScript is a superset of JavaScript that compiles to clean JavaScript output.

  • TensorFlow photo TensorFlow

    An Open Source Machine Learning Framework for Everyone

  • Django photo Django

    The Web framework for perfectionists with deadlines.

  • D3 photo D3

    Bring data to life with SVG, Canvas and HTML. 📊📈🎉

Recommend Topics

  • javascript

    JavaScript (JS) is a lightweight interpreted programming language with first-class functions.

  • web

    Some thing interesting about web. New door for the world.

  • server

    A server is a program made to process requests and deliver data to clients.

  • Machine learning

    Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.

  • Game

    Some thing interesting about game, make everyone happy.

Recommend Org

  • Facebook photo Facebook

    We are working to build community through open source technology. NB: members must have two-factor auth.

  • Microsoft photo Microsoft

    Open source projects and samples from Microsoft.

  • Google photo Google

    Google ❤️ Open Source for everyone.

  • D3 photo D3

    Data-Driven Documents codes.