CogVideoX1.5-5B is an open-source text-to-video generation model analogous to the commercial model QingYing. It is designed to create video based on text prompts, supports the English language, and also offers image-to-video generation (version CogVideoX1.5-5B-I2V). The model is available on platforms such as Hugging Face, ModelScope, and WiseModel.
Features:
Inference Requirements:
The model is a component of the video generation pipeline, consisting of:
Total: ~10.5B parameters
| Model Name | Context | Type | GPU | Status | Link |
|---|
There are no public endpoints for this model yet.
Rent your own physically dedicated instance with hourly or long-term monthly billing.
We recommend deploying private instances in the following scenarios:
| Name | GPU | Generation time, sec. | |||
|---|---|---|---|---|---|
| 1 | $0.53 | Launch | |||
| 1 | $0.83 | Launch | |||
| 1 | $1.02 | Launch | |||
| 1 | $1.59 | Launch | |||
| 1 | $2.37 | Launch | |||
| 1 | $3.83 | Launch | |||
| 1 | $4.11 | Launch | |||
| 1 | $4.74 | Launch | |||
| Name | GPU | Generation time, sec. | |||
|---|---|---|---|---|---|
| 1 | $0.33 | Launch | |||
| 1 | $0.38 | Launch | |||
| 1 | $0.53 | Launch | |||
| 1 | $0.83 | Launch | |||
| 1 | $1.02 | Launch | |||
| 1 | $1.59 | Launch | |||
| 1 | $2.37 | Launch | |||
| 1 | $3.83 | Launch | |||
| 1 | $4.11 | Launch | |||
| 1 | $4.74 | Launch | |||
| Name | GPU | Generation time, sec. | |||
|---|---|---|---|---|---|
| 1 | $0.33 | Launch | |||
| 1 | $0.38 | Launch | |||
| 1 | $0.38 | Launch | |||
| 1 | $0.53 | Launch | |||
| 1 | $0.57 | Launch | |||
| 1 | $0.83 | Launch | |||
| 1 | $1.02 | Launch | |||
| 1 | $1.59 | Launch | |||
| 1 | $2.37 | Launch | |||
| 1 | $3.83 | Launch | |||
| 1 | $4.11 | Launch | |||
| 1 | $4.74 | Launch | |||
Contact our dedicated neural networks support team at nn@immers.cloud or send your request to the sales department at sale@immers.cloud.