LLM Reference

CogVideo Models by Zhipu AI

Zhipu AIProprietary
3 models2022

Details

ResearcherZhipu AI
LicenseProprietary
Commercial useCommercial use: conditional
Models3
Released2022

Capabilities

Vision2 of 3 models
Multimodal2 of 3 models

About

CogVideo is a family of 3 AI models by Zhipu AI, released in 2022.

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

3 in view
CogVideoCurrent

Use when the workload needs 9.4 parameters.

2022-129.4 parameters

Use when the workload needs video generation, multimodal inputs, and video.

Unknown releasevideo generationmultimodal inputsvideo

Use when the workload needs video generation, multimodal inputs, and video.

Unknown releasevideo generationmultimodal inputsvideo

Release Timeline

2 release groups
2022-12
1 current
CogVideo
9.4 parameters
Current
Unknown release
2 current
CogVideoX-3
video generationmultimodal inputsvideo
Current
CogVideoX-Flash
video generationmultimodal inputsvideo
Current

Specifications(3 models)

CogVideo model specifications comparison
ModelReleasedParametersVisionMultimodal
CogVideo2022-129.4NoNo
CogVideoX-3YesYes
CogVideoX-FlashYesYes

Available From(1 provider)

Frequently Asked Questions

What is CogVideo used for?
CogVideo is used for video generation, video, and vision and multimodal work. The family description and listed model capabilities point to those workloads as the best fit.
How does CogVideo compare to GLM-5?
CogVideo by Zhipu AI is strongest where you need video generation, while GLM-5 by Zhipu AI is the closest related family to check for coding. CogVideo has 3 listed variants, while GLM-5 reaches up to 1m context, so compare the specs and pricing tables before choosing a production model.
Which CogVideo model should I use?
If price is the main constraint, use the pricing table first because CogVideo does not have complete provider pricing in the local data. For the most capable/latest local choice, evaluate CogVideoX-3 with multimodal inputs.