LLM Reference
Hugging Face Inference Endpoints

MOVA 720p on Hugging Face Inference Endpoints

MOVA · MOSI AI

Open Source

Last refreshed 2026-06-29. Next refresh: weekly.

Why use MOVA 720p on Hugging Face Inference Endpoints?

Hugging Face Inference Endpoints offers MOVA 720p with competitive pricing. Hugging Face is a leading AI community and platform dedicated to democratizing artificial intelligence.

Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: OpenMOSS-Team/MOVA-720p
Model ID
OpenMOSS-Team/MOVA-720p

Request example

Curated snippets for this provider are not sourced yet. Use Hugging Face Inference Endpoints documentation with model ID OpenMOSS-Team/MOVA-720p.

Gotchas

  • Use provider model ID "OpenMOSS-Team/MOVA-720p", not the LLMReference slug "mova-720p".

Capabilities

VisionMultimodalAudio

About MOVA 720p

MOVA 720p is the higher-resolution open-weight MOVA checkpoint for synchronized video-audio generation. MOSI AI and the OpenMOSS Team describe MOVA as a 32B-parameter mixture-of-experts model with 18B active parameters during inference, designed for native image-to-video-audio and text-to-video-audio generation with synchronized audio, lip sync, and sound effects.

Get Started

Model Specs

Released2026-01-29
Parameters32B total / 18B active
ArchitectureMixture of Experts

Related Models on Hugging Face Inference Endpoints