MOVA
A foundation model that generates synchronized video and audio in a single model
About
MOVA (MOSS Video and Audio) generates video and synchronized audio in a single model. This repository provides model weights, inference, training, LoRA fine-tuning, and evaluation workflows.
Open Source Health
- Stars
- 1,111
- Forks
- 91
- License
- Apache-2.0
- Last commit
- 27 days ago
Resources & Links
Related Categories
Vendor
OpenMOSS (SII)
Open-source projects on GitHub: MOSS-Transcribe-Diarize, MOSS-TTSD and MOVA
Quick Links
Open Source
More by OpenMOSS (SII)
Related Products
Peft
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
Top category match
Sophia Script For Windows
zap: The most powerful PowerShell module for fine-tuning Windows 10 & Windows 11 on GitHub
Top category match
KTransformers
A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
Top category match
LongWriter
[ICLR 2025] LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs
Top category match
Regress Lm
Library for sequence-to-sequence numeric prediction, applicable to any tokenizable input, and allows pretraining and fine-tuning over multiple tasks.
Top category match