MusicGen MetaAudioCraft is Meta AI's open-source framework for generative audio, including MusicGen (text-to-music), AudioGen (text-to-sound), and EnCodec (neural audio codec).(0)0
A AIOZ-GDANCEAIOZ-GDANCE is a large-scale open-source dataset and AI model for generating coherent group dance choreographies from music, covering 7 dance styles and 16 music genres. Published at CVPR 2023.(0)0
GenmoGenmo builds open-source, state-of-the-art video generation models. Try Mochi 1 in the browser or run it locally via GitHub and HuggingFace.(0)0
Qwen AIQwen is Alibaba's open-source family of foundation models covering LLMs, image generation, translation, and AI safety. Available on Hugging Face, GitHub, ModelScope, and via API.(0)0
A AnimateDiffAnimateDiff is an open-source framework that injects a motion modeling module into personalized Stable Diffusion models to generate animated images from text prompts—no fine-tuning required.(0)0
M MagicDanceMagicDance uses identity-aware diffusion to transfer human poses and facial expressions onto any target in video — zero fine-tuning required. Open-source ICML 2024 research.(0)0
Genmo AIGenmo is an AI research lab building open, state-of-the-art video generation models. Try Mochi 1, their cutting-edge text-to-video model, in the browser or run it locally for free.(0)0
InsightFaceInsightFace is a state-of-the-art open-source library for face detection, recognition, alignment, 3D reconstruction, and attribute analysis. MIT licensed and production-ready.(0)0
BrightbandBrightband builds open-source AI Earth System models for probabilistic weather and climate forecasting, providing tools and benchmarks for academia, government, and enterprise.(0)0
Text Generation WebUIRun large language models locally with Text Generation WebUI — supports text, vision, tool-calling, and fine-tuning. 100% offline, open source, and private.(0)0