M MagicDanceMagicDance uses identity-aware diffusion to transfer human poses and facial expressions onto any target in video — zero fine-tuning required. Open-source ICML 2024 research.(0)0
A AnimateDiffAnimateDiff is an open-source framework that injects a motion modeling module into personalized Stable Diffusion models to generate animated images from text prompts—no fine-tuning required.(0)0
CogVideo AICogVideo AI is an open-source framework for generating videos from text and images, featuring CogVideoX (2024) and CogVideo (ICLR 2023) with fine-tuning support via CogKit.(0)0
OpenCollarOpenCollar is an open-source initiative providing customizable hardware and software for wildlife tracking collars, enabling affordable conservation monitoring for a wide range of species.(0)0
Milvus AI Vector DBMilvus is an open-source vector database built for GenAI applications. Perform high-speed similarity searches and scale to tens of billions of vectors with minimal performance loss.(0)0
MLflowMLflow is the largest open source AI engineering platform. Debug, evaluate, monitor, and deploy AI agents, LLMs, and ML models with 30M+ monthly downloads.(0)0
NVIDIA FLARENVIDIA FLARE is an open-source, domain-agnostic federated learning SDK that enables privacy-preserving distributed AI model training across multiple data sources without sharing raw data.(0)0
Nari Labs DiaDia is a 1.6B parameter open-source text-to-speech model that generates ultra-realistic dialogue with emotion control, audio conditioning, and nonverbal sounds in one pass.(0)0
NVIDIA FourCastNetNVIDIA FourCastNet uses Spherical Fourier Neural Operators to deliver global weather forecasts 1,000x faster than traditional NWP models, powered by PyTorch and the ERA5 dataset.(0)0
OmniParserOmniParser by Microsoft is an open-source tool that parses UI screenshots into structured elements, enabling accurate vision-based GUI automation agents powered by GPT-4V and other VLMs.(0)0