GenmoGenmo builds open-source, state-of-the-art video generation models. Try Mochi 1 in the browser or run it locally via GitHub and HuggingFace.(0)0
Soundscape CommunityOpen-source iOS app by Microsoft Research that uses spatialized audio to help visually impaired users navigate their surroundings with confidence.(0)0
Genmo AIGenmo is an AI research lab building open, state-of-the-art video generation models. Try Mochi 1, their cutting-edge text-to-video model, in the browser or run it locally for free.(0)0
NVIDIA FourCastNetNVIDIA FourCastNet uses Spherical Fourier Neural Operators to deliver global weather forecasts 1,000x faster than traditional NWP models, powered by PyTorch and the ERA5 dataset.(0)0
Coqui AICoqui AI was an open-source TTS and voice cloning platform featuring the XTTS model, supporting 17+ languages and voice cloning from just 3 seconds of audio.(0)0
Tortoise TTSTortoise TTS is a free, open-source multi-voice TTS system emphasizing realistic prosody and intonation. Clone voices and generate high-quality speech with this research-grade Python library.(0)0
OpenVoiceOpenVoice is a free, open-source voice cloning AI by MIT and MyShell that supports zero-shot cross-lingual cloning, tone color replication, and granular voice style control.(0)0
Nari Labs DiaDia is a 1.6B parameter open-source text-to-speech model that generates ultra-realistic dialogue with emotion control, audio conditioning, and nonverbal sounds in one pass.(0)0
Comma AI openpilotopenpilot is an open-source AI driver assistance system supporting 325+ car models. Get lane centering, adaptive cruise, automated lane changes, and dashcam — plug in and drive.(0)0
Qwen AIQwen is Alibaba's open-source family of foundation models covering LLMs, image generation, translation, and AI safety. Available on Hugging Face, GitHub, ModelScope, and via API.(0)0