Running on Zero Agents 227 MiniMax H3 Turbo LoRA 🎬 227 Video generation with a synchronized soundtrack
ibm-granite/granite-speech-5.0-470m-turboctc Automatic Speech Recognition • 0.5B • Updated 13 days ago • 30k • 63
firdhokk/speech-emotion-recognition-with-facebook-wav2vec2-large-xlsr-53 Audio Classification • 0.3B • Updated Dec 15, 2024 • 855 • 2
nDimensional/Qwen3.5-9B-Uncensored-Safetensors Image-Text-to-Text • 9B • Updated Apr 2 • 1.43k • • 10
Running on Zero MCP Featured 84 Breeze TTS 2 🎙 84 Bilingual TTS with voice design, cloning, and direction
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation Paper • 2502.03930 • Published Feb 6, 2025 • 3