DataoceanAI/Chinese_Male_Speech_Synthesis_Corpus_Live_Streaming_for_Sales Updated Jan 10, 2025 • 47 • 5
Baghdad99/saad-speech-recognition-hausa-audio-to-text Automatic Speech Recognition • 0.2B • Updated Nov 5, 2023 • 17 • 12
InternW0-Δ: A World Action Model Bridging Predictive Dynamics and Actions with 20K+ Hours of Open Data Paper • 2609.31394 • Published 6 days ago • 27
Lajavaness/wav2vec2-lg-xlsr-fr-speech-emotion-recognition Audio Classification • 0.3B • Updated Apr 25, 2024 • 458 • 18
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 7 days ago • 32
DataoceanAI/American_English_Male_Speech_Synthesis_Corpus_Gentle_and_Mature_Aged_30_40 Updated Jan 10, 2025 • 36 • 7
DataoceanAI/Chinese_Female_Speech_Synthesis_Corpus_Live_Streaming_for_Sales_with_Multi_Styles Updated Jan 10, 2025 • 48 • 8
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 8 days ago • 18
MemoryAthena: Adaptive Routing over Latent and Generated Memories Paper • 2609.25853 • Published 9 days ago • 11
JEV-as-a-Judge: Accept When Confident, Escalate When Unsure Paper • 2609.26550 • Published 9 days ago • 42