Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
LMMs-Lab-Audio
community
Activity Feed
Request to join this org
Follow
5
AI & ML interests
Feeling and building the multimodal intelligence
Recent Activity
kcz358
authored
a paper
4 days ago
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
kcz358
authored
a paper
2 months ago
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning
mwxely
submitted
a paper
2 months ago
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning
View all activity
Team members
4
models
0
None public yet
datasets
23
Sort: Recently updated
lmms-lab-audio/timit-tts
Updated
Feb 15
•
18
lmms-lab-audio/song-describer
Viewer
•
Updated
Feb 13
•
1.85k
•
64
lmms-lab-audio/europal-asr
Viewer
•
Updated
Feb 13
•
215
•
17
lmms-lab-audio/WenetSpeech
Updated
Sep 23, 2025
•
291
lmms-lab-audio/voicebench
Viewer
•
Updated
Aug 29, 2025
•
20.6k
•
1.23k
lmms-lab-audio/StepEval-Audio-Paralinguistic
Viewer
•
Updated
Aug 25, 2025
•
550
•
33
lmms-lab-audio/Librispeech-concat
Viewer
•
Updated
Apr 6, 2025
•
177
•
116
lmms-lab-audio/Omni_Bench_fix
Viewer
•
Updated
Mar 31, 2025
•
1.13k
•
255
•
1
lmms-lab-audio/mmau
Viewer
•
Updated
Mar 17, 2025
•
10k
•
2.52k
•
1
lmms-lab-audio/fleurs
Viewer
•
Updated
Feb 5, 2025
•
2.41k
•
351
View 23 datasets