Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
LMMs-Lab-Audio
community
Activity Feed
Request to join this org
Follow
5
AI & ML interests
Feeling and building the multimodal intelligence
Recent Activity
kcz358
authored
a paper
4 days ago
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
mwxely
authored
a paper
5 days ago
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
kcz358
authored
a paper
28 days ago
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
View all activity
Team members
4
models
0
None public yet
datasets
23
Sort: Recently updated
lmms-lab-audio/timit-tts
Updated
Feb 15
•
8
lmms-lab-audio/song-describer
Viewer
•
Updated
Feb 13
•
1.85k
•
38
lmms-lab-audio/europal-asr
Viewer
•
Updated
Feb 13
•
215
•
25
lmms-lab-audio/WenetSpeech
Updated
Sep 23, 2025
•
212
lmms-lab-audio/voicebench
Viewer
•
Updated
Aug 29, 2025
•
20.6k
•
587
•
1
lmms-lab-audio/StepEval-Audio-Paralinguistic
Viewer
•
Updated
Aug 25, 2025
•
550
•
210
lmms-lab-audio/Librispeech-concat
Viewer
•
Updated
Apr 6, 2025
•
177
•
40
lmms-lab-audio/Omni_Bench_fix
Viewer
•
Updated
Mar 31, 2025
•
1.13k
•
352
•
1
lmms-lab-audio/mmau
Viewer
•
Updated
Mar 17, 2025
•
10k
•
1.96k
•
1
lmms-lab-audio/fleurs
Viewer
•
Updated
Feb 5, 2025
•
2.41k
•
383
View 23 datasets