Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
LMMs-Lab-Audio
community
Activity Feed
Request to join this org
Follow
5
AI & ML interests
Feeling and building the multimodal intelligence
Recent Activity
kcz358
authored
a paper
3 days ago
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
mwxely
authored
a paper
4 days ago
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
kcz358
authored
a paper
27 days ago
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
View all activity
Team members
4
models
0
None public yet
datasets
23
Sort: Recently updated
lmms-lab-audio/timit-tts
Updated
Feb 15
•
8
lmms-lab-audio/song-describer
Viewer
•
Updated
Feb 13
•
1.85k
•
38
lmms-lab-audio/europal-asr
Viewer
•
Updated
Feb 13
•
215
•
24
lmms-lab-audio/WenetSpeech
Updated
Sep 23, 2025
•
181
lmms-lab-audio/voicebench
Viewer
•
Updated
Aug 29, 2025
•
20.6k
•
604
•
1
lmms-lab-audio/StepEval-Audio-Paralinguistic
Viewer
•
Updated
Aug 25, 2025
•
550
•
215
lmms-lab-audio/Librispeech-concat
Viewer
•
Updated
Apr 6, 2025
•
177
•
40
lmms-lab-audio/Omni_Bench_fix
Viewer
•
Updated
Mar 31, 2025
•
1.13k
•
356
•
1
lmms-lab-audio/mmau
Viewer
•
Updated
Mar 17, 2025
•
10k
•
1.87k
•
1
lmms-lab-audio/fleurs
Viewer
•
Updated
Feb 5, 2025
•
2.41k
•
372
View 23 datasets