Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Inference Optimization
community
Activity Feed
Follow
105
AI & ML interests
None defined yet.
Recent Activity
nm-research
updated
a model
2 days ago
inference-optimization/Qwen3-8B-speculator.dspark.swa.gammatv.2048anc.muon.regencollection.selfdistill-v3-epoch2
nm-research
published
a model
2 days ago
inference-optimization/Qwen3-8B-speculator.dspark.swa.gammatv.2048anc.muon.regencollection.selfdistill-v3-epoch2
nm-research
updated
a model
2 days ago
inference-optimization/Qwen3-8B-speculator.dspark.swa.gammatv.2048anc.muon.regencollection.selfdistill-v3-epoch1
View all activity
Team members
16
inference-optimization
's models
431
Sort: Recently updated
inference-optimization/DeepSeek-V3-debug-multiply
1B
•
Updated
Jan 23
•
12
inference-optimization/Qwen3-0.6B-debug-add-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
3
inference-optimization/Qwen3-0.6B-debug-multiply-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
3
inference-optimization/Qwen3-0.6B-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
178
inference-optimization/Qwen3-0.6B-debug-add-W4A16-G128
0.6B
•
Updated
Jan 23
•
3
inference-optimization/Qwen3-0.6B-debug-multiply-W4A16-G128
0.6B
•
Updated
Jan 23
•
7
inference-optimization/Qwen3-0.6B-W4A16-G128
0.6B
•
Updated
Jan 23
•
483
inference-optimization/Qwen3-0.6B-debug-add
0.6B
•
Updated
Jan 23
•
6
inference-optimization/Qwen3-0.6B-debug-multiply
0.6B
•
Updated
Jan 23
•
4
inference-optimization/DeepSeek-V3-debug-empty
1B
•
Updated
Jan 23
•
1.13k
inference-optimization/Qwen3-Next-80B-A3B-Thinking-quantized.w8a8
Updated
Dec 24, 2025
Previous
1
...
13
14
15
Next