Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Inference Optimization
community
Activity Feed
Follow
61
AI & ML interests
None defined yet.
Recent Activity
nm-research
updated
a model
about 3 hours ago
inference-optimization/Qwen3-8B-speculator.dflash.swa.dpace.fullvocab.adamw.2048anc-qwen235b-instr-bs16-v3-avg121314
nm-research
published
a model
about 3 hours ago
inference-optimization/Qwen3-8B-speculator.dflash.swa.dpace.fullvocab.adamw.2048anc-qwen235b-instr-bs16-v3-avg121314
nm-research
updated
a model
about 9 hours ago
inference-optimization/Qwen3-8B-speculator.dflash.swa.dpace.fullvocab.adamw.2048anc-qwen235b-instr-bs16-v3-10ep-ckpt14
View all activity
Team members
20
inference-optimization
's models
281
Sort: Recently updated
inference-optimization/gpt-oss-120b-ckpt3-speculator.eagle3
0.9B
•
Updated
Mar 12
•
5
inference-optimization/Qwen3-Coder-Next.w4a16
Text Generation
•
12B
•
Updated
Mar 12
•
468
•
1
inference-optimization/sarvam-105b-FP8-Dynamic
Text Generation
•
106B
•
Updated
Mar 9
•
8
inference-optimization/sarvam-30b-FP8-Dynamic
Text Generation
•
32B
•
Updated
Mar 9
•
7
•
1
inference-optimization/sarvam-30b-NVFP4
Text Generation
•
19B
•
Updated
Mar 9
•
16
•
1
inference-optimization/sarvam-105b-NVFP4
61B
•
Updated
Mar 9
•
14
•
1
inference-optimization/Qwen3.5-35B-A3B-FP8-Dynamic
35B
•
Updated
Mar 6
•
2
inference-optimization/gpt-oss-20b-FP8-Dynamic
21B
•
Updated
Mar 5
•
5
•
1
inference-optimization/Qwen3-30B-A3B-Instruct-2507-FP8-Block
31B
•
Updated
Mar 4
•
3
inference-optimization/Ministral-3-14B-Instruct-2512.w8a8
Updated
Feb 4
•
3
inference-optimization/Ministral-3-14B-Instruct-2512.w4a16
Updated
Feb 3
•
191
inference-optimization/Meta-Llama-3-8B-Instruct-NVFP4-GPTQ-Quant
5B
•
Updated
Jan 29
•
4
inference-optimization/Meta-Llama-3-8B-Instruct-NVFP4-GPTQ-MSE
5B
•
Updated
Jan 29
•
5
inference-optimization/DeepSeek-V3-debug-multiply-FP8_DYNAMIC
1B
•
Updated
Jan 24
•
7
inference-optimization/DeepSeek-V3-debug-add-FP8_DYNAMIC
1B
•
Updated
Jan 24
•
6
inference-optimization/DeepSeek-V3-debug-empty-FP8_DYNAMIC
1B
•
Updated
Jan 23
•
34k
inference-optimization/DeepSeek-V3-debug-multiply-NVFP4A16
0.9B
•
Updated
Jan 23
•
4
inference-optimization/DeepSeek-V3-debug-add-NVFP4A16
0.9B
•
Updated
Jan 23
•
1
inference-optimization/DeepSeek-V3-debug-empty-NVFP4A16
0.9B
•
Updated
Jan 23
•
706
inference-optimization/DeepSeek-V3-debug-add
1B
•
Updated
Jan 23
•
4
inference-optimization/DeepSeek-V3-debug-multiply
1B
•
Updated
Jan 23
•
3
inference-optimization/Qwen3-0.6B-debug-add-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
4
inference-optimization/Qwen3-0.6B-debug-multiply-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
2
inference-optimization/Qwen3-0.6B-FP8_BLOCK
0.6B
•
Updated
Jan 23
•
534
inference-optimization/Qwen3-0.6B-debug-add-W4A16-G128
0.6B
•
Updated
Jan 23
•
2
inference-optimization/Qwen3-0.6B-debug-multiply-W4A16-G128
0.6B
•
Updated
Jan 23
•
2
inference-optimization/Qwen3-0.6B-W4A16-G128
0.6B
•
Updated
Jan 23
•
783
inference-optimization/Qwen3-0.6B-debug-add
0.6B
•
Updated
Jan 23
•
4
inference-optimization/Qwen3-0.6B-debug-multiply
0.6B
•
Updated
Jan 23
•
5
inference-optimization/DeepSeek-V3-debug-empty
1B
•
Updated
Jan 23
•
1.86k
Previous
1
...
7
8
9
10
Next