Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
153.4
TFLOPS
G
MadGoatHaz
7
2
3
Follow
0 followers
·
3 following
AI & ML interests
None yet
Recent Activity
new
activity
3 days ago
Vishva007/Qwen3.8-27B-W4A16-AutoRound-GPTQ:
Issue & Fix: vLLM Speculative Decoding (MTP) Fails Due to Quantized Draft Head
updated
a collection
3 days ago
Qwen3.8-27b
new
activity
3 days ago
Frozenlock/Qwen3.8-27B-int4-AutoRound:
Great stuff on Ampere
View all activity
Organizations
None yet
MadGoatHaz
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
Vishva007/Qwen3.8-27B-W4A16-AutoRound-GPTQ
3 days ago
Issue & Fix: vLLM Speculative Decoding (MTP) Fails Due to Quantized Draft Head
👍
1
#1 opened 3 days ago by
MadGoatHaz
New activity in
Frozenlock/Qwen3.8-27B-int4-AutoRound
3 days ago
Great stuff on Ampere
🚀
1
#2 opened 3 days ago by
MadGoatHaz
New activity in
bottlecapai/ThinkingCap-Qwen3.6-27B-FP8
about 1 month ago
Deployment Note: vLLM JIT Path Resolution -nvidia/cu13/lib/-
3
#5 opened about 2 months ago by
MadGoatHaz
New activity in
bottlecapai/ThinkingCap-Qwen3.6-27B-FP8
about 2 months ago
Fix for vLLM Loading Crash on Ampere/Ada GPUs (TP=8, RuntimeError: size_n = 12)
4
#2 opened about 2 months ago by
MadGoatHaz
vllm loaded with error
4
#1 opened about 2 months ago by
wac81
New activity in
Xingyu-Zheng/Qwen3.6-27B-INT8-FOEM
4 months ago
MTP not accepting tokens
2
#1 opened 4 months ago by
MadGoatHaz
MTP not accepting tokens
2
#1 opened 4 months ago by
MadGoatHaz
Load more