Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
GadflyII
GadflyII
23
1
Follow
AlexR7's profile picture
21world's profile picture
Michalea's profile picture
30 followers
·
1 following
AI & ML interests
None yet
Recent Activity
new
activity
7 days ago
Motif-Technologies/Motif-3:
Custom vLLM merge request to main vLLM?
new
activity
7 days ago
Motif-Technologies/Motif-3-NVFP4:
Please submit a PR to vLLM for upstream model support?
liked
a model
4 months ago
ibm-granite/granite-4.1-8b-fp8
View all activity
Organizations
GadflyII
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
Motif-Technologies/Motif-3
7 days ago
Custom vLLM merge request to main vLLM?
👍
1
7
#6 opened 8 days ago by
pathosethoslogos
New activity in
Motif-Technologies/Motif-3-NVFP4
7 days ago
Please submit a PR to vLLM for upstream model support?
👍
1
2
#2 opened 8 days ago by
GadflyII
liked
a model
4 months ago
ibm-granite/granite-4.1-8b-fp8
Text Generation
•
9B
•
Updated
Apr 29
•
29.9k
•
14
New activity in
GadflyII/GLM-4.6V-NVFP4
4 months ago
Well done nvfp4 quant
3
#1 opened 7 months ago by
josephbreda
New activity in
GadflyII/Qwen3-Coder-Next-NVFP4
4 months ago
Why Your NVFP4 Model Is Slower Than FP8 on the GB10 (NVIDIA Spark) — And How to Fix It
🤯
👍
5
6
#5 opened 6 months ago by
scottgl
New activity in
GadflyII/GLM-4.7-Flash-MTP-NVFP4
5 months ago
SGLang and MTP
1
#2 opened 6 months ago by
Michalea
New activity in
GadflyII/Qwen3-Coder-Next-NVFP4
6 months ago
Model requests?
12
#4 opened 6 months ago by
pathosethoslogos
New activity in
GadflyII/GLM-4.6V-NVFP4
6 months ago
Fails on a single DGX spark with errors below
1
#2 opened 6 months ago by
Adrian1234
New activity in
GadflyII/GLM-4.7-Flash-MXFP4
6 months ago
Update MXFP4 format to compressed-tensors
1
#3 opened 6 months ago by
mgoin
New activity in
lukealonso/MiniMax-M2.5-NVFP4
6 months ago
Here's the vLLM recipe I'm using with 2x RTX Pro 6000
👍
3
17
#1 opened 6 months ago by
zenmagnets
New activity in
GadflyII/Qwen3-Coder-Next-NVFP4
6 months ago
MMLU PRO Benchmark
3
#3 opened 6 months ago by
sevapru
vLLM 0.16?
1
#2 opened 6 months ago by
MMaxHugg
New activity in
GadflyII/Qwen3-Coder-Next-NVFP4
7 months ago
Memory
1
#1 opened 7 months ago by
struxx
New activity in
GadflyII/GLM-4.7-Flash-NVFP4
7 months ago
confused response
7
#8 opened 7 months ago by
jiangyizhi
updated
a model
7 months ago
GadflyII/Qwen3-Coder-Next-NVFP4
Text Generation
•
Updated
Feb 4
•
59.7k
•
45
published
a model
7 months ago
GadflyII/Qwen3-Coder-Next-NVFP4
Text Generation
•
Updated
Feb 4
•
59.7k
•
45
New activity in
GadflyII/GLM-4.7-Flash-NVFP4
7 months ago
MTP quality, 47 layer
3
#7 opened 7 months ago by
Michalea
updated
a model
7 months ago
GadflyII/GLM-4.7-Flash-MTP-NVFP4
Text Generation
•
19B
•
Updated
Feb 2
•
147
•
5
New activity in
GadflyII/GLM-4.7-Flash-MTP-NVFP4
7 months ago
Upload folder using huggingface_hub
#1 opened 7 months ago by
GadflyII
published
a model
7 months ago
GadflyII/GLM-4.7-Flash-MTP-NVFP4
Text Generation
•
19B
•
Updated
Feb 2
•
147
•
5
Load more