Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
2140.8
TFLOPS
Michael Goin
mgoin
103
21
22
Follow
smilex1's profile picture
pierrci's profile picture
blobpenguin's profile picture
55 followers
·
16 following
mgoin_
mgoin
AI & ML interests
LLM inference optimization, compression, quantization, pruning, distillation
Recent Activity
updated
a model
18 days ago
mgoin/Qwen3-4B-speculator.dflash2
published
a model
18 days ago
mgoin/Qwen3-4B-speculator.dflash2
updated
a model
25 days ago
mgoin/Qwen3.8-2.4T-A95B-NVFP4-pruned75
View all activity
Organizations
mgoin
's models
113
Sort: Recently updated
mgoin/Qwen3-4B-speculator.dflash2
1B
•
Updated
17 days ago
•
292
•
2
mgoin/Qwen3.8-2.4T-A95B-NVFP4-pruned75
Text Generation
•
349B
•
Updated
24 days ago
•
201
•
1
mgoin/Qwen3.8-2.4T-A95B-NVFP4-pruned94
Text Generation
•
124B
•
Updated
24 days ago
•
171
mgoin/Kimi-K3-pruned75
Image-Text-to-Text
•
737B
•
Updated
Jul 30
•
10.8k
mgoin/Kimi-K3-pruned50
Image-Text-to-Text
•
1.4T
•
Updated
Jul 30
•
231
•
1
mgoin/GLM-5.2-speculator.dspark-preview
Text Generation
•
4B
•
Updated
Jul 7
•
948
•
61
mgoin/GLM-5.2-speculator.dspark-block16
Text Generation
•
4B
•
Updated
Jul 2
•
75
•
2
mgoin/Qwen3-8B-speculator.dspark-reasoning
Text Generation
•
2B
•
Updated
Jun 28
•
40
mgoin/Qwen3-8B-speculator.dflash
Text Generation
•
2B
•
Updated
Jun 28
•
16
mgoin/Qwen3-8B-speculator.dspark
Text Generation
•
2B
•
Updated
Jun 28
•
80
mgoin/Qwen3.6-35B-A3B-2Bit-GSQ-ct
Image-Text-to-Text
•
35B
•
Updated
May 27
•
8
mgoin/Qwen3-0.6B-MXFP8
0.6B
•
Updated
Feb 16
•
1.63k
mgoin/GLM-4.6-FP8-BLOCK
Text Generation
•
357B
•
Updated
Feb 10
•
24
mgoin/Qwen3-0.6B-NVFP4
0.6B
•
Updated
Aug 26, 2025
•
54
mgoin/mlperf-inference-llama3.1-8b-data
Updated
Jul 15, 2025
mgoin/Llama-3.1-8B-Instruct-FP8-BLOCK
8B
•
Updated
Jul 1, 2025
•
4
mgoin/SEMIKONG-70B-W4A16-G128
71B
•
Updated
Jun 16, 2025
•
3
mgoin/llama-4-tiny-random
Text Generation
•
6.69M
•
Updated
May 14, 2025
•
7
mgoin/Qwen1.5-14B-Chat-GPTQ
Text Generation
•
Updated
Mar 5, 2025
•
9
mgoin/pixtral-12b
Image-Text-to-Text
•
13B
•
Updated
Feb 7, 2025
•
644
•
1
mgoin/Llama-3.2-1B-Instruct-FP8-ATTN
1B
•
Updated
Dec 23, 2024
•
5
mgoin/Llama-3.2-1B-Instruct-FP8-dynamic-ATTN
1B
•
Updated
Dec 23, 2024
•
5
mgoin/Pixtral-Large-Instruct-2411
Updated
Nov 19, 2024
•
2
mgoin/Qwen2.5-Coder-32B-Instruct-fp8
Updated
Nov 13, 2024
mgoin/nemotron-3-8b-chat-4k-sft-hf
Text Generation
•
9B
•
Updated
Nov 13, 2024
•
53
mgoin/llava-onevision-qwen2-7b-ov-hf-bnb-full-4bit
Image-Text-to-Text
•
8B
•
Updated
Nov 5, 2024
•
9
mgoin/MiniCPM-Llama3-V-2_5-int4
Visual Question Answering
•
9B
•
Updated
Oct 31, 2024
•
11
mgoin/DeepSeek-Coder-V2-Lite-Instruct-FP8
16B
•
Updated
Sep 20, 2024
•
30
mgoin/Mixtral-8x7B-Instruct-v0.1-FP8
47B
•
Updated
Sep 20, 2024
•
5
mgoin/Nemotron-nemo-checkpoints
Updated
Aug 30, 2024
Previous
1
2
3
4
Next