Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
2140.8
TFLOPS
Michael Goin
mgoin
105
21
22
Follow
brwoodside's profile picture
olka-fi's profile picture
htdung167's profile picture
51 followers
·
16 following
mgoin_
mgoin
AI & ML interests
LLM inference optimization, compression, quantization, pruning, distillation
Recent Activity
published
a model
1 day ago
mgoin/Kimi-K3-pruned75
updated
a model
1 day ago
mgoin/Kimi-K3-pruned75
published
a model
1 day ago
mgoin/Kimi-K3-pruned50
View all activity
Organizations
mgoin
's models
109
Sort: Recently updated
mgoin/Kimi-K3-pruned75
Image-Text-to-Text
•
737B
•
Updated
1 day ago
•
267
mgoin/Kimi-K3-pruned50
Image-Text-to-Text
•
1.4T
•
Updated
1 day ago
•
7
mgoin/GLM-5.2-speculator.dspark-block16
Text Generation
•
4B
•
Updated
29 days ago
•
422
•
2
mgoin/Qwen3-8B-speculator.dspark-reasoning
Text Generation
•
2B
•
Updated
Jun 28
•
94
mgoin/Qwen3-8B-speculator.dflash
Text Generation
•
2B
•
Updated
Jun 28
•
14
mgoin/Qwen3-8B-speculator.dspark
Text Generation
•
2B
•
Updated
Jun 28
•
95
mgoin/Qwen3.6-35B-A3B-2Bit-GSQ-ct
Image-Text-to-Text
•
35B
•
Updated
May 27
•
11
mgoin/Qwen3-0.6B-MXFP8
0.6B
•
Updated
Feb 16
•
7.3k
mgoin/GLM-4.6-FP8-BLOCK
Text Generation
•
357B
•
Updated
Feb 10
•
12
mgoin/Qwen3-0.6B-NVFP4
0.6B
•
Updated
Aug 26, 2025
•
475
mgoin/mlperf-inference-llama3.1-8b-data
Updated
Jul 15, 2025
mgoin/Llama-3.1-8B-Instruct-FP8-BLOCK
8B
•
Updated
Jul 1, 2025
•
17
mgoin/SEMIKONG-70B-W4A16-G128
71B
•
Updated
Jun 16, 2025
•
4
mgoin/llama-4-tiny-random
Text Generation
•
6.69M
•
Updated
May 14, 2025
•
11
mgoin/Qwen1.5-14B-Chat-GPTQ
Text Generation
•
Updated
Mar 5, 2025
•
4
mgoin/pixtral-12b
Image-Text-to-Text
•
13B
•
Updated
Feb 7, 2025
•
522
•
1
mgoin/Llama-3.2-1B-Instruct-FP8-ATTN
1B
•
Updated
Dec 23, 2024
•
6
mgoin/Llama-3.2-1B-Instruct-FP8-dynamic-ATTN
1B
•
Updated
Dec 23, 2024
•
5
mgoin/Pixtral-Large-Instruct-2411
Updated
Nov 19, 2024
mgoin/Qwen2.5-Coder-32B-Instruct-fp8
Updated
Nov 13, 2024
mgoin/nemotron-3-8b-chat-4k-sft-hf
Text Generation
•
9B
•
Updated
Nov 13, 2024
•
115
mgoin/llava-onevision-qwen2-7b-ov-hf-bnb-full-4bit
Image-Text-to-Text
•
8B
•
Updated
Nov 5, 2024
•
8
mgoin/MiniCPM-Llama3-V-2_5-int4
Visual Question Answering
•
9B
•
Updated
Oct 31, 2024
•
5
mgoin/DeepSeek-Coder-V2-Lite-Instruct-FP8
16B
•
Updated
Sep 20, 2024
•
9
mgoin/Mixtral-8x7B-Instruct-v0.1-FP8
47B
•
Updated
Sep 20, 2024
•
4
mgoin/Nemotron-nemo-checkpoints
Updated
Aug 30, 2024
mgoin/Minitron-4B-Base-FP8
Text Generation
•
4B
•
Updated
Aug 16, 2024
•
17
•
3
mgoin/Nemotron-4-340B-Base-hf
Text Generation
•
341B
•
Updated
Aug 8, 2024
•
11
•
1
mgoin/Nemotron-4-340B-Instruct-hf-FP8
Text Generation
•
341B
•
Updated
Aug 8, 2024
•
12
•
3
mgoin/Nemotron-4-340B-Base-hf-FP8
Text Generation
•
341B
•
Updated
Aug 8, 2024
•
332
•
2
Previous
1
2
3
4
Next