Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
24.1
TFLOPS
Xuan-Son Nguyen
ngxson
86
120
65
P(doom)
100%
Follow
drl13's profile picture
sacphu1207's profile picture
Quazim0t0's profile picture
672 followers
·
80 following
https://blog.ngxson.com
ngxson
ngxson
ngxson
ngxson.hf.co
AI & ML interests
Doing AI for fun, not for profit
Recent Activity
updated
a collection
3 days ago
Decision models
updated
a collection
3 days ago
Decision models
updated
a Space
3 days ago
ngxson/wllama
View all activity
Organizations
ngxson
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
ggml-org/gguf-my-repo
4 months ago
Fix Git Cloning
#209 opened 4 months ago by
usermma
New activity in
ggml-org/gemma-4-12B-it-GGUF
4 months ago
new model?
😔
👀
2
#1 opened 4 months ago by
DaBaiTu
GGUF advertises wrong max context?
6
#2 opened 4 months ago by
coder543
New activity in
ggml-org/Qwen3-1.7B-GGUF
5 months ago
Qwen3-1.7B-f16.gguf unsafe ?
1
#1 opened 5 months ago by
jwa56
New activity in
zai-org/GLM-4.7-Flash
9 months ago
llama.cpp inference - 20 times (!) slower than OSS 20 on a RTX 5090
➕
1
9
#12 opened 9 months ago by
cmp-nct
New activity in
ngxson/GLM-4.7-Flash-GGUF
9 months ago
Model is... funny lol
❤️
1
3
#2 opened 9 months ago by
ramarivera
New activity in
zai-org/GLM-4.6V-Flash
10 months ago
llama.cpp support
🚀
9
#21 opened 10 months ago by
ngxson
New activity in
moonshotai/Kimi-K2-Thinking
11 months ago
Python script to decompress tensors?
➕
🔥
4
15
#2 opened 11 months ago by
ubergarm
New activity in
ggml-org/gguf-my-lora
about 1 year ago
Fixing the error ModuleNotFoundError: No module named \'mistral_common\'\n
#4 opened about 1 year ago by
Moaazsoliman
New activity in
huihui-ai/Huihui-gpt-oss-20b-BF16-abliterated
about 1 year ago
Unable to GGUF quant: Errors out.
21
#2 opened about 1 year ago by
DavidAU
New activity in
ggml-org/gpt-oss-20b-GGUF
about 1 year ago
Update README.md
#2 opened about 1 year ago by
m18coppola
New activity in
ikawrakow/Qwen3-30B-A3B
about 1 year ago
GitHub account and ik_llama.cpp are down?!
😔
👍
4
60
#2 opened about 1 year ago by
ubergarm
New activity in
ggml-org/SmolLM3-3B-GGUF
about 1 year ago
Error loading chat template
1
#1 opened about 1 year ago by
rockerBOO
New activity in
tencent/Hunyuan-A13B-Instruct
over 1 year ago
attention_mask bug
😎
3
2
#18 opened over 1 year ago by
ngxson
New activity in
unsloth/gemma-3n-E4B-it-GGUF
over 1 year ago
Gemma 3n fixes for GGUFs and Ollama
❤️
2
21
#6 opened over 1 year ago by
shimmyshimmer
New activity in
Qwen/Qwen3-Embedding-0.6B-GGUF
over 1 year ago
Missing LAST pooling setting
👍
1
#5 opened over 1 year ago by
ngxson
New activity in
XiaomiMiMo/MiMo-VL-7B-SFT
over 1 year ago
Remove base_model
#1 opened over 1 year ago by
ngxson
New activity in
XiaomiMiMo/MiMo-VL-7B-RL
over 1 year ago
Correct base_model
#2 opened over 1 year ago by
ngxson
New activity in
ggml-org/ultravox-v0_5-llama-3_2-1b-GGUF
over 1 year ago
fix: add barebones README to link this quanitzed model to its base_model
#2 opened over 1 year ago by
aviallon
New activity in
Qwen/Qwen2-Audio-7B-Instruct
over 1 year ago
Poor result - probably model is corrupted?
👀
3
1
#22 opened over 1 year ago by
ngxson
Load more