🏗️ Building on HF
Rodri Mora
bullerwins
AI & ML interests
Gen AI, MLOps, Fine tuning and Quantizing models
Recent Activity
updated a model 27 days ago
bullerwins/GLM-5.3-Flash-exl3-4bpw-ablit published a model 27 days ago
bullerwins/GLM-5.3-Flash-exl3-4bpw-ablit liked a model about 1 month ago
zai-org/GLM-5.3Organizations
KLD Benchmarks + why Q8_K_XL vs MXFP4 naming
❤️👍 11
6
#11 opened 2 months ago
by
danielhanchen
KeyError: 'layers.0.ffn.experts.w13_qweight' when running on A100 GPU
2
#1 opened 2 months ago
by
JC1DA
Smaller quants
👀 1
1
#1 opened 2 months ago
by
Cathariel
Imatrix calibration file
4
#8 opened 2 months ago
by
bullerwins
code
1
#13 opened 4 months ago
by
erichartford
Puts spaces behind slashes in file paths
2
#1 opened 4 months ago
by
NRade
Compatible version with Ampere? SM8.6
1
#6 opened 4 months ago
by
bullerwins
Update README.md
1
#2 opened 6 months ago
by
bullerwins
Incorrect chat template
#1 opened 6 months ago
by
bullerwins
Can you add more details to the model card?
6
#1 opened 6 months ago
by
bullerwins
PPL test
1
#2 opened 8 months ago
by
bullerwins
Update README.md
1
#2 opened 9 months ago
by
bullerwins
llm-compressor recipe
👍 1
1
#1 opened 10 months ago
by
bullerwins
Does the Ktranformers deployment support non-AVX-512?
1
#16 opened 11 months ago
by
bullerwins
Definitely interested in this one!
🚀 2
25
#1 opened 11 months ago
by
mtcl
Interleaved Thinking, minimax:tool_call parsing
👍 1
1
#29 opened 11 months ago
by
0xSero
Wrong output
4
#2 opened 11 months ago
by
bullerwins
Is it compatible with vLLM?
4
#1 opened 11 months ago
by
bullerwins
Quantization code
1
#1 opened 11 months ago
by
bullerwins