01/
Open Models
huggingface.co/Mohaaxa ↗
Open weights →
Published GPTQ & AWQ quantized Qwen2.5-1.5B — with model cards documenting calibration data & methodology.
Benchmarked quantized SmolVLM on Jetson Nano via llama.cpp — real edge numbers, not extrapolations.
GPTQAWQQwen2.5-1.5BSmolVLMJetson Nanollama.cpp
Every released quantization ships with a model card covering calibration data, method, and measured trade-offs — the same standard the research holds benchmarks to. Related reading: F-03 · Vision on 1-bit LLMs.