Non-uniform GGUF quantizations via GSQ + RCO: per-tensor mixed precision in standard GGUF form
AI & ML interests
None defined yet.
Recent Activity
View all activity
Papers
GPTQ-2D: Cubic-Time Two-Sided Adaptive Rounding
MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling, https://huggingface.co/papers/2604.18556
-
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling
Paper • 2604.18556 • Published • 12 -
ISTA-DASLab/Kimi-K2.6-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 49 • 1 -
ISTA-DASLab/Kimi-K2.5-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 39 • 1 -
ISTA-DASLab/Llama-3.1-70B-Instruct-2Bit-GSQ
Text Generation • 7B • Updated • 24
Non-uniform GGUF quantizations via GSQ + RCO: per-tensor mixed precision in standard GGUF form
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling, https://huggingface.co/papers/2604.18556
-
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling
Paper • 2604.18556 • Published • 12 -
ISTA-DASLab/Kimi-K2.6-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 49 • 1 -
ISTA-DASLab/Kimi-K2.5-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 39 • 1 -
ISTA-DASLab/Llama-3.1-70B-Instruct-2Bit-GSQ
Text Generation • 7B • Updated • 24
models 164
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF
Image-Text-to-Text • 27B • Updated • 297k • 425
ISTA-DASLab/Qwen3.8-27B-3Bit-GSQ
Image-Text-to-Text • 27B • Updated • 2.25k • 9
ISTA-DASLab/Qwen3.6-35B-A3B-2Bit-GSQ
Image-Text-to-Text • 36B • Updated • 2.08k • 4
ISTA-DASLab/Kimi-K2.5-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 39 • 1
ISTA-DASLab/Kimi-K2.6-2Bit-GSQ
Image-Text-to-Text • 84B • Updated • 49 • 1
ISTA-DASLab/Qwen3.5-4B-GGUF-GSQ
Text Generation • 4B • Updated • 255 • 3
ISTA-DASLab/Kimi-K2.5-P48-NVFP4-W4A4-Preview
Image-Text-to-Text • Updated • 45 • 1
ISTA-DASLab/Llama-3.1-70B-Instruct-2Bit-GSQ
Text Generation • 7B • Updated • 24
ISTA-DASLab/Llama-3.1-70B-Instruct-3Bit-GSQ
Text Generation • 9B • Updated • 18
ISTA-DASLab/Qwen3-8B-GGUF-GSQ
Text Generation • 8B • Updated • 51 • 3