gitaskhub

Mixed-precision quantization for LLMs. Every layer refracts into a different format based on its sensitivity. Native compressed-tensors export, validated on Qwen3.6-35B-A3B MoE with MTP speculative decoding.

Language · Python
License · NOASSERTION
Ask anything about this repo to start.

By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).