--- license: apache-2.0 base_model: Qwen/Qwen3-30B-A3B-Thinking-2507 tags: - quantized - rtn - 3-bit - thinking - reasoning --- # Qwen__Qwen3-30B-A3B-Thinking-2507_RTN_w3g128 This is a 3-bit RTN (Round-To-Nearest) quantized version of [Qwen/Qwen3-30B-A3B-Thinking-2507](https://huggingface.co/Qwen/Qwen3-30B-A3B-Thinking-2507). ## Quantization Details - **Method**: RTN (Round-To-Nearest) - **Bits**: 3-bit - **Group Size**: 128 - **Base Model**: [Qwen/Qwen3-30B-A3B-Thinking-2507](https://huggingface.co/Qwen/Qwen3-30B-A3B-Thinking-2507) ## Usage ```python from transformers import AutoModelForCausalLM, AutoTokenizer model_id = "quantpa/Qwen__Qwen3-30B-A3B-Thinking-2507_RTN_w3g128" tokenizer = AutoTokenizer.from_pretrained(model_id) model = AutoModelForCausalLM.from_pretrained(model_id) # Use the model for inference ``` ## Model Details - **Quantization**: RTN 3-bit - **Original Model**: Qwen/Qwen3-30B-A3B-Thinking-2507 - **Quantized by**: quantpa