GPTQ models for nn.Linear 4bit weight & 16bit activation (symmetric) integer quantization
Karamjot Singh
karamjotsingh
AI & ML interests
Computer Vision, LLM Optimization
Recent Activity
updated a model 23 days ago
karamjotsingh/Llama-3.2-3B-Instruct updated a collection 2 months ago
GPTQ DeepCompressor updated a model 2 months ago
karamjotsingh/gemma-4-E4B