cd/entity/GPTQΒ· homeβ€Ί entitiesβ€Ί GPTQ
grep -l @gptq /news/*.json | wc -l β†’ 9

GPTQ

mentions 9 type Organization feed RSS

// recent coverage 9 mentions

00:00
2026-07-03
deepresearch.ninja
large-language-models

LLM Quantization Methods: A Comprehensive Comparative Analysis

A comprehensive analysis of 15+ large language model quantization methods categorizes them into four paradigms: CPU-optimized GGUF-based approaches, GPU-native weight-only kernels, NVIDIA's floating-p…

10:53
2026-06-27
dev.to
machine-learning

How I Implemented GPTQ from Scratch (and What I Learned)

A developer implemented GPTQ quantization from scratch on a nanoGPT model, achieving only 1.1% perplexity degradation across 61 quantized layers. The implementation uses second-order optimization to r…

13:01
2026-06-24
gist.github.com
large-language-models

NVIDIA GenAI LLM Certification Lab

NVIDIA has released a GenAI LLM Certification Lab that guides developers through building a production-ready fine-tuning and optimization pipeline. The lab covers data preparation, LoRA fine-tuning wi…

14:58
2026-06-06
vettedconsumer.com
large-language-models

GGUF vs. GPTQ vs. AWQ: The Plain-English Guide to LLM Quantization

GGUF, GPTQ, and AWQ are the three dominant formats for running quantized large language models locally, each optimized for different hardware and use cases. GGUF, the format used by llama.cpp and its …

// co-occurs with top 8 entities
// topics top 6 topics