|
About the quantization category
|
|
0
|
2558
|
October 2, 2019
|
|
XNNPACKQuantizer.set_module_name() not working as expected
|
|
3
|
63
|
March 5, 2026
|
|
[ROCm][CI] fp8 acceptable accuracy threshold
|
|
2
|
71
|
February 20, 2026
|
|
Extracting int8 weights and other quant params after convert_pt2e
|
|
1
|
53
|
February 17, 2026
|
|
Limitations of Int8 QAT for Linear Layers
|
|
1
|
48
|
February 17, 2026
|
|
Post training quantized model gets the error "Copying from quantized Tensor to non-quantized Tensor is not allowed" even though I'm not copying tensor
|
|
5
|
72
|
February 12, 2026
|
|
Huge accuracy drop from QAT model after convert_pt2e
|
|
1
|
56
|
January 27, 2026
|
|
Variable-bit (sub 8-bits) quantization for custom hardware deployment with power-of-two (pot) scales
|
|
10
|
1511
|
January 6, 2026
|
|
Ost Training Quantization fails on SPAN model with type_as
|
|
1
|
41
|
December 4, 2025
|
|
Difference of IntxWeightOnlyConfig/UIntxWeightOnlyConfig/Int8WeightOnlyConfig/Int4WeightOnlyConfig/
|
|
2
|
65
|
December 4, 2025
|
|
Quantize convolution layer
|
|
1
|
71
|
December 4, 2025
|
|
PT2E quantization doesn't reduce the model size
|
|
2
|
108
|
December 4, 2025
|
|
How to use quantized weights for manual implementation of the model in FPGA?
|
|
2
|
1100
|
September 28, 2025
|
|
[pt2e][quant] Quantization of operators with multiple outputs (RNN, LSTM)
|
|
4
|
335
|
September 15, 2025
|
|
GPU MEM% allocation vs batch size and temporal dimension
|
|
3
|
102
|
September 13, 2025
|
|
TorchAO Migration
|
|
0
|
88
|
September 11, 2025
|
|
Does export support quantized models with torchAo
|
|
1
|
92
|
September 11, 2025
|
|
Should I perform quantization after activation functions like sigmoid and SiLU?
|
|
0
|
82
|
September 9, 2025
|
|
Quantization of Hybrid Pytorch Model
|
|
0
|
60
|
September 8, 2025
|
|
Error while converting quantized Torch model to ONNX
|
|
0
|
96
|
September 5, 2025
|
|
My model is taking too much time in calculating FFT to find top k
|
|
1
|
81
|
September 2, 2025
|
|
FX mode static_quantization for YOLOv7
|
|
16
|
1040
|
August 4, 2025
|
|
Could not run 'aten::quantize_per_tensor' with arguments from the 'QuantizedCPU' backend
|
|
7
|
4276
|
July 17, 2025
|
|
RuntimeError: quantized::conv2d_prepack() is missing value for argument 'stride'
|
|
1
|
73
|
July 1, 2025
|
|
Why is there such a significant difference between floating-point convolution and quantized integer convolution results?
|
|
2
|
93
|
June 30, 2025
|
|
[MPS] When device='mps', aten.linear.default op is not decomposed
|
|
1
|
80
|
June 5, 2025
|
|
Logits mismatch between PyTorch inference and manual implementation
|
|
1
|
119
|
April 29, 2025
|
|
QAT model drops accuracy after converting with torch.ao.quantization.convert
|
|
1
|
131
|
April 29, 2025
|
|
Qint8 Activations in PyTorch
|
|
1
|
263
|
April 25, 2025
|
|
How to do qat after ptq in PyTorch2 quantization?
|
|
1
|
181
|
April 25, 2025
|