pytorch-quantization

Pytorch Quantization

PyTorch-Quantization is a toolkit for training and evaluating PyTorch models with simulated quantization. Quantization can be added to the model automatically, or manually, allowing the model to be tuned for accuracy and performance. Quantization is compatible with NVIDIAs high performance integer kernels which leverage integer Tensor Cores. The quantized model can be exported to ONNX and imported to an upcoming version of TensorRT.

Install

Binaries

pip install pytorch-quantization --extra-index-url https://pypi.ngc.nvidia.com

From Source

git clone https://github.com/NVIDIA/TensorRT.git
cd tools/pytorch-quantization

Install prerequisites

pip install -r requirements.txt
pip install torch

Build and install pytorch-quantization

python setup.py install

NGC Container

pytorch-quantization is preinstalled in NVIDIA NGC PyTorch container since 20.12, e.g. nvcr.io/nvidian/pytorch:20.12-py3

Resources

Pytorch Quantization Toolkit userguide
Quantization Basics whitepaper

Name		Name	Last commit message	Last commit date
parent directory ..
docs		docs
examples		examples
pytorch_quantization		pytorch_quantization
scripts		scripts
src		src
tests		tests
.coveragerc		.coveragerc
.gitignore		.gitignore
.pylintrc		.pylintrc
.style.yapf		.style.yapf
CONTRIBUTING.md		CONTRIBUTING.md
LICENSE		LICENSE
MANIFEST.in		MANIFEST.in
README.md		README.md
VERSION		VERSION
requirements.txt		requirements.txt
setup.cfg		setup.cfg
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

README.md

Pytorch Quantization

Install

Binaries

From Source

NGC Container

Resources

FilesExpand file tree

pytorch-quantization

Directory actions

More options

Directory actions

More options

Latest commit

History

pytorch-quantization

Folders and files

parent directory

README.md

Pytorch Quantization

Install

Binaries

From Source

NGC Container

Resources