Skip to content

Automating Optimization of Quantized Deep Learning Models on CUDA

tvm.ai
13 pointscrowwork2 comments
On HN

Comments

Nice work accelerating convolutional models! It might be better to see (or cite papers about) the trade-off how model performance (accuracy, etc) changes w.r.t. how it is quantized.

With learning-based program optimizer, we can competitive performance on benchmark models and significant boost on emerging models against TensorRT(int8).

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.