does quantising a model reduce its performance ?[R]
Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.
If I were to quantise a fp32 model to fp8(or any other), would the information loss be drastic ?
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.