</>
Skill
Recall
RECALL WHAT YOU KNOW
Library
Planner
Why SkillRecall
More
About
News
Contact
Sign in
Explore Library
Create free account
Create free account
Explore Library
Quiz
Advanced
Save
Inference Optimization
How quantization reduces model size and speeds inference.
What does quantization do to a model?
A
Reduces numerical precision of weights (e.g., float32 to int8) to shrink size and speed inference
B
Adds more layers to improve accuracy
C
Encrypts the model weights
D
Increases the training dataset size