Small models are overconfident because they're distilled from large models
Small models are trained to copy big models' answers and their confidence. If they are trained to know their own limits, will this make them smarter? submitted by /u/TinyDetective110 [link] [comments]
📄
This source provides headlines only. Use the button below to read the complete article on the original site.
📰 Read the original article on r/LocalLLaMA
Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.