Do small language models know what they don't know?
Researchers are testing whether entropy-based confidence signals can improve the accuracy of small language models containing fewer than 3 billion parameters. By evaluating seven different methodologies, the study aims to determine if models optimized for consumer hardware can better identify when they lack sufficient information to answer accurately.
Covered by 1 source
- AarXiv CS.AI↗Prashant Mudgal1d ago