10 ms·
>The findings reveal that even slight changes in the phrasing of questions can cause major discrepancies in model performance that can undermine their reliabili
by zero-sharp 2y ago
>The findings reveal that even slight changes in the phrasing of questions can cause major discrepancies in model performance that can undermine their reliability in scenarios requiring logical consistency.
Relevant conversation with Yann Lecun:
https://www.youtube.com/watch?v=5t1vTLU7s40&t=4189s https://www.youtube.com/watch?v=5t1vTLU7s40&t=4189s
- mgh2 2y agoApple is generally anti-hype, more truth grounded than competitors