Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes
DEV Community
Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes
This is a submission for the Kaggle Benchmarking Challenge What I Benchmarked A while...
0 comments
No comments yet.