The Illusion of Thinking in Large Reasoning Models (LRM)

カートのアイテムが多すぎます

ご購入は五十タイトルがカートに入っている場合のみです。

カートに追加できませんでした。

しばらく経ってから再度お試しください。

ウィッシュリストに追加できませんでした。

しばらく経ってから再度お試しください。

ほしい物リストの削除に失敗しました。

しばらく経ってから再度お試しください。

ポッドキャストのフォローに失敗しました

ポッドキャストのフォロー解除に失敗しました

The Illusion of Thinking in Large Reasoning Models (LRM)

無料で聴く

ポッドキャストの詳細を見る

このコンテンツについて

This episode investigates the reasoning capabilities of Large Reasoning Models (LRMs), a new generation of language models designed for complex problem-solving. The authors evaluate LRMs using controllable puzzle environments to systematically analyze how performance changes with problem complexity, unlike traditional benchmarks that often suffer from data contamination. Key findings reveal three performance regimes: standard LLMs surprisingly excel at low complexity, LRMs gain an advantage at medium complexity, and both models experience complete collapse at high complexity, often exhibiting a counter-intuitive decline in reasoning effort despite having a sufficient token budget. The analysis also examines the internal reasoning traces, uncovering patterns like "overthinking" on simpler tasks and highlighting limitations in LRMs' ability to follow explicit algorithms or maintain consistent reasoning across different puzzle types.

Send us a text

Support the show

Podcast:
https://kabir.buzzsprout.com

YouTube:
https://www.youtube.com/@kabirtechdives

Please subscribe and share.