Back

Controlling Reasoning Effort in LLMs

45 points7 hoursmagazine.sebastianraschka.com
simonw3 hours ago

I'm amused by how the whole reasoning model thing feels like a formalization of the old "think step by step" prompting hack, which was discovered against GPT-3 two years after that model was first released.

My favorite trick for controlling the reasoning level is the hack where you look at the output token stream and spot the token for "the model has concluded reasoning"... and then replace that with the tokens for "wait, but" and force it to keep going!

sva_3 hours ago

I recommend his book, "Build a Reasoning Model (From Scratch)", which is also linked in the article.

https://sebastianraschka.com/books/#build-a-reasoning-model-...

kimonsodu2 hours ago

[flagged]

connerpro4 hours ago

[flagged]