跳到主要内容
Great tips on working with reasoning models.
Normally you would append what the model figured out after the document and ask again.
It turns out that where you put the reasoning trace changes long-context accuracy by 50 points.
Transformers process causally, so a task state