14 ms·
> When you actually put them to the test you see 500k tokens of reasoning with "Actually..." and "Wait..." in every third paragraph of their reasoning trace. I
by philipbjorge 24d ago
> When you actually put them to the test you see 500k tokens of reasoning with "Actually..." and "Wait..." in every third paragraph of their reasoning trace.
I've wondered if this is part of why we don't see the reasoning traces for Anthropic's models before -- Open models might just be accurately surfacing how the sausage is made.
- lnenad 24d agoI'm assuming it's definitely part of the equation, but considering that I'm getting more tps but still waiting a lot more time for code to come out I'd assume it's not a 1:1 comparison. Plus I'm running quants, maybe with full precision it's better.