← Back to KHAO

Anthropic · Claude ·

First, users might notice that the revealed thinking is more detached and less personal-sounding than Claude’s default outputs

2 min read

Compiled by KHAO Editorial — aggregated from 2 sources. See llms.txt for citation guidance.

✓ KHAO Verified

The performance of Claude 3.7 Sonnet versus its predecessor model on the OSWorld evaluation, testing multimodal computer use skills. “Pass @ 1”: the model has only a single attempt to solve a particular problem for it to count as having passed.

That’s because Anthropic didn’t perform Anthropic’s standard character training on the model’s thought process.

Key facts

Summary

Some things come to them nearly instantly: “what day is it today?” Others take much more mental stamina, like solving a cryptic crossword or debugging a complex piece of code. Now, Claude has that same flexibility. Extended thinking mode isn’t an option that switches to a different model with a separate strategy. Claude's new extended thinking capability gives it an impressive boost in intelligence. As well as giving Claude the ability to think for longer and thus answer tougher questions, they've decided to make its thought process visible in raw form.

#Anthropic #Claude