"Also notable: 4.7 now defaults to NOT including a human-readable reasoning toke...

motoboi · 2026-04-16T16:20:53 1776356453

The reasoning is the secret sauce. They don't output that. But to let you have some feedback about what is going on, they pass this reasoning through another model that generates a human friendly summary (that actively destroys the signal, which could be copied by competition).

XenophileJKO · 2026-04-16T16:29:48 1776356988

Don't or can't.

My assumption is the model no longer actually thinks in tokens, but in internal tensors. This is advantageous because it doesn't have to collapse the decision and can simultaneously propogate many concepts per context position.

ainch · 2026-04-16T17:24:44 1776360284

I would expect to see a significant wall clock improvement if that was the case - Meta's Coconut paper was ~3x faster than tokenspace chain-of-thought because latents contain a lot more information than individual tokens.

Separately, I think Anthropic are probably the least likely of the big 3 to release a model that uses latent-space reasoning, because it's a clear step down in the ability to audit CoT. There has even been some discussion that they accidentally "exposed" the Mythos CoT to RL [0] - I don't see how you would apply a reward function to latent space reasoning tokens.

[0]: https://www.lesswrong.com/posts/K8FxfK9GmJfiAhgcT/anthropic-...

clbrmbr · 2026-04-17T00:31:56 1776385916

There’s also a paper [0] from many well known researchers that serves as a kind of informal agreement not to make the CoT unmonitorable via RL or neuralese. I also don’t think Anthropic researchers would break this “contract”.

[0] https://arxiv.org/abs/2507.11473

haellsigh · 2026-04-16T16:34:21 1776357261

If that's true, then we're following the timeline of https://ai-2027.com/

magicalist · 2026-04-16T19:44:55 1776368695

> If that's true, then we're following the timeline

Literally just a citation of Meta's Coconut paper[1].

Notice the 2027 folk's contribution to the prediction is that this will have been implemented by "thousands of Agent-2 automated researchers...making major algorithmic advances".

So, considering that the discussion of latent space reasoning dates back to 2022[2] through CoT unfaithfulness, looped transformers, using diffusion for refining latent space thoughts, etc, etc, all published before ai 2027, it seems like to be "following the timeline of ai-2027" we'd actually need to verify that not only was this happening, but that it was implemented by major algorithmic advances made by thousands of automated researchers, otherwise they don't seem to have made a contribution here.

[1] https://ai-2027.com/#:~:text=Figure%20from%20Hao%20et%20al.%...

[2] https://arxiv.org/html/2412.06769v3#S2

butlike · 2026-04-16T19:32:30 1776367950

Hilariously, I clicked back a bunch and got a client side error. We have a long way to go. I wouldn't worry about it.

matltc · 2026-04-16T16:59:51 1776358791

Care to expound on that? Maybe a reference to the relevant section?

ACCount37 · 2026-04-16T17:10:00 1776359400

Ctrl-F "neuralese" on that page.

9991 · 2026-04-16T17:08:37 1776359317

You should just read the thing, whether or not you believe it, to have an informed opinion on the ongoing debate.

matltc · 2026-04-17T02:38:25 1776393505

I did read it a while back. Was curious what parent was referring to specifically

9991 · 2026-04-17T17:59:50 1776448790

March 2027 -> Neuralese recurrence and memory

> For example, perhaps models will be trained to think in artificial languages that are more efficient than natural language but difficult for humans to interpret.

9991 · 2026-04-16T17:07:28 1776359248

That's not supposed to happen til 2027. Ruh roh.

literalAardvark · 2026-04-16T17:37:36 1776361056

Only if you ignore context and just ctrl-f in the timeline.

What are you, Haiku?

But yeah, in many ways we're at least a year ahead on that timeline.

JoshuaDavid · 2026-04-16T18:36:00 1776364560

Don't.

The first 500 or so tokens are raw thinking output, then the summarizer kicks in for longer thinking traces. Sometimes longer thinking traces leak through, or the summarizer model (i.e. Claude Haiku) refuses to summarize them and includes a direct quote of the passage which it won't summarize. Summarizer prompt can be viewed [here](https://xcancel.com/lilyofashwood/status/2027812323910353105...), among other places.

WhitneyLand · 2026-04-16T17:15:41 1776359741

No, there is research in that direction and it shows some promise but that’s not what’s happening here.

XenophileJKO · 2026-04-16T17:40:44 1776361244

Are you sure? It would be great to get official/semi-official validation that thinking is or is not resolved to a token embedding value in the context.

astrange · 2026-04-16T18:46:07 1776365167

You can read the model cards. Claude thinks in regular text, but the summarizer is to hide its tool use and other things (web searches, coding).

alex7o · 2026-04-16T16:42:06 1776357726

Most likely, would be cool yes see a open source Nivel use diffusion for thinking.

motoboi · 2026-04-16T17:25:19 1776360319

Don't. thinking right now is just text. Chain of though, but just regular tokens and text being output by the model.

dheera · 2026-04-16T17:22:05 1776360125

Although it's more likely they are protecting secret sauce in this case, I'm wondering if there is an alternate explanation that LLMs reason better when NOT trying to reason with natural language output tokens but rather implement reasoning further upstream in the transformer.

HarHarVeryFunny · 2026-04-17T15:54:25 1776441265

I would doubt it. They are mostly trained on natural language. They may be getting some visual reasoning capability from multi-modal training on video, but their reasoning doesn't seem to generalize much from one domain to another.

Some future AGI, not LLM based, that learns from it's own experience based on sensory feedback (and has non-symbolic feedback paths) presumably would at least learn some non-symbolic reasoning, however effective that may be.

dheera · 2026-04-19T17:34:47 1776620087

My argument for this is mostly that we don't use language for all forms of reasoning, and are likely doing so on some internal representations or embeddings. Animals also demonstrate abilities to reason with situations without actually having a language.

I see language more as a protocol for inter-agent communication (including human-human communication) but it contains a lot of inefficiencies and historical baggage and is not necessarily the optimal representation of ideas within a brain.

boomskats · 2026-04-16T16:23:59 1776356639

'Hey Claude, these tokens are utter unrelated bollocks, but obviously we still want to charge the user for them regardless. Please construct a plausible explanation as to why we should still be able to do that.'