upvote
I agree with everything you say except this:

> You can make any modern LLM explain its reasoning

You can make any modern LLM create a plausible, self-consistent explanation that looks like reasoning, but it's not "the reasoning it used to arrive at that answer".

reply
Tangent: This is often true of humans as well.

We often make a decision based on a gut feeling, and then backfill a logical reason supporting our feeling, without even realizing we're doing it -- rationalization.

reply
And we all know some people that rationalise poor choices and misbehavior, hide their mistakes, etc, to an unacceptable degree. Sometimes the individual knows they are rationalising but continues anyway, other times they seem incapable of seeing that.

When you ask people who are rationalising poor behaviour about the scenario, but it is someone else doing it, they may arrive at a better answer. Can we use multiple LLMs to achieve self criticism and critical thinking?

reply
isn't this happening already? There's the concept called "thinking" where the models talks with itself before giving you the final answer
reply
Tangent on the tangent: I think that's true in a minority of cases and in a majority of AI cases. Though in principle I think it should be possible for an LLM to have access to and faithfully represent its own reasoning.
reply
On the contrary, I would argue that it's true in a totality of AI cases.

To your point, I agree that nominally there should be a way to give conceptual names to paths of weights, and when answering a question, notice which weights were and were not applied and retrospect on that.

That's not what reasoning traces as they currently exist are, though.

reply
This was beautifully shown by asking a model to explain how it added two numbers together (something like 45+21), and it told a plausible story, when in fact they showed it was some rotation on a helix living in some internal manifold.

Like asking a human "how did you catch that fast ball coming at you?"

reply
It could be that the rotation in the helix manifold whatever is a low level representation of the logical steps (carry the 2, add the next column,...) it's describing. The point stands that the explanation it generates doesn't necessarily in all cases reflect what it "actually did" but your counterexample doesn't hold.
reply
Seriously, anybody with a passing knowledge of LLMs knows thats not how they function. You can't encode logic in them because that's not how they work. It's a statistical model with useful emergent properties. It doesn't think, it doesn't reason, it isn't aware of facts or the rules of logic.
reply
> It's a statistical model with useful emergent properties. It doesn't think, it doesn't reason, it isn't aware of facts or the rules of logic.

What makes you so sure your own brain doesn't work the same way?

reply
> This AI generated post (100% on Pangram) is pretty out of date.

Quite ironic given the topic. It seems that the author’s model indeed contained too much knowledge about old Gemini releases, and did not do enough tool calling.

reply
>>if a model is factually wrong a claim with a source is checkable and a claim from weights isn't.

>Why? If a model's weights claim that Bart Simpson became President in 2020, why does this fact suddenly become uncheckable?

Because in one case you have a source you can use to validate the fact, and in the other you don't. Though, as you explain earlier in your comment, the premise is misguided/hallucinated.

reply
Yea the "When the fact lives outside the model, a wrong answer has an address" sentence seems aggressively AI written. Saw that and my senses went off.
reply
Senses of what? LOL. The whole Internet is AI generated by now and we all contribute to that on daily basis. get used to it or dull your senses ...
reply
>>You can make any modern LLM explain its reasoning and find sources for its claims.

>The internet is full of wrong information and I cannot magically edit it to make it all correct, so this doesn't help me.

My favorite RAG experience was asking Bart (or whatever they were calling Gemini back then) an answer to a question I knew.

It gave me the opposite of the truth (as was common with LLMs at the time).

But weirdly, it had cited sources for this "fact."

I checked the sources. Two of them, both AI SEO slop.

In this moment, andai was enlightened...

reply
A true camper doesn't need to check Pangram, Jimbo. He goes by pure animal instinct!
reply