Anthropic has developed an AI 'brain scanner' to understand how LLMs work and it turns out the reason why chatbots are terrible at simple math and hallucinate is weirder than you thought

[email protected]

I mean neural networks are modeled after biological neurons/brains after all. Kind of makes sense...

[email protected]

This is what the ARC-AGI test by Chollet has also shows regarding current AI / LLMs. They have a tendency to approach problems with this trial and error method and can be extremely inefficient (in their current form) with anything involving abstract / deductive reasoning.

Most LLMs do terribly at the test with the most recent breakthrough being with reasoning models. But even the reasoning models struggle.

ARC-AGI is simple, but it demands a keen sense of perception and, in some sense, judgment. It consists of a series of incomplete grids that the test-taker must color in based on the rules they deduce from a few examples; one might, for instance, see a sequence of images and observe that a blue tile is always surrounded by orange tiles, then complete the next picture accordingly. It’s not so different from paint by numbers.

The test has long seemed intractable to major AI companies. GPT-4, which OpenAI boasted in 2023 had “advanced reasoning capabilities,” didn’t do much better than the zero percent earned by its predecessor. A year later, GPT-4o, which the start-up marketed as displaying “text, reasoning, and coding intelligence,” achieved only 5 percent. Gemini 1.5 and Claude 3.7, flagship models from Google and Anthropic, achieved 5 and 14 percent, respectively.

[email protected]

'is weirder than you thought '

I am as likely to click a link with that line as much as if it had

'this one weird trick' or 'side hussle'.

I would really like it if headlines treated us like adults and got rid of click baity lines.

[email protected]

I don't think it knows the full sentence, it just doesn't search for the words in the order they will be in the sentence. It finds the end-words first to make the poem rhyme, than looks for the rest of the words. I do it this way as well just like many other people trying to create any kind of rhyming text.

[email protected]

interestingly, too, this is a technique when you're improvising songs, it's called Target Rhyming.

The most effective way is to do A / B^1 / C / B^2 rhymes. You pick the B^2 rhyme, let's say, "ibruprofen" and you get all of A and B^1 to think of a rhyme

Oh its Christmas time
And I was up on my roof when
I heard a jolly old voice
Ask me for ibuprofen

And the audience thinks you're fucking incredible for complex rhymes.

[email protected]

you can't trust its explanations as to what it has just done.

I might have had a lucky guess, but this was basically my assumption. You can't ask LLMs how they work and get an answer coming from an internal understanding of themselves, because they have no 'internal' experience.

Unless you make a scanner like the one in the study, non-verbal processing is as much of a black box to their 'output voice' as it is to us.

[email protected]

They do it because it works on the whole. If straight titles were as effective they'd be used instead.

[email protected]

But you wouldn't multiply, say, 74*14 to get the answer.

[email protected]

Don't tell me that my thoughts aren't weird enough.

[email protected]

The one weird trick that makes clickbait work

[email protected]

But then you wouldn't need to click on thir Ad infested shite website where 1-2 paragraphs worth of actual information is stretched into a giant essay so that they can show you more Ads the longer you scroll

[email protected]

It really is quite unfortunate, I wish titles do what titles are supposed to do instead of being baits.but you are right, even consciously trying to avoid clicking sometimes curiosity gets the best of me. But I am improving.

[email protected]

The problem with common core math isn’t that rounding is inherently bad, it’s that you don’t start with that as a framework.

[email protected]

I might. Then I can subtract 74 to get 74*14, and subtract 28 to get 72*13.

I don't generally do that to 'weird' numbers, I usually get closer to multiples of 5, 9, 10, or 11.

But a computer stores information differently. Perhaps it moves closer to numbers with simpler binary addresses.

[email protected]

This is what I do, except I would add 700 and 236 at the end.

Well except I would probably add 700 and 116 or something, because my working memory fucking sucks and my brain drops digits very easily when there's more than 1

[email protected]

Maybe you're right. Maybe it's Markov chains all the way down.

The only way I can think to test this would be to "poison" the training data with faulty arithmetic to see if it is just recalling precedent or actually implementing an algorithm.

[email protected]

But you're doing two calculations now, an approximate one and another one on the last digits, since you're going to do the approximate calculation you might act as well just do the accurate calculation and be done in one step.

This solution, while it works, has the feeling of evolution. No intelligent design, which I suppose makes sense considering the AI did essentially evolve.

[email protected]

Appreciate the advice on how my brain should work.

[email protected]

Not, but I'd do 7510 + 754, then subtract the extra.

The LLM method of doing it with multiple numbers without proper interpolation though makes it extra weird

[email protected]

People are generally shit at understanding probabilities and even when they have a fairly strong math background tend to explain probablistic outcomes through anthropomorphism rather than doing the more difficult and "think-painy" statistical analysis that would be required to know if there was anything more to it.

I myself start to have thoughts that balatro is purposefully screwing me over or feeding me outcomes when it's just randomness and probability as stated.

Ultimately, it's easier (and more fun) for us to reason that way and it largely serves us better in everyday life.

But these things are entire casinos' worth of probability and statistics in and of themselves, and the people developing them want desperately to believe that they are something more than pseudorandom probabilistic fancy autocomplete engines.

Add the difficulty of getting someone to understand how something works when their salary depends on them not understanding it to the existing inability of humans to reason probabilistically and the AGI from LLM delusion becomes near impossible to shake for some folks.

I wouldn't be surprised if this AI hype bubble yields a cult in the end.

agnos.is Forums

Anthropic has developed an AI 'brain scanner' to understand how LLMs work and it turns out the reason why chatbots are terrible at simple math and hallucinate is weirder than you thought