Discussion about this post

User's avatar
Asa Boxer's avatar

Great stuff as usual. I was happily surprised to discover that Claude could make an exclusionary list, which Chomsky said it couldn't do. Not that long ago it couldn't; ChatGPT couldn't just a couple of years back. The task I asked for was to provide a list of all the words in the English language that have the vowel 'a' to the exclusion of all other vowels. The vowel 'a' could be repeated, but no other vowels could be included. I also specified that it should not invent or fabricate. It proceeded to work using a self-corrective method. Here is Claude:

"I need to be transparent with you: this task is harder than it appears, and I risk fabricating or including errors at scale. Here is the honest problem:

My knowledge of the full English lexicon isn't exhaustive enough to produce a complete list reliably.

Generating a long list risks false inclusions (words I misremember) or missed exclusions (words where I forget a vowel).

What I can offer with confidence — a solid, verified-by-reasoning sample:"

The complete list it gave included examples of errors with a large red 'x' beside them. This is the sort of feedback that mimics self-correction and convinces folks that there's some level of sentience going on. But convincing mimicry doesn't hold up under real pressure. Humour me a moment...

(I find it interesting that this sort of exclusion-based pattern finding is "difficult" enough to break Claude.) Next, I asked Claude to write five poems using the list and only the list. The poems didn't have to rhyme, but they had to make sense. (The basis for this is an awful book called Eunoia that was written many years back. It was a very mechanical book as one would expect. So I've been curious to know when AI has reached that level.) Claude failed to perform the task on several levels, and I was ready to be lenient, since the bar was set low to begin with. From the technical perspective it failed by including two vowels that it claimed not to have used in its own final analysis. And from a sense perspective it just failed where a human would have seen the obvious. Ex: "camp dark" and "path vast." A lot of patterns were clearly filched cliche, like "band marched" and "branch snapped." But "snapped branch" would have been better. In short, when pressed to make sense with restrictions like the one I provided, the Claude's methods are denuded and the lack of understanding what it's doing becomes glaringly evident.

PS: Note carefully Claude's response about the difficulties. It's worth considering how words might be "misremembered." Claude can't deal with words as both semantic entities and data at the same time. Question: can it be trained to do this?

Francis Turner's avatar

You have put some intellectual rigor behind what I was trying to say here - https://ombreolivier.substack.com/p/ai-bicycle-or-train - about how AI is, on its own, intrinsically limited and that if you don't put in the effort you get recycled slop

13 more comments...

No posts

Ready for more?