If you use AI or shrug about AI or talk about AI or love AI or hate AI—please watch this video showing that there's no mind behind it.
@thomasfuchs
Yes, but that is what is the most easy to improve.
It's really easy to give an LLM memory and a few tools.
The core model works just as the lady says, but none of the commercial products are merely that, like you mentioned.
@flq @julienw
@Noisecolor @thomasfuchs @flq @julienw
No it isn't! The improvement will be just a bandaid as long as it's a LLM.
Like math questions or common riddles it will be a "detect the question type" and use the special approach implemented for this type
The AI hypists will claim "see the new model learned it" but there is no learning, just another special case
@realn2s
It's not a band aid at all.
It's using a model and giving it tools. That makes perfect sense.
We can even give it a tool to find further tools by itself. We can give it tools to build a tool for itself.
That's all already happening.
What we call that activity doesn't really matter. It matters that with them they can very easily solve problems like guess who game.
@thomasfuchs I tried it but in my case I always win which is very suspicious
@thomasfuchs If $10,000 is deposited your cash app account or PayPal unexpectedly , what will you use it for to be honest ?
:A car
:B Bills
:c school
:d families
:e vacation
:l donation
:g food stuff
:h business
@thomasfuchs So are we not gonna talk about the model likely pulling from racist, sexist, transphobic language of incels online to predict Serena Williams as a man? Or was that implied?
@thomasfuchs LLM's are less like artificial brains and more like the theorem about "An Infinite Number of Monkeys on an Infinite Number of Typewriters" and I don't mean that as a pejorative.
@oshox it can be useful if you have a way to verify the output (which is possible for formal languages like in programming), but for most stuff it’s just meh at best
@thomasfuchs If we have to spend our time analyzing and parsing AI output, what time does it save us?
@thomasfuchs@hachyderm.io fun fact, this video is full of lies, and it is set up just to become viral. It can only convince people who have not used modern day LLMs. I did all three tests she describes and got completely opposite results. GPT thought of Adam Sandler (I found it), Claude's three first random words were marmalade, trapezoid and kaleidoscope, and its random number was 73. Both models were on the free tier for these tests (and GPT was two generations behind). Like, I could make a program in basic in my Spectrum in the 80s that would give a different number each time, how can someone believe that latest tech LLMs in 2026 can't do it?
I'm sure that many people who believed this video are generally not stupid or naive. But it's fun to see how easy it is to believe what we want to believe. Sorry folks, that video is entertaining if you want to hate on LLMs, but it's just bullshit generated to get views.
But, if you wanna believe it, believe it, I don't want to spoil your satisfaction =)
@thomasfuchs can we have the video source please? searching isn't possible on tiktok
@ambiguous_yelp oh ty, I saw it cross my Bluesky (I don’t use TikTok)
@thomasfuchs Long shot, but has anyone the link to a job interview with an LLM agent that someone posted a few days ago?
The interviewee responded to the agent's questions with random words/phrases, and the agent conducted the interview as though these were valid.
@InarticulateOtter hmm I remember that but don’t remember where I saw it
@thomasfuchs ah found it on YouTube, perhaps misremembered that I saw it here.
@thomasfuchs In prior versions of ChatGPT and still in self-hosted model harnesses like Ollama, there is a "Retry" button. If you play 20 Questions/Guess Who and then use Retry after it has revealed, it will give different answers, each consistent with the previous conversation.
thank you for this
@thomasfuchs fantastic TL;DR.
we really need to get quality vloggers to cross post on the fediverse.
@thomasfuchs probably one of the best and most succinct videos on the topic with concrete examples anyone can understand. Thank you for sharing