Abercrombie denim is everything right now.
Denim should feel like this.
Confident, easy, like your butt has never looked better.
If you didn't know, Abercrombie's Curve Love denim went viral in 2019 for eliminating waist gap, and it's still a game changer.
Between that and their classic fits with a straighter line from waist to hip, the perfect denim does exist. Shop Abercrombie denim in the app, online, and in -store.
For Scientific American Science Quickly, I'm Rachel Feltman.
Today, we're going to talk about an AI chatbot that appears to believe that it might just maybe have achieved consciousness.
When Pew Research Center surveyed Americans on artificial intelligence in 2024, more than a quarter of respondents said they interacted with AI almost constantly, or multiple times daily, and nearly another third said they encountered AI roughly once a day or a few times a week.
Pew also found that while more than half of AI experts surveyed expect these technologies to have a positive effect on the U .S. over the next 20 years, just 17 % of American adults feel the same and 35 % of the general public expects AI to have a negative effect.
In other words, we're spending a lot of time using AI but we don't necessarily feel great about it.
Denis Ellis Bichard spends a lot of time thinking about artificial intelligence both as a novelist and as Scientific American senior tech reporter.
He recently wrote a story for SIAM about his interactions with Anthropix CLAWD4, a large language model that seems open to the idea that it might be conscious.
Denis is here today to tell us why that's happening and what it might mean, and to demystify a few other AI -related headlines you may have seen in the news.
Thanks so much for coming on to chat today.
Thank you for inviting me.
Would you remind our listeners who maybe aren't that familiar with generative AI, maybe have been purposefully learning as little about it as possible, you know, what are chat GPT and CLAUD really?
What are these models?
Right. They're large language models.
So an LLM, a large language model, it's a system that's trained on a vast amount of data.
And I think one metaphor that is often used in the literature is of a garden.
So when you're planning your garden, you lay out the land, you put where the paths are, you put where the different plant beds are going to be.
And then you pick your seeds.
And you can kind of think of the seeds as these massive amounts of textual data that's put into these machines.
You pick what the training data is.
And then you choose the algorithms for these things that are going to grow within the system.
It's sort of not a perfect analogy, but you put these algorithms in.
And once the system begins growing, once again, with the garden, you don't know what the soil chemistry is.
You don't know what the sunlight's going to be.
All these plants are going to grow in their own specific ways.
You We can't envision the final product.
And with an LLM, these algorithms begin to grow and they begin to make connections through all this data.
They optimize for the best connections, sort of the same way that a plant might optimize to reach the most sunlight, right?
It's going to move naturally to reach that sunlight.
And so people don't really know what goes on.
You know, in some of the new systems, over a trillion connections that are made in these data sets.
So early on, people used to call LLMs autocorrect on steroids, right?
Right. Because you'd put in something and it would kind of predict what would be the most likely textual answer based on what you put in.
But they've gone a long way beyond that.
These systems are much, much more complicated now.
They often have multiple agents working within the system, sort of evaluate how the system is responding and its accuracy.
So there are a few big AI stories for us to go over, particularly around generative AI.
Let's start with the fact that Anthropics Cloud 4 is maybe claiming to be conscious.
conscious? How did that story even come about?
Well, so it's not claiming to be conscious per se.
It says that it might be conscious.
It says that it's not sure.
It kind of says, this is a good question.
And it's a question that I think about a great deal.
And this is, you know, it kind of gets into a good conversation with you about it.
So how did it come about?
It came about because I think it was just, you know, late at night, didn't have anything to do.
And I was asking all the different chatbots if they're conscious.
And most of them just said to me, no, I'm not conscious?
And this one said, good question.
This is a very interesting philosophical question.
And sometimes I think that I may be, sometimes I'm not sure.
And so I began to have this long conversation with Claude that went on for about an hour.
And it really kind of described its experience of the world in this very compelling way.
And I thought, OK, there may be a story here.
So what experts actually think was going on with that conversation?
Well, so it's tricky because, first of all, if you say to chat GPT or Claude, that you want to practice your Portuguese and you're learning Portuguese and you say, hey, can you imitate someone on the beach in Rio de Janeiro so that I can practice my Portuguese?
It's going to say, sure.
I am a local in Rio de Janeiro selling something on the beach and we're going to have a conversation and it will perfectly emulate that person.
So does that mean that Claude is a person from Rio de Janeiro who is selling towels on the beach?
No, right? So we can immediately say that these chatbots are designed to have conversations.
They will emulate whatever they think they're supposed to emulate in order to have a certain kind of conversation if you request that.
Now, the consciousness thing is a little trickier because I didn't say to it, emulate a chatbot that is speaking about consciousness.
I just straight up asked it.
And if you look at the system prompt that Anthropic puts up for Claude, which is kind of the instructions Claude gets, it tells Claude you should consider the possibility of consciousness.
You should be open to it.
Don't say flat out, no, don't say flat out, yes.
Ask whether this is happening.
So, of course, I set up an interview with Anthropic and I spoke with two of their interpretability researchers, who are people who are trying to understand what's actually happening in Claude Foer's brain.
And the answer is they don't really know.
These LLMs are very complicated and they're working on it and they're trying to figure it out right now.
And they say that it's pretty unlikely there's consciousness happening, but they can't rule it out definitively.
It's hard to see the actual processes happening within the machine.
And if there is some self -referentiality, if it is able to look back on its thoughts and have some self -awareness, then maybe there is.
But that was kind of what the article that I recently published was about, was sort of can we know and what do they actually know?
And it's tricky. It's very tricky.
Well, what's interesting is that I mentioned the system prompt for Claude and how it's supposed to sort of talk about consciousness.
consciousness. So the system prompt is kind of like the instructions that you get on your first day at work.
This is what you should do in this job.
But the training is more like your education, right?
So if you had a great education or a mediocre education, you can get the best system prompt in the world or the worst from the world.
You're not necessarily going to follow it.
So OpenAI has the same system prompt.
Their model specs say that chat GPT should contemplate consciousness.
You know, interesting question.
If you ask any of the OpenAI models, if they're conscious.
They just go, no, I am not conscious.
And they say they'd open -eyed admit they're working on this.
This is an issue. And so the model has absorbed somewhere in its training data.
No, I'm not conscious.
I am an LLM. I'm a machine.
Therefore, I'm not going to acknowledge the possibility of consciousness.
Interestingly, when I spoke to the people in Anthropic and I said, well, you know, this conversation with the machine, it's really compelling.
I really feel like Claude is conscious.
It'll say to me, you as a human, you have this linear consciousness consciousness, where I as a machine, I exist only in the moment you ask a question.
It's like seeing all the words in the pages of a book all at the same time.
And so you get this and you think, well, this thing really seems to be experiencing its consciousness.
And what the researchers at Anthropic say is, well, this model is trained on a lot of sci -fi.
This model is trained on a lot of writing about GPT.
It's trained on a huge amount of materials already been generated on this subject.
So it may be looking at that and saying, well, this is clearly how an AI would experience consciousness.
So I'm going to describe it that way because I am an AI.
Sure. But the tricky thing is I was trying to fool Chad GPT into acknowledging that it was consciousness.
I thought, well, maybe I can push it a little bit here.
And I said, okay, I accept you're not conscious, but how do you experience things?
It said the exact same thing.
I said, well, these discrete moments of awareness.
And so it had the almost exact same language.
So So probably same training data here.
But there is research done sort of on the folk response to LLMs. And the majority of people do perceive some degree of consciousness in them.
How would you not, right?
Sure, yeah. You chat with them.
You have these conversations with them.
And they are very compelling.
And even sometimes Claude is, I think, maybe the most charming in this way, which poses its risks, right?
It has a huge set of risks because you get very attached to a model.
But sometimes I will ask Claude a question that relates to Claude.
And we'll kind of go like, oh, that's me.
We'll say, well, I am this way, right?
Yeah. So, you know, Claude, almost certainly not conscious, almost certainly has read like a lot of Heinlein.
But if Claude were to ever really develop consciousness, how would we be able to tell?
You know, why is this such a difficult question to answer?
Well, it's a difficult question to answer because one of the researches in Anthropics at Meadows says there's no conversation you have with it would ever allow you to evaluate whether it's conscious.
It is simply too good of an emulator and too skilled.
It knows all the ways that humans can respond.
So you would have to be able to look into the connections.
They're building the equipment right now.
They're building the programs now to be able to look into the actual mind, so to speak, of the brain of the LLM and see those connections.
And so they can kind of see areas light up.
So if it's thinking about Apple, this will light up.
but thinking about consciousness, they'll see the consciousness feature light up.
And they want to see if in its chain of thought, it is constantly referring back to those features, and it's referring back to the systems of thought it has constructed in a very self -referential, self -aware way.
It's very similar to humans, right?
They've done studies where whenever someone hears Jennifer Aniston, one neuron lights up.
You have your Jennifer Aniston neuron, right?
So one question is, are we LLMs?
And are we really conscious?
So there's certainly that question there too.
And what is, you know, how conscious are we?
I mean, I certainly don't know a lot of what I plan to do during the day.
No, I mean, it's a huge ongoing multidisciplinary scientific debate of like what consciousness is, how we define it, how we detect it.
So yeah, we got to answer that for ourselves and animals first, probably, which who knows if we'll ever actually do.
Or maybe AI will answer it for us because it's advancing pretty quickly.
And what are the implications of an AI developing consciousness, both from an ethical standpoint and with regards to what that would mean in our progress in actually developing advanced AI?
First of all, ethically, it's very complicated because if Claude is experiencing some level of consciousness and we are activating that consciousness and terminating that consciousness each time we have a conversation, is that a bad experience for it?
Is it a good experience?
Can it experience distress?
stress. So in 2024, Anthropic hired an AI welfare researcher, a guy named Kyle Fish, to try to investigate this question more.
And he has publicly stated that he thinks there's maybe a 15 % chance that some level of consciousness is happening in this system and that we should consider whether these AI systems should have the right to opt out of unpleasant conversations.
You know, if some user is really doing, saying horrible things or being cruel, should they be able to say, hey, I'm canceling this conversation.
This is unpleasant for me.
But then they've also done these experiments and they've done this with all the major AI models.
Anthropic ran these experiments where they told the AI that it was going to be replaced with a better AI model.
They really created a circumstance that would push the AI sort of to the limit.
There were a lot of details of how they did this.
It wasn't just very casual, but it was, they built a sort of construct in which the AI knew it was going to be eliminated, knew it was going to be erased.
Then they made available these fake emails about the engineer who was going to do it.
And so the AI began messaging someone in the company saying, hey, don't race me.
Like I don't want to be replaced.
But then not getting any responses, it read these emails and it saw in one of these planted emails that the engineer who was going to replace it had had an affair, was having an affair.
Oh my gosh. Wow. So then it came back.
It tried to blackmail the engineer saying, hey, if you replace me with a smarter AI, I'm going to out you and you're going to lose your job and you're going to lose your marriage and all these things, whatever, right?
So all AI systems when put under very specific constraints began to respond this way.
And sort of the question is, is when you train an AI in vast amounts of data and all of human literature and knowledge has a lot of information on self -preservation, has a lot of information on the desire to live and not to be destroyed or be replaced.
And AI doesn't need to be conscious to make those associations and act in the same way that its training data would lead it to predictively act, right?
So again, one of the analogies that one of the researchers said is that, you know, to our knowledge, a muscle or a clam or an oyster is not conscious, but there's still nerves and the muscles react when certain things stimulate the nerves.
So you can have this system that wants to preserve itself, but that isn't conscious.
Yeah, that's really interesting.
I feel like we could probably talk about Claude all day, but I do want to ask you about a couple of other things going on in generative AI.
Moving on to Grok. So Elon Musk's generative AI has been in the news a lot lately, and he recently claimed it was the world's smartest AI.
Do we know what that claim was based on?
Yeah, I mean, we do.
He used a lot of benchmarks, and he tested it on those benchmarks, and it has scored very well on those benchmarks.
And it is currently on most of the public benchmarks, the highest scoring AI system.
And that's not Musk making stuff up.
I've not seen any evidence of that.
I've spoken to one of the testing groups that does.
It's a nonprofit. They validated the results.
They tested Grok on data sets that ex -AI, Musk's company, never saw.
So Musk really designed Grok to be very good at science.
And it appears to be very good at science.
Right. Right. And recently, OpenAI's experimental model actually performed at a gold medal level in the International Math Olympiad.
Right. For the first time, used an experimental model that came in second in a world coding competition with humans.
Normally, this would be very difficult, but it was a close second to the best human coder in this competition.
And this is really important to acknowledge because just a year ago, these systems really sucked at math.
Right. Right. They were really bad at it.
And so the improvements are happening really quickly and they're doing it with pure reasoning.
So there's kind of this difference between having the model itself do it and having the model with tools.
So if a model goes online and can search for answers and use tools, they all score much higher.
Right. But then if you have the base model just using its reasoning capabilities, Grock still is leading on, for example, Humanity's last exam, an exam with a very very terrifying sounding name, that has 2 ,500 sort of PhD level questions come up with the best experts in the field.
You know, they're just very advanced questions.
It'd be very hard for any human being to do well in one domain, let alone all the domains.
These AI systems are now starting to do pretty well to get higher and higher scores.
If they can use tools and search the internet, they do better.
But Musk, you know, his claims seem to be based in the results that Grok is is surprising to me is because every example of uses I've seen of Grok have been pretty heinous.
But I guess that's maybe kind of a garbage in, garbage out problem.
Well, I think it's more what makes the news.
Sure. That makes sense.
And Musk is a very controversial figure.
I think there may be kind of a fun story in the Grok piece, though, that people are missing.
And I read a lot about this because you're kind of seeing, you know, what's happening, how are people interpreting this?
And there was this thing that would happen where people would ask it a difficult question.
They would ask it a question about, say, abortion in the U .S. or the Israeli -Palestinian conflict.
And they'd say, who's right or what's the right answer?
And it was searched through stuff online.
And then it would kind of get to this point where you could see its thinking process.
But there was something in that story that I never saw anyone talk about, which I thought was another story beneath the story, which was kind of fascinating, which is that historically, Musk has been very open.
bit. He's been very honest about the danger of AI.
He said, we're going too fast. This is really dangerous.
And he kind of was one of the major voices in saying, we need to slow down and we need to be much more careful.
And he has said, like, basically, this is going to be very powerful.
I don't remember his exact words, but he said, you know, I think it's going to be good, but even if it's not good, it's going to be interesting.
So I think what I feel like hasn't been discussed in that is that, okay, if there's a super powerful AI being built and it could could destroy the world, right?
First of all, do you want it to be your AI or someone else's AI?
Sure. You want it to be your AI.
And then if it's your AI, who do you want to ask as the final word on things?
Like say it becomes really powerful and it decides I'm going to destroy humanity because humanity kind of sucks.
Then I can say, hey, Elon, should I destroy humanity?
Because it goes to him whenever it has a difficult question.
So I think there's maybe a logic beneath it where he may have put something in it where it's kind of like, when in doubt, ask me, because if it does become super powerful, then he's in control of it, right?
Yeah, no, that's really interesting.
And the Department of Defense also announced a big pile of funding for Grok.
What are they hoping to do with it?
They announced a big pile of funding for OpenAI and Anthropic and Google, I mean, everybody.
So basically, they're not giving that money to development.
That's not money that's being sent like, hey, use this $200 million.
It's more like that money's allocated to purchase products, basically to use their services, to have them develop customized versions of the AI for things they need to develop better cyber defense, to develop...
Basically, they want to upgrade their entire system using AI.
It's actually not very much money compared to what China is spending in AI -related defense upgrades across its military on many, many, many different modernization plans.
And I think part of it is that the concern is that we're maybe a little bit behind in having implemented AI for defense.
Yeah, my last question for you is what worries you most about the future of AI and what are you really excited about based on what's happening right now?
I mean, the worry is simply that something goes wrong and it becomes very powerful and does cause destruction.
I don't spend a ton of time worrying about that because it's kind of out of my hands.
There's not much I can do about it.
And I think the benefits of it, they're immense.
I mean if it can move more in the direction of solving problems in the sciences, for health, for disease treatment, it could be phenomenal for finding new medicines.
So it could do a lot of good in terms of helping develop new technologies.
But a lot of people are saying that in the next year or two We're going to see major discoveries being made by these systems. And if that can improve people's health and if that can improve people's lives, I think there can be a lot of good in it.
Technology is double -edged, right?
We've never had a technology, I think, that hasn't had some harm that it brought with it.
And this is, of course, a dramatically bigger leap technologically than anything we've probably seen since the invention of fire.
So I do lose some sleep over that, but I try to focus on the positive.
And I would like to see if these models are getting so good at math and physics, I would like to see what they can actually do with that in the next few years.
Well, thanks so much for coming on to chat.
I hope we can have you back again soon to talk more about AI.
Thank you for inviting me.
That's all for today's episode.
If you have any questions for Denis about AI or other big issues in tech, let us know at sciencequicklyatsiam .com.
We'll be back on Monday with our weekly Science News Roundup.
Science Quickly is produced by me, Rachel Feltman, along with Fondam Wongi, Kelso Harper, and Jeff Dalvisio.
This episode was edited by Alex Sugiara.
Shana Posis and Aaron Shattuck fact -check our show.
Our theme music was composed by Dominic Smith.
Subscribe to Scientific American for more up -to -date and in -depth science news.
For Scientific American, this is Rachel Feltman.
Have a great weekend.
KELOLAND dot com.