Welcome to the ai hustle podcast.
Today on the show we're talking about a brand new um product that is coming out of thinking machines.
It is a ai model that essentially listens while it talks, so similar to a human.
A lot of times we've had this latency issue with things like 11 labs or other of these kind of AI voice receptionists.
There's so many different industries that this applies to, but there's a latency where you speak and it has to try to take in all the data, compute it, think of its response and spit it out to as fast as possible.
And there's kind of a little latency gap.
So they're trying to solve this by essentially having the AI model listen and actively taking information while it's also talking, and be able to do this kind of simultaneously like a human would.
We're going to get into all of this.
It's fascinating also because this is coming out of Thinking Machine, which is from Miri Mirati, who's raised a ton of money and hasn't had a ton of products come out.
So anyways, a lot to get into here.
But before we do, I wanted to mention.
If you haven't already joined our AI Hustle School community, we would love to have you as a member.
Every single week we record a bonus video where we break down a tutorial on the tools, strategies and things we're using to grow and scale our businesses with AI.
We show you the actual products and projects we're actively working on.
We show you revenue numbers.
We show you deep dives, everything we can't share publicly.
We show you the tools we're using.
And so it's an amazing place.
You can go join.
It's only $19 a month.
We try to make this affordable for everybody.
We'll keep this price low.
And if you lock in this price, if we raise it in the future, it won't be raised on you.
There's over 200 members in here.
So lots of people that are actively building and sharing what they're working on.
And we'd love... to have you be one of them.
So I'll leave a link in the description to that.
Jamie, kicking over to this story with Thinking Machine Labs, I guess what was the most interesting part of this software to you?
Because I know you personally are working on and you have tools and you help businesses with kind of these AI, receptionist kind of things.
So for you, what are your first thoughts on this?
I mean, I think this is a great idea.
I think that's a real problem that they need to solve, because even if you have really quick latency with, let's say, an AI receptionist, there's still a little bit of unnatural lag that happens between the time you finish speaking and then it starts talking back to you.
And sometimes, even if you pause... but it's a natural pause.
Like I just stopped talking for a moment.
It will start responding, even though you haven't finished your thoughts.
So I, you know, I think there's up to this point, this AI, you know, receptionist AI, voice agents have gotten really good, but they're still not quite to that, that human level yet, where you know, as I'm talking right now, you could be, you are thinking and processing what I'm saying.
A traditional, you know agent will.
It has to first turn all of your what your sentence into a text, then figure out the context and respond.
And it's kind of like a a chain process where this is.
This is trying to make it more conversational.
Anyways, I think it's a great idea.
I also think it's interesting because we've talked about thinking machines lab on this podcast before and how it's kind of funny because their website was very cryptic for a long time.
Last year they raised 2 billion in last July, giving them a 12 billion valuation, yet they had no product.
So it was basically, from my perspective, all based around Miriam Marotti's name and her credibility and things like that.
All that to say is we finally have a product, or at least this is.
They're calling it a research project.
They haven't officially released a product yet, but I think it's going to actually solve a really a big problem.
And I think, you know, the evaluation maybe will be legit.
So what are your thoughts on it, Jayden?
Yeah, I mean I do think one of the interesting parts about this whole business.
I want to get into some of the technical stuff here because I do think it's really fascinating.
But as far as the whole business, because we've talked about Thinking Machine before and we're, like you know, like
They've raised so much money, and it's all kind of off of Maria Moratti.
One interesting thing about her though, that I will say is she is probably a multi-multi-billionaire just from her open AI stake.
So when she goes and raises a couple billion dollars, people aren't just giving it to someone.
I guess that is just pure kind of clout and credibility.
There was recently an article – that I talked about how Ilya Seskovor, one of the other co-founders of OpenAI, he went up on trial to kind of defend why he kicked out Sam Altman.
Anyways, there's all the drama between Elon Musk and Sam Altman trial.
My favorite thing about the trial though, is that we're getting all of the like nitty gritty details because of discovery and because it's like a legal case that you don't get anywhere else.
And in all of that, he had to reveal that he has about a $7 billion OpenAI stake.
So he's one of the co-founders with Mario Moratti, one of the originals.
So you can imagine Mario Moratti maybe she may have joined like a little bit later and maybe he was a little bit more of an OG.
But I mean she's got to have somewhere between 4 and 7 billion in assets, uh equity or a stake in opening eyes, especially as it's kind of growing and stuff.
So she's over there doing her thinking machine labs.
She's, you know, raising it like a $2 billion valuation or she's raising $2 billion for it.
Um, and get a higher valuation.
But she also has like a massive fortune just tied to opening eye.
And at the end of the day, like that's where she can make a lot of her money.
So anyways, I do think that's a very interesting part of the whole story.
But as far as the tech of what they've released, I actually love it.
I think I was... you know, I was really curious what she was going to come out with.
Um, because obviously, if she was just trying to make another competing LLM to chat GPT which they very well might, who knows, but like it didn't seem like it was, I don't know.
You know, you go, leave open AI and just make another open AI clone.
This is cool because it's something completely unique, which is kind of what I expected.
Her and even Ilya's working on something.
And they're making these kind of completely unique platforms and solutions that didn't exist before.
And so specifically what they have done with this, it's called a full duplex.
That's the technical term.
And, like you mentioned, voice assistants.
Right now they run on this what's called a sequential loop.
So I'm talking, it listens, it transcribes it, it sends it back.
But when you have an actual conversation with someone, you'll know it's quite a bit different, right.
Like especially imagine like a heated argument.
I'm thinking of like a podcast.
But maybe you have a heated argument with someone you know and like you kind of talk over each other and they say something and you say no and then whatever right, or like a negotiation that's like really active, or people like talking really fast and they're all talking over each other.
That is incredibly hard slash, impossible for an am model.
I'm just thinking about myself having conversations with chat gpt on their voice mode, which of course, the voice mode.
It sounds natural.
It sounds good.
But the problem with chat GP voice mode is it's like giving me a long response and it kind of already answered the question in the first half of its response.
And so I'm like, okay, anyways, like I don't really need that.
Like move on.
Tell me more about this.
As soon as I start talking, it's like, it like freezes.
And it's like, uh, sorry, I missed that.
What did you say?
Like yeah, you know, it just like cuts itself off and it's like has this awkward moment where if I was just talking and I had kind of maybe it would finish it sentence and listen to me at the same time and then immediately start talking um to me throughout it.
And it's kind of reminds me of what we're actually getting out of, something like Claude cowork or maybe even opening eyes kind of doing this too, where mid response you can actually ping it with new information, as it's kind of writing on its response and it will change dynamically the response, which for quick little responses isn't that important because it doesn't pretty fast.
But when it's doing like something like cloud cowork, it's doing like this massive multi-step project or building a website or a webpage for you.
And you know it takes five minutes, and halfway through you send it a message and then it can kind of pivot and change.
I don't know.
It feels like the same concept, but they've done this in a technical way where it's not seconds or minutes.
It is milliseconds that they're able to get the turnaround and latency down, which is incredible.
Yeah, I mean, what do you think the implications are for this?
Because right now, I think people are still turned off by the idea of, for example, AI receptionists.
So it takes a lot of ability.
It takes a lot of...
You have to be very creative in how you present that to somebody as far as something that they would want for their business, especially because people are thinking back two years to the last time they called tech support somewhere and they had to talk to a very robotic person.
Yeah.
Answering service.
And they're like, I absolutely don't want that.
It's changed a lot already, but this is kind of going to take it to the next level.
So you know, how do we?
What are your thoughts as far as like?
How long will this take to be adopted by by the masses, or at least by by companies, so that they feel comfortable with this?
So, yeah, I mean, I think there is a limited research preview that's coming in the next few months.
They're going to have a wider release later this year.
I think, until this piece of the puzzle gets solved, and I'd be so curious if OpenAI just clones this, which very possibly they can, or if people integrate with them to help them do it.
I'm not sure how technical it actually gets.
But I think until this piece of the puzzle is solved, people are still gonna feel like it's quite robotic, because if you're talking and it's talking and then it pauses and you have this glitch, it's just this awkward moment and it just feels like a robot and we don't like it.
I think at the end of the day, the AI receptionist idea will be fully realized and people won't mind or care if it sounds perfect, acts perfect.
It accomplishes everything an AI receptionist could accomplish.
Maybe it has all of the same authorizations to issue refunds or whatever.
Like for me when I talk to chatbots on a company's website, if it can issue me my refund or help me with my problem or whatever You know.
Whatever the issue is, if it has authorization to actually solve my problem for me, I don't care if it's a human.
In fact, I'd rather it not be a human.
I'd rather just talk to a bot that can solve my problem.
But the problem is a lot of times we talk to these bots and they ask all these questions.
They're like, let me route you to a human.
And then it's just annoying.
You just feel like you kind of wasted your time.
Like there was a screener kind of like screening you.
So I think that's probably part of the issue.
And then there's like the whole latency problem, which is what they are basically solving here.
And I think once this rolls out and with the right authorization, these bots will be awesome.
The other thing that I do think is interesting, because you mentioned this a couple times of how...
People don't really like the concept of AI.
And even when you talk about a receptionist, they have these bad experiences they think back to.
One thing that I think would give people a better, more positive connotation to all of this is less of replacing a person with the AI so like replacing your receptionist with the AI and more of hey, you can do more than you did before.
Before your customer support people.
They would get around within six to 12 hours to customer support requests.
Or maybe you're getting super bogged down.
You get days behind.
You know that's a very bad experience.
Or maybe you're really bad at calling people back and it just doesn't happen.
So I think pitching things as like you were unable to do this before and now you're able versus like you can you know?
Now I just can replace that person or do it a little faster.
Like a little better isn't as good as like you are dropping the ball in this whole area.
And now this is going to fix that problem.
So I don't know.
That's a kind of a perspective shift that I think might be a better pitch for people.
For sure.
I have one more question for you, because I just thought of this, but someone who has seemed to expand beyond just voice agents.
Recently that I've noticed has been 11 Labs.
They have They've adopted on some actual agent functionality.
They have, you know, a little music production.
They're really trying to expand on beyond voice just to audio in general.
Do you think?
I mean, I feel like that would be a perfect pairing with thinking machines product.
Here the full duplex.
Do you think they're going to partner with someone like that?
Or do you think they're going to try to?
I'm just trying to think of how they can market this and actually make money off this new development they've.
If it's incredibly hard and technical and like they solved something that's very difficult for other people to solve, then I would see companies like Eleven Labs 100 partnering with them on this.
If it's just like they figured out something cool, but they figured out at first and everyone else can kind of figure it out, which is what it feels like happens so often in AI, right?
Like as soon as opening eyes, like we have a reasoning model and it's just running through a whole bunch of prompts.
Everyone could kind of see how the reasoning work.
They reverse engineered it.
All the Chinese labs cloned it.
And then now we all have the reasoning models, because they weren't like a crazy technical thing to figure out.
If that's the case, then I think you'll probably see companies like OpenAI and Google roll their own version of it.
You might see OpenAI and Google roll their own version regardless, because they don't really like integrating or working remotely.
With other people, they can just try to figure out how to figure it out.
But yeah, for a lot of smaller companies I mean, 11 Labs is getting to be one of the bigger companies now, but for ones that aren't opening AI and Google I think you'll see a lot of partnerships with companies, with them to be able to get stuff down.
I mean they're able to get the response time in 04 seconds, which is faster than OpenAI and Google.
And I mean, they solve a real problem.
But my question is when they say look, limited research previews coming in the next few months, wider rollout later this year, it's like yeah, it's later this year.
Is Google and OpenAI not going to have figured this out and already rolled this out?
So I think that'll be the big question is timing.
For sure.
Yeah, no, thanks for your input.
Hey, if you enjoyed this episode, we'd really appreciate a rating or review, wherever you're listening.
Those help us reach more people and we really appreciate them.
Also, be sure to check out our AI Hustle School community.
We release bonus content over there each week about how you can actually use some of these tools.
We try to stay up to date on the latest ones that are released. and give you guys tutorials.
Try them out.
See if they can actually make you money.
You're going to want to check out our school community.
We'd love to have you be a part of it.
Thanks for listening, and we'll see you next time.