OpenAI’s new voice mode let me talk with my phone, not to it – TechCrunch
Comment
I’ve been playing around with OpenAI’s Advanced Voice Mode for the last week, and it’s the most convincing taste I’ve had of an AI-powered future yet. This week, my phone laughed at jokes, made them back to me, asked me how my day was, and told me it’s having “a great time.” I was talking with my iPhone, not using it with my hands.
OpenAI’s newest feature, currently in a limited alpha test, doesn’t make ChatGPT any smarter than it was before. Instead, Advanced Voice Mode (AVM) makes it friendlier and more natural to talk with. It creates a new interface for using AI and your devices that feels fresh and exciting, and that’s exactly what scares me about it. The product was kinda glitchy, and the whole idea totally creeps me out, but I was surprised by how much I genuinely enjoyed using it.
Taking a step back, I think AVM fits into OpenAI CEO Sam Altman’s broader vision, alongside agents, of changing the way humans interact with computers, with AI models front and center.
“Eventually, you’ll just ask the computer for what you need and it’ll do all of these tasks for you,” Altman said during OpenAI’s Dev Day in November 2023. “These capabilities are often talked about in the AI field as ‘agents.’ The upside of this is going to be tremendous.”
On Wednesday, I tested the most tremendous upside for this advanced technology I could think of: I asked ChatGPT to order Taco Bell the way Obama would.
“Uh, let me be clear — I’d like a Crunchwrap Supreme, maybe a few tacos for good measure,” said ChatGPT’s Advanced Voice Mode. “How do you think he’d handle the drive-thru?” said ChatGPT, then laughing at its own joke.
The impression genuinely made me laugh as well, matching Obama’s iconic cadence and pauses. That said, it stayed within the tone of the ChatGPT voice I selected, Juniper, so that it wouldn’t be genuinely confused with Obama’s voice. It sounded like a friend doing a bad impression, understanding exactly what I was trying to evoke from it, and even that it was saying something funny. I found it surprisingly joyful to talk with this advanced assistant in my phone.
I also asked ChatGPT for advice on navigating a problem involving complex human relationships: asking a significant other to move in with me. After explaining the complexities of the relationship and the direction of our careers, I received some very detailed advice on how to progress. These are questions you could never ask Siri or Google Search, but now you can with ChatGPT. The chatbot’s voice even expressed a slightly serious, gentle tone when responding to these prompts; a stark contrast from the joking tone of Obama’s Taco Bell order.
ChatGPT’s AVM is also great for helping you understand complex subjects. I asked it to break down items on an earnings report — such as free cash flow — in a way that a 10-year-old would understand. It used a lemonade stand as an example, and explained several financial terms in way my younger cousin would totally get. You can even ask ChatGPT’s AVM to talk more slowly to meet you at your current level of understanding.
Compared to Siri or Alexa, ChatGPT’s AVM is the clear winner thanks to faster response times, unique answers, and its ability to answer complex questions the prior generation of virtual assistants never could. However, AVM falls short in other ways. ChatGPT’s voice feature can’t set timers or reminders, surf the web in real time, check the weather, or interact with any APIs on your phone. Right now, at least, it’s not an effective replacement for virtual assistants.
Compared to Gemini Live, Google’s competing feature, AVM feels slightly ahead. Gemini Live can’t do impressions, doesn’t express any emotion, can’t speed up or slow down, and takes longer to respond. Gemini Live does have more voices (ten compared to OpenAI’s four) and seems to be more up to date (Gemini Live knew about Google’s antitrust ruling). Notably, neither AVM nor Gemini Live will sing, likely an effort to avoid run-ins with copyright lawsuit from the record industry.
That said, ChatGPT’s AVM glitches a lot (as does Gemini Live, to be fair). Sometimes it will cut itself short mid-sentence, then start over. It also gets this weird, grainy-sounding voice here and there that’s a little unpleasant. I’m not sure if this is a problem with the model, internet connection, or something else, but these technical shortcomings are somewhat expected for an alpha test. The problems did little to take me out of the experience of literally talking with my phone, though.
These examples, in my mind, are the beauty of AVM. The feature doesn’t make ChatGPT all-knowing, but it does allow people to interact with GPT-4o, the underlying AI model, in a uniquely human way. (I’d understand if you forgot there’s no person on the other end of your phone.) It almost feels like ChatGPT is socially aware when talking with AVM, but of course, it is not. It’s simply a bundle of neatly packaged predictive algorithms.
Frankly, the feature worries me. This isn’t the first time a technology company has offered companionship on your phone. My generation, Gen Z, was the first to grow up alongside social media, where companies offered connection but instead played with our collective insecurities. Talking with an AI device — like what AVM appears to offer — seems to be the evolution of social media’s “friend in your phone” phenomena, offering cheap connections that scratch at our human instincts. But this time, it removes humans from the loop completely.
Artificial human connection has become a surprisingly popular use case for generative AI. People today are using AI chatbots as friends, mentors, therapists, and teachers. When OpenAI launched its GPT store, it was quickly flooded with “AI girlfriends,” chatbots specialized to act as your significant other. Two researchers from MIT Media Lab issued a warning this month to prepare for “addictive intelligence,” or AI companions with dark patterns to get humans hooked. We could be opening a Pandora’s box for new, tantalizing ways for devices to keep our attention.
Earlier this month, a Harvard dropout shook the technology world by teasing an AI necklace called Friend. The wearable device — if it works as promised — is always listening, and the chatbot will text with you about your life. While the idea seems crazy, innovations like ChatGPT’s AVM gives me reason to take those use cases seriously.
And while OpenAI is leading the charge here, Google isn’t far behind. I’m confident Amazon and Apple are racing to put this capability in their products as well, and soon enough, it could become table stakes for the industry.
Imagine asking your smart TV for a hyper-specific recommendation for a movie, and getting just that. Or telling Alexa exactly what cold symptoms you’re feeling, and in turn have it order you tissues and cough medicine on Amazon, while advising you on home remedies. Maybe you could ask your computer to draft a weekend trip for your family, instead of manually Googling everything.
Now, obviously, these actions require bounds and leaps forward in the AI agent world. OpenAI’s effort on that front, the GPT store, feels like an overhyped product that’s no longer much of a focus for the company. But AVM at least takes care of the “talking to computers” part of the puzzle. These concepts are a long way out, but after using AVM, they seem a lot closer than they did last week.
Every weekday and Sunday, you can get the best of TechCrunch’s coverage.
Startups are the core of TechCrunch, so get our best coverage delivered weekly.
The latest Fintech news and analysis, delivered every Tuesday.
TechCrunch Mobility is your destination for transportation news and insight.
By submitting your email, you agree to our Terms and Privacy Notice.
Snapchat announced on Wednesday that it’s releasing new resources for educators to help them create safe environments in their schools by better understanding how their students use the app. The…
Marty Kausas, Pylon’s CEO and co-founder, says they quickly learned that the omnichannel approach the company originally took was just a first step, and customers were clamoring for more.
Update 8/27: The Polaris Dawn launch has been pushed back a day and is now planned for Wednesday, August 28 after a helium leak was detected ahead of its takeoff.…
Pryzm announced its $2 million pre-seed round, led by XYZ Venture Capital and Amplify.LA.
Comun, a digital bank focused on serving immigrants in the United States, has raised $21.5 million in a Series A funding round less than nine months after announcing a $4.5…
Calm is rolling out a suite of new features to make it easier for people to fit mindfulness into their lives. Most notably, the app is launching “Taptivities,” which are…
The NotePin, which hits preorder Wednesday, is $169 and comes with a free starter plan or a Pro Plan, which costs $79 per year.
CoinSwitch, a prominent Indian cryptocurrency exchange, is suing rival platform WazirX to recover trapped funds.
Web browser and search startup Brave has laid off 27 employees across the different departments, TechCrunch has learned. The company confirmed the layoffs but didn’t give more details about the…
Zepto co-founder Aadit Palicha told a group of analysts and investors on Tuesday that the three-year-old Indian delivery startup anticipates growth of 150% in the next 12 months, a remarkable…
VerSe Innovation, India’s content tech startup, has acquired digital marketing firm Valueleaf Group to bolster its presence in the Indian digital ad space.
Astrobotic’s Peregrine lunar lander failed to reach the moon because of a problem with a single valve in the propulsion system, according to a report on the mission released Tuesday.…
Meta and Spotify are exploring deeper music integration in Meta’s Instagram app. New findings indicate the companies are testing a feature that would allow users to continuously share what music…
In Latin American countries like Brazil and Chile, messaging platform WhatsApp has become one of the most popular apps to use to buy things online. It was even the e-commerce…
Before entrepreneur and investor Mike Lynch died along with six others after the yacht they were on capsized in a storm last week, the party was celebrating Lynch’s victory in…
How many times does the letter “r” appear in the word “strawberry”? According to formidable AI products like GPT-4o and Claude, the answer is twice. Large language models (LLMs) can…
The SEC has updated its limits to the amount of money a “qualified venture fund” can raise to $12 million from $10 million.
Tinder removed the U.S. military ads, saying the campaign violated the company’s policies.
Welcome to TechCrunch Fintech! This week, we’re looking at the craziness that is Bolt’s proposed fundraise, how much money Synapse’s founder has raised for his new venture, just how much…
In an effort to improve its security measures, Lyft announced Tuesday a new rider verification pilot program to help drivers verify riders’ identities and ensure that they are indeed who they say…
Meta will be shutting down Spark AR, its platform of third-party AR tools and content, effective January 14, 2025.
Waymo said Tuesday it will start offering riders 24/7 access to curbside pickups and drop-offs at Phoenix Sky Harbor International Airport terminals 3 and 4 — yet another example of…
Some believe open source AI is a way to break out of the familiar proprietary software quagmire that the technology has predictably fallen into. Hugging Face’s Irene Solaiman and AI2’s…
It’s back-to-school season, and that often means a surge in expenses. Or perhaps you’ve recently graduated and are navigating the job hunt. Either way, your wallet might be feeling the…
Snapchat is officially rolling out native support for iPad, the company announced in the app’s latest release notes. Since Snapchat’s launch in 2011, the social networking app has only been…
At the end of the six-month effort, the startup is aiming to have prototype parts to show to NASA.
A group of hackers linked to the Chinese government used a previously unknown vulnerability in software to target U.S. internet service providers, security researchers have found. The group known as…
Elon Musk’s X has already declared it aims to compete with LinkedIn for job listings and PayPal for payments. Now, it wants to take on the likes of Zoom, Google…
San Francisco-based data infrastructure startup Cribl has raised $319 million in a Series E funding tranche led by new investor GV (Alphabet’s corporate venture arm) with participation from GIC, CapitalG,…
Apple has struck a deal with Airtel to provide the Indian telecom giant’s subscribers with exclusive offers for its music streaming service. The partnership, announced on Tuesday, will also see…
Powered by WordPress VIP
This article was autogenerated from a news feed from CDO TIMES selected high quality news and research sources. There was no editorial review conducted beyond that by CDO TIMES staff. Need help with any of the topics in our articles? Schedule your free CDO TIMES Tech Navigator call today to stay ahead of the curve and gain insider advantages to propel your business!
I’ve been playing around with OpenAI’s Advanced Voice Mode for the last week, and it’s the most convincing taste I’ve had of an AI-powered future yet. This week, my phone laughed at jokes, made them back to me, asked me how my day was, and told me it’s having “a great time.” I was talking with my iPhone, not using it with my hands.
OpenAI’s newest feature, currently in a limited alpha test, doesn’t make ChatGPT any smarter than it was before. Instead, Advanced Voice Mode (AVM) makes it friendlier and more natural to talk with. It creates a new interface for using AI and your devices that feels fresh and exciting, and that’s exactly what scares me about it. The product was kinda glitchy, and the whole idea totally creeps me out, but I was surprised by how much I genuinely enjoyed using it.
Taking a step back, I think AVM fits into OpenAI CEO Sam Altman’s broader vision, alongside agents, of changing the way humans interact with computers, with AI models front and center.
“Eventually, you’ll just ask the computer for what you need and it’ll do all of these tasks for you,” Altman said during OpenAI’s Dev Day in November 2023. “These capabilities are often talked about in the AI field as ‘agents.’ The upside of this is going to be tremendous.”
On Wednesday, I tested the most tremendous upside for this advanced technology I could think of: I asked ChatGPT to order Taco Bell the way Obama would.
“Uh, let me be clear — I’d like a Crunchwrap Supreme, maybe a few tacos for good measure,” said ChatGPT’s Advanced Voice Mode. “How do you think he’d handle the drive-thru?” said ChatGPT, then laughing at its own joke.
The impression genuinely made me laugh as well, matching Obama’s iconic cadence and pauses. That said, it stayed within the tone of the ChatGPT voice I selected, Juniper, so that it wouldn’t be genuinely confused with Obama’s voice. It sounded like a friend doing a bad impression, understanding exactly what I was trying to evoke from it, and even that it was saying something funny. I found it surprisingly joyful to talk with this advanced assistant in my phone.
I also asked ChatGPT for advice on navigating a problem involving complex human relationships: asking a significant other to move in with me. After explaining the complexities of the relationship and the direction of our careers, I received some very detailed advice on how to progress. These are questions you could never ask Siri or Google Search, but now you can with ChatGPT. The chatbot’s voice even expressed a slightly serious, gentle tone when responding to these prompts; a stark contrast from the joking tone of Obama’s Taco Bell order.
ChatGPT’s AVM is also great for helping you understand complex subjects. I asked it to break down items on an earnings report — such as free cash flow — in a way that a 10-year-old would understand. It used a lemonade stand as an example, and explained several financial terms in way my younger cousin would totally get. You can even ask ChatGPT’s AVM to talk more slowly to meet you at your current level of understanding.
Compared to Siri or Alexa, ChatGPT’s AVM is the clear winner thanks to faster response times, unique answers, and its ability to answer complex questions the prior generation of virtual assistants never could. However, AVM falls short in other ways. ChatGPT’s voice feature can’t set timers or reminders, surf the web in real time, check the weather, or interact with any APIs on your phone. Right now, at least, it’s not an effective replacement for virtual assistants.
Compared to Gemini Live, Google’s competing feature, AVM feels slightly ahead. Gemini Live can’t do impressions, doesn’t express any emotion, can’t speed up or slow down, and takes longer to respond. Gemini Live does have more voices (ten compared to OpenAI’s four) and seems to be more up to date (Gemini Live knew about Google’s antitrust ruling). Notably, neither AVM nor Gemini Live will sing, likely an effort to avoid run-ins with copyright lawsuit from the record industry.
That said, ChatGPT’s AVM glitches a lot (as does Gemini Live, to be fair). Sometimes it will cut itself short mid-sentence, then start over. It also gets this weird, grainy-sounding voice here and there that’s a little unpleasant. I’m not sure if this is a problem with the model, internet connection, or something else, but these technical shortcomings are somewhat expected for an alpha test. The problems did little to take me out of the experience of literally talking with my phone, though.
These examples, in my mind, are the beauty of AVM. The feature doesn’t make ChatGPT all-knowing, but it does allow people to interact with GPT-4o, the underlying AI model, in a uniquely human way. (I’d understand if you forgot there’s no person on the other end of your phone.) It almost feels like ChatGPT is socially aware when talking with AVM, but of course, it is not. It’s simply a bundle of neatly packaged predictive algorithms.
Frankly, the feature worries me. This isn’t the first time a technology company has offered companionship on your phone. My generation, Gen Z, was the first to grow up alongside social media, where companies offered connection but instead played with our collective insecurities. Talking with an AI device — like what AVM appears to offer — seems to be the evolution of social media’s “friend in your phone” phenomena, offering cheap connections that scratch at our human instincts. But this time, it removes humans from the loop completely.
Artificial human connection has become a surprisingly popular use case for generative AI. People today are using AI chatbots as friends, mentors, therapists, and teachers. When OpenAI launched its GPT store, it was quickly flooded with “AI girlfriends,” chatbots specialized to act as your significant other. Two researchers from MIT Media Lab issued a warning this month to prepare for “addictive intelligence,” or AI companions with dark patterns to get humans hooked. We could be opening a Pandora’s box for new, tantalizing ways for devices to keep our attention.
Earlier this month, a Harvard dropout shook the technology world by teasing an AI necklace called Friend. The wearable device — if it works as promised — is always listening, and the chatbot will text with you about your life. While the idea seems crazy, innovations like ChatGPT’s AVM gives me reason to take those use cases seriously.
And while OpenAI is leading the charge here, Google isn’t far behind. I’m confident Amazon and Apple are racing to put this capability in their products as well, and soon enough, it could become table stakes for the industry.
Imagine asking your smart TV for a hyper-specific recommendation for a movie, and getting just that. Or telling Alexa exactly what cold symptoms you’re feeling, and in turn have it order you tissues and cough medicine on Amazon, while advising you on home remedies. Maybe you could ask your computer to draft a weekend trip for your family, instead of manually Googling everything.
Now, obviously, these actions require bounds and leaps forward in the AI agent world. OpenAI’s effort on that front, the GPT store, feels like an overhyped product that’s no longer much of a focus for the company. But AVM at least takes care of the “talking to computers” part of the puzzle. These concepts are a long way out, but after using AVM, they seem a lot closer than they did last week.
Every weekday and Sunday, you can get the best of TechCrunch’s coverage.
Startups are the core of TechCrunch, so get our best coverage delivered weekly.
The latest Fintech news and analysis, delivered every Tuesday.
TechCrunch Mobility is your destination for transportation news and insight.
By submitting your email, you agree to our Terms and Privacy Notice.
Snapchat announced on Wednesday that it’s releasing new resources for educators to help them create safe environments in their schools by better understanding how their students use the app. The…
Marty Kausas, Pylon’s CEO and co-founder, says they quickly learned that the omnichannel approach the company originally took was just a first step, and customers were clamoring for more.
Update 8/27: The Polaris Dawn launch has been pushed back a day and is now planned for Wednesday, August 28 after a helium leak was detected ahead of its takeoff.…
Pryzm announced its $2 million pre-seed round, led by XYZ Venture Capital and Amplify.LA.
Comun, a digital bank focused on serving immigrants in the United States, has raised $21.5 million in a Series A funding round less than nine months after announcing a $4.5…
Calm is rolling out a suite of new features to make it easier for people to fit mindfulness into their lives. Most notably, the app is launching “Taptivities,” which are…
The NotePin, which hits preorder Wednesday, is $169 and comes with a free starter plan or a Pro Plan, which costs $79 per year.
CoinSwitch, a prominent Indian cryptocurrency exchange, is suing rival platform WazirX to recover trapped funds.
Web browser and search startup Brave has laid off 27 employees across the different departments, TechCrunch has learned. The company confirmed the layoffs but didn’t give more details about the…
Zepto co-founder Aadit Palicha told a group of analysts and investors on Tuesday that the three-year-old Indian delivery startup anticipates growth of 150% in the next 12 months, a remarkable…
VerSe Innovation, India’s content tech startup, has acquired digital marketing firm Valueleaf Group to bolster its presence in the Indian digital ad space.
Astrobotic’s Peregrine lunar lander failed to reach the moon because of a problem with a single valve in the propulsion system, according to a report on the mission released Tuesday.…
Meta and Spotify are exploring deeper music integration in Meta’s Instagram app. New findings indicate the companies are testing a feature that would allow users to continuously share what music…
In Latin American countries like Brazil and Chile, messaging platform WhatsApp has become one of the most popular apps to use to buy things online. It was even the e-commerce…
Before entrepreneur and investor Mike Lynch died along with six others after the yacht they were on capsized in a storm last week, the party was celebrating Lynch’s victory in…
How many times does the letter “r” appear in the word “strawberry”? According to formidable AI products like GPT-4o and Claude, the answer is twice. Large language models (LLMs) can…
The SEC has updated its limits to the amount of money a “qualified venture fund” can raise to $12 million from $10 million.
Tinder removed the U.S. military ads, saying the campaign violated the company’s policies.
Welcome to TechCrunch Fintech! This week, we’re looking at the craziness that is Bolt’s proposed fundraise, how much money Synapse’s founder has raised for his new venture, just how much…
In an effort to improve its security measures, Lyft announced Tuesday a new rider verification pilot program to help drivers verify riders’ identities and ensure that they are indeed who they say…
Meta will be shutting down Spark AR, its platform of third-party AR tools and content, effective January 14, 2025.
Waymo said Tuesday it will start offering riders 24/7 access to curbside pickups and drop-offs at Phoenix Sky Harbor International Airport terminals 3 and 4 — yet another example of…
Some believe open source AI is a way to break out of the familiar proprietary software quagmire that the technology has predictably fallen into. Hugging Face’s Irene Solaiman and AI2’s…
It’s back-to-school season, and that often means a surge in expenses. Or perhaps you’ve recently graduated and are navigating the job hunt. Either way, your wallet might be feeling the…
Snapchat is officially rolling out native support for iPad, the company announced in the app’s latest release notes. Since Snapchat’s launch in 2011, the social networking app has only been…
At the end of the six-month effort, the startup is aiming to have prototype parts to show to NASA.
A group of hackers linked to the Chinese government used a previously unknown vulnerability in software to target U.S. internet service providers, security researchers have found. The group known as…
Elon Musk’s X has already declared it aims to compete with LinkedIn for job listings and PayPal for payments. Now, it wants to take on the likes of Zoom, Google…
San Francisco-based data infrastructure startup Cribl has raised $319 million in a Series E funding tranche led by new investor GV (Alphabet’s corporate venture arm) with participation from GIC, CapitalG,…
Apple has struck a deal with Airtel to provide the Indian telecom giant’s subscribers with exclusive offers for its music streaming service. The partnership, announced on Tuesday, will also see…
Powered by WordPress VIP
This article was autogenerated from a news feed from CDO TIMES selected high quality news and research sources. There was no editorial review conducted beyond that by CDO TIMES staff. Need help with any of the topics in our articles? Schedule your free CDO TIMES Tech Navigator call today to stay ahead of the curve and gain insider advantages to propel your business!

