OpenAI’s GPT-4o-Powered ChatGPT Is Now More (Terrifyingly) Conversational and Life-Like

OpenAI on Monday announced GPT-4o, a new flagship generative AI model that expands on the capabilities of its predecessor, GPT-4. The “o” in GPT-4o stands for “omni,” reflecting the model’s ability to handle multiple modalities, including text, speech, and video. 

One of the key improvements in GPT-4o is its ability to reason across voice, text, and vision, which OpenAI CTO Mira Murati believes is crucial for the future of human-machine interaction. The model’s integration into OpenAI’s AI-powered chatbot, ChatGPT, allows users to interact with the platform more naturally, as if it were a personal assistant. 

Users can now interrupt ChatGPT while it’s answering, and the model can respond in real-time, even picking up on nuances in the user’s voice and generating responses in various emotive styles — even singing.

GPT-4o also enhances ChatGPT’s vision capabilities, enabling it to analyze photos or desktop screens and answer related questions on topics ranging from software code to clothing brands. Murati suggests that future iterations of the model could allow ChatGPT to watch live events, such as sports games, and explain the rules to users.

In addition to its improved ease of use, GPT-4o boasts enhanced performance in around 50 languages and faster processing times at lower costs compared to GPT-4 Turbo. However, due to the risk of misuse, OpenAI plans to initially launch GPT-4o’s new audio capabilities to “a small group of trusted partners.”

Alongside the release of GPT-4o, OpenAI announced a refreshed ChatGPT UI on the web, a desktop version for macOS, and the expansion of previously paywalled features to free users, such as the ability to upload files and photos and search the web for answers.


Information for this story was found via Tech Crunch, Ars Technica, and the sources and companies mentioned. The author has no securities or affiliations related to the organizations discussed. Not a recommendation to buy or sell. Always do additional research and consult a professional before purchasing a security. The author holds no licenses.

Video Articles

Can the World Actually Supply $6 Copper? | Greg Ferron – PTX Metals

1911 Gold: The Power Of A Mine Restart

Is Gold Repeating the 2005 Setup Before The Big Run? | Geordie Mark

Recommended

Nord Precious Metals Hits Multiple Intervals Of Mineralization In Latest Drill Hole At Castle East

Goliath Resources Sees 13% Grade Boost As Stifel Draws Parallels To Great Bear

Related News

Shopify Employee Breaks NDA To Reveal Firm Quietly Replacing Laid Off Workers With AI

In a Twitter thread, a Shopify (TSX: SHOP) employee has broken their non-disclosure agreement (NDA)...

Thursday, July 20, 2023, 10:30:28 AM

xAI Confirms Layoffs During Reorganization as Half of Founding Team Departs

Elon Musk’s artificial intelligence startup xAI laid off employees this week as part of a...

Friday, February 13, 2026, 12:09:00 PM

OpenAI Board Found Out About the Launch of ChatGPT on Twitter

Former OpenAI board member Helen Toner revealed that the board was not informed about the...

Friday, May 31, 2024, 08:03:44 AM

Hype Over? ChatGPT’s Worldwide Traffic Is Down For The First Time Since It Launched

It appears that OpenAI’s popular large language model ChatGPT has already peaked.  Traffic for the...

Thursday, July 6, 2023, 03:06:00 PM

AI Music Company Attracts High-Profile Investors in $125M Round

Suno, a generative AI music company, has secured $125 million in its latest funding round,...

Wednesday, May 22, 2024, 10:07:54 AM