OpenAI’s GPT-4o-Powered ChatGPT Is Now More (Terrifyingly) Conversational and Life-Like

OpenAI on Monday announced GPT-4o, a new flagship generative AI model that expands on the capabilities of its predecessor, GPT-4. The “o” in GPT-4o stands for “omni,” reflecting the model’s ability to handle multiple modalities, including text, speech, and video. 

One of the key improvements in GPT-4o is its ability to reason across voice, text, and vision, which OpenAI CTO Mira Murati believes is crucial for the future of human-machine interaction. The model’s integration into OpenAI’s AI-powered chatbot, ChatGPT, allows users to interact with the platform more naturally, as if it were a personal assistant. 

Users can now interrupt ChatGPT while it’s answering, and the model can respond in real-time, even picking up on nuances in the user’s voice and generating responses in various emotive styles — even singing.

GPT-4o also enhances ChatGPT’s vision capabilities, enabling it to analyze photos or desktop screens and answer related questions on topics ranging from software code to clothing brands. Murati suggests that future iterations of the model could allow ChatGPT to watch live events, such as sports games, and explain the rules to users.

In addition to its improved ease of use, GPT-4o boasts enhanced performance in around 50 languages and faster processing times at lower costs compared to GPT-4 Turbo. However, due to the risk of misuse, OpenAI plans to initially launch GPT-4o’s new audio capabilities to “a small group of trusted partners.”

Alongside the release of GPT-4o, OpenAI announced a refreshed ChatGPT UI on the web, a desktop version for macOS, and the expansion of previously paywalled features to free users, such as the ability to upload files and photos and search the web for answers.


Information for this story was found via Tech Crunch, Ars Technica, and the sources and companies mentioned. The author has no securities or affiliations related to the organizations discussed. Not a recommendation to buy or sell. Always do additional research and consult a professional before purchasing a security. The author holds no licenses.

Video Articles

Why Canada Has So Few Projects That Can Be Built Before 2030 | Dan Wilton – First Mining

Guanajuato Silver: Q3 Results Overshadowed By Silver Ripping

I Went to See the Highest Grade Silver on Earth | Nord Precious Metals

Recommended

Steadright Locks Up Goundafa Polymetallic Mine Under Binding MOU

Emerita Resources Awards Contract For Pre-Feasibility Study On Iberian Belt West Project

Related News

C3.ai Erases Gains With Poor Revenue Outlook

Artificial intelligence software developer C3.ai (NYSE: AI) published fourth-quarter results after the bell on Wednesday,...

Friday, June 2, 2023, 10:59:08 AM

Apple Testing AI Search Tools for Safari, Executive Tells Antitrust Trial

Apple‘s (Nasdaq: AAPL) senior services executive testified on Wednesday that the company is considering multiple...

Saturday, May 10, 2025, 09:28:00 AM

Canadian News Giants Sues OpenAI For Exploiting Journalism for Profit

A coalition of Canadian news organizations—including The Canadian Press, Torstar, The Globe and Mail, Postmedia,...

Monday, December 2, 2024, 02:54:00 PM

OpenAI Power Struggle: Employees Vs. Board Vs. Sam Altman, Explained

While the real reason behind Sam Altman’s sudden dismissal as CEO of OpenAI seems to...

Tuesday, November 21, 2023, 09:54:00 AM

Google Considers Nuclear Power for AI Data Centers

Google (Nasdaq: GOOG) CEO Sundar Pichai has hinted at the possibility of using nuclear energy...

Sunday, October 6, 2024, 07:39:00 AM