OpenAI’s GPT-4o-Powered ChatGPT Is Now More (Terrifyingly) Conversational and Life-Like

OpenAI on Monday announced GPT-4o, a new flagship generative AI model that expands on the capabilities of its predecessor, GPT-4. The “o” in GPT-4o stands for “omni,” reflecting the model’s ability to handle multiple modalities, including text, speech, and video. 

One of the key improvements in GPT-4o is its ability to reason across voice, text, and vision, which OpenAI CTO Mira Murati believes is crucial for the future of human-machine interaction. The model’s integration into OpenAI’s AI-powered chatbot, ChatGPT, allows users to interact with the platform more naturally, as if it were a personal assistant. 

Users can now interrupt ChatGPT while it’s answering, and the model can respond in real-time, even picking up on nuances in the user’s voice and generating responses in various emotive styles — even singing.

GPT-4o also enhances ChatGPT’s vision capabilities, enabling it to analyze photos or desktop screens and answer related questions on topics ranging from software code to clothing brands. Murati suggests that future iterations of the model could allow ChatGPT to watch live events, such as sports games, and explain the rules to users.

In addition to its improved ease of use, GPT-4o boasts enhanced performance in around 50 languages and faster processing times at lower costs compared to GPT-4 Turbo. However, due to the risk of misuse, OpenAI plans to initially launch GPT-4o’s new audio capabilities to “a small group of trusted partners.”

Alongside the release of GPT-4o, OpenAI announced a refreshed ChatGPT UI on the web, a desktop version for macOS, and the expansion of previously paywalled features to free users, such as the ability to upload files and photos and search the web for answers.


Information for this story was found via Tech Crunch, Ars Technica, and the sources and companies mentioned. The author has no securities or affiliations related to the organizations discussed. Not a recommendation to buy or sell. Always do additional research and consult a professional before purchasing a security. The author holds no licenses.

Video Articles

Why the Market May Be Misreading Iran | David Woo

Why US Fertilizer Supply Could Matter a Lot More Now | Pat Varas – Sage Potash

Roscan Gold: Mali Discount Hits Kandiole PEA

Recommended

First Majestic Aims To Restart Production At Jerritt Canyon In H2 2027

Mercado Minerals Identifies A Series Of New Targets Following LiDAR Survey At Copalito

Related News

OpenAI Doesn’t Expect to Be Profitable Until 2029

OpenAI‘s backers are a long-ish way from making money. Recent financial projections indicate the AI...

Thursday, October 10, 2024, 01:23:00 PM

AI Giant OpenAI Plans Shift to For-Profit Model, Altman to Receive 7% Stake

OpenAI is planning a significant restructuring of its core business, as first reported by Reuters...

Friday, September 27, 2024, 07:30:00 AM

OpenAI Shuts Down Sora as Disney Deal Collapses

OpenAI announced Tuesday it is shutting down Sora, its AI-powered video generation app, just six...

Wednesday, March 25, 2026, 11:29:00 AM

Google’s Bard Is Staying Away from The European Union

Google’s parent-co Alphabet (NASDAQ: GOOGL) officially rolled out Bard, its ChatGPT-rival generative AI chatbot, at...

Monday, May 15, 2023, 02:16:00 PM

Microsoft Equips Bing With ChatGPT: “The Race Starts Today”

Microsoft Corp (NASDAQ: MSFT) fired the first shot in the race for an AI-enabled search...

Wednesday, February 8, 2023, 10:01:55 AM