OpenAI’s GPT-4o-Powered ChatGPT Is Now More (Terrifyingly) Conversational and Life-Like

OpenAI on Monday announced GPT-4o, a new flagship generative AI model that expands on the capabilities of its predecessor, GPT-4. The “o” in GPT-4o stands for “omni,” reflecting the model’s ability to handle multiple modalities, including text, speech, and video. 

One of the key improvements in GPT-4o is its ability to reason across voice, text, and vision, which OpenAI CTO Mira Murati believes is crucial for the future of human-machine interaction. The model’s integration into OpenAI’s AI-powered chatbot, ChatGPT, allows users to interact with the platform more naturally, as if it were a personal assistant. 

Users can now interrupt ChatGPT while it’s answering, and the model can respond in real-time, even picking up on nuances in the user’s voice and generating responses in various emotive styles — even singing.

GPT-4o also enhances ChatGPT’s vision capabilities, enabling it to analyze photos or desktop screens and answer related questions on topics ranging from software code to clothing brands. Murati suggests that future iterations of the model could allow ChatGPT to watch live events, such as sports games, and explain the rules to users.

In addition to its improved ease of use, GPT-4o boasts enhanced performance in around 50 languages and faster processing times at lower costs compared to GPT-4 Turbo. However, due to the risk of misuse, OpenAI plans to initially launch GPT-4o’s new audio capabilities to “a small group of trusted partners.”

Alongside the release of GPT-4o, OpenAI announced a refreshed ChatGPT UI on the web, a desktop version for macOS, and the expansion of previously paywalled features to free users, such as the ability to upload files and photos and search the web for answers.


Information for this story was found via Tech Crunch, Ars Technica, and the sources and companies mentioned. The author has no securities or affiliations related to the organizations discussed. Not a recommendation to buy or sell. Always do additional research and consult a professional before purchasing a security. The author holds no licenses.

Video Articles

First Majestic Q3 Earnings: Another RECORD Quarter!

Barrick Q3 Earnings: Juicing Shareholder Returns Amid Declining Production

Wheaton Q3 Earnings: Cash Operating Margins Skyrocket

Recommended

Goliath Resources Extends High Grade Zone To 580 Metres In Latest Assays

Emerita Resources Hits 2.7% Copper, 1.85 g/t Gold Over 9.6 Metres At El Cura

Related News

No More Freeloading: Reddit Wants To Be Paid For AI Training

Reddit, with its vast data pool of human conversations about every imaginable topic collected over...

Wednesday, April 19, 2023, 02:14:00 PM

Authors Sue Anthropic for Alleged ‘Large-Scale Theft’ of Copyrighted Books

AI startup Anthropic is facing a class-action lawsuit alleging copyright infringement. Filed on Monday in...

Wednesday, August 21, 2024, 04:14:00 PM

Meta’s New AI Chatbot Said That Mark Zuckerberg’s Company ‘Exploits People For Money’

BlenderBot 3, Meta Platforms’ (NASDAQ: META) latest artificial intelligence-powered chatbot was recently released for a...

Monday, August 15, 2022, 04:35:00 PM

A Week Before Trump Transition, Biden Team Proposes Sweeping AI Chip Controls

The Biden administration proposed new controls Monday over the export of advanced AI computer chips,...

Tuesday, January 14, 2025, 12:59:00 PM

Google Considers Nuclear Power for AI Data Centers

Google (Nasdaq: GOOG) CEO Sundar Pichai has hinted at the possibility of using nuclear energy...

Sunday, October 6, 2024, 07:39:00 AM