Google launches Gemini – an AI model 'stronger than GPT-4'

07.12.2023

geminiera |

Gemini — AI model synthesizing knowledge of 57 subjects

Gemini — AI model synthesizing knowledge of 57 subjects

Google Gemini synthesizes knowledge from 57 subjects to solve problems, being the first AI to surpass humans at the expert level.

Gemini launched on the evening of December 6, is Google's most advanced and general AI model to date, competing with OpenAI's GPT-4.

Unlike other popular large language models in recent times, Gemini is built in a multimodal direction, meaning it can generalize, operate and combine on many different types of information including text, code, audio, images and video.

Gemini's outstanding test results

To meet flexible usage needs, from data centers to mobile devices, Google said Gemini 1.0 will be offered in three different sizes, including: Gemini Ultra, Gemini Pro and Gemini Nano.

Gemini's outstanding test results

Special capabilities of the new model

Gemini FlexibleModel 8675 1701887275 |
Correlating the three dimension versions of the Gemini AI model. Image: Google

According to test results published by Google, Gemini Ultra scored 90% in the Massive Multitask Language Understanding (MMLU) test.

Special capabilities of the new model

With this result, Gemini is the first AI to surpass humans at the expert level, which scored 89.8% on the same test.

Additionally, this strongest version of Gemini also exceeded 30 out of 32 benchmarks in major language modeling research and development, scoring 59.4% in MMMU (massive multimodal cross-disciplinary understanding) capability, which covers multimodal tasks spanning different domains that require deliberative reasoning.

Gemini Ultra and Pro versions

Demis Hassabis, CEO of Google DeepMind, representative of Gemini Team, said the company wants to build a new generation of AI models inspired by the human way of recognizing and interacting with the world.

“Today, we are one step closer to this vision by introducing Gemini, the most advanced and general AI model ever developed by Google,” said Hassabis.

In addition to strong performance, Google says Gemini 1.0 is trained to recognize text, images, sounds and more at the same time, helping it better understand nuanced information and answer questions related to complex topics.

Gemini Ultra and Pro versions

gemini mm 02 1200 1701887275 1 |
Illustration of the types of information that Gemini can process, such as: text, photos, sounds, videos. Image: Google

According to Google, these features help Gemini read and extract information from hundreds of thousands of documents, thereby opening up the ability to create new breakthroughs in many fields, from science to finance in a short time.

During the launch, Google said that Gemini Ultra version is the version for the most complex tasks and is in the process of completing safety testing before officially launching.

Meanwhile, the Pro version is now used in the Bard chatbot.

This is also the biggest upgrade for Bard since its launch.

Luu Quy – VnExpress

Topic AI & Technology
Tags AI Technology Google

Read more

meta ra m t g i plus 1 | GUDJOB Meta Launches Plus Plans for Facebook, Instagram, and WhatsApp Meta rolls out Plus subscriptions Meta rolls out Plus subscriptions Meta is rolling out subscription plans for… 08.06.2026 google marketing live 2 | GUDJOB Google Marketing Live 2026: Gemini Powers the Next Generation of AI-Driven Advertising and Commerce On May 21, 2026, the Google Marketing Live (GML) 2026 event officially introduced the new generation… 25.05.2026 AI Website traffic | GUDJOB AI changes the traffic game: When clicks are no longer everything For years, traffic and clicks have been the central metrics for evaluating digital marketing performance. However,… 16.03.2026 AI Traffic | GUDJOB Why brands need to optimize their mobile apps and websites for AI Until now, brands have built their digital experiences around Google's search algorithm, but experts say that… 04.03.2026