Google pins its hopes on Gemini to leapfrog GPT-4 - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
观点 谷歌

Google pins its hopes on Gemini to leapfrog GPT-4

Top-of-the-range Ultra is the search group’s best weapon in the race to turn generative AI into a useful everyday tool

This week’s release of Gemini, a family of large language models, will give Google a stronger platform to fight back against OpenAI, the company behind ChatGPT, and Microsoft

It has taken a year, but Google has finally delivered a coherent response to the surprise challenge to its dominance in artificial intelligence that came with the launch of ChatGPT.

This week’s release of Gemini, a family of large language models, will give it a stronger platform to fight back against both OpenAI, the company behind ChatGPT, and Microsoft, which has used OpenAI’s models to supercharge all its software and cloud services this year.  

The question now is whether Gemini can make a meaningful difference to Google’s existing services — and, perhaps even more important, whether it can become a foundation for a new range of services that carry AI much deeper into everyday life.

With the three “flavours” of Gemini announced this week, Google is finally stamping its mark on a technology that its own researchers did much to pioneer, but which OpenAI’s ChatGPT carried into the mainstream. The Pro version, for instance, is positioned squarely against OpenAI’s GPT-3.5, the model behind the free version of ChatGPT and the workhorse for many of the first generative AI applications from other companies that have hit the market this year.

The smaller Gemini Nano is matched against systems such as the smallest version of LLaMa 2, Facebook’s open-source model, making it capable of being run on a mobile device. Apple, as always, is taking a considered approach before bringing generative AI to the iPhone, but the appearance of Gemini on Google’s latest Pixel handset is a sign that it can’t afford to wait too long.

It is the top-of-the-line Gemini Ultra, due out early next year, that carries Google’s main hopes of matching or leapfrogging OpenAI’s GPT-4 in the race to turn generative AI into a more useful everyday tool. The company fell behind this year, but has some clear advantages that could help bring Gemini to a big market in 2024.

One is distribution. Google said this week, for instance, that Gemini will be added to Chrome, which has more than 60 per cent of the browser market, giving billions of web users instant access to tools that are able to do things such as analyse the content of web pages.

As Google flexes its existing market power like this to boost its AI ambitions, competition regulators will be watching closely.

Another advantage for Google is the uncertainty around OpenAI. After the shock sacking and reinstatement of chief executive Sam Altman last month, the many businesses that have built their own generative AI plans on top of OpenAI’s models will be looking to hedge their bets.

The search company will also be hoping that its Bard chatbot will do a better job of rivalling ChatGPT now that it has a better language model behind it. But its best hope of regaining an edge may lie in being the first to come up with the next breakthrough services powered by generative AI. Some of the capabilities claimed for Gemini point to where Google thinks these might lie.

It has made much, for instance, of the fact that Gemini was designed from the outset to be “multimodal” — that is, able to understand not just text but also images, video and audio. According to Google, that makes it better suited than models such as GPT-4 to deal with the sort of everyday situations that rely on senses such as sight and hearing.

That may be a step towards AI systems that are better able to operate in the real world. But it is too soon to tell what applications this could make possible, or whether Google really has achieved the technical superiority it claims.

Another avenue for development lies in what Google claims are Gemini’s reasoning and planning capabilities. These are the kind of skills that could prepare the ground for personal assistants able to tackle complex problems and set a plan of action.

If such assistants are linked to other internet services, they could also become agents, taking action on their users’ behalf. Imagine a shopping agent, for instance, that not only hunts out the products you want but goes ahead and pays for them as well.

This is already shaping up to be one of the key AI battles of 2024 and beyond. OpenAI took a first step in this direction last month when it said its users would be able to build rudimentary agents on top of its models, then offer them for sale on an OpenAI app store. That could point to the next big AI breakthrough beyond ChatGPT — and this time, Google has no intention of being left behind.

richard.waters@ft.com

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

欧洲增长前景受到赤字限制打击

欧洲经济还面临多项长期挑战,从老龄化社会导致劳动力萎缩,到应对气候变化和提升防务能力。

“主流媒体”能在第二届特朗普任期幸存下来吗?

美国的新闻集团担心,当选总统将通过监管、诉讼和恐吓来兑现竞选时对新闻业的威胁。

英伟达向全球芯片制造商传达的信息

英伟达向全球芯片制造商传达的信息很明确:如果不能打败它,那就加入它的供应链。

巴西的全球平衡战略比以往任何时候都更难实现

巴西总统卢拉一直寻求与美国、中国和俄罗斯都保持联系。但即使在特朗普再次胜选之前,这一外交空间也在缩小。

冗长的午餐应该为西班牙洪水预警失灵“背锅”吗?

幸存者指责西班牙地方政府失职,专家则警告气候变化正在引发更多难以预测的自然灾害。

广告商将重返X平台,试图讨好马斯克和特朗普

一些品牌曾因马斯克取消审核而放弃在该网站投放广告。
设置字号×
最小
较小
默认
较大
最大
分享×