AI

Google brings Gemini Pro to Vertex AI

Kommentar

Google Cloud logo on the side of.a building with a one way sign in the foreground.
Image Credits: 400tmax / Getty Images

After coming to Bard and the Pixel 8 Pro last week, Gemini, Google’s recently announced flagship GenAI model family, is launching for Google Cloud customers using Vertex AI.

Gemini Pro, a lightweight version of a more capable Gemini model, Gemini Ultra, currently in private preview for a “select set” of customers, is now accessible in public preview in Vertex AI, Google’s fully managed AI dev platform, via the new Gemini Pro API. The API is free to use “within limits” for the time being (more on what that means later) and supports 38 languages and regions including Europe, as well as features like chat functionality and filtering.

“Gemini’s a state-of-the-art natively multimodal model that has sophisticated reasoning advanced coding skills,” Google Cloud CEO Thomas Kurian said during a press briefing on Tuesday. “[Now,] developers will be able to build their own applications against it.”

Gemini Pro API

By default, the Gemini Pro API in Vertex accepts text as input and generates text as output, similar to generative text model APIs like Anthropic’s, AI21’s and Cohere’s. An additional endpoint, Gemini Pro Vision, also launching today in preview, can process text and imagery — including photos and video — and output text along the lines of OpenAI’s GPT-4 with Vision model.

Image processing addresses one of the major criticisms of Gemini following its unveiling last Wednesday — namely that the version of Gemini powering Bard, a fine-tuned Gemini Pro model, can’t accept images despite technically being “multimodal” (i.e. trained on a range of data including text, images, videos and audio). Questions linger around Gemini’s image analysis performance and skills, especially in light of a misleading product demo. But now, at least, users will be able to take the model and its image comprehension for a spin themselves.

Within Vertex AI, developers can customize Gemini Pro to specific contexts and use cases leveraging the same fine-tuning tools available for other Vertex-hosted models, like Google’s PaLM 2. Gemini Pro can also be connected to external APIs to perform particular actions or “grounded” to improve the accuracy and relevance of the model’s responses, either with third-party data from an app or database or with data from the web and Google Search.

Citation checking — another existing Vertex AI capability, now with support for Gemini Pro — serves as an additional fact-checking measure by highlighting the sources of information Gemini Pro used to arrive at a response.

“Grounding allows us to take an answer that Gemini’s generated and compare that with a set of data that sits within a company’s own systems … or web sources,” Kurian said. “[T]his comparison allows you to improve the quality of the model’s answers.”

Kurian spent a fair chunk of time spotlighting Gemini Pro’s control, moderation and governance options — seemingly pushing back against coverage implying that Gemini Pro isn’t the strongest model out there. Will the reassurances be enough to convince developers? Maybe. But if they aren’t, Google’s sweetening the pot with discounts.

Input for Gemini Pro on Vertex AI will cost $0.0025 per character while output will cost $0.00005 per character. (Vertex customers pay per 1,000 characters and, in the case of models like Gemini Pro Vision, per image.) That’s reduced 4x and 2x, respectively, from the pricing for Gemini Pro’s predecessor. And for a limited time — until early next year — Gemini Pro is free to try for Vertex AI customers.

“Our goal is to attract developers with attractive pricing,” Kurian said with candor.

Beefing up Vertex

Google’s bringing other new features to Vertex AI in the hopes of dissuading developers from rival platforms like Bedrock.

Several pertain to Gemini Pro. Soon, Vertex customers will be able to tap Gemini Pro to power custom-built conversational voice and chat agents, providing what Google describes as “dynamic interactions … that support advanced reasoning.” Gemini Pro will also become an option for driving search summarization, recommendation and answer generation features in Vertex AI, drawing on documents across modalities (e.g. PDFs, images) from different sources (e.g. OneDrive, Salesforce) to satisfy queries. 

Kurian says that he expects the Gemini Pro-powered conversational and search features to arrive “very early” in 2024.

Elsewhere in Vertex, there’s now Automatic Side by Side (Auto SxS). An answer to AWS’ recently announced Model Evaluation on Bedrock, Auto SxS lets developers evaluate models in an “on-demand,” “automated” fashion; Google claims Auto SxS is both faster and more cost-efficient than manually evaluated models (although the jury’s out on that pending independent testing). 

Google’s also adding models to Vertex from third parties including, Mistral and Meta, and introducing “step-by-step” distillation, a technique that creates smaller, specialized and low-latency models from larger models. In addition, Google’s extending its indemnification policy to include outputs from PaLM 2 and its Imagen models, meaning the company will legally defend eligible customers implicated in lawsuits over IP disputes involving those models’ outputs.

Generative AI models have a tendency to regurgitate training data — an obvious concern for corporate customers. If it’s one day discovered that a vendor like Google used copyrighted data to train a model without first obtaining the proper licensing, that vendor’s customers could end up on the hook for incorporating IP-infringing work into their projects.

Some vendors claim fair use as a defense. But — cognizant of enterprises’ wariness — an increasing number are expanding their indemnification policies around GenAI offerings.

Google’s stopping short of expanding its Vertex AI indemnification policy to cover customers using the Gemini Pro API. The company says, however, that it’ll do so once the Gemini Pro API launches publicly.

More TechCrunch

In early 2018, VC Mike Moritz wrote in the FT that “Silicon Valley would be wise to follow China’s lead,” noting the pace of work at tech companies was “furious”…

This is how bad China’s startup scene looks now

Fei-Fei Li, the Stanford professor many deem the “Godmother of AI,” has raised $230 million for her new startup, World Labs, from backers including Andreessen Horowitz, NEA, and Radical Ventures.…

Fei-Fei Li’s World Labs comes out of stealth with $230M in funding

Bolt says it has settled its long-standing lawsuit with its investor Activant Capital. One-click payments startup Bolt is settling the suit by buying out the investor’s stake “after which Activant…

Fintech Bolt is buying out the investor suing over Ryan Breslow’s $30M loan

The rise of neobanks has been fascinating to witness, as a number of companies in recent years have grown from merely challenging traditional banks to being massive players in and…

Dave and Varo Bank execs are coming to TechCrunch Disrupt 2024

OpenAI released its new o1 models on Thursday, giving ChatGPT users their first chance to try AI models that pause to “think” before they answer. There’s been a lot of…

First impressions of OpenAI o1: An AI designed to overthink it

Featured Article

Investors rebel as TuSimple pivots from self-driving trucks to AI gaming

TuSimple, once a buzzy startup considered a leader in self-driving trucks, is trying to move its assets to China to fund a new AI-generated animation and video game business. The pivot has not only puzzled and enraged several shareholders, but also threatens to pull the company back into a legal…

Investors rebel as TuSimple pivots from self-driving trucks to AI gaming

Welcome to Startups Weekly — your weekly recap of everything you can’t miss from the world of startups. Want it in your inbox every Friday? Sign up here. This week…

Some startups and investors are more risk-averse than others

Silicon Valley startup accelerator Y Combinator will expand the number of cohorts it runs each year from two to four starting in 2025, Bloomberg reported Thursday, and TechCrunch confirmed today.…

Y Combinator expanding to four cohorts a year in 2025

Telegram has had a tough few weeks. The messaging app’s founder, Pavel Durov, was arrested in late August and later released on a €5 million bail in France, charged with…

Telegram CEO Durov’s arrest hasn’t dampened enthusiasm for its TON blockchain

Martin Casado, a general partner at Andreessen Horowitz, will tackle one of the most pressing issues facing today’s tech world — AI regulation — only at TechCrunch Disrupt 2024, taking…

A fireside chat with Andreessen Horowitz partner Martin Casado at TechCrunch Disrupt 2024

Christina Cacioppo, CEO and co-founder of Vanta, will be on the SaaS Stage at TechCrunch Disrupt 2024 to reveal how Vanta is redefining security and compliance automation and driving innovation…

Vanta’s Christina Cacioppo takes the stage at TechCrunch Disrupt 2024

On Thursday, cybersecurity giant Fortinet disclosed a breach involving customer data.  In a statement posted online, Fortinet said an individual intruder accessed “a limited number of files” stored on a…

Fortinet confirms customer data breach

Meta has confirmed that it’s restarting efforts to train its AI systems using public Facebook and Instagram posts from its U.K. userbase. The company claims it has “incorporated regulatory feedback” into a…

Meta reignites plans to train AI using UK users’ public Facebook and Instagram posts

Following the moves of other tech giants, Spotify announced on Friday it’s introducing in-app parental controls in the form of “managed accounts” for listeners under the age of 13. The…

Spotify begins piloting parent-managed accounts for kids on family plans

Uber users in Austin and Atlanta will be able to hail Waymo robotaxis through the app in early 2025 as part of a partnership between the two companies. 

Waymo robotaxis to become available on Uber in Austin, Atlanta in early 2025

There are plenty of calendar and scheduling apps that take care of your professional life and help you slot in meetings with your teammates and work collaborators. Howbout is all…

Howbout raises $8M from Goodwater to build a calendar that you can share with your friends

Delhivery claims Ecom Express has inaccurately represented Delhivery’s business metrics when drawing comparisons in its IPO filing. 

SoftBank-backed Delhivery contests metrics in rival Ecom Express’ IPO filing

It was a matter of time, but Apple is going to allow third-party app stores on the iPad starting next week, on September 16. This change will occur with the…

Alternative app stores will be allowed on Apple iPad in the EU from September 16

The U.K.’s antitrust regulator has delivered its provisional ruling in a longstanding battle to combine two of the country’s major telecommunication operators. The Competition and Markets Authority (CMA) says that…

Three and Vodafone’s $19B merger hits the skids as UK rules the deal would adversely impact customers and MVNOs

Late Thursday evening, Oprah Winfrey aired a special on AI, appropriately titled “AI and the Future of Us.” Guests included OpenAI CEO Sam Altman, tech influencer Marques Brownlee, and current…

Oprah just had an AI special with Sam Altman and Bill Gates — here are the highlights

Antonio Moraes, the grandson of a late prominent Brazilian billionaire, was never interested in joining the family-owned conglomerate of construction companies and a bank. Shortly after graduating from college, he…

XP Health grabs $33M to bring employees more affordable vision care

A crew of four private astronauts made history in the early hours of Thursday when they opened the hatch of their SpaceX Dragon capsule and conducted the first commercial spacewalk. …

Polaris Dawn astronauts perform historic private spacewalk while wearing SpaceX-made suits

Keith Rabois, managing director of Khosla Ventures, was having dinner with a “very successful CEO” in October 2018 when the CEO asked him a question: How many people does it…

Keith Rabois says Miami is still a great place for startups, even as a16z leaves

By making the AI info label harder to find, it might be easier for users to be deceived by content that was edited with AI, especially as editing tools become…

Meta is making its AI info label less visible on content edited or modified by AI tools

Cohost, a would-be X rival launched to the public in June 2022, is shutting down, the company announced via the social network’s staff account earlier this week. The service had…

Cohost, the X rival founded with an anti-Big Tech manifesto, is running out of money and will shut down

At the MTV Video Music Awards (VMAs) on Wednesday night, new technology allowed fans to shop their favorite artists’ styles as they appeared on the screen. Though the drama from…

Shopsense AI lets music fans buy dupes inspired by red-carpet looks at the VMAs

Featured Article

A comprehensive list of 2024 tech layoffs

A complete list of all the known layoffs in tech, from Big Tech to startups, broken down by month throughout 2024.

A comprehensive list of 2024 tech layoffs

Working away on his PhD in Munich only a few years ago, Stephan Herrmann (now a doctor) couldn’t have conceived of a time when his idea for a carbon-negative power…

This startup is making manure out of other biogas power plants and now has $62M to play with

ChatGPT, OpenAI’s text-generating AI chatbot, has taken the world by storm since its launch in November 2022. What started as a tool to hyper-charge productivity through writing essays and code…

ChatGPT: Everything you need to know about the AI-powered chatbot

Faraday Future is doling out big raises and bonuses to its CEO and its founder, despite having delivered just 13 cars in its 10-year history and recently laying off or…

Faraday Future gives CEO and founder raises and bonuses after delivering 13 cars