{"id":279384,"date":"2025-04-11T07:04:53","date_gmt":"2025-04-11T07:04:53","guid":{"rendered":"https:\/\/news.talkwithrattan.com\/index.php\/2025\/04\/11\/gemini-2-5-flash-unveiled-as-a-faster-and-cheaper-ai-model-for-developers\/"},"modified":"2025-04-11T07:04:54","modified_gmt":"2025-04-11T07:04:54","slug":"gemini-2-5-flash-unveiled-as-a-faster-and-cheaper-ai-model-for-developers","status":"publish","type":"post","link":"https:\/\/news.talkwithrattan.com\/index.php\/2025\/04\/11\/gemini-2-5-flash-unveiled-as-a-faster-and-cheaper-ai-model-for-developers\/","title":{"rendered":"Gemini 2.5 Flash Unveiled as a Faster and Cheaper AI Model for Developers"},"content":{"rendered":"<div style=\"text-align:center\"><img decoding=\"async\" src=\"https:\/\/i1.wp.com\/i.gadgets360cdn.com\/large\/gemini_25_1744350955269.jpg?ssl=1\" class=\"attachment-post-thumbnail size-post-thumbnail wp-post-image\" alt=\"Gemini 2.5 Flash Unveiled as a Faster and Cheaper AI Model for Developers\" title=\"Gemini 2.5 Flash Unveiled as a Faster and Cheaper AI Model for Developers\" \/><\/div><p> <br \/>\n<\/p>\n<div>\n<p><a href=\"https:\/\/www.gadgets360.com\/tags\/google\">Google<\/a> released its second artificial intelligence (AI) model in the Gemini 2.5 family on Thursday. Dubbed Gemini 2.5 Flash, it is a cost-efficient low-latency model which is designed for tasks requiring real-time inference, conversations at scale, and those which are generalistic in nature. The Mountain View-based tech giant will soon make the AI model available on both the Google AI Studio as well as Vertex AI to help users and developers access the Gemini 2.5 Flash, and build applications and agents using it.<\/p>\n<h2 id=\"gemini-2-5-flash-is-now-available-on-vertex-ai\">Gemini 2.5 Flash Is Now Available on Vertex AI<\/h2>\n<p>In a <a href=\"https:\/\/cloud.google.com\/blog\/products\/ai-machine-learning\/gemini-2-5-pro-flash-on-vertex-ai\" target=\"_blank\" rel=\"nofollow noopener\">blog post<\/a>, the tech giant detailed its latest large language model (LLM). Alongside announcing the debut of the Flash model, the post also confirmed that the Gemini 2.5 Pro model is now available on Vertex AI. Differentiating between the use cases of the two models, Google said the Pro model is ideal for tasks that require intricate knowledge, multi-step analyses, and making nuanced decisions.<\/p>\n<p>On the other hand, the Flash model prioritises speed, low latency, and cost efficiency. Calling it a workhorse model, the tech giant said it is an \u201cideal engine for responsive virtual assistants and real-time summarisation tools where efficiency at scale is key.\u201d<\/p>\n<p>While <a href=\"https:\/\/www.gadgets360.com\/ai\/news\/google-gemini-2-5-pro-ai-model-release-availability-specifications-performance-features-openai-8014917\">launching<\/a> the 2.5 Pro model,\u00a0Google had specified\u00a0that all LLMs in this series would feature natively built reasoning or \u201cthinking\u201d capability. This means the 2.5 Flash also comes with \u201cdynamic and controllable reasoning.\u201d Developers can adjust the processing time for a query based on the complexity, enabling them to get a granular control over the response generation times.<\/p>\n<p>For its enterprise clients, Google is also introducing the Vertex AI Model Optimiser tool. Available as an experimental feature within the platform, it takes away the confusion of choosing a specific model when users are not sure. The feature can automatically generate the highest-quality response for each prompt based on factors such as quality and cost.<\/p>\n<p>Google did not release a technical paper or model information card alongside the release, so information about its architecture, pre- and post-training processes, and benchmark scores are not known. The company might release it at a later time while making the model available to end consumers.<\/p>\n<p>Meanwhile, the tech giant is also adding new tools to support agentic application building on Vertex AI. The company is adding a new Live application programming interface (API) for <a href=\"https:\/\/www.gadgets360.com\/tags\/gemini\">Gemini<\/a> models that will allow AI agents to process streaming audio, video, and text with low latency to let it complete tasks in real-time.<\/p>\n<p>The Live API, which is powered by Gemini 2.5 Pro, also supports resumable sessions longer than 30 minutes, multilingual audio output, time-stamped transcripts for analysis, tool integration, and more.<\/p>\n<\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/www.gadgets360.com\/ai\/news\/google-gemini-2-5-flash-ai-model-cost-efficiency-latency-released-vertex-ai-8138222#rss-gadgets-news\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Google released its second artificial intelligence (AI) model in the Gemini 2.5 family on Thursday. Dubbed Gemini 2.5 Flash, it is a cost-efficient low-latency model which is designed for tasks requiring real-time inference, conversations at scale, and those which are generalistic in nature. The Mountain View-based tech giant will soon make the AI model available [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":279385,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"tdm_status":"","tdm_grid_status":"","fifu_image_url":"https:\/\/i.gadgets360cdn.com\/large\/gemini_25_1744350955269.jpg","fifu_image_alt":"","footnotes":""},"categories":[607],"tags":[1274,17725,890,3476,15571,8623,21784,64,216115,216114,16497,3958,5729],"amp_enabled":true,"_links":{"self":[{"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/posts\/279384"}],"collection":[{"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/comments?post=279384"}],"version-history":[{"count":1,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/posts\/279384\/revisions"}],"predecessor-version":[{"id":279386,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/posts\/279384\/revisions\/279386"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/media\/279385"}],"wp:attachment":[{"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/media?parent=279384"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/categories?post=279384"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/news.talkwithrattan.com\/index.php\/wp-json\/wp\/v2\/tags?post=279384"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}