Google Introduces Gemini 3.1 Flash-Lite as Its Fastest and Most Cost-Efficient AI Model

The Gemini 3.1 Flash-Lite is rolling out in preview via the Gemini API in AI Studio and Vertex AI.

Advertisement
Written by Akash Dutta, Edited by Rohan Pal | Updated: 5 March 2026 15:32 IST
Highlights
  • The AI model costs $0.25 per million input tokens
  • Gemini 3.1 Flash-Lite costs $1.50 per million output tokens
  • Google claims the new model outperforms 2.5 Flash in response speed

Gemini 3.1 Flash-Lite achieved an Elo score of 1432 on the Arena.ai Leaderboard

Photo Credit: Google

Google introduced the Gemini 3.1 Flash-Lite artificial intelligence (AI) model on Thursday. Calling it the fastest and the most cost-efficient AI model in the Gemini 3 series, the Mountain View-based tech giant said it is designed for high-volume developer workloads. The model is currently not available to end users and has been reserved for developers and enterprises via specific channels. The company also claimed that the model's output speed is higher than that of the 2.5 series. Notably, the Gemini 3.1 Flash-Lite is currently only available in preview.

Gemini 3.1 Flash-Lite Is Here

In a blog post, the tech giant announced and detailed its latest Gemini 3.1 series large language model (LLM). Currently, the Gemini 3.1 Flash-Lite can be accessed in preview via the Gemini application programming interface (API) in Google AI Studio, and via Vertex AI for enterprises.

Advertisement

Coming to capabilities, the company said the 3.1 Flash-Lite outperforms 2.5 Flash with a “2.5X faster Time to First Answer Token,” and a 45 percent increase in output speed, citing the Artificial Analysis benchmark. It is also said to have achieved an Elo score of 1432 on the Arena.ai leaderboard. It is also claimed to outperform GPT-5 mini, Claude 4.5 Haiku, and Grok 4.1 Fast in terms of output speed.

In AI Studio and Vertex AI, developers will be able to access the LLM in standard and thinking modes, with the latter allowing users to control the thinking time for a task. Highlighting some use cases, Google said the model can handle high-volume translation and content moderation, and can also be used for complex tasks, such as generating user interfaces and dashboards, creating simulations, or just following instructions.

Advertisement

The company also claimed that the Gemini 3.1 Flash-Lite is a cost-efficient AI model, with one million input tokens priced at $0.25 (roughly Rs. 23) and output tokens priced at $1.5 (roughly Rs. 137) per million tokens. In comparison, the Gemini 2.5 Flash costs $0.3 (roughly Rs. 27.5) per million input and $2.5 (roughly Rs. 229) per million output tokens.

 

Get your daily dose of tech news, reviews, and insights, in under 80 characters on Gadgets 360 Turbo. Connect with fellow tech lovers on our Forum. Follow us on X, Facebook, WhatsApp, Threads and Google News for instant updates. Catch all the action on our YouTube channel.

Advertisement

Related Stories

Popular Mobile Brands
  1. Poco M8 Power India Launch Appears Imminent as Teaser Goes Live
  2. Xiaomi 18 Series Leak Reveals Potential India Launch Window
  3. Samsung Galaxy Z Fold 8 Leads Samsung's Foldable Production Plan: Report
  4. Apple Hikes Apple Music, Apple One Subscription Prices in India
  5. Samsung Galaxy Z Fold 8 Series Leak Reveals Key Specs Ahead of Launch
  6. OnePlus 16 and Oppo Find X10 Ultra May Not Launch in India
  7. Best Smartphones With AI Camera Features
  8. Apple's M2 Extreme, M3 Extreme Chips Never Saw the Light of Day
  9. Google Pixel 11a Could Get a Connectivity Upgrade With New MediaTek Modem
  10. Motorola Edge 70 Max Now on Sale in India With Up to Rs. 5,000 Discount
  1. Poco M8 Power India Launch Teased; Flipkart Availability Confirmed: Here's What You Need to Know
  2. Xiaomi 18 Series Tipped to Launch Early in India; Ultra Model Said to Be Replaced By Pro Max Variant
  3. Motorola Edge 70 Max Goes on Sale in India With Snapdragon 8 Gen 5 SoC: Price, Offers
  4. Samsung Galaxy Z Fold 8, Z Fold 8 Ultra, and Z Flip 8 Marketing Images Leaked Ahead of July 22 Launch
  5. Apple Cancelled High-End M2 Extreme and M3 Extreme Chips Over Cost Concerns: Report
  6. Honor X7e Plus 5G Launched With Snapdragon 4 Gen 4 SoC, 8,100mAh Battery: Price, Specifications
  7. Samsung Plans to Manufacture 2.8 Million Galaxy Z Fold 8 Units This Year: Report
  8. OnePlus 16 and Oppo Find X10 Ultra May Skip India Launch, Leak Suggests
  9. Apple Music, Apple One Subscription Prices Increased in India; Individual Plan Now Starts at Rs. 139
  10. Google Pixel 11a Key Specifications, Colour Options Leak; May Get Tensor G6 Chip, MediaTek Modem
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2026. All rights reserved.