Live data from Hacker News

Gemini 3.1 Flash-Lite: Built for intelligence at scale

blog.google

11–20 of 34 posts

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#12
You can test Gemini 3.1 Lite transcription capabilities in https://ottex.ai — the only dictation app supporting Gemini models with native audio input.

We benchmarked it for real-life voice-to-text use cases:

                
Key takeaways:

- 1.8x faster than Gemini 3 Flash on average

- ~1.4 sec transcription time for short to medium recordings

- ~$0.50/mo for heavy users (10h+ transcription)

- Close to SOTA audio understanding and formatting instruction following

- Multilingual: one model, 100+ languages

Gemini is slowly making $15/month voice apps obsolete.

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#13

For the last 2 years, startup wisdom has been that models will continue to get cheaper and better. Claude first, and now Gemini has shown that it's not the case. We priced an enterprise contract using Flash 1.5 pricing last summer, and today that contract would be unit economic negative if we used Flash 3. Flash 2.5 and now Flash 3.1 Lite barely breaks even. I predict open-source models and fine-tuning are going to m…

Not true. You just measure cost by amount of money spent per task. I would argue that this lite version is equivalent to older flash.

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#14
post #12

You can test Gemini 3.1 Lite transcription capabilities in https://ottex.ai — the only dictation app supporting Gemini models with native audio input. We benchmarked it for real-life voice-to-text use cases: Key takeaways: - 1.8x faster than Gemini 3 Flash on average - ~1.4 sec transcription time for short to medium recordings - ~$0.50/mo for heavy users (10h+ transcription) - Close to SOTA audio understanding and fo…

You know what would be great? A light weight wrapper model for voice that can use heavier ones in the background.

That much is easy but what if you could also speak to and interrupt the main voice model and keep giving it instructions? Like speaking to customer support but instead of putting you on hold you can ask them several questions and get some live updates

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#16

For the last 2 years, startup wisdom has been that models will continue to get cheaper and better. Claude first, and now Gemini has shown that it's not the case. We priced an enterprise contract using Flash 1.5 pricing last summer, and today that contract would be unit economic negative if we used Flash 3. Flash 2.5 and now Flash 3.1 Lite barely breaks even. I predict open-source models and fine-tuning are going to m…

Not true. You just measure cost by amount of money spent per task. I would argue that this lite version is equivalent to older flash.

Yea but there is a whole world of tasks for which Flash 2.5-lite was sufficiently intelligent. Given Google's depreciation policy, there will soon be no way to get that intelligence at that price.

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#17
post #10

For the last 2 years, startup wisdom has been that models will continue to get cheaper and better. Claude first, and now Gemini has shown that it's not the case. We priced an enterprise contract using Flash 1.5 pricing last summer, and today that contract would be unit economic negative if we used Flash 3. Flash 2.5 and now Flash 3.1 Lite barely breaks even. I predict open-source models and fine-tuning are going to m…

I mean the same level of intelligence does get cheaper. People just care about being on the frontier. But if you track a single level of intelligence the price just drops and drops.

What's the cheaper alternative from Gemini for Flash-2.5-lite level intelligence when it gets deprecated on 22nd July 2026?

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#19
post #12

You can test Gemini 3.1 Lite transcription capabilities in https://ottex.ai — the only dictation app supporting Gemini models with native audio input. We benchmarked it for real-life voice-to-text use cases: Key takeaways: - 1.8x faster than Gemini 3 Flash on average - ~1.4 sec transcription time for short to medium recordings - ~$0.50/mo for heavy users (10h+ transcription) - Close to SOTA audio understanding and fo…

Can you show some comparisons for WER and other ASR models? Especially for non english.

Re: Gemini 3.1 Flash-Lite: Built for intelligence at scale

#20

For the last 2 years, startup wisdom has been that models will continue to get cheaper and better. Claude first, and now Gemini has shown that it's not the case. We priced an enterprise contract using Flash 1.5 pricing last summer, and today that contract would be unit economic negative if we used Flash 3. Flash 2.5 and now Flash 3.1 Lite barely breaks even. I predict open-source models and fine-tuning are going to m…

Opus 4.5 became significantly cheaper than Opus 4.1
Post reply on HN