← Back to feed

Gemini-3.5-Transcribe

L4 · DeveloperModels & ReleasesHacker News· 8/27/2026

Provides developers with enterprise-grade transcription APIs for building voice agents and real-time applications.

AI Summary

Google launches Gemini 3.5 Transcribe, a speech-to-text model with 4.0% WER for and 2.6% for non-streaming use cases, now available via API.

Read Original
0 upvotes · 0 downvotes · 1 min read

Related Articles

L4 · DeveloperModels & Releases@GoogleDeepMind
Google DeepMind (@GoogleDeepMind): We’re rolling out Gemini Omni 1.1 Flash to make generative video highly controllable, faster to iterate on, and more polished for production-grade use. Here’s how you can try it in @FlowbyGoogle and…

Google DeepMind is releasing Gemini Omni 1.1 Flash, a generative video model focused on controllability, faster iteration, and production readiness.

L3 · BuilderModels & ReleasesHacker News
Gemini Omni 1.1 Flash

Google released Gemini Omni 1.1 Flash, a production-ready video generation model with new controls for scene extension, frame interpolation, 4K upscaling, and faster prototyping.

L5 · ResearcherModels & ReleasesHugging Face Blog
Granite 4.2 LLMs: How They're Built

IBM releases Granite 4.2 reasoning LLMs in 3B, 8B, and 30B sizes with 512K context, agentic RL training, and Apache 2.0 license.

L4 · DeveloperModels & Releases@GoogleAI
Google AI (@GoogleAI): https://t.co/LJGIbM3Ksy

Google's WeatherNext Cyclones model, used in real-time by the U.S. National Hurricane Center, provided 5-day advance warning for a Category 5 hurricane and offers accuracy improvements equivalent to a decade of meteorological progress.