voice technology

AppWizard
August 27, 2026
Gemini 3.5 Transcribe is a new speech-to-text model designed for intelligent voice interactions, overcoming challenges faced by traditional speech recognition systems, such as background noise and complex jargon. It transforms raw audio into accurate, formatted text and is available to users of the Gemini app and Android devices. Developers can access its capabilities through the Gemini API in Google AI Studio and the Gemini Enterprise Agent Platform. The model supports two APIs: real-time streaming for interactive applications with sub-second latency and pre-recorded audio processing that includes speaker attribution and word-level timestamps. It enhances transcription accuracy by managing self-corrections, eliminating filler words, and auto-formatting text. The model has a Word Error Rate of 4.0% for streaming and 2.6% for non-streaming scenarios, effectively capturing alphanumeric entities. It adapts to custom vocabulary, supports over 85 languages, and can identify multiple speakers in pre-recorded audio.
AppWizard
June 5, 2026
Google is preparing to introduce the Rambler feature for Gboard, its new AI-powered voice typing capability, as part of the enhancements for Android 17. Rambler can understand natural speech, remove filler words, and detect self-corrections during dictation. A hidden toggle for Rambler has been found in the latest Gboard beta, indicating that preparations for its rollout are in progress. The feature may initially be exclusive to select flagship devices, such as the Samsung Galaxy S26 Ultra or the Pixel 10 series.
Search