跳到正文
Vercel Blog· Zachary Chen·· 8 天前精選AI 評分66

Gemini 3.8 TTS模型现已在AI Gateway上线

Gemini 3.8 text-to-speech models now available on AI Gateway

AI 導讀

Gemini 3.8 Flash‑Lite TTS 与 Gemini 3.8 Flash TTS 现已在 AI Gateway 上线,支持 100 多种语言的文本转语音,支持长篇朗读、交付控制和双人对话。

推薦理由

Gemini 3.8 TTS模型已在AI Gateway上线,支持100多种语言、长篇朗读和双人对话,开发者可通过AI SDK轻松生成语音。

正文 · 原文

Gemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS from Google are now available on AI Gateway.

Both models generate speech from text in more than 100 languages and support long-form narration, control over delivery, and two-speaker dialogue.

  • google/gemini-3.8-flash-lite-tts is optimized for high-volume speech generation, with controls for tone, pacing, and line-by-line delivery.

  • google/gemini-3.8-flash-tts is optimized for expressive voice generation and custom character design, with controls for acting cues, pacing, accents, and conversational reactions.

To generate speech, pass either model ID to the AI SDK (7 or later):

import { generateSpeech } from 'ai';

import { gateway } from '@ai-sdk/gateway';

import { writeFile } from 'node:fs/promises';

const result = await generateSpeech({

model: gateway.speechModel('google/gemini-3.8-flash-lite-tts'),

text: 'Welcome to the audio edition.',

voice: 'Kore',

instructions: 'Warm, relaxed, and speaking slowly',

outputFormat: 'wav',

});

await writeFile('speech.wav', result.audio.uint8Array);

Generate a WAV file with Gemini 3.8 Flash-Lite TTS.

Try Flash-Lite TTS or Flash TTS in the model playground. See the text-to-speech guide for setup instructions, two-speaker dialogue, and more examples.

來源:Vercel Blog · vercel.com