Lyria 3 Clip Preview
Lyria 3 Clip Preview is an audio model — speech-to-text, text-to-speech, or a unified audio-in/audio-out variant depending on the provider. Available via Google. 1,048,576-token context. Free tier on OpenRouter. Currently in preview — capabilities and pricing may shift.
Lyria 3 Clip Preview is a multimodal AI model from Google. It costs $0.000 per million input tokens and $0.000 per million output tokens (blended $0.000/M), with a 1,048,576-token context window.
- Speech transcription
- Multilingual coverage
- Realtime or batch ingest depending on provider
- Multimodal: handles images alongside text
- Meeting transcription
- Voice interfaces
- Subtitle pipelines
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Lyria 3 Clip Preview cost?
Lyria 3 Clip Preview costs $0.000 per million input tokens and $0.000 per million output tokens, for a blended reference rate of $0.000 per million tokens.
What is Lyria 3 Clip Preview's context window?
Lyria 3 Clip Preview supports up to 1,048,576 tokens of context in a single request.
What is Lyria 3 Clip Preview best for?
Lyria 3 Clip Preview is well suited to Speech transcription, Multilingual coverage and Realtime or batch ingest depending on provider.
Who makes Lyria 3 Clip Preview?
Lyria 3 Clip Preview is developed and served by Google.