Google: Lyria 3 Clip Preview
Google: Lyria 3 Clip Preview is a text-and-image input model with a 1-million-token context window and a 65,536-token output ceiling. It does not support tool use, reasoning modes, or structured output, so its utility is limited to straightforward generative tasks that fit within those input modalities. The model is currently free, which makes it worth shortlisting for developers and researchers who want to experiment with large-context multimodal prompting at no cost. The significant caveat is that there is no independent benchmark coverage yet, so actual quality relative to paid alternatives is unverified. Users who need proven, measurable performance for production work should treat this as an exploratory option until third-party evaluations become available.
- Model ID
- google/lyria-3-clip-preview
- Vendor
- Released
- March 2026
- Tokenizer
- Other
- Input Modalities
- text, image
- Output Modalities
- text, audio
- Max Output
- 65,536 tokens
- Tool Calling
- not supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
Similar models
Google: Lyria 3 Pro Preview
Google: Gemma 3 27B
Google: Gemma 3 12B
Google: Nano Banana 2 (Gemini 3.1 Flash Image)
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Google: Gemma 3 4B
Quick answers
- Is Google: Lyria 3 Clip Preview free to use?
- Yes. It is currently listed at no cost per token via OpenRouter, subject to provider rate limits.
- What is Google: Lyria 3 Clip Preview's context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Lyria 3 Clip Preview support tool calling?
- It supports structured output.
- Can Google: Lyria 3 Clip Preview process images?
- Yes, it accepts image input.