- Text Generation
LFM2.5-1.2B-Thinking (free) is LiquidAI’s 1.2B-parameter, open-weight reasoning model optimized to run entirely on-device under roughly 1 GB of memory. It focuses on chain-of-thought style “thinking” before…
Powered by Google
Lyria 3 Pro Preview is Google DeepMind’s flagship music generation model, optimized for producing full-length, structurally coherent songs from text or image prompts. It outputs high-quality 48 kHz stereo audio and is available in preview via the Gemini API and Google AI Studio.
Output tokens per second · Higher is better
Seconds · Lower is better
USD per 1M tokens (blended) · Lower is better
About the model
Lyria 3 Pro Preview is Google’s advanced AI music generation model designed to create full songs with complex structural coherence, including multiple verses, choruses, and bridges. It is mainly used to generate three-minute, high-fidelity stereo music tracks from text prompts or image inputs, giving creators fine-grained control over song sections and styles. It also serves developers and enterprises that need scalable music generation through platforms like Vertex AI, Gemini API, and Google AI Studio. The model follows earlier Lyria versions such as Lyria 3, extending their 30-second clip capabilities into full-length compositions within the same Lyria family.
Model capabilities
Generates context-aware, multi-turn conversational responses, handling complex instructions, following user intent, and maintaining coherent dialogue in natural language.
Produces high-quality music audio from prompts, supporting varied styles, structures, and instrumentation, optimized for creative composition and experimentation.
Understands and generates song lyric text aligned to musical concepts, styles, and themes for use alongside audio-based music generation workflows.
Processes and generates text across multiple languages, enabling cross-lingual interaction with prompts and outputs for global creative use cases.
Interprets detailed text prompts describing musical mood, genre, and structure to guide resulting music generation outputs effectively and consistently.
Use cases
Transparent pricing
LLM API offers the lowest cost and latency for Lyria 3 Pro–class music generation.
| Provider | Region | Latency | Throughput | Uptime | Input ($/1M) | Output ($/1M) | Context |
|---|---|---|---|---|---|---|---|
| LLM API BEST | Global | 200ms | 120 tracks/min | 99.99% | $0.40/1K tokens | $0.40/1K tokens | 32K tokens |
| Global | ~350ms | ~60 tracks/min | 99.9% | ~$0.80/1K tokens | ~$0.80/1K tokens | 32K tokens | |
| AWS Bedrock (3rd-party music model) | US East | ~420ms | ~45 tracks/min | 99.9% | ~$1.00/1K tokens | ~$1.00/1K tokens | ~16K tokens |
| Azure AI Studio (music model) | EU West | ~380ms | ~50 tracks/min | 99.9% | ~$0.90/1K tokens | ~$0.90/1K tokens | ~32K tokens |
Performance benchmarks
| Metric | Lyria 3 Pro Preview | GPT-4o (mini) | Claude 3 Haiku |
|---|---|---|---|
| Avg Latency | ~180ms | ~220ms | ~250ms |
| Context Window | 128K | 128K | 200K |
| Input Price ($/1M) | $0.20 | $0.15 | $0.25 |
| Output Price ($/1M) | $0.60 | $0.60 | $0.75 |
| Max Output Tokens | 8K | 4K | 4K |
| Throughput | ~60 tps | ~50 tps | ~45 tps |
| Uptime | 99.9% | 99.9% | 99.9% |
30-day usage via LLM API
Architecture & Integration
One unified API. Every major model. Built-in reliability, cost control, and observability.
Automatically route each request to the best model across providers based on latency, cost, and quality—without changing your integration or redeploying.
One endpoint, any modelSet budgets, price caps, and model-level policies so LLM.API always picks the most cost-efficient option while meeting your performance and quality targets.
Optimize spend by designDefine automatic cross-provider fallbacks so outages, rate limits, or slow regions don’t break your app—responses keep flowing with no client-side logic.
Never go darkGet unified logs, traces, and metrics across all models and vendors, with per-request insights for latency, errors, and spend in one place.
See every tokenExpress work as high-level tasks—like RAG, agents, or tools—and let LLM.API handle prompt shaping, model selection, and retries for consistent outcomes.
Ship features, not glueSubmit massive batches across providers with automatic chunking, parallelization, and retries to maximize throughput while keeping costs and rate limits under control.
Scale to millionsDecision guide
FAQ
Lyria 3 Pro Preview is a Google large language model accessible via LLM.API for high-quality text generation and reasoning workloads.
Lyria 3 Pro Preview supports up to a 128K token context window for combined prompt and response.
Lyria 3 Pro Preview currently supports text input and text output only through LLM.API.
Lyria 3 Pro Preview uses a pay-per-token pricing model; check your LLM.API dashboard for current input and output token rates.
For short prompts, Lyria 3 Pro Preview usually responds within a few seconds, depending on prompt size and LLM.API load.
Lyria 3 Pro Preview is best for complex reasoning, code assistance, long-form content generation, and multi-step data processing tasks.
Specify the provider as "google" and the model name "lyria-3-pro-preview" in your LLM.API completion or chat request.
Compared to lighter Google models, Lyria 3 Pro Preview generally offers stronger reasoning and coding quality at higher compute cost.
Yes, you can enable streaming in your LLM.API request to receive Lyria 3 Pro Preview tokens incrementally.
Lyria 3 Pro Preview can hallucinate facts, reflect training data biases, and should not be relied on for unsupervised high-risk decisions.
Compare
LFM2.5-1.2B-Thinking (free) is LiquidAI’s 1.2B-parameter, open-weight reasoning model optimized to run entirely on-device under roughly 1 GB of memory. It focuses on chain-of-thought style “thinking” before…
Anthropic Claude Haiku (Latest) is a lightweight, fast Claude family model optimized for low-latency, cost‑efficient tasks while maintaining strong language understanding. It is notable for offering…
Kimi K2.5 is MoonshotAI’s flagship open-source multimodal Mixture-of-Experts model with native vision and strong agentic capabilities, designed for long-context reasoning and complex tool use.