- Text Generation
Orpheus 3B is a 3-billion-parameter English text-to-speech model from Canopy Labs, optimized for natural prosody, expressive delivery, and real-time streaming speech generation. It is notable for…
Powered by Sourceful
Riverflow V2 Fast Preview is Sourceful’s fastest preview variant in the Riverflow V2 lineup, offering high-throughput text-to-image and image-to-image generation with an 8K token context window.
Output tokens per second · Higher is better
Seconds · Lower is better
USD per 1M tokens (blended) · Lower is better
About the model
Riverflow V2 Fast Preview is a preview-stage Sourceful model that unifies text-to-image and image-to-image capabilities for fast, production-oriented visual generation. It is mainly used for quickly generating design, branding, and illustration assets from text prompts, and for rapidly transforming or iterating on existing images in latency-sensitive workflows. It also supports 8K context for complex prompt specifications and multimodal inputs where both textual and visual instructions guide the output. As part of the Riverflow V2 preview family, it succeeds and outperforms the earlier Riverflow 1 models while sitting alongside the Riverflow V2 Standard Preview and Riverflow V2 Max Preview variants.
Model capabilities
Generates high-quality images directly from natural language prompts using Sourceful’s unified text-to-image generation pipeline.
Transforms or refines existing images based on text instructions, leveraging unified image-to-image generation capabilities.
Accepts image inputs, including via URLs within size limits, enabling multimodal workflows combining visual and textual information.
Optimized as the fastest Riverflow V2 preview variant, suitable for interactive, latency-sensitive image applications and rapid experimentation.
Supports prompts in multiple languages for controlling image generation, allowing localized visual content creation from diverse text inputs.
Use cases
Transparent pricing
LLM API offers the lowest cost and latency for Riverflow V2 Fast Preview–class models.
| Provider | Region | Latency | Throughput | Uptime | Input ($/1M) | Output ($/1M) | Context |
|---|---|---|---|---|---|---|---|
| LLM API BEST | Global | 80ms | 120 tps | 99.99% | $0.20 | $0.60 | 128K |
| Sourceful | Global | ~220ms | ~45 tps | ~99.9% | ~$0.30 | ~$0.90 | ~64K |
| OpenRouter | Global | ~240ms | ~40 tps | ~99.9% | ~$0.32 | ~$0.95 | ~64K |
| Together AI | US East | ~250ms | ~38 tps | ~99.9% | ~$0.34 | ~$1.00 | ~64K |
Performance benchmarks
| Metric | Riverflow V2 Fast Preview | OpenAI GPT-4.1 Mini | Anthropic Claude 3 Haiku |
|---|---|---|---|
| Avg Latency | ~180ms | ~220ms | ~250ms |
| Context Window | 128K | 128K | 200K |
| Input Price ($/1M) | $0.05 | $0.15 | $0.25 |
| Output Price ($/1M) | $0.10 | $0.60 | $1.25 |
| Max Output Tokens | 4K | 4K | 4K |
| Throughput | 80 tps | 60 tps | 50 tps |
| Uptime | 99.9% | 99.9% | 99.9% |
30-day usage via LLM API
Architecture & Integration
One unified API. Every major model. Built-in reliability, cost control, and observability.
Dynamically route each request to the optimal model or provider based on cost, latency, or quality—no client changes, just smarter traffic from a single endpoint.
One endpoint, every modelAutomatically balance premium and budget models, enforce spend controls, and track per-team usage so you ship faster without surprise invoices or manual tuning.
More performance per dollarDefine fallback chains across providers so outages, rate limits, or bad responses transparently fail over—keeping your AI features up without incident pages.
Never go darkGet per-request traces, metrics, and logs across all models and vendors in one place, making debugging, optimization, and compliance reviews actually manageable.
One pane for all callsDescribe tasks—chat, extract, classify, generate—not vendor APIs. LLM.API normalizes schemas, tools, and prompts so you can swap models without rewrites.
Code to tasks, not vendorsRun large-scale jobs over millions of inputs with automatic chunking, retries, rate-limit handling, and progress tracking, all via a simple, consistent batch API.
Scale from 10 to millionsDecision guide
FAQ
Riverflow V2 Fast Preview is a Sourceful language model accessible via LLM.API, optimized for fast, low-latency text generation in development and prototyping scenarios.
It is best for high-throughput chatbots, lightweight agents, and iterative application development where fast responses and low cost matter more than peak quality.
Riverflow V2 Fast Preview supports a context window of up to 8,000 tokens per request via LLM.API.
The model is tuned for low first-token latency and high tokens-per-second throughput, making it suitable for interactive applications and streaming responses.
Riverflow V2 Fast Preview supports text-in, text-out generation only, without native image, audio, or tool-calling modalities.
Pricing is usage-based per 1,000 tokens, with lower rates than Sourceful’s higher-tier models to favor experimentation and high-volume workloads.
You specify the model name "sourceful/riverflow-v2-fast-preview" in your LLM.API request, using the standard chat or completion endpoint.
It is generally cheaper and faster but may produce slightly lower-quality reasoning, coding, and long-form outputs than larger Sourceful models.
It can struggle with very long multi-step reasoning, highly specialized domain questions, and tasks requiring strict factual accuracy without external tools.
Yes, the model supports token streaming over LLM.API, allowing partial results to be sent as they are generated.
Compare
Orpheus 3B is a 3-billion-parameter English text-to-speech model from Canopy Labs, optimized for natural prosody, expressive delivery, and real-time streaming speech generation. It is notable for…
Qwen3 VL 30B A3B Thinking is a large multimodal Qwen model with around 30 billion parameters, designed for vision-language reasoning with extended “thinking” capabilities. It is…
Riverflow V2 Max Preview is Sourceful’s most powerful Riverflow V2 preview model, a unified text-to-image and image-to-image generator. It is designed to exceed the performance of…