- Text Generation
MiniMax M2.1 is a second-generation, open-weight Mixture-of-Experts large language model from MiniMax, optimized for real-world coding, tool use, and long-horizon agentic workflows. It is notable for…
Powered by Sourceful
Riverflow V2 Max Preview is Sourceful’s most powerful Riverflow V2 preview model, a unified text-to-image and image-to-image generator. It is designed to exceed the performance of the earlier Riverflow 1 family while supporting rich, high-context visual generation workflows.
Output tokens per second · Higher is better
Seconds · Lower is better
USD per 1M tokens (blended) · Lower is better
About the model
Riverflow V2 Max Preview is a high-performance image generation model from Sourceful that unifies text-to-image and image-to-image capabilities. It is mainly used for creating high-quality images from text prompts and for transforming or editing images with additional textual guidance. It also targets professional creative workflows such as brand asset production and other controlled visual design tasks. Riverflow V2 Max Preview belongs to the Riverflow V2 family and is positioned as the flagship successor to the Riverflow 1 model family.
Model capabilities
Generates high-quality images from detailed natural language prompts using Sourceful’s unified Riverflow V2 Max Preview image model.
Transforms or enhances existing input images based on guidance prompts, leveraging Riverflow V2’s unified image-to-image capabilities.
Supports chat-style prompting via OpenRouter’s chat completions API for iterative refinement of image generation instructions.
Handles image requests up to Sourceful’s 4.5MB limit, with best practices recommending image URLs instead of Base64 payloads.
Accepts prompts in multiple languages through OpenRouter, enabling image generation workflows for international users and content.
Use cases
Transparent pricing
Save up to ~70% vs other Riverflow V2 Max-compatible APIs
| Provider | Region | Latency | Throughput | Uptime | Input ($/1M) | Output ($/1M) | Context |
|---|---|---|---|---|---|---|---|
| LLM API BEST | Global | 120ms | 120 tps | 99.99% | $0.20 | $0.40 | 256K |
| Sourceful | Global | ~260ms | ~40 tps | ~99.9% | ~$0.60 | ~$1.20 | ~128K |
| OpenRouter | Global | ~280ms | ~35 tps | ~99.8% | ~$0.70 | ~$1.40 | ~128K |
| Together AI | US East | ~240ms | ~50 tps | ~99.9% | ~$0.55 | ~$1.10 | ~128K |
Performance benchmarks
| Metric | Riverflow V2 Max Preview | OpenAI GPT-4.1 Mini | Anthropic Claude 3.5 Haiku |
|---|---|---|---|
| Avg Latency | ~180ms | ~220ms | ~250ms |
| Context Window | 128K | 128K | 200K |
| Input Price ($/1M) | $0.20 | $0.15 | $0.25 |
| Output Price ($/1M) | $0.60 | $0.60 | $0.80 |
| Max Output Tokens | 4K | 4K | 4K |
| Throughput | ~60 tps | ~50 tps | ~45 tps |
| Uptime | 99.9% | 99.9% | 99.9% |
30-day usage via LLM API
Architecture & Integration
One unified API. Every major model. Built-in reliability, cost control, and observability.
Automatically route each request to the best-fit model across providers using rules, metadata, and performance signals — no client changes, just smarter traffic.
One endpoint, any modelControl spend with per-route pricing rules, hard caps, and automated downshifts to cheaper models while keeping SLA and quality constraints intact.
Predictable, tunable spendDesign multi-step failover chains across providers with timeouts and retries so user-facing features stay online even when a model or region fails.
Never ship single points of failureTrace every request across providers with logs, metrics, and structured events to debug latency spikes, errors, and quality regressions in one place.
One pane for every callDefine high-level tasks—chat, extract, classify, generate—and let LLM.API pick tools, models, and prompts so you ship features, not glue code.
Describe tasks, not wiringRun millions of inferences in parallel with automatic chunking, retries, and cost tracking—no custom workers, no rate-limit headaches.
Batch at platform scaleDecision guide
FAQ
Riverflow V2 Max Preview is a large language model from Sourceful focused on high-quality text generation and reasoning, accessible through the LLM.API platform.
Riverflow V2 Max Preview is best for building chatbots, agents, data analysis assistants, and complex reasoning workflows where accuracy and controllability matter.
Riverflow V2 Max Preview supports a large-context window suitable for multi-turn conversations and long documents, but exact token limits depend on the LLM.API configuration.
Riverflow V2 Max Preview currently supports text input and text output; image, audio, and video inputs are not available via LLM.API for this model.
Typical responses arrive within a few seconds, with latency varying based on prompt size, output length, and current LLM.API load.
Pricing for Riverflow V2 Max Preview is usage-based per input and output token, with exact rates shown in the LLM.API pricing dashboard.
Select the Sourceful provider and the Riverflow V2 Max Preview model name in your LLM.API request, then send standard chat or completion-style payloads.
Riverflow V2 Max Preview targets a balance of quality, cost, and speed, often suited as a mid-to-high tier alternative to flagship frontier models.
Riverflow V2 Max Preview can produce hallucinations, lacks real-time knowledge, and should not be relied on for critical decisions without external verification.
Riverflow V2 Max Preview can hallucinate facts, lacks real-time internet access, and may struggle with highly specialized domain knowledge without grounding tools.
Compare
MiniMax M2.1 is a second-generation, open-weight Mixture-of-Experts large language model from MiniMax, optimized for real-world coding, tool use, and long-horizon agentic workflows. It is notable for…
Qwen3.5-9B is a 9‑billion‑parameter multimodal language model from Qwen that supports long-context reasoning over text and images. It is designed to offer strong reasoning, coding, and…
Cogito v2.1 671B is Deep Cogito’s flagship 671B-parameter open-weight Mixture-of-Experts language model optimized for efficient hybrid reasoning. It delivers frontier-level performance while using significantly shorter reasoning…