- Text Generation
MiMo-V2-Flash is an open-source Mixture-of-Experts language model from Xiaomi optimized for fast, long-context reasoning and coding. It combines a 309B-parameter MoE architecture with only 15B active…
Powered by Black Forest Labs
FLUX.2 Flex is Black Forest Labs’ developer-tunable FLUX.2 image generation and editing model, offering flexible control over resolution, speed–quality tradeoffs, and highly accurate text rendering.
Output tokens per second · Higher is better
Seconds · Lower is better
USD per 1M tokens (blended) · Lower is better
About the model
FLUX.2 Flex is a flexible, parameter-tunable variant of Black Forest Labs’ FLUX.2 image generation family that exposes direct control over key generation settings such as inference steps, guidance scale, and output resolution. It is mainly used for text-to-image and image-to-image generation where creators need fine-grained control over typography, layout, and visual detail, such as UI mockups, infographics, and branded assets. It is also suited to multi-reference image editing workflows that require consistent identity and style across iterations while balancing cost, speed, and fidelity. FLUX.2 Flex belongs to the FLUX.2 model family alongside variants such as FLUX.2 Max and FLUX.2 Klein, sharing the same core image-generation architecture with different optimization targets.
Model capabilities
Generates high-quality images from natural language prompts, with tunable quality-speed tradeoffs via adjustable inference steps and guidance.
Edits and enhances existing images up to multi-megapixel resolutions based on text instructions while preserving global coherence and composition.
Supports multiple reference images to maintain character identity, product appearance, and stylistic consistency across generated or edited outputs.
Produces clean, readable typography and complex text layouts, suitable for posters, UI mockups, infographics, and other text-heavy visuals.
Accepts prompts in different languages and reliably interprets them for image generation, enabling multilingual creative workflows.
Use cases
Transparent pricing
LLM API offers the lowest cost per image and fastest global latency for FLUX.2-class models.
| Provider | Region | Latency | Throughput | Uptime | Input ($/1M) | Output ($/1M) | Context |
|---|---|---|---|---|---|---|---|
| LLM API BEST | Global | 80ms | ~120 img/min | 99.99% | $0.005/img | $0.00/img | 1 img |
| Black Forest Labs (Official API) | EU West | ~180ms | ~60 img/min | 99.9% | ~$0.020/img | $0.00/img | 1 img |
| Hugging Face Inference Endpoints | Global | ~220ms | ~40 img/min | 99.5% | ~$0.030/img | $0.00/img | 1 img |
| Replicate | US East | ~250ms | ~35 img/min | 99.5% | ~$0.035/img | $0.00/img | 1 img |
Performance benchmarks
| Metric | FLUX.2 Flex (Black Forest Labs) | Stable Diffusion 3 Medium (Stability AI) | DALL·E 3 (OpenAI) |
|---|---|---|---|
| Latency per Image | ~900ms | ~1100ms | ~1200ms |
| Throughput | ~45 img/s | ~35 img/s | ~30 img/s |
| Max Resolution | 1536×1536 | 2048×2048 | 2048×2048 |
| Price per Image | ~$0.020 | ~$0.018 | ~$0.040 |
| Supported Formats | PNG, JPEG, WEBP | PNG, JPEG, WEBP | PNG, JPEG |
| Uptime | 99.9% | 99.5% | 99.9% |
30-day usage via LLM API
Architecture & Integration
One unified API. Every major model. Built-in reliability, cost control, and observability.
Dynamically route requests across providers and models based on latency, cost, or quality. Swap vendors without code changes or client-side orchestration.
One endpoint, every modelSet per-project budgets, caps, and policies while automatically choosing the most cost-efficient model. Keep spend predictable without sacrificing performance or flexibility.
Optimize every tokenRecover from provider outages, rate limits, and timeouts with built-in failover logic. Your critical flows stay up, even when individual APIs go down.
Resilience by defaultTrace every request across providers with unified logs, metrics, and latency breakdowns. Debug failures and tune prompts using one consistent telemetry layer.
See every token hopDefine higher-level tasks that chain tools, models, and workflows behind a single call. Keep your app code clean while LLM.API handles the coordination.
Ship workflows, not glueSend massive batches of prompts in one request with automatic chunking, retries, and aggregation. Maximize throughput and minimize overhead for large-scale workloads.
Scale tokens, not codeDecision guide
FAQ
FLUX.2 Flex is an image generation model by Black Forest Labs designed for high-quality, flexible text-to-image generation via API.
FLUX.2 Flex currently supports text-to-image generation and image-to-image transformations, but does not process or generate pure text responses.
On LLM.API, FLUX.2 Flex is billed per image generation request, with cost depending on resolution and optional advanced parameters like steps or guidance.
Typical FLUX.2 Flex generations complete within a few seconds, varying with image resolution, generation steps, and overall LLM.API backend load.
FLUX.2 Flex accepts reasonably long text prompts, but is constrained by maximum prompt length in characters and image upload size limits for image-to-image.
You call FLUX.2 Flex using the LLM.API unified endpoint, specifying the provider as Black Forest Labs and the model name as flux-2-flex.
FLUX.2 Flex is best for fast, high-quality image generation where you need flexible style control and cost-effective inference at scale.
FLUX.2 Flex trades some peak fidelity and specialization for greater speed, flexibility, and lower cost compared to heavier FLUX.2 variants.
FLUX.2 Flex can struggle with rendering accurate text in images, complex small details, and may produce biased or inappropriate content without careful prompting.
Yes, you can orchestrate FLUX.2 Flex with other vision or language models in a single LLM.API integration by switching the model parameter per request.
Compare
MiMo-V2-Flash is an open-source Mixture-of-Experts language model from Xiaomi optimized for fast, long-context reasoning and coding. It combines a 309B-parameter MoE architecture with only 15B active…
Riverflow V2 Standard Preview is the standard variant of Sourceful's Riverflow 2.0 preview lineup, offering unified text-to-image and image-to-image generation focused on production-grade creative workflows.
Qwen3 ASR Flash is Qwen’s high-accuracy, multilingual automatic speech recognition (ASR) service optimized for real-time transcription of short audio. It is built on the Qwen3-Omni foundation…