- Text Generation
FLUX.2 Pro is a professional-grade image generation and editing model from Black Forest Labs, optimized for photorealistic quality, strong prompt adherence, and reliable production use. It…
Powered by Openrouter
Body Builder (beta) is an OpenRouter model that converts natural language descriptions into structured OpenRouter API request objects, enabling automated construction of complex, multi-model calls.
Output tokens per second · Higher is better
Seconds · Lower is better
USD per 1M tokens (blended) · Lower is better
About the model
Body Builder (beta) is a text-only OpenRouter model that transforms free-form natural language instructions into structured OpenRouter API request payloads. It is mainly used to generate multi-model requests, build custom routing logic, and programmatically create chat or completion calls from human-readable task descriptions. Teams also employ it to simplify orchestration layers in applications that coordinate several AI models through the OpenRouter API. As an in-house OpenRouter model released in December 2025 with a 128K context window, it belongs to the broader family of OpenRouter system and helper models alongside tools like Auto Router and Fusion.
Model capabilities
Supports interactive conversational exchanges for refining fitness-related prompts and understanding generated workout or body-transformation programs.
Designed for use via OpenRouter, integrating into monitoring or orchestration pipelines that track requests, responses, and performance.
May interact with image-related or visual fitness content in toolchains, though no specific official image-processing specialization is documented.
Can help rephrase, structure, or systematize body-building prompts and results text within OpenRouter-based workflows.
Potentially usable in pipelines that extract or structure fitness-related text data; no dedicated OCR functionality is officially specified.
Use cases
Transparent pricing
LLM API offers the lowest cost and highest performance for Body Builder-class models.
| Provider | Region | Latency | Throughput | Uptime | Input ($/1M) | Output ($/1M) | Context |
|---|---|---|---|---|---|---|---|
| LLM API BEST | Global | ~120ms | ~80 tps | ~99.99% | ~$0.05 | ~$0.15 | ~128K |
| Openrouter | Global | ~220ms | ~40 tps | ~99.9% | ~$0.10 | ~$0.30 | ~64K |
| OpenAI-compatible Proxy A | US East | ~200ms | ~50 tps | ~99.9% | ~$0.08 | ~$0.25 | ~64K |
| Cloud ML Gateway B | EU West | ~250ms | ~35 tps | ~99.5% | ~$0.12 | ~$0.35 | ~32K |
| Serverless LLM Host C | Global | ~260ms | ~30 tps | ~99.0% | ~$0.15 | ~$0.40 | ~32K |
Performance benchmarks
| Metric | Body Builder (beta) | OpenAI o3-mini | Anthropic Claude 3.5 Sonnet |
|---|---|---|---|
| Avg Latency | ~180ms | ~220ms | ~250ms |
| Context Window | 128K | 200K | 200K |
| Input Price ($/1M) | $0.40 | $0.15 | $3.00 |
| Output Price ($/1M) | $0.80 | $0.60 | $15.00 |
| Max Output Tokens | 8K | 16K | 8K |
| Throughput | 60 tps | 100 tps | 80 tps |
| Uptime | 99.0% | 99.9% | 99.9% |
30-day usage via LLM API
Architecture & Integration
One unified API. Every major model. Built-in reliability, cost control, and observability.
Automatically route each request to the optimal model across providers based on latency, price, and quality—without changing your integration or redeploying code.
One endpoint, every model.Control and optimize spend with per-route pricing rules, cheaper model fall-throughs, and real-time usage insights so you never overpay for commodity tokens again.
Cut AI costs, not value.Define automatic failover chains so if a provider degrades or times out, traffic transparently retries on backups—no user-visible errors or manual intervention.
Stay up when others fail.Get request-level logs, latency traces, and model-by-model success metrics in one place so you can debug prompts, tune routes, and prove reliability to stakeholders.
See every token’s journey.Call high-level tasks like chat, tools, rerank, or embeddings through one consistent interface, while LLM.API picks the right model and parameters under the hood.
Think tasks, not models.Submit massive batches of prompts or embeddings in a single call with automatic chunking, retry, and aggregation for efficient large-scale workloads.
Scale from dozens to millions.Decision guide
FAQ
Body Builder (beta) is an Openrouter model accessible via LLM.API, designed for AI-assisted code and text generation workflows.
Body Builder (beta) supports a context window whose exact size depends on the provider’s configuration exposed through LLM.API.
Pricing for Body Builder (beta) is determined by Openrouter and is passed through LLM.API’s metered billing for input and output tokens.
Latency and throughput for Body Builder (beta) depend on Openrouter’s infrastructure and current load, typically suitable for interactive application use.
Body Builder (beta) supports text input and text output; it does not natively handle image, audio, or video content.
Specify the model identifier for Body Builder (beta) in your LLM.API request along with your API key, messages, and any desired generation parameters.
Body Builder (beta) trades some raw capability and stability for experimental features, making it less predictable than mature Openrouter text models.
As a beta model, Body Builder (beta) can be less stable, occasionally produce inconsistent outputs, and may change behavior as the provider updates it.
If LLM.API supports streaming for this model, you can enable it by setting the appropriate streaming flag in your request payload.
Because it is in beta, Body Builder (beta) is better suited for experimentation and prototyping than strict, latency-sensitive production workloads.
Compare
FLUX.2 Pro is a professional-grade image generation and editing model from Black Forest Labs, optimized for photorealistic quality, strong prompt adherence, and reliable production use. It…
Nemotron 3 Super is NVIDIA’s open-weight, 120B-parameter hybrid Mamba-Transformer Mixture-of-Experts language model optimized for high-throughput agentic reasoning workloads. It is notable for combining LatentMoE experts, long-context…
Gemini 3.1 Flash TTS Preview is Google’s low-latency text‑to‑speech model that generates natural, expressive speech with fine-grained control via style prompts and audio tags. It is…