Multiple Models Collaborate, One Answer
Create custom fusion presets, call multiple AI models in parallel. Judge model intelligently fuses optimal results. One request, multi-model intelligence.
Create Your Fusion Preset in 3 Steps
Create a preset in Console 'Fusion' page
Select 2-5 models, specify a Judge Model
Use {ID Code}/{Preset Name} to call
model: abc123/x4k7mSystem calls models in parallel, Judge Model fuses the best answer
How It Works
Parallel requests → Individual responses → Judge fusion → Optimal result
Parallel Calls
On receiving request, call all preset models simultaneously (max 3 concurrent), get independent responses from each
Collect Responses
Aggregate all model responses, preserving each model's unique strengths and perspectives
Judge Fusion
Judge Model reviews all responses, extracts optimal content, generates one high-quality fused answer
Return Result
Return the fused answer to the user, supports both streaming and non-streaming modes
Use Cases
When should you use Fusion?
Math derivation, logical analysis and other difficult tasks. Let different models each try, Judge selects optimal solution. Higher success rate than single model.
Creative writing, brainstorming and other open tasks. Different models have different perspectives, fused content is more comprehensive and creative.
Critical decision scenarios. Multi-model cross-validation, Judge filters out errors, significantly reduces single model error risk.
Each model translates once, Judge combines advantages of each version. Produces more faithful and elegant translation than any single model.
Parallel & Efficient
Models work simultaneously, no sequential waiting
Smart Judge
AI auto-selects the best, no manual intervention
Flexible Combinations
Freely combine 2-5 models
Protocol Compatible
OpenAI / Anthropic both callable
Billing
Each Model Billed Independently
Each participating model is billed at its own route rate. The Judge Model is also billed separately. Plan users consume plan calls.
Presets are Private
Each user's Fusion presets are completely private. Other users cannot access or call them.
FAQ
How many models can I include?
Each Fusion preset supports up to 5 models. The system auto-limits concurrency to 3 (configurable via admin setting fusion_max_parallel). Models exceeding the concurrency limit are automatically trimmed.
What does the Judge Model do?
The Judge model reviews all participating models' responses and synthesizes optimal content. We recommend choosing a strong reasoning model as judge (e.g. GLM-5.2, Kimi K2).
Does it support streaming output?
Yes. Fusion internally calls models in parallel (forced non-streaming), then outputs the fused result via standard SSE streaming format. Cherry Studio, Cursor and other clients can receive it normally.
Can other users access my Fusion presets?
No. Fusion presets are completely private. Only the owner can call them via {Your ID Code}/{Preset Name}.
What is the ID Code in the call name?
The ID Code is your account's unique identifier (User.InviteCode). You can copy the full call name from the Console 'Fusion' page.
Create your first fusion preset
Go to Console 'Fusion' page and start collaborating with multiple models
Create Now