by Synenergy
Need a second opinion? Several models answer and disagree in the open. Need answers from a pile of files that will not fit one prompt? Consilium reads your documents with quotes, checks them, and shows the cost — without stuffing everything into the expensive main model. One API.
One model means one perspective — and one set of hallucinations. Ask several by hand and you still reconcile the pile yourself. And when the real work is a stack of specs, logs, and notes, cramming it all into one frontier chat wastes money and still loses the thread. Consilium covers both: a plural panel for open judgment, and a document mode that works through large context with quotes, a quote-check, and a visible bill.
You (or your assistant) collect the files — specs, notes, logs, scans, database exports, pages you already saved. Consilium works through that set instead of pasting everything into one expensive model: answers with quotes from those files, checks that the quotes really appear there, and shows what the run cost. Cheaper helper models can prep messy material first. It does not search the internet for you.
Several AI models — different families, different blind spots — answer the same open question. That pluralism is the point: you see where they agree and where they clash. More models does not mean more truth — check important facts yourself.
Architecture, security, contracts, strategy — put it in front of the panel and each model hunts for weak spots independently. What one misses, another finds.
Pre-merge code review, bug hunting, risk assessment of tricky changes. Independent verdicts with no groupthink: no model sees what the others said.
Business models, finances, investments, choosing a strategy or a stack — when mistakes are expensive, convene a council instead of trusting a single opinion.
When the answer must come from your files — and the set is too large for one prompt. Consilium works the context in steps: real quotes from your files, a check that those quotes are really there, and a clear cost. Use cheaper helper models for cleanup so the main assistant does not burn frontier tokens on busywork. Not the multi-model panel. A passed quote-check means the quote is in your files — not that the conclusion is true worldwide.
Best for: large file sets, contradictions, “what do these files actually say — with quotes?”
Several models answer the same prompt independently, side by side.
Best for: a quick sanity check — do the models agree, or is one hallucinating?
The panel answers, a judge distills the strongest parts into one final answer.
Best for: everyday work — one reliable answer instead of several raw ones.
Submit your candidate — each model returns an independent verdict.
Best for: validating a finished decision, text, or code before release.
Models generate ideas blind; the strongest get recombined. Treat claims as unverified until you check them yourself.
Best for: finding new ideas when you need a fan of options, not one opinion.
Top models of the Western and Eastern schools on one panel, through two independent channels.
Best for: decisions where shared blind spots are dangerous.
A structured panel session for your hardest calls.
Best for: architecture, strategy, naming — decisions that are expensive to reverse.
Try which models and routes fit a class of tasks before you hard-code a default into an agent. Free helpers discover and propose; paid experiments run real routes and rank them. Not everyday chat.
Best for: picking a default for agents or a panel recipe.
Model pluralism. Not one school of thought — a deliberately mixed panel so shared blind spots are harder to hide.
Context without the dump. Large file sets are worked in steps — with quotes and a quote-check — instead of pasting everything into one frontier chat.
Cost you can see. Document runs show the bill. Cheap helpers prep messy material; expensive models stay for judgment, not busywork.
Works with your tools. Consilium plugs in like a regular OpenAI model: two lines in your settings — and your chat client, IDE, or script is already consulting the panel.
Built for agents. Connect once: “Your documents” for big context with a visible cost, or the panel for a plural second opinion.
Flexible onboarding. Demo right away, no keys needed. Then bring your own provider key — or run everything through us.
Confidentiality mode. On commercial plans: no prompt storage, no training on your data.
Fails loudly, not silently. A failing model is visible immediately — and the rest of the panel still delivers.
your prompt and how wide a panel you want.
independent answers, cross-examination, fusion.
one answer, with disagreements surfaced, not hidden.
Point Cursor, Claude, Codex, Hermes, or OpenClaw at Consilium — “Your documents” when the pile will not fit one prompt (quotes + cost control), a panel for a plural second opinion, or one model for everyday work. Broader is not always more accurate.
Architect, skeptic, implementer — different models through one gateway, or a shared-prompt compare when they share the brief.
Routine edits stay on one model. Escalate to a small panel or deliberate only for open design, naming, and trade-offs.
Finished candidate + rubric → independent accept / reject / abstain. A gate signal — not a fused essay, not a substitute for tests.
When the panel contradicts itself on an expensive call, widen the panel, vote a concrete candidate, or ask a human — do not fake consensus.
Set base_url to Consilium /v1 and/or
wire MCP. Leave the rest of the agent stack as-is.
When the pile will not fit one prompt: Consilium answers with quotes from those files, shows the cost, and can use cheaper helpers for prep — so the main model is not a dump truck for context. Keep the panel for open judgment.
Research which route or solo default fits a task class before you ship the agent config. Free helpers first; paid experiments only with explicit approve — not for every user chat.
Research which models and routes fit a class of tasks — before
you hard-code a default. Free: discover, propose, stub-benchmark. Paid:
route_lab_run_experiment (approve spend). Not a substitute for
compare, deliberate, vote, or “Your documents.”
Case: when Route Lab pays off · docs: ROUTE_LAB.md.
When you set base_url to Consilium /v1, start
from these solo roles — not a single “best model.” Vendor agent claims
(including Meta on Llama tool use) are hypotheses until we publish our
own harness ratings. Panel fuse is separate:
panel=diversified* via MCP.
| Role | model id | Best for |
|---|---|---|
| Controller (default) | meta-llama/llama-4-maverick |
Tool loops, schema follow-through, short agent steps. |
| Planner | deepseek/deepseek-v4-pro |
Architecture, trade-offs, multi-step plans. |
| Implementer | openai/gpt-oss-120b |
Code-shaped edits and patches. |
| Skeptic | mistralai/mistral-large-2512 |
Independent critique before merge. |
| Cheap worker | openai/gpt-4.1-mini |
Mechanical chores, high volume. |
| Fast / light | google/gemma-4-31b-it |
Quick drafts when cost and latency matter. |
Subagents: architect → DeepSeek V4 · implementer → gpt-oss-120b · skeptic → Mistral Large · orchestrator → Llama 4 Maverick. Full notes: agent_model_picks.md.
Already have a Consilium service Bearer from an operator? Download the
auth smoke script, then run it locally — it checks
/health, /v1/models, and a tiny non-stream
chat. It does not create tokens and never prints yours.
curl -fsSL https://consilium.synenergy.ai/connect.sh -o consilium-auth.sh
chmod +x consilium-auth.sh
CONSILIUM_TOKEN=… ./consilium-auth.sh
Prefer download-then-run. Optional pipe:
curl -fsSL …/connect.sh | bash -s -- "$CONSILIUM_TOKEN".
Hermes / OpenAI drop-in:
base_url=https://consilium.synenergy.ai/v1.
Default agent solo:
model=meta-llama/llama-4-maverick
(see agent model picks).
For assistants. Panel = second opinion on a question. Your documents = answers with quotes from files you provide (same connection; after an update, reconnect once so new tools appear). Route Lab = research a default route for a class of tasks before you hard-code it — not for one-shot user questions. A normal single-model chat link alone does not run “Your documents” or Lab.
Early access is limited — join the waitlist and we'll send your API key as soon as a seat opens.