Model showdown · live

Stop paying frontier prices for every part of the prompt.

TokenGrill breaks a prompt into sub-prompts, sends each one to the right model for that job, then merges the results back into a single answer. Run it head-to-head against any single model and watch the tokens, spend and time side by side.

Decompose

One prompt is split into the sub-tasks it actually contains.

Route

Each sub-task goes to the cheapest model that can do it well.

Synthesize

Sub-answers are merged into one clean, coherent response.

Save

You pay frontier prices only where frontier reasoning is needed.

How the orchestrator earns its savings

A single frontier model charges its top rate for the boring 80% of your prompt too. The orchestrator only spends that rate where the work actually requires it.

01

Intelligent prompt decomposition

A fast, low-cost planner reads the request and turns it into at most four self-contained sub-tasks, flagging which ones genuinely need deep reasoning or drafting.

02

Proactive token & cost prediction

Before anything expensive runs, each sub-task is sized: expected input tokens, projected output, and the cost of running it on each candidate model.

03

Model routing & parallel execution

Hard sub-tasks go to a strong model; extraction, summarizing and formatting go to a cheap one. Independent sub-tasks run at the same time instead of one long serial call.

04

Contextual aggregation & synthesis

A final pass merges every sub-answer into one non-repetitive response, streamed back to you with the full token and spend tally across every sub-call.

Every number in the demo below is measured from the real run — token counts come back from the providers, and spend is calculated from published list pricing per 1K tokens, summed across the planner, every worker call and the final synthesis.

01Live demo

One prompt. Two contenders. Real numbers.

Pick any model on the left and the TokenGrill Orchestrator on the right — or two single models against each other. Attach a document, run them together, and compare tokens, spend and speed from the same run.

Orchestrator vs. any single modelAttach a documentToken & spend comparison
One prompt · two models
Shared public demo · 12 runs left this session0 / 8,000

Sample prompts

Model ATokenGrill

Blended cost — planner, parallel workers and synthesis, billed per sub-call

Pick a model, enter a prompt above, and run the showdown.

In

—

Out

—

Cost

—

First token

—

Model BOpenAI

$1.25 / 1M in · $10.00 / 1M out

Pick a model, enter a prompt above, and run the showdown.

In

—

Out

—

Cost

—

First token

—