An illustrated walkthrough · 36 secondsEvery question is different.

01 / 06
FORGE routes evidence form and thinking depth together. An illustrative question about the 1900 and 1924 Summer Olympics is routed to summarized evidence and a low thinking budget. A frozen host answers Paris. An easy question instead takes the Direct and NoThink route. These are explanatory examples, not measured routing predictions. YOUR QUESTIONWhich city hosted boththe 1900 and 1924Summer Olympics?Whichcityhosted…A query, ready to route. needs evidence? A whole corpus.How much is enough? ① What evidence?Direct?Query onlySummaryKey evidenceRawFull passagesThe support choice informs the thinking head. ② How much thought?NoThinkCoT promptLow budgetHigh budget One joint action: Summary + Low Retrieved passages1900 → Paris, France1924 → Paris, France Keep what matters. QUERY-AWARE EXTRACT1900 → Paris, France1924 → Paris, FranceRelevant evidence, fewer tokens. Question tokensSummary tokensThinking: LowA tailored request Frozen hostWeights unchanged. ANSWER TOKENSParis[end] QUESTIONWhich city hosted both the 1900 and 1924 Summer Olympics?Summary evidence+Low thinkingParis.The right support. The right amount of thought.Adapt the request around the model — keep the model frozen. A DIFFERENT QUESTIONHow many daysare in a week?No retrieved evidence needed. FORGESupport: DirectThinking: NoThink Same frozen host Answer: 7 days.Different query. Different route.

Questions differ in what they need. FORGE decides what evidence to provide and how much reasoning to use.

Illustrative queries and choices; feature acquisition omitted. Full uses four pre-route host probes; Lite uses none. Low/High require a host with thinking-budget controls.