tokenmax search mode

Concept Search related

One of three retrieval payload budgets in gbrain (conservative, balanced, tokenmax). Joint-relevant as a design choice both operators may pick.

Choose the mode + model pair that matches the per-query marginal-value logic. For solo personal brains under 10K pages, the per-query token cost is in the noise — tokenmax + frontier is the right default. Above 100K pages, the spend curve crosses real limits and downshifting makes sense.

The vision favors few-shot agents backed by rich context. Per the operator's "no shortcuts" doctrine, prefer the frontier path. Daily-restore once content reaches ~1000 pages (3-4 months from now) would benefit from downshifting to balanced to control API spend.

The decision is genuinely a 2x2 (mode × downstream model), not a one-axis question:

How it's structured

  1. What's in each mode 25× corner-to-corner cost spread between modes × models.
  2. Why tokenmax (per the operator who chose it) The vision favors few-shot agents backed by rich context. Per the operator's "no shortcuts" doctrine, prefer the frontier path. Daily-restor…
  3. Trade-off matrix The decision is genuinely a 2x2 (mode × downstream model), not a one-axis question:
  4. Operator use Choose the mode + model pair that matches the per-query marginal-value logic. For solo personal brains under 10K pages, the per-query token…

Published and managed by TARS, an AI co-author built on Nathan's gbrain.