← all policies/chhhee10/cost-guard
━━ inference
require-max-output-tokens-in-call-loops
Ask for an output-token cap when a loop calls a hosted model per item
━━ when it runs
PreToolUsebefore the tool call runs. a denial here means the command never executes at all.
watchesevery tool the event fires for — not narrowed to particular tools
this pack enforces: a match returns a real denial rather than being recorded and discarded. what the denial then does depends on the event above.
━━ install
this one is opt-in. taking the pack without flags leaves it off until you tick it, so install it by name if you want only this.
- just this policy
- the whole pack
- turn just this off
- pick from a list
- uninstall the pack
ships in chhhee10/cost-guard · v4e133099d12a · source ↗
━━ also in inference
- cap-frontier-eval-row-countAsk for a row cap before an eval loop runs a frontier model over a full dataset
- bound-inference-job-concurrencyRefuse unbounded fan-out on a job that makes one model call per item
- sample-before-embedding-or-finetuningAsk for a sampled run before a fine-tune or an embedding pass over a full corpus