BehavTest

BehavTest › Reference

Cost

Cost is computed from the provider's reported token usage, priced per category: regular input, cache reads, cache writes (5-minute and 1-hour), and output. If a model has no known price, or usage is missing, the cost is unknown (shown as such), never guessed.

Prices ship in src/pricing/prices.json (dated 2026-09-23): current Anthropic models, and OpenAI's GPT-6, GPT-5.x, GPT-4.1, GPT-4o and o4-mini families. Where OpenAI shows no cache-read or cache-write price for a model, a call that uses one has unknown cost. Things to know:

Add or override prices in the suite:

"pricing": [{ "provider": "openai", "model": "my-model", "inputPerMTok": 2.5, "outputPerMTok": 10, "cachedInputPerMTok": 1.25, "validUntil": "2027-01-31" }]

or with --prices prices.json (an array or {entries: [...]}; validUntil is optional). Check the provider's pricing page: BehavTest's table is a convenience, not a bill.