Commit Graph
4 Commits
Author SHA1 Message Date
ImBenjiandClaude Opus 5 9b490c4d39 fix: honour OPEN_ROUTER_CHEAP_MODEL, it was set and never read
The env var has been set to google/gemma-4-31b-it all along and nothing ever
mapped it onto openRouter.cheapModel, so graphWorker's fallback chain silently
used the main model instead. Graph entity resolution is the highest volume llm
call in the system and its entire job is to reply with the number of a match.

Measured on the live key, same prompt:

  deepseek v4 flash   7564 completion tokens, 7558 of them reasoning  $0.0013708
  gemma-4-31b-it        14 completion tokens, 0 reasoning             $0.0000121

113x. Gemma is actually the more expensive model per token, which is why this
was worth measuring rather than reasoning about prices: the cost is not the
price of the tokens, it is a reasoning model spending seven thousand tokens
thinking about a multiple choice question.

This also explains the reasoning tokens dominating the usage dashboard, and why
the daily spend roughly doubled today rather than yesterday. Removing the token
ceiling let a trivial prompt reason without bound. The ceiling was never the
right control for that, the model choice is.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01WnNxwxfXSbeNtjvtz5gayb
2026-09-04 21:57:17 +01:00
ImBenji c4028cc394 feat: add autonomous paper-trading and calibration pipeline 2026-08-03 14:03:27 +01:00
ImBenji 7ceaaf2401 refactor: load environment variables from .env file and update openRouter configuration 2026-04-27 17:43:48 +01:00
ImBenji 04966fac55 refactor: update worker commands and add new scripts for API rebuilding and queue feeding 2026-04-27 15:03:15 +01:00