YuJitang
417416087b
fix: use backend thread token usage for header total ( #2800 )
...
* fix: use backend thread token usage for header total
* Refactor thread token usage fetch
2026-05-09 19:40:32 +08:00
YuJitang
d02f762ab0
feat: refine token usage display modes ( #2329 )
...
* feat: refine token usage display modes
* docs: clarify token usage accounting semantics
* fix: avoid duplicate subtask debug keys
* style: format token usage tests
* chore: address token attribution review feedback
* Update test_token_usage_middleware.py
* Update test_token_usage_middleware.py
* chore: simplify token attribution fallback
* fix token usage metadata follow-up handling
---------
Co-authored-by: Willem Jiang <willem.jiang@gmail.com>
2026-05-04 09:56:16 +08:00
YuJitang
105db00987
feat: show token usage per assistant response ( #2270 )
...
* feat: show token usage per assistant response
* fix: align client models response with token usage
* fix: address token usage review feedback
* docs: clarify token usage config example
---------
Co-authored-by: Willem Jiang <willem.jiang@gmail.com>
2026-04-16 08:56:49 +08:00
Matt Van Horn
b40b05f623
feat(frontend): display token usage per conversation turn ( #1229 )
...
Surface the usage_metadata that PR #1218 added to the streaming API.
A compact indicator in the chat header shows cumulative tokens consumed
per thread, with a tooltip breakdown of input/output/total counts.
Co-authored-by: Matt Van Horn <455140+mvanhorn@users.noreply.github.com>
Co-authored-by: Claude Opus 4.6 <noreply@anthropic.com>
Co-authored-by: Willem Jiang <willem.jiang@gmail.com>
2026-03-24 08:59:35 +08:00