Files
OpenJarvis/frontend
Jon Saad-Falcon 3cf4bbae30 feat: display thinking tokens in X-Ray footer and stream usage data
- Ollama engine: capture eval_count and prompt_eval_count from the
  streaming final chunk (includes thinking/reasoning tokens).
- Routes: include usage data in the SSE finish chunk so the frontend
  gets accurate token counts even in streaming mode.
- X-Ray footer: show estimated thinking tokens when completion_tokens
  significantly exceeds visible output (e.g. "1695 generated · 11 prompt
  · ~1680 thinking").
2026-03-14 14:55:42 -07:00
..
2026-03-12 17:29:39 +00:00
2026-03-12 17:29:39 +00:00