Conversation turn latency, p50, across two optimisations

Two changes on 11 September: the acknowledgement stopped blocking the provider call, then only task-relevant tools were offered to the model. Median turn time fell from 3.08 s to 1.27 s.

Conversation turn latency, p50, across two optimisations

Baseline, pre-optimisation 3,083 Acknowledgement no longer blocks … 1,723 Plus task-relevant tool projection 1,270
Scale: ms.

Two changes on 11 September: the acknowledgement stopped blocking the provider call, then only task-relevant tools were offered to the model. Median turn time fell from 3.08 s to 1.27 s.

Test: Latency Lab (tools/latency/lab.py): 21 scripted journeys (status, time, recall, project, hud, latency, chat; three passes) through the normal launcher; end of input to end of turn.

Protocol

Test definition

Latency Lab (tools/latency/lab.py): 21 scripted journeys (status, time, recall, project, hud, latency, chat; three passes) through the normal launcher; end of input to end of turn.

Limits

Limitations

  • Typed inlet (no microphone/ASR audio); hosted cognition model over the network; 21 turns per run (1 cold, 20 warm), p99 not computable; same-day optimisation loop, 21/21 journeys correct; not a certified release figure.

Log

Recorded values

Variant Value Evidence Samples Source
Baseline, pre-optimisation 11 Sept 2026 3,083 ms MEASURED 21 docs/analysis/latency/lab-20260911T221912Z.json
Acknowledgement no longer blocks provider call 11 Sept 2026 1,723 ms MEASURED 21 docs/analysis/latency/lab-20260911T222249Z.json
Plus task-relevant tool projection 11 Sept 2026 1,270 ms MEASURED 21 docs/analysis/latency/lab-20260911T223508Z.json

Source paths refer to the private JARVIS repository and runtime records. They are listed for audit and are not published.