Reasoning (Thinking) & Telemetry

Learn how to control model deliberation effort and supervise context window consumption and token metrics in real time.

1. Reasoning-Capable Models

Next-generation models (such as DeepSeek-R1, OpenAI o1/o3-mini, Claude 3.7 Sonnet with thinking, and Qwen QwQ) generate internal thought chains prior to answering.

ZeroChat streams these thinking chunks in real time, rendering them inside collapsible thought blocks with timers.

The Reasoning Effort Selector

The composer toolbar includes a speedometer indicator. Its needle reflects the selected intensity; click it to open a compact panel with an intensity slider and an X close button.

Agent Checkpoint (agent_checkpoint): Enable this tool from the reasoning menu to let models store intermediate conclusions before complex tool loops or RAG searches.

2. Unified Context Hub

Next to the profile selector, the telemetry chip (e.g., 1.4k / 128k) reflects context health via color-coded status:

3. Context Hub Popover Diagnostics

Click the token pill to open the comprehensive telemetry diagnostic view:

Section 1: Global Context Window

Section 2: Prompt Cache Savings (KV Cache)

Section 3: Last Turn Performance