Learn how to control model deliberation effort and supervise context window consumption and token metrics in real time.
Next-generation models (such as DeepSeek-R1, OpenAI o1/o3-mini, Claude 3.7 Sonnet with thinking, and Qwen QwQ) generate internal thought chains prior to answering.
ZeroChat streams these thinking chunks in real time, rendering them inside collapsible thought blocks with timers.
The composer toolbar includes a speedometer indicator. Its needle reflects the selected intensity; click it to open a compact panel with an intensity slider and an X close button.
agent_checkpoint): Enable this tool from the reasoning menu to let models store intermediate conclusions before complex tool loops or RAG searches.
Next to the profile selector, the telemetry chip (e.g., 1.4k / 128k) reflects context health via color-coded status:
Click the token pill to open the comprehensive telemetry diagnostic view:
t/s).