Reasoning models emit intermediate thinking tokens before the final answer. Hide that entirely and the wait feels featureless; show it with the same visual weight as the answer and the actual conclusion gets buried. This pattern splits the difference: while it's running, a shimmering summary line ("Comparing search results…") shows something is happening, and once done it collapses into a one-line summary ("Thought for 4s") that's optional to expand.
It differs from a typing indicator in purpose — three dots are a "starting soon" signal with no content, while a thinking indicator carries actual (if brief) intermediate reasoning text that partly explains why it's taking a while.
When expanded, it should show real model output, or at least reflect its actual order — flashing decorative phrases unrelated to what's really happening is the kind of thing users eventually notice was fake, and trust doesn't come back easily.
When to use
Use it for models or agent flows where reasoning genuinely takes a few seconds or more. On a question that could be answered instantly, it reads as manufactured drama.