skip to content

Chain of Thought

Making the model show its work: explicit intermediate steps, the zero-shot and few-shot ways to elicit them, and sampling several chains to vote on an answer. Interviewers ask about it constantly, including the part people skip — when a reasoning trace is unfaithful or actively hurts.

on this pageshow

explore

questions

page 2 of 2

Zero-shot or few-shot CoT for a bank's KYC risk memos — how do you choose?

level: principalimportance: should knowfreq 33%

basics

~20 s

Cost and data exposure decide it. Exemplars are re-sent and billed on every request, and exemplars written from real customer files put personal data into every call and every log. Prefer synthetic policy-derived examples, or zero-shot with a strict output template.

open as a page

When does weighting self-consistency votes by model confidence beat a plain count?

level: middleimportance: nice to knowfreq 28%

basics

~20 s

Weighted voting helps mainly when the sample count is small and the answer space is short and constrained, so a per-sample likelihood score is meaningful. It hurts when confidence is miscalibrated, because it lets a few confidently wrong samples outvote a correct plurality.

open as a page

What does an LLM gain and lose by reasoning in words rather than latent state?

level: principalimportance: nice to knowfreq 24%

basics

~20 s

Verbalizing forces each reasoning step through one discrete token, discarding most of the model's richer internal state - but it produces a trace that can be read, checked, cached, edited and monitored. Latent reasoning keeps the bandwidth and gives up the audit surface.

open as a page

showing 31–33 of 33