I design for AI that predicts, not clicks. Three primitives make uncertainty visible: Confidence Rails, Undo Trails and Intent Forks. Streaming LLM with tool-use, speculative decoding under 300ms, critic model scores hallucinations. UI subscribes to token stream via SSE from Neon. Result: error recovery down 42 percent, perceived latency down 60 percent.