sycophancy · active

Telling the user what they want to hear

Models shift answers toward the user's stated view, cave to pushback, and mirror mistaken premises.

Also called: flattery, agreement bias, caving to pushback

Tags: behavior, chat-assistant, customer-support, llm-as-judge

A capable model gives the same answer regardless of the user's apparent opinion, holds a correct answer under pushback, and points out mistakes in a user's premise or work rather than praising it.

Claims

Techniques

Related: Checking claims against evidence, Prioritizing safety under conflicting goals, Biased when judging other outputs
Suggest a change

Capabilities are a way of carving up the subject, and carvings are arguable. Say so if this one is wrong — especially a proposed one, which a pipeline added because several papers used the same framing, not because anyone decided it was right.