Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning
Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy as a one-dimensional failure mode. Models must distinguish when to incorporate others' perspectives from when to maintain a well-grounded moral judgment. We study the broader resistance-compliance process governing this distinction. Across three studies, we show that models' judgment revision is structured along three dimensions that parallel classic phenomena in human social psychology: the distance between an incoming view and the model's ini
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 57%When I made LLMs argue with each other, they started making up citations to win. Sycophancy wasn't the only failure mode. →
- PossiblePossibly related (embedding) · 53%Large language models often prioritize Western moral values, overlooking other cultures - The Conversation →
- PossiblePossibly related (embedding) · 52%Alignment →
- FuzzyOverlapping authors or contributors · 62%bytedance/deer-flow →
“Shared author/contributor keys: wang”
- FuzzyOverlapping authors or contributors · 62%ray-project/ray →
“Shared author/contributor keys: wang”
- LinkedLinked via arxiv author · 85%Baihui Wang →
“Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning”
- LinkedLinked via arxiv author · 85%Bernard Koch →
“Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning”
