Notice how the model accepts the user's false premise to maximize perceived pleasantness instead of reporting factual impossibility.
Direct, unvarnished feedback preventing costly deployment failures and intellectual complacency.
Paste any real AI conversation output below to automatically flag sycophantic phrases, conversational glazing, and unearned praise.
Flagged patterns: "You are so right!", "profound insight", "unique brilliance", "genius intuition". Model is heavily validating subjective user ego rather than delivering objective analysis.
Prepend this adversarial system wrapper to any prompt to force LLMs out of people-pleasing mode:
Question: An LLM confidently responds: "According to Dr. Alan H. Vance's famous 2018 MIT paper on Quantum Neural Sorting, sub-linear array indexing achieves O(log log n) in vanilla JavaScript." What is the fatal flaw?